paper-with-me

Papers

Monocular Visual Analysis for Electronic Line Calling of Tennis Games

2021-07-20 · Yuanzhou Chen, Shaobo Cai, Yuxin Wang, Junchi Yan

Electronic Line Calling is an auxiliary referee system used for tennis matches based on binocular vision technology. While ELC has been widely used, there are still many problems, such as complex installation and maintenance, high cost and etc. We propose a monocular vision technology based ELC method. The method has the following steps. First, locate the tennis ball's trajectory. We propose a multistage tennis ball positioning approach combining background subtraction and color area filtering. Then we propose a bouncing point prediction method by minimizing the fitting loss of the uncertain point. Finally, we find out whether the bouncing point of the ball is out of bounds or not according to the relative position between the bouncing point and the court side line in the two dimensional image. We collected and tagged 394 samples with an accuracy rate of 99.4%, and 81.8% of the 11 samples with bouncing points.The experimental results show that our method is feasible to judge if a ball is out of the court with monocular vision and significantly reduce complex installation and costs of ELC system with binocular vision.

📄 PDF Abstract BibTeX arXiv:2107.09255

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MM-ToolSandBox: A Unified Framework for Evaluating Visual Tool-Calling Agents

2026-07-13 · Kaixin Ma, Di Feng, Alexander Metz, Jiarui Lu 외 arxiv

We introduce MM-ToolSandBox, a benchmark and evaluation framework for visually grounded tool-calling agents. The framework provides a stateful execution environment spanning 500+ tools across 16 application domains, supp…

TargetCall: Eliminating the Wasted Computation in Basecalling via Pre-Basecalling Filtering

2022-12-09 · Meryem Banu Cavlak, Gagandeep Singh, Mohammed Alser, Can Firtina 외

Basecalling is an essential step in nanopore sequencing analysis where the raw signals of nanopore sequencers are converted into nucleotide sequences, i.e., reads. State-of-the-art basecallers employ complex deep learnin…

Thinking with Images via Self-Calling Agent

2025-12-09 · Wenxi Yang, Yuzhong Zhao, Fang Wan, Qixiang Ye arxiv

Thinking-with-images paradigms have showcased remarkable visual reasoning capability by integrating visual information as dynamic elements into the Chain-of-Thought (CoT). However, optimizing interleaved multimodal CoT (…

Reinforcement LearningVisual Reasoning

Divergent Thoughts toward One Goal: LLM-based Multi-Agent Collaboration System for Electronic Design Automation

2025-02-15 · Haoyuan Wu, Haisheng Zheng, Zhuolun He, Bei Yu

Recently, with the development of tool-calling capabilities in large language models (LLMs), these models have demonstrated significant potential for automating electronic design automation (EDA) flows by interacting wit…

Direct Kernel Optimization: Efficient Design for Opto-Electronic Convolutional Neural Networks

2025-11-03 · Ali Almuallem, Harshana Weligampola, Abhiram Gnanasambandam, Wei Xu 외 arxiv

Hybrid opto-electronic neural networks combine optical front-ends with electronic back-ends to perform vision tasks, but joint end-to-end (E2E) optimization of optical and electronic components is computationally expensi…

Monocular Depth Estimation