paper-with-me

Papers

EvaluationNet: Can Human Skill be Evaluated by Deep Networks?

2017-05-31 · Seong Tae Kim, Yong Man Ro

With the recent substantial growth of media such as YouTube, a considerable number of instructional videos covering a wide variety of tasks are available online. Therefore, online instructional videos have become a rich resource for humans to learn everyday skills. In order to improve the effectiveness of the learning with instructional video, observation and evaluation of the activity are required. However, it is difficult to observe and evaluate every activity steps by expert. In this study, a novel deep learning framework which targets human activity evaluation for learning from instructional video has been proposed. In order to deal with the inherent variability of activities, we propose to model activity as a structured process. First, action units are encoded from dense trajectories with LSTM network. The variable-length action unit features are then evaluated by a Siamese LSTM network. By the comparative experiments on public dataset, the effectiveness of the proposed method has been demonstrated.

📄 PDF Abstract BibTeX arXiv:1705.11077

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Sigmoid Activation 설명 없음
Tanh Activation 설명 없음
LSTM An LSTM is a type of recurrent neural network that addresses the vanishing gradient problem in vanilla…

Similar Papers 제목 키워드 기반

Revisiting Text-to-Image Evaluation with Gecko: On Metrics, Prompts, and Human Ratings

2024-04-25 · Olivia Wiles, Chuhan Zhang, Isabela Albuquerque, Ivana Kajić 외

While text-to-image (T2I) generative models have become ubiquitous, they do not necessarily generate images that align with a given prompt. While previous work has evaluated T2I alignment by proposing metrics, benchmarks…

Exploring and Testing Skill-Based Behavioral Profile Annotation: Human Operability and LLM Feasibility under Schema-Guided Execution

2026-04-16 · Yufeng Wu arxiv

Behavioral Profile (BP) annotation is difficult to automate because it requires simultaneous coding across multiple linguistic dimensions. We treat BP annotation as a bundle of annotation skills rather than a single task…

MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills

2026-04-22 · Yingyong Hou, Xinyuan Lao, Huimei Wang, Qianyu Yao 외 arxiv

Background: Agent skills are increasingly deployed as modular, reusable capability units in AI agent systems. Medical research agent skills require safeguards beyond general-purpose evaluation, including scientific integ…

SKID RAW: Skill Discovery from Raw Trajectories

2021-03-26 · Daniel Tanneberg, Kai Ploeger, Elmar Rueckert, Jan Peters

Integrating robots in complex everyday environments requires a multitude of problems to be solved. One crucial feature among those is to equip robots with a mechanism for teaching them a new task in an easy and natural w…

Variational Inference

Simultaneous Estimation of Manipulation Skill and Hand Grasp Force from Forearm Ultrasound Images

2025-02-01 · Keshav Bimbraw, Srikar Nekkanti, Daniel B. Tiller II, Mihir Deshmukh 외

Accurate estimation of human hand configuration and the forces they exert is critical for effective teleoperation and skill transfer in robotic manipulation. A deeper understanding of human interactions with objects can …