Few-Shot Human Motion Prediction via Meta-Learning
Human motion prediction, forecasting human motion in a few milliseconds conditioning on a historical 3D skeleton sequence, is a long-standing problem in computer vision and robotic vision. Existing forecasting algorithms rely on extensive annotated motion capture data and are brittle to novel actions. This paper addresses the problem of few-shot human motion prediction, in the spirit of the recent progress on few-shot learning and meta-learning. More precisely, our approach is based on the insight that having a good generalization from few examples relies on both a generic initial model and an effective strategy for adapting this model to novel tasks. To accomplish this, we propose proactive and adaptive meta-learning (PAML) that introduces a novel combination of model-agnostic meta-learning and model regression networks and unifies them into an integrated, end-to-end framework. By doing so, our meta-learner produces a generic model initialization through aggregating contextual information from a variety of prediction tasks, while this model can effectively adapt to a specific task by leveraging learning-to-learn knowledge about how to transform few-shot model parameters to many-shot model parameters. The resulting PAML predictor model significantly improves the prediction performance on the heavily benchmarked H3.6M dataset in the small-sample size regime.
Code (0)
등록된 구현이 없습니다.
Tasks
Few-Shot LearningHuman motion predictionMeta-Learningmotion predictionPredictionSimilar Papers 제목 키워드 기반
Off-the-shelf ChatGPT is a Good Few-shot Human Motion Predictor
To facilitate the application of motion prediction in practice, recently, the few-shot motion prediction task has attracted increasing research attention. Yet, in existing few-shot motion prediction works, a specific mod…
Human motion predictionIn-Context LearningLanguage ModelingLanguage Modelling+2Meta-PerSER: Few-Shot Listener Personalized Speech Emotion Recognition via Meta-learning
This paper introduces Meta-PerSER, a novel meta-learning framework that personalizes Speech Emotion Recognition (SER) by adapting to each listener's unique way of interpreting emotion. Conventional SER systems rely on ag…
Emotion RecognitionMeta-LearningSpeech Emotion RecognitionFew-shot human motion prediction for heterogeneous sensors
Human motion prediction is a complex task as it involves forecasting variables over time on a graph of connected sensors. This is especially true in the case of few-shot learning, where we strive to forecast motion seque…
Few-Shot LearningHuman motion predictionmotion predictionPrediction+3MoML: Online Meta Adaptation for 3D Human Motion Prediction
In the academic field the research on human motion prediction tasks mainly focuses on exploiting the observed information to forecast human movements accurately in the near future horizon. However a significant gap a…
Bilevel OptimizationHuman motion predictionMeta-Learningmotion predictionVisual Affect Analysis: Predicting Emotions of Image Viewers with Vision-Language Models
Vision-language models (VLMs) show promise as tools for inferring affect from visual stimuli at scale; it is not yet clear how closely their outputs align with human affective ratings. We benchmarked nine VLMs, ranging f…
Emotion Classification