paper-with-me

홈 › Papers

OSCAR: Data-Driven Operational Space Control for Adaptive and Robust Robot Manipulation

2021-10-02 · Josiah Wong, Viktor Makoviychuk, Anima Anandkumar, Yuke Zhu

Learning performant robot manipulation policies can be challenging due to high-dimensional continuous actions and complex physics-based dynamics. This can be alleviated through intelligent choice of action space. Operational Space Control (OSC) has been used as an effective task-space controller for manipulation. Nonetheless, its strength depends on the underlying modeling fidelity, and is prone to failure when there are modeling errors. In this work, we propose OSC for Adaptation and Robustness (OSCAR), a data-driven variant of OSC that compensates for modeling errors by inferring relevant dynamics parameters from online trajectories. OSCAR decomposes dynamics learning into task-agnostic and task-specific phases, decoupling the dynamics dependencies of the robot and the extrinsics due to its environment. This structure enables robust zero-shot performance under out-of-distribution and rapid adaptation to significant domain shifts through additional finetuning. We evaluate our method on a variety of simulated manipulation problems, and find substantial improvements over an array of controller baselines. For more results and information, please visit https://cremebrule.github.io/oscar-web/.

📄 PDF Abstract BibTeX arXiv:2110.00704

Code (1)

NVlabs/oscar pytorch

Tasks

Robot Manipulation

Methods 이 논문이 사용한 방법론

OSCAR OSCAR is a new learning method that uses object tags detected in images as anchor points to ease the learning of image-text alignment. The model take a triple as input…

Similar Papers 제목 키워드 기반

OSCAR: Operating System Control via State-Aware Reasoning and Re-Planning

2024-10-24 · Xiaoqiang Wang, Bang Liu

Large language models (LLMs) and large multimodal models (LMMs) have shown great potential in automating complex tasks like web browsing and gaming. However, their ability to generalize across diverse applications remain…

Navigate

OSCAR: Orchestrated Self-verification and Cross-path Refinement

2026-04-02 · Yash Shah, Abhijit Chakraborty, Naresh Kumar Devulapally, Vishnu Lokhande 외 arxiv

Diffusion language models (DLMs) expose their denoising trajectories, offering a natural handle for inference-time control; accordingly, an ideal hallucination mitigation framework should intervene during generation usin…

How could Neural Networks understand Programs?

2021-05-10 · Dinglan Peng, Shuxin Zheng, Yatao Li, Guolin Ke 외

Semantic understanding of programs is a fundamental problem for programming language processing (PLP). Recent works that learn representations of code based on pre-training techniques in NLP have pushed the frontiers in …

valid

OSCaR: Orthogonal Subspace Correction and Rectification of Biases in Word Embeddings

2020-06-30 · EMNLP 2021 11 · Sunipa Dev, Tao Li, Jeff M. Phillips, Vivek Srikumar

Language representations are known to carry stereotypical biases and, as a result, lead to biased predictions in downstream tasks. While existing methods are effective at mitigating biases by linear projection, such meth…

Word Embeddings

Group-sparse Matrix Recovery

2014-02-20 · Xiangrong Zeng, Mário A. T. Figueiredo

We apply the OSCAR (octagonal selection and clustering algorithms for regression) in recovering group-sparse matrices (two-dimensional---2D---arrays) from compressive measurements. We propose a 2D version of OSCAR (2OSCA…

Clusteringregression