paper-with-me

Robot Manipulation Generalization

2개 벤치마크 · 논문 19편 · 이 태스크의 논문 보기 →

Benchmarks

The COLOSSEUM

결과 9개

GEMBench

결과 6개

Most implemented

Segment Anything

2023-04-05 · 구현 32개

Papers

Towards Generalizable Vision-Language Robotic Manipulation: A Benchmark and LLM-guided 3D Policy

2024-10-02 · Ricardo Garcia, ShiZhe Chen, Cordelia Schmid

Generalizing language-conditioned robotic policies to new tasks remains a significant challenge, hampered by the lack of suitable simulation benchmarks. In this paper, we address this gap by introducing GemBench, a novel…

Motion PlanningRobot ManipulationRobot Manipulation GeneralizationTask Planning

Sam2Rad: A Segmentation Model for Medical Images with Learnable Prompts

2024-09-10 · Assefa Seyoum Wahd, Banafshe Felfeliyan, Yuyue Zhou, Shrimanti Ghosh 외

Foundation models like the segment anything model require high-quality manual prompts for medical image segmentation, which is time-consuming and requires expertise. SAM and its variants often fail to segment structures …

Image SegmentationMedical Image Segmentationparameter-efficient fine-tuningPrompt Learning+3

SAM2-Adapter: Evaluating & Adapting Segment Anything 2 in Downstream Tasks: Camouflage, Shadow, Medical Image Segmentation, and More

2024-08-08 · Tianrun Chen, Ankang Lu, Lanyun Zhu, Chaotao Ding 외

The advent of large models, also known as foundation models, has significantly transformed the AI research landscape, with models like Segment Anything (SAM) achieving notable success in diverse image segmentation scenar…

Image SegmentationMedical Image Segmentationobject-detectionObject Detection+4

SAM 2: Segment Anything in Images and Videos

2024-08-01 · Nikhila Ravi, Valentin Gabeur, Yuan-Ting Hu, Ronghang Hu 외

We present Segment Anything Model 2 (SAM 2), a foundation model towards solving promptable visual segmentation in images and videos. We build a data engine, which improves model and data via user interaction, to collect …

Image SegmentationRobot Manipulation GeneralizationSegmentationSemantic Segmentation+5

Segment Anything for Videos: A Systematic Survey

2024-07-31 · Chunhui Zhang, Yawen Cui, Weilin Lin, Guanjie Huang 외

The recent wave of foundation models has witnessed tremendous success in computer vision (CV) and beyond, with the segment anything model (SAM) having sparked a passion for exploring task-agnostic visual foundation model…

Image SegmentationRobot Manipulation GeneralizationSemantic SegmentationSurvey+4

Generative Image as Action Models

2024-07-10 · Mohit Shridhar, Yat Long Lo, Stephen James

Image-generation diffusion models have been fine-tuned to unlock new capabilities such as image-editing and novel view synthesis. Can we similarly unlock image-generation models for visuomotor control? We present GENIMA,…

Image GenerationRobot ManipulationRobot Manipulation Generalization

전체 19편 보기 →