Robot Manipulation Generalization
2개 벤치마크 · 논문 19편 · 이 태스크의 논문 보기 →
Benchmarks
The COLOSSEUM
GEMBench
Most implemented
Segment Anything
SAM 2: Segment Anything in Images and Videos
Sam2Rad: A Segmentation Model for Medical Images with Learnable Prompts
Segment Anything for Videos: A Systematic Survey
Instruction-driven history-aware policies for robotic manipulations
Papers
Towards Generalizable Vision-Language Robotic Manipulation: A Benchmark and LLM-guided 3D Policy
Generalizing language-conditioned robotic policies to new tasks remains a significant challenge, hampered by the lack of suitable simulation benchmarks. In this paper, we address this gap by introducing GemBench, a novel…
Motion PlanningRobot ManipulationRobot Manipulation GeneralizationTask PlanningSam2Rad: A Segmentation Model for Medical Images with Learnable Prompts
Foundation models like the segment anything model require high-quality manual prompts for medical image segmentation, which is time-consuming and requires expertise. SAM and its variants often fail to segment structures …
Image SegmentationMedical Image Segmentationparameter-efficient fine-tuningPrompt Learning+3SAM2-Adapter: Evaluating & Adapting Segment Anything 2 in Downstream Tasks: Camouflage, Shadow, Medical Image Segmentation, and More
The advent of large models, also known as foundation models, has significantly transformed the AI research landscape, with models like Segment Anything (SAM) achieving notable success in diverse image segmentation scenar…
Image SegmentationMedical Image Segmentationobject-detectionObject Detection+4SAM 2: Segment Anything in Images and Videos
We present Segment Anything Model 2 (SAM 2), a foundation model towards solving promptable visual segmentation in images and videos. We build a data engine, which improves model and data via user interaction, to collect …
Image SegmentationRobot Manipulation GeneralizationSegmentationSemantic Segmentation+5Segment Anything for Videos: A Systematic Survey
The recent wave of foundation models has witnessed tremendous success in computer vision (CV) and beyond, with the segment anything model (SAM) having sparked a passion for exploring task-agnostic visual foundation model…
Image SegmentationRobot Manipulation GeneralizationSemantic SegmentationSurvey+4Generative Image as Action Models
Image-generation diffusion models have been fine-tuned to unlock new capabilities such as image-editing and novel view synthesis. Can we similarly unlock image-generation models for visuomotor control? We present GENIMA,…
Image GenerationRobot ManipulationRobot Manipulation Generalization