paper-with-me

SuperRS-VQA, HighRS-VQA

홈페이지 · 논문 1편

We introduce SuperRS-VQA (avg. 8,376×8,376) and HighRS-VQA (avg. 2,000×1,912), the highest-resolution vision-language datasets in RS to date, covering 22 real-world dialogue tasks