Scene Classification
2개 벤치마크 · 논문 496편 · 이 태스크의 논문 보기 →
Benchmarks
Places365-Standard
Most implemented
Spatial Information Considered Network for Scene Classification
Remote Sensing Image Scene Classification: Benchmark and State of the Art
RSMamba: Remote Sensing Image Classification with State Space Model
Vision-Language Models in Remote Sensing: Current Progress and Future Trends
Papers
Semi-Supervised Adaptation of Vision-Language Models for Image Classification
Vision-language models like CLIP have shown sig- nificant potential in handling natural images, yet their perfor- mance is often limited by the distinct characteristics of satellite imagery. While parameter-efficient ada…
Scene ClassificationImage ClassificationLabel-Decoupled Style Augmentation for Domain Generalization in Multi-Label Remote Sensing Scene Classification
Multi-label classification assigns several co-occurring labels to each aerial scene, yet deployed models often encounter data distributions different from their training. Feature-statistics augmentation such as MixStyle,…
Multi-Label ClassificationDomain GeneralizationScene ClassificationPromptable Concept Segmentation from Above: Evaluating SAM 3's Zero-Shot and One-Shot Capabilities in Remote Sensing
The deployment of large-scale foundation models, such as the Segment Anything Model 3 (SAM 3), promises a transition toward open-vocabulary, training-free computer vision. However, their capacity to generalize out-of-dis…
Instance SegmentationScene ClassificationObject DetectionTerraDiT-$Ω$: Unified Spatial Control for Satellite Image Synthesis with Any Geospatial Primitive
Generative models have achieved remarkable progress, yet applying them to satellite imagery remains challenging. Unlike natural imagery, satellite scenes are structured by spatially complex and semantically distinct geom…
Scene ClassificationData AugmentationObject DetectionMSNN-LINet: Cross-Modal Learning via Continuous Linear Integration
We present LINet (Linear Integration Network), a Multi-Stream Neural Network (MSNN) for RGB-D scene classification. Current multi-modal architectures treat feature fusion as a discrete, ad-hoc event: early fusion entangl…
Scene ClassificationWeGenBench: A Multidimensional Diagnostic Benchmark towards Text-to-Image Model Optimization
Recent text-to-image generation models have demonstrated remarkable capabilities in synthesizing highly realistic images from text inputs alone. Although existing benchmarks can evaluate the generation capabilities of va…
Text-to-Image GenerationScene Classification