paper-with-me

홈 › Papers

A Large Scale Open-Source Image and Video Dataset for Robust Wildfire Detection and Classification

2026-06-08 · Emadeldeen Hamdan, Yingyi Luo, B. Ugur Toreyin, Erdem Koyuncu, Adam J. Watts, Ugur Gudukbay, Ahmet Enis Cetin arxiv

Wildfire detection and monitoring are critical for mitigating fire spread and reducing environmental and infrastructural damage. In this work, we introduce GWFP (Global Wildfire Prevention Dataset), a large-scale, open-source dataset of wildfire images and videos designed to support early fire and smoke detection research. GWFP contains geographically diverse wildfire scenes, including flames, smoke, Waterdog/Fog environmental conditions, Near Infrared (NIR) imagery, Ember, and challenging negative samples collected from real-world scenarios worldwide. To evaluate dataset robustness and cross-domain generalization, we benchmark multiple convolutional and transformer-based architectures across both in-domain and cross-dataset settings. Additionally, we explore lightweight frequency--spatial feature interaction using Hadamard-enhanced residual connections (HTE-ResNet) to analyze representation robustness under domain-shift conditions. Experimental results demonstrate strong cross-dataset generalization and practical utility for real-world wildfire monitoring applications. The dataset and source code will be publicly released upon acceptance.

📄 PDF Abstract BibTeX arXiv:2606.10174

Code (0)

등록된 구현이 없습니다.

Tasks

Domain Generalization

Similar Papers 제목 키워드 기반

OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

2025-12-08 · Haoyang He, Jie Wang, Jiangning Zhang, Zhucun Xue 외 arxiv

The quality and diversity of instruction-based image editing datasets are continuously increasing, yet large-scale, high-quality datasets for instruction-based video editing remain scarce. To address this gap, we introdu…

Image Editing

LAION-BVD: A 10-Million-Hour Open Video Dataset for Multimodal Pre-training

2026-08-25 · Andreas Hochlehnert, Marianna Nezhurina, Mehdi Cherti, Andrej Radonjic 외 hf

We present LAION-BVD, a large-scale open video dataset for multimodal learning, which contains 1.3B platform-specific video URLs collected from CommonCrawl. From these, we download 80M videos with a total duration of 10 …

Text Retrieval

MUG-V 10B: High-efficiency Training Pipeline for Large Video Generation Models

2025-10-20 · Yongshun Zhang, Zhongyi Fan, Yonghang Zhang, Zhangzikang Li 외 arxiv

In recent years, large-scale generative models for visual content (\textit{e.g.,} images, videos, and 3D objects/scenes) have made remarkable progress. However, training large-scale video generation models remains partic…

Video GenerationVideo Alignment

Wan: Open and Advanced Large-Scale Video Generative Models

2025-03-26 · Team Wan, Ang Wang, Baole Ai, Bin Wen 외

This report presents Wan, a comprehensive and open suite of video foundation models designed to push the boundaries of video generation. Built upon the mainstream diffusion transformer paradigm, Wan achieves significant …

Video EditingVideo Generation

CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers

2022-05-29 · Wenyi Hong, Ming Ding, Wendi Zheng, Xinghan Liu 외

Large-scale pretrained transformers have created milestones in text (GPT-3) and text-to-image (DALL-E and CogView) generation. Its application to video generation is still facing many challenges: The potential huge compu…

Text-to-Video GenerationVideo Generation