FW-GAN: Frequency-Driven Handwriting Synthesis with Wave-Modulated MLP Generator
Labeled handwriting data is often scarce, limiting the effectiveness of recognition systems that require diverse, style-consistent training samples. Handwriting synthesis offers a promising solution by generating artificial data to augment training. However, current methods face two major limitations. First, most are built on conventional convolutional architectures, which struggle to model long-range dependencies and complex stroke patterns. Second, they largely ignore the crucial role of frequency information, which is essential for capturing fine-grained stylistic and structural details in handwriting. To address these challenges, we propose FW-GAN, a one-shot handwriting synthesis framework that generates realistic, writer-consistent text from a single example. Our generator integrates a phase-aware Wave-MLP to better capture spatial relationships while preserving subtle stylistic cues. We further introduce a frequency-guided discriminator that leverages high-frequency components to enhance the authenticity detection of generated samples. Additionally, we introduce a novel Frequency Distribution Loss that aligns the frequency characteristics of synthetic and real handwriting, thereby enhancing visual fidelity. Experiments on Vietnamese and English handwriting datasets demonstrate that FW-GAN generates high-quality, style-consistent handwriting, making it a valuable tool for augmenting data in low-resource handwriting recognition (HTR) pipelines. Official implementation is available at https://github.com/DAIR-Group/FW-GAN
Code (0)
등록된 구현이 없습니다.
Tasks
Handwriting RecognitionSimilar Papers 제목 키워드 기반
SpiS-GAN: Spiral-Modulated Handwriting Synthesis with Star Operation
Training robust handwriting recognition (HTR) systems requires massive amounts of annotated data, which is often difficult to acquire. While synthetic handwriting generation offers a practical solution to expand training…
Handwriting RecognitionMIMO Beampattern Synthesis using Adaptive Frequency Modulated Waveforms
This paper demonstrates a method that synthesizes narrowband Multiple-Input Multiple-Output (MIMO) beampatterns using the Multi-Tone Sinusoidal Frequency Modulated (MTSFM) waveform model. MIMO arrays transmit unique wave…
FREAK: Frequency-modulated High-fidelity and Real-time Audio-driven Talking Portrait Synthesis
Achieving high-fidelity lip-speech synchronization in audio-driven talking portrait synthesis remains challenging. While multi-stage pipelines or diffusion models yield high-quality results, they suffer from high computa…
Audio-Visual SynchronizationCDMA and Non-Uniform Multiplexing: Dynamic Range in MIMO Radar Waveforms
This paper presents a performance comparison of various MIMO radar multiplexing approaches where the increasing number of transmitters adversely affects the dynamic range of the resultant MIMO system. The investigated mu…
High-precision terahertz frequency modulated continuous wave imaging method using continuous wavelet transform
Inspired by the extensive application of terahertz imaging technologies in the field of aerospace, we exploit a terahertz frequency modulated continuous wave imaging method with continuous wavelet transform algorithm to …