WaveOp-LiteFM: Lightweight Neural-Operator Flow Matching for Satellite-to-Radar Precipitation Retrieval
Chunlei Shi, Yecheng Zhang, Yufeng Zhu, Dan Niu, Yichao Dong, Yongchao Feng, Junming Hou
Abstract
Satellite-to-radar (S2R) retrieval refers to estimating ground-based radar precipitation from geostationary satellite observations, enabling precipitation monitoring in regions with limited radar coverage. While recent generative flow matching models have greatly advanced retrieval quality, they face a critical trade-off: pixel-space formulations suffer from the prohibitive computational costs of attention-based U-Net velocity networks, whereas latent-space modeling often sacrifices fine precipitation details or struggles with sparse targets. To address this dilemma, we propose WaveOp-LiteFM, a lightweight neural operator flow matching framework for S2R retrieval. Our approach introduces a novel velocity backbone built upon the spectral-local-wavelet (SLW) block, enabling efficient and stable flow matching in pixel space. Specifically, the SLW block disentangles precipitation features into three distinct frequency regimes: (i) the spectral branch captures large-scale stratiform organization; (ii) the local branch models short-range interactions; and (iii) the wavelet branch enhances sharp structures while suppressing noisy high-frequency responses. Building on this design, an input-adaptive gating mechanism dynamically fuses features from the three functional branches. Furthermore, a skip gate efficiently reintegrates encoder features through additive fusion within the decoder, avoiding the costly channel concatenation used in conventional U-Net architectures. Experiments on the SEVIR and Southeast China datasets show that WaveOp-LiteFM achieves state-of-the-art retrieval performance while substantially reducing computational costs. Beyond benchmark evaluation, large-area inference over China, including a recent Typhoon Bavi case, demonstrates that WaveOp-LiteFM maintains reliable retrieval quality in large-scale real-world scenarios.
Create a lesson
Related papers
How AI Experiences Art: Emergent Aesthetic Structure in a Self-Supervised Multimodal Embedding Space
Corey D. C. Heath
Self-Reflective Multi-modal Reasoning for Short-Video Fake News Detection
Pinjie Xu, Yuzhou Yang, Zhikai Tan et al.
Emotion Understanding in Streaming Video with Trajectory-Aware Reliability
Qingsong Wang, Qigong Lei, Zitong Wang et al.
Learning to Prefer Reliably: Error-Augmented Emotion Preference Optimization with Calibrated Fusion
Zilong Huang, Junyi Peng, Junjie Li et al.
EVEREST:Endogenous Vision-Language Reinforcement Reasoning Exploration for Urban Socio-Semantic Segmentation
Qixiu Li, Zhongzhi He, Xiang Zhu et al.
Task-disentangled Low-Rank Adaptation for Versatile Audio-visual Multi-modal Learning Tasks within a Unified Framework
Hanyu Xuan, Mengqi Zhang, Junjun Mao et al.