Preprint2026
StreamOPD: A Post-Training Recipe with Spatio-Temporal Cue Gating for Streaming Video Understanding
Keming Wu, Baoyi Wang, Kaichen Zhang, et al.
StreamOPD is a post-training recipe combining verifiable streaming-video data, thinking-mode on-policy distillation, and instruct-mode deployment, improving streaming video understanding without inference-time memory or retrieval.
0Aug 17, 2026Reinforcement LearningRetrieval Augmented Generation
arXiv