TT-VidT: Decoupling the Temporal Axis for Efficient Motion-Centric Video Pretraining Paper • 2609.33419 • Published 3 days ago • 13
ReactVAU: A Slow-Fast Decoupled Framework for Streaming Video Anomaly Understanding Paper • 2609.07941 • Published 23 days ago • 19
TT-VidT: Decoupling the Temporal Axis for Efficient Motion-Centric Video Pretraining Paper • 2609.33419 • Published 3 days ago • 13
ImagenWorld: Stress-Testing Image Generation Models with Explainable Human Evaluation on Open-ended Real-World Tasks Paper • 2603.27862 • Published Mar 29 • 32
Navigating Text-To-Image Customization:From LyCORIS Fine-Tuning to Model Evaluation Paper • 2309.14859 • Published Sep 26, 2023 • 6
TIPO: Text to Image with Text Presampling for Prompt Optimization Paper • 2411.08127 • Published Nov 12, 2024 • 4
Graph-Based Captioning: Enhancing Visual Descriptions by Interconnecting Region Captions Paper • 2407.06723 • Published Jul 9, 2024 • 11
Clearer Frames, Anytime: Resolving Velocity Ambiguity in Video Frame Interpolation Paper • 2311.08007 • Published Nov 14, 2023 • 1
Navigating Text-To-Image Customization:From LyCORIS Fine-Tuning to Model Evaluation Paper • 2309.14859 • Published Sep 26, 2023 • 6