Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Paper • 2607.18789 • Published 16 days ago • 3
Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models Paper • 2607.12463 • Published 23 days ago • 107
phonsobon/Images_captioning_fine_tune_Florence-2-base Image-to-Text • 0.2B • Updated 20 days ago • 68 • 1
Learning A Unified Risk Map for Autonomous Driving in Partially Observable Environments Paper • 2605.22189 • Published May 21 • 8
Geometry Matters: 3D Foundation Priors for Learning Semantic Correspondence Paper • 2605.30093 • Published May 28 • 15
Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players Paper • 2605.28816 • Published May 27 • 433