GAE: Learning a Geometry-Native Latent Space for 3D-Consistent World Generation Paper • 2609.24981 • Published 3 days ago • 35
Paint-Anything: Unified Any-Color Control for Image Generation and Editing Paper • 2609.20816 • Published 7 days ago • 37
OmniVBench: A Benchmark and Large-Scale Dataset for Omni Reference-to-Video Generation Paper • 2609.22069 • Published 6 days ago • 28
DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression Paper • 2609.19969 • Published 7 days ago • 171
Video DeltaNet: A Video-Native Hybrid Attention for Livestream Video Generation Paper • 2609.20744 • Published 7 days ago • 49
PhysBrain 1.5: From Vision-Language Models to Physical Foundation Models Paper • 2609.14973 • Published 10 days ago • 172
Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation Paper • 2609.11638 • Published 14 days ago • 697
FastVideo/FastVideo-FastH3-4-step-Preview-v1-VSA-DataFree Text-to-Video • 35B • Updated 19 days ago • 534k • 310
The Missing Temporal Link: Temporal Context Routing for Script-Driven Audio-Video Generation Paper • 2609.02367 • Published 22 days ago • 37
LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes Paper • 2609.03796 • Published 21 days ago • 186
H3-World: Turning Language Understanding into World Control Paper • 2609.01560 • Published 23 days ago • 52
Ring Forcing: Towards Precise Long-Term Memory for Autoregressive Video Diffusion Paper • 2608.26794 • Published 28 days ago • 17
LayerRecall: A State-Conditioned Memory Router for Long-Horizon Consistency in Video Generation Paper • 2608.28460 • Published 27 days ago • 30
WithEveryone: Unified Planning and Identity Grounding for Group Image Generation Paper • 2608.20336 • Published Aug 20 • 42