ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search Paper • 2609.13356 • Published 7 days ago • 311
Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation Paper • 2609.11638 • Published 8 days ago • 714
FreeFlow: A Bias-free Hierarchical Transformer for Optical Flow Estimation Paper • 2609.11486 • Published 8 days ago • 34
SenseNova-U1.5: Towards Native Unified Visual Intelligence Paper • 2609.11929 • Published 8 days ago • 262
Marigold V2: Revisiting Diffusion Transformers for Monocular Depth Estimation Paper • 2609.08084 • Published 10 days ago • 72
One Symptom, Three Levers: A Critical Review of On-Policy Self-Distillation Paper • 2608.25936 • Published 23 days ago • 17
ZipTok3D: High-Fidelity 3D Tokenization with Compact Token Prefixes Paper • 2609.01740 • Published 17 days ago • 28
On the Design Fundamentals of Pixel Text Representation Learning Paper • 2609.01147 • Published 17 days ago • 32
It Takes Two to Match: Co-Evolving Generative Retriever with Reinforcement Learning Paper • 2609.00638 • Published 17 days ago • 75
What Makes Good Agentic Data? An ACE Lens on Data Generation for LLM Agents Paper • 2608.27260 • Published 22 days ago • 73
From Seeing to Acting: Smart Glasses as First-Person Intelligence Platforms Paper • 2608.24877 • Published 24 days ago • 11
Annotations as Rollouts: Efficient and Scalable Reinforcement Learning for Video MLLMs Paper • 2608.20492 • Published 29 days ago • 111
InfinityEdit: Infinite Video Editing with a Lightweight Edit-Ignition Adapter Paper • 2608.20910 • Published 28 days ago • 39
Graph Engineering in the Era of LLM Agents: From Individual Intelligence to System Intelligence Paper • 2608.21156 • Published 28 days ago • 63
SimpleOPD: Simple Tokenizer-Agnostic On-Policy Distillation for Long-Context Reasoning Paper • 2608.14277 • Published Aug 14 • 36