AURORA-LM: Autoencoding Unified Representation for Continuous-Latent Diffusion Language Modeling Paper • 2608.02602 • Published 16 days ago • 80
ACE-Data-0: Human-Centric Ambient Capture as Embodied Data Engine Paper • 2607.28625 • Published 20 days ago • 45
HumanCLAW: Can Vision-Language Models Act Through a Body? Paper • 2607.27180 • Published 21 days ago • 76
Apple-π: Benchmarking Thinking with Video Towards Law-Grounded Physical Intelligence Paper • 2607.16401 • Published Jul 17 • 44
VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding Paper • 2607.14935 • Published Jul 16 • 172
Running Agents Featured 46 SenseNova Vision 📚 46 Analyze or generate images with AI-powered vision tasks
SenseNova-Vision Collection Vision as Unified Multimodal Generation • 5 items • Updated Jul 10 • 34
S-Agent: Spatial Tool-Use Elicits Reasoning for Spatial Intelligence Paper • 2606.20515 • Published Jun 18 • 41