-
Self-Distilled Agentic Reinforcement Learning
Paper • 2605.15155 • Published • 117 -
MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory
Paper • 2605.15128 • Published • 64 -
Orchard: An Open-Source Agentic Modeling Framework
Paper • 2605.15040 • Published • 20 -
MMSkills: Towards Multimodal Skills for General Visual Agents
Paper • 2605.13527 • Published • 122
Collections
Discover the best community collections!
Collections including paper arxiv:2605.15128
-
microsoft/VibeVoice-1.5B
Text-to-Speech • 3B • Updated • 684k • 2.49k -
pythontech9/AGENTIC-AI
Updated -
MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory
Paper • 2605.15128 • Published • 64 -
upstage/Solar-Open2-250B
Text Generation • 250B • Updated • 5.34k • 756
-
Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team
Paper • 2506.14234 • Published • 41 -
MoTE: Mixture of Ternary Experts for Memory-efficient Large Multimodal Models
Paper • 2506.14435 • Published • 7 -
Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory
Paper • 2504.19413 • Published • 74 -
MemOS: A Memory OS for AI System
Paper • 2507.03724 • Published • 170
-
Refusal in Language Models Is Mediated by a Single Direction
Paper • 2406.11717 • Published • 16 -
Self-Distilled Agentic Reinforcement Learning
Paper • 2605.15155 • Published • 117 -
MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models
Paper • 2605.14906 • Published • 77 -
MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory
Paper • 2605.15128 • Published • 64
-
YuanLabAI/Yuan3.0-Ultra-int4
1T • Updated • 19 • 7 -
yujiepan/qwen3.5-moe-tiny-random
Image-Text-to-Text • 4.9M • Updated • 1.4k • 3 -
meta-llama/Llama-3.2-1B
Text Generation • 1B • Updated • 697k • • 2.68k -
moonshotai/Kimi-K2.6
Image-Text-to-Text • 1T • Updated • 425k • • 1.61k
-
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning
Paper • 2506.22434 • Published • 10 -
VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning
Paper • 2507.13348 • Published • 80 -
RewardDance: Reward Scaling in Visual Generation
Paper • 2509.08826 • Published • 73 -
Grasp Any Region: Towards Precise, Contextual Pixel Understanding for Multimodal LLMs
Paper • 2510.18876 • Published • 37
-
Self-Distilled Agentic Reinforcement Learning
Paper • 2605.15155 • Published • 117 -
MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory
Paper • 2605.15128 • Published • 64 -
Orchard: An Open-Source Agentic Modeling Framework
Paper • 2605.15040 • Published • 20 -
MMSkills: Towards Multimodal Skills for General Visual Agents
Paper • 2605.13527 • Published • 122
-
Refusal in Language Models Is Mediated by a Single Direction
Paper • 2406.11717 • Published • 16 -
Self-Distilled Agentic Reinforcement Learning
Paper • 2605.15155 • Published • 117 -
MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models
Paper • 2605.14906 • Published • 77 -
MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory
Paper • 2605.15128 • Published • 64
-
YuanLabAI/Yuan3.0-Ultra-int4
1T • Updated • 19 • 7 -
yujiepan/qwen3.5-moe-tiny-random
Image-Text-to-Text • 4.9M • Updated • 1.4k • 3 -
meta-llama/Llama-3.2-1B
Text Generation • 1B • Updated • 697k • • 2.68k -
moonshotai/Kimi-K2.6
Image-Text-to-Text • 1T • Updated • 425k • • 1.61k
-
microsoft/VibeVoice-1.5B
Text-to-Speech • 3B • Updated • 684k • 2.49k -
pythontech9/AGENTIC-AI
Updated -
MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory
Paper • 2605.15128 • Published • 64 -
upstage/Solar-Open2-250B
Text Generation • 250B • Updated • 5.34k • 756
-
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning
Paper • 2506.22434 • Published • 10 -
VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning
Paper • 2507.13348 • Published • 80 -
RewardDance: Reward Scaling in Visual Generation
Paper • 2509.08826 • Published • 73 -
Grasp Any Region: Towards Precise, Contextual Pixel Understanding for Multimodal LLMs
Paper • 2510.18876 • Published • 37
-
Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team
Paper • 2506.14234 • Published • 41 -
MoTE: Mixture of Ternary Experts for Memory-efficient Large Multimodal Models
Paper • 2506.14435 • Published • 7 -
Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory
Paper • 2504.19413 • Published • 74 -
MemOS: A Memory OS for AI System
Paper • 2507.03724 • Published • 170