-
Self-Distilled Agentic Reinforcement Learning
Paper • 2605.15155 • Published • 118 -
MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory
Paper • 2605.15128 • Published • 65 -
Orchard: An Open-Source Agentic Modeling Framework
Paper • 2605.15040 • Published • 21 -
MMSkills: Towards Multimodal Skills for General Visual Agents
Paper • 2605.13527 • Published • 124
Collections
Discover the best community collections!
Collections including paper arxiv:2605.15128
-
microsoft/VibeVoice-1.5B
Text-to-Speech • 3B • Updated • 378k • 2.48k -
pythontech9/AGENTIC-AI
Updated -
MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory
Paper • 2605.15128 • Published • 65 -
upstage/Solar-Open2-250B
Text Generation • 250B • Updated • 13.6k • 753
-
Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team
Paper • 2506.14234 • Published • 41 -
MoTE: Mixture of Ternary Experts for Memory-efficient Large Multimodal Models
Paper • 2506.14435 • Published • 7 -
Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory
Paper • 2504.19413 • Published • 71 -
MemOS: A Memory OS for AI System
Paper • 2507.03724 • Published • 170
-
Refusal in Language Models Is Mediated by a Single Direction
Paper • 2406.11717 • Published • 15 -
Self-Distilled Agentic Reinforcement Learning
Paper • 2605.15155 • Published • 118 -
MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models
Paper • 2605.14906 • Published • 79 -
MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory
Paper • 2605.15128 • Published • 65
-
YuanLabAI/Yuan3.0-Ultra-int4
1T • Updated • 12 • 6 -
yujiepan/qwen3.5-moe-tiny-random
Image-Text-to-Text • 4.9M • Updated • 1.06k • 3 -
meta-llama/Llama-3.2-1B
Text Generation • 1B • Updated • 1.12M • 2.59k -
moonshotai/Kimi-K2.6
Image-Text-to-Text • 1T • Updated • 538k • • 1.61k
-
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning
Paper • 2506.22434 • Published • 10 -
VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning
Paper • 2507.13348 • Published • 80 -
RewardDance: Reward Scaling in Visual Generation
Paper • 2509.08826 • Published • 73 -
Grasp Any Region: Towards Precise, Contextual Pixel Understanding for Multimodal LLMs
Paper • 2510.18876 • Published • 37
-
Self-Distilled Agentic Reinforcement Learning
Paper • 2605.15155 • Published • 118 -
MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory
Paper • 2605.15128 • Published • 65 -
Orchard: An Open-Source Agentic Modeling Framework
Paper • 2605.15040 • Published • 21 -
MMSkills: Towards Multimodal Skills for General Visual Agents
Paper • 2605.13527 • Published • 124
-
Refusal in Language Models Is Mediated by a Single Direction
Paper • 2406.11717 • Published • 15 -
Self-Distilled Agentic Reinforcement Learning
Paper • 2605.15155 • Published • 118 -
MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models
Paper • 2605.14906 • Published • 79 -
MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory
Paper • 2605.15128 • Published • 65
-
YuanLabAI/Yuan3.0-Ultra-int4
1T • Updated • 12 • 6 -
yujiepan/qwen3.5-moe-tiny-random
Image-Text-to-Text • 4.9M • Updated • 1.06k • 3 -
meta-llama/Llama-3.2-1B
Text Generation • 1B • Updated • 1.12M • 2.59k -
moonshotai/Kimi-K2.6
Image-Text-to-Text • 1T • Updated • 538k • • 1.61k
-
microsoft/VibeVoice-1.5B
Text-to-Speech • 3B • Updated • 378k • 2.48k -
pythontech9/AGENTIC-AI
Updated -
MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory
Paper • 2605.15128 • Published • 65 -
upstage/Solar-Open2-250B
Text Generation • 250B • Updated • 13.6k • 753
-
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning
Paper • 2506.22434 • Published • 10 -
VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning
Paper • 2507.13348 • Published • 80 -
RewardDance: Reward Scaling in Visual Generation
Paper • 2509.08826 • Published • 73 -
Grasp Any Region: Towards Precise, Contextual Pixel Understanding for Multimodal LLMs
Paper • 2510.18876 • Published • 37
-
Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team
Paper • 2506.14234 • Published • 41 -
MoTE: Mixture of Ternary Experts for Memory-efficient Large Multimodal Models
Paper • 2506.14435 • Published • 7 -
Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory
Paper • 2504.19413 • Published • 71 -
MemOS: A Memory OS for AI System
Paper • 2507.03724 • Published • 170