-
Demystifying Agent Skills: Why They Work-Until They Don't
Paper • 2608.14036 • Published • 169 -
Harness the Memory: A Holistic Evaluation of Memory Substrates in Memory Agents
Paper • 2608.15008 • Published • 15 -
Oracle Agent Memory as an Enterprise Memory Substrate for Long-Horizon AI Agents
Paper • 2607.13157 • Published -
SelfMem: Self-Optimizing Memory for AI Agents
Paper • 2607.03726 • Published
Collections
Discover the best community collections!
Collections including paper arxiv:2604.04804
-
SkillEvolBench: Benchmarking the Evolution from Episodic Experience to Procedural Skills
Paper • 2605.24117 • Published • 22 -
MUSE-Autoskill: Self-Evolving Agents via Skill Creation, Memory, Management, and Evaluation
Paper • 2605.27366 • Published • 30 -
SkillGrad: Optimizing Agent Skills Like Gradient Descent
Paper • 2605.27760 • Published • 27 -
Skill0.5: Joint Skill Internalization and Utilization for Out-of-Distribution Generalization in Agentic Reinforcement Learning
Paper • 2605.28424 • Published • 32
-
Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills
Paper • 2603.25158 • Published • 56 -
SkillClaw: Let Skills Evolve Collectively with Agentic Evolver
Paper • 2604.08377 • Published • 294 -
How Well Do Agentic Skills Work in the Wild: Benchmarking LLM Skill Usage in Realistic Settings
Paper • 2604.04323 • Published • 41 -
SkillX: Automatically Constructing Skill Knowledge Bases for Agents
Paper • 2604.04804 • Published • 35
-
XSkill: Continual Learning from Experience and Skills in Multimodal Agents
Paper • 2603.12056 • Published • 34 -
Memento-Skills: Let Agents Design Agents
Paper • 2603.18743 • Published • 59 -
SWE-Skills-Bench: Do Agent Skills Actually Help in Real-World Software Engineering?
Paper • 2603.15401 • Published • 20 -
Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills
Paper • 2603.25158 • Published • 56
-
BitNet: Scaling 1-bit Transformers for Large Language Models
Paper • 2310.11453 • Published • 108 -
Self-RAG: Learning to Retrieve, Generate, and Critique through Self-Reflection
Paper • 2310.11511 • Published • 79 -
In-Context Learning Creates Task Vectors
Paper • 2310.15916 • Published • 43 -
Matryoshka Diffusion Models
Paper • 2310.15111 • Published • 46
-
Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills
Paper • 2603.25158 • Published • 56 -
SkillX: Automatically Constructing Skill Knowledge Bases for Agents
Paper • 2604.04804 • Published • 35 -
Dive into Claude Code: The Design Space of Today's and Future AI Agent Systems
Paper • 2604.14228 • Published • 25 -
OpenMobile: Building Open Mobile Agents with Task and Trajectory Synthesis
Paper • 2604.15093 • Published • 30
-
SkillX: Automatically Constructing Skill Knowledge Bases for Agents
Paper • 2604.04804 • Published • 35 -
XSkill: Continual Learning from Experience and Skills in Multimodal Agents
Paper • 2603.12056 • Published • 34 -
AutoAgent: Evolving Cognition and Elastic Memory Orchestration for Adaptive Agents
Paper • 2603.09716 • Published • 1 -
SkillOS: Learning Skill Curation for Self-Evolving Agents
Paper • 2605.06614 • Published • 48
-
dLLM: Simple Diffusion Language Modeling
Paper • 2602.22661 • Published • 154 -
OpenSeeker: Democratizing Frontier Search Agents by Fully Open-Sourcing Training Data
Paper • 2603.15594 • Published • 150 -
Qianfan-OCR: A Unified End-to-End Model for Document Intelligence
Paper • 2603.13398 • Published • 155 -
Penguin-VL: Exploring the Efficiency Limits of VLM with LLM-based Vision Encoders
Paper • 2603.06569 • Published • 120
-
lusxvr/nanoVLM-222M
Image-Text-to-Text • 0.2B • Updated • 1.35k • 103 -
Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Paper • 2503.09516 • Published • 41 -
AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time
Paper • 2505.24863 • Published • 98 -
QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning
Paper • 2505.17667 • Published • 89
-
Demystifying Agent Skills: Why They Work-Until They Don't
Paper • 2608.14036 • Published • 169 -
Harness the Memory: A Holistic Evaluation of Memory Substrates in Memory Agents
Paper • 2608.15008 • Published • 15 -
Oracle Agent Memory as an Enterprise Memory Substrate for Long-Horizon AI Agents
Paper • 2607.13157 • Published -
SelfMem: Self-Optimizing Memory for AI Agents
Paper • 2607.03726 • Published
-
Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills
Paper • 2603.25158 • Published • 56 -
SkillX: Automatically Constructing Skill Knowledge Bases for Agents
Paper • 2604.04804 • Published • 35 -
Dive into Claude Code: The Design Space of Today's and Future AI Agent Systems
Paper • 2604.14228 • Published • 25 -
OpenMobile: Building Open Mobile Agents with Task and Trajectory Synthesis
Paper • 2604.15093 • Published • 30
-
SkillEvolBench: Benchmarking the Evolution from Episodic Experience to Procedural Skills
Paper • 2605.24117 • Published • 22 -
MUSE-Autoskill: Self-Evolving Agents via Skill Creation, Memory, Management, and Evaluation
Paper • 2605.27366 • Published • 30 -
SkillGrad: Optimizing Agent Skills Like Gradient Descent
Paper • 2605.27760 • Published • 27 -
Skill0.5: Joint Skill Internalization and Utilization for Out-of-Distribution Generalization in Agentic Reinforcement Learning
Paper • 2605.28424 • Published • 32
-
SkillX: Automatically Constructing Skill Knowledge Bases for Agents
Paper • 2604.04804 • Published • 35 -
XSkill: Continual Learning from Experience and Skills in Multimodal Agents
Paper • 2603.12056 • Published • 34 -
AutoAgent: Evolving Cognition and Elastic Memory Orchestration for Adaptive Agents
Paper • 2603.09716 • Published • 1 -
SkillOS: Learning Skill Curation for Self-Evolving Agents
Paper • 2605.06614 • Published • 48
-
Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills
Paper • 2603.25158 • Published • 56 -
SkillClaw: Let Skills Evolve Collectively with Agentic Evolver
Paper • 2604.08377 • Published • 294 -
How Well Do Agentic Skills Work in the Wild: Benchmarking LLM Skill Usage in Realistic Settings
Paper • 2604.04323 • Published • 41 -
SkillX: Automatically Constructing Skill Knowledge Bases for Agents
Paper • 2604.04804 • Published • 35
-
dLLM: Simple Diffusion Language Modeling
Paper • 2602.22661 • Published • 154 -
OpenSeeker: Democratizing Frontier Search Agents by Fully Open-Sourcing Training Data
Paper • 2603.15594 • Published • 150 -
Qianfan-OCR: A Unified End-to-End Model for Document Intelligence
Paper • 2603.13398 • Published • 155 -
Penguin-VL: Exploring the Efficiency Limits of VLM with LLM-based Vision Encoders
Paper • 2603.06569 • Published • 120
-
XSkill: Continual Learning from Experience and Skills in Multimodal Agents
Paper • 2603.12056 • Published • 34 -
Memento-Skills: Let Agents Design Agents
Paper • 2603.18743 • Published • 59 -
SWE-Skills-Bench: Do Agent Skills Actually Help in Real-World Software Engineering?
Paper • 2603.15401 • Published • 20 -
Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills
Paper • 2603.25158 • Published • 56
-
lusxvr/nanoVLM-222M
Image-Text-to-Text • 0.2B • Updated • 1.35k • 103 -
Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Paper • 2503.09516 • Published • 41 -
AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time
Paper • 2505.24863 • Published • 98 -
QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning
Paper • 2505.17667 • Published • 89
-
BitNet: Scaling 1-bit Transformers for Large Language Models
Paper • 2310.11453 • Published • 108 -
Self-RAG: Learning to Retrieve, Generate, and Critique through Self-Reflection
Paper • 2310.11511 • Published • 79 -
In-Context Learning Creates Task Vectors
Paper • 2310.15916 • Published • 43 -
Matryoshka Diffusion Models
Paper • 2310.15111 • Published • 46