-
DataComp-VLM: Improved Open Datasets for Vision-Language Models
Paper • 2606.28551 • Published • 51 -
SkillCoach: Self-Evolving Rubrics for Evaluating and Enhancing Agentic Skill-Use
Paper • 2607.01874 • Published • 21 -
PACE: A Proxy for Agentic Capability Evaluation
Paper • 2607.02032 • Published • 18 -
Measuring the Gap Between Human and LLM Research Ideas
Paper • 2607.01233 • Published • 18
Collections
Discover the best community collections!
Collections including paper arxiv:2605.18747
-
A Survey of On-Policy Distillation for Large Language Models
Paper • 2604.00626 • Published • 11 -
World Action Models: A Survey
Paper • 2606.20781 • Published • 59 -
Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models
Paper • 2505.04921 • Published • 187 -
The Landscape of Agentic Reinforcement Learning for LLMs: A Survey
Paper • 2509.02547 • Published • 238
-
Code as Agent Harness
Paper • 2605.18747 • Published • 224 -
SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture
Paper • 2605.12500 • Published • 198 -
From Context to Skills: Can Language Models Learn from Context Skillfully?
Paper • 2604.27660 • Published • 73 -
PhysBrain 1.0 Technical Report
Paper • 2605.15298 • Published • 61
-
Graph Neural Network Training with Data Tiering
Paper • 2111.05894 • Published -
Graph Neural Networks are Dynamic Programmers
Paper • 2203.15544 • Published • 1 -
Graph Neural Networks for Jamming Source Localization
Paper • 2506.03196 • Published -
Code as Agent Harness
Paper • 2605.18747 • Published • 224
-
Harnessing LLM Agents with Skill Programs
Paper • 2605.17734 • Published • 34 -
Code as Agent Harness
Paper • 2605.18747 • Published • 224 -
Kimi K3: Open Frontier Intelligence
Paper • 2607.24653 • Published • 522 -
SemaPLC: A Project-Grounded, Verification-Gated Agent Harness for PLC Code Generation
Paper • 2608.18565 • Published • 74
-
DataComp-VLM: Improved Open Datasets for Vision-Language Models
Paper • 2606.28551 • Published • 51 -
SkillCoach: Self-Evolving Rubrics for Evaluating and Enhancing Agentic Skill-Use
Paper • 2607.01874 • Published • 21 -
PACE: A Proxy for Agentic Capability Evaluation
Paper • 2607.02032 • Published • 18 -
Measuring the Gap Between Human and LLM Research Ideas
Paper • 2607.01233 • Published • 18
-
A Survey of On-Policy Distillation for Large Language Models
Paper • 2604.00626 • Published • 11 -
World Action Models: A Survey
Paper • 2606.20781 • Published • 59 -
Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models
Paper • 2505.04921 • Published • 187 -
The Landscape of Agentic Reinforcement Learning for LLMs: A Survey
Paper • 2509.02547 • Published • 238
-
Code as Agent Harness
Paper • 2605.18747 • Published • 224 -
SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture
Paper • 2605.12500 • Published • 198 -
From Context to Skills: Can Language Models Learn from Context Skillfully?
Paper • 2604.27660 • Published • 73 -
PhysBrain 1.0 Technical Report
Paper • 2605.15298 • Published • 61
-
Graph Neural Network Training with Data Tiering
Paper • 2111.05894 • Published -
Graph Neural Networks are Dynamic Programmers
Paper • 2203.15544 • Published • 1 -
Graph Neural Networks for Jamming Source Localization
Paper • 2506.03196 • Published -
Code as Agent Harness
Paper • 2605.18747 • Published • 224
-
Harnessing LLM Agents with Skill Programs
Paper • 2605.17734 • Published • 34 -
Code as Agent Harness
Paper • 2605.18747 • Published • 224 -
Kimi K3: Open Frontier Intelligence
Paper • 2607.24653 • Published • 522 -
SemaPLC: A Project-Grounded, Verification-Gated Agent Harness for PLC Code Generation
Paper • 2608.18565 • Published • 74