-
SkillEvolBench: Benchmarking the Evolution from Episodic Experience to Procedural Skills
Paper • 2605.24117 • Published • 22 -
MUSE-Autoskill: Self-Evolving Agents via Skill Creation, Memory, Management, and Evaluation
Paper • 2605.27366 • Published • 30 -
SkillGrad: Optimizing Agent Skills Like Gradient Descent
Paper • 2605.27760 • Published • 27 -
Skill0.5: Joint Skill Internalization and Utilization for Out-of-Distribution Generalization in Agentic Reinforcement Learning
Paper • 2605.28424 • Published • 32
Collections
Discover the best community collections!
Collections including paper arxiv:2602.02474
-
Balancing Specialized and General Skills in LLMs: The Impact of Modern Tuning and Data Strategy
Paper • 2310.04945 • Published • 1 -
Skill-it! A Data-Driven Skills Framework for Understanding and Training Language Models
Paper • 2307.14430 • Published • 3 -
SoK: Agentic Skills -- Beyond Tool Use in LLM Agents
Paper • 2602.20867 • Published • 2 -
MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents
Paper • 2602.02474 • Published • 63
-
CAR-bench: Evaluating the Consistency and Limit-Awareness of LLM Agents under Real-World Uncertainty
Paper • 2601.22027 • Published • 85 -
Reinforcement World Model Learning for LLM-based Agents
Paper • 2602.05842 • Published • 28 -
Accurate Failure Prediction in Agents Does Not Imply Effective Failure Prevention
Paper • 2602.03338 • Published • 26 -
MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents
Paper • 2602.02474 • Published • 63
-
STEP3-VL-10B Technical Report
Paper • 2601.09668 • Published • 196 -
Advancing Open-source World Models
Paper • 2601.20540 • Published • 135 -
Youtu-Agent: Scaling Agent Productivity with Automated Generation and Hybrid Policy Optimization
Paper • 2512.24615 • Published • 119 -
SkillRL: Evolving Agents via Recursive Skill-Augmented Reinforcement Learning
Paper • 2602.08234 • Published • 76
-
dLLM: Simple Diffusion Language Modeling
Paper • 2602.22661 • Published • 154 -
OpenSeeker: Democratizing Frontier Search Agents by Fully Open-Sourcing Training Data
Paper • 2603.15594 • Published • 150 -
Qianfan-OCR: A Unified End-to-End Model for Document Intelligence
Paper • 2603.13398 • Published • 155 -
Penguin-VL: Exploring the Efficiency Limits of VLM with LLM-based Vision Encoders
Paper • 2603.06569 • Published • 120
-
GLM-5: from Vibe Coding to Agentic Engineering
Paper • 2602.15763 • Published • 205 -
Recurrent-Depth VLA: Implicit Test-Time Compute Scaling of Vision-Language-Action Models via Latent Iterative Reasoning
Paper • 2602.07845 • Published • 71 -
LLaDA2.1: Speeding Up Text Diffusion via Token Editing
Paper • 2602.08676 • Published • 71 -
MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents
Paper • 2602.02474 • Published • 63
-
Agentic Reasoning for Large Language Models
Paper • 2601.12538 • Published • 207 -
From Code Foundation Models to Agents and Applications: A Practical Guide to Code Intelligence
Paper • 2511.18538 • Published • 306 -
Agent Learning via Early Experience
Paper • 2510.08558 • Published • 276 -
Weak-Driven Learning: How Weak Agents make Strong Agents Stronger
Paper • 2602.08222 • Published • 290
-
VisMem: Latent Vision Memory Unlocks Potential of Vision-Language Models
Paper • 2511.11007 • Published • 15 -
O-Mem: Omni Memory System for Personalized, Long Horizon, Self-Evolving Agents
Paper • 2511.13593 • Published • 28 -
General Agentic Memory Via Deep Research
Paper • 2511.18423 • Published • 172 -
MemEvolve: Meta-Evolution of Agent Memory Systems
Paper • 2512.18746 • Published • 31
-
SkillEvolBench: Benchmarking the Evolution from Episodic Experience to Procedural Skills
Paper • 2605.24117 • Published • 22 -
MUSE-Autoskill: Self-Evolving Agents via Skill Creation, Memory, Management, and Evaluation
Paper • 2605.27366 • Published • 30 -
SkillGrad: Optimizing Agent Skills Like Gradient Descent
Paper • 2605.27760 • Published • 27 -
Skill0.5: Joint Skill Internalization and Utilization for Out-of-Distribution Generalization in Agentic Reinforcement Learning
Paper • 2605.28424 • Published • 32
-
dLLM: Simple Diffusion Language Modeling
Paper • 2602.22661 • Published • 154 -
OpenSeeker: Democratizing Frontier Search Agents by Fully Open-Sourcing Training Data
Paper • 2603.15594 • Published • 150 -
Qianfan-OCR: A Unified End-to-End Model for Document Intelligence
Paper • 2603.13398 • Published • 155 -
Penguin-VL: Exploring the Efficiency Limits of VLM with LLM-based Vision Encoders
Paper • 2603.06569 • Published • 120
-
Balancing Specialized and General Skills in LLMs: The Impact of Modern Tuning and Data Strategy
Paper • 2310.04945 • Published • 1 -
Skill-it! A Data-Driven Skills Framework for Understanding and Training Language Models
Paper • 2307.14430 • Published • 3 -
SoK: Agentic Skills -- Beyond Tool Use in LLM Agents
Paper • 2602.20867 • Published • 2 -
MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents
Paper • 2602.02474 • Published • 63
-
GLM-5: from Vibe Coding to Agentic Engineering
Paper • 2602.15763 • Published • 205 -
Recurrent-Depth VLA: Implicit Test-Time Compute Scaling of Vision-Language-Action Models via Latent Iterative Reasoning
Paper • 2602.07845 • Published • 71 -
LLaDA2.1: Speeding Up Text Diffusion via Token Editing
Paper • 2602.08676 • Published • 71 -
MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents
Paper • 2602.02474 • Published • 63
-
Agentic Reasoning for Large Language Models
Paper • 2601.12538 • Published • 207 -
From Code Foundation Models to Agents and Applications: A Practical Guide to Code Intelligence
Paper • 2511.18538 • Published • 306 -
Agent Learning via Early Experience
Paper • 2510.08558 • Published • 276 -
Weak-Driven Learning: How Weak Agents make Strong Agents Stronger
Paper • 2602.08222 • Published • 290
-
CAR-bench: Evaluating the Consistency and Limit-Awareness of LLM Agents under Real-World Uncertainty
Paper • 2601.22027 • Published • 85 -
Reinforcement World Model Learning for LLM-based Agents
Paper • 2602.05842 • Published • 28 -
Accurate Failure Prediction in Agents Does Not Imply Effective Failure Prevention
Paper • 2602.03338 • Published • 26 -
MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents
Paper • 2602.02474 • Published • 63
-
STEP3-VL-10B Technical Report
Paper • 2601.09668 • Published • 196 -
Advancing Open-source World Models
Paper • 2601.20540 • Published • 135 -
Youtu-Agent: Scaling Agent Productivity with Automated Generation and Hybrid Policy Optimization
Paper • 2512.24615 • Published • 119 -
SkillRL: Evolving Agents via Recursive Skill-Augmented Reinforcement Learning
Paper • 2602.08234 • Published • 76
-
VisMem: Latent Vision Memory Unlocks Potential of Vision-Language Models
Paper • 2511.11007 • Published • 15 -
O-Mem: Omni Memory System for Personalized, Long Horizon, Self-Evolving Agents
Paper • 2511.13593 • Published • 28 -
General Agentic Memory Via Deep Research
Paper • 2511.18423 • Published • 172 -
MemEvolve: Meta-Evolution of Agent Memory Systems
Paper • 2512.18746 • Published • 31