-
Multi-Agent Computer Use
Paper • 2606.01533 • Published • 7 -
OpenSkill: Open-World Self-Evolution for LLM Agents
Paper • 2606.06741 • Published • 29 -
Socratic-SWE: Self-Evolving Coding Agents via Trace-Derived Agent Skills
Paper • 2606.07412 • Published • 12 -
Bayesian-Agent: Posterior-Guided Skill Evolution for LLM Agent Harnesses
Paper • 2606.08348 • Published • 16
Collections
Discover the best community collections!
Collections including paper arxiv:2607.01131
-
EvoMaster: A Foundational Agent Framework for Building Evolving Autonomous Scientific Agents at Scale
Paper • 2604.17406 • Published • 6 -
Solvita: Enhancing Large Language Models for Competitive Programming via Agentic Evolution
Paper • 2605.15301 • Published • 22 -
Learning to Foresee: Unveiling the Unlocking Efficiency of On-Policy Distillation
Paper • 2605.11739 • Published • 61 -
SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution
Paper • 2605.18401 • Published • 131
-
GARDO: Reinforcing Diffusion Models without Reward Hacking
Paper • 2512.24138 • Published • 30 -
DiffThinker: Towards Generative Multimodal Reasoning with Diffusion Models
Paper • 2512.24165 • Published • 53 -
Dynamic Large Concept Models: Latent Reasoning in an Adaptive Semantic Space
Paper • 2512.24617 • Published • 67 -
Coupling Experts and Routers in Mixture-of-Experts via an Auxiliary Loss
Paper • 2512.23447 • Published • 100
-
Large Language Models Generate Harmful Content Using a Distinct, Unified Mechanism
Paper • 2604.09544 • Published • 7 -
OmniJigsaw: Enhancing Omni-Modal Reasoning via Modality-Orchestrated Reordering
Paper • 2604.08209 • Published • 27 -
Training a Student Expert via Semi-Supervised Foundation Model Distillation
Paper • 2604.03841 • Published • 11 -
How to Fine-Tune a Reasoning Model? A Teacher-Student Cooperation Framework to Synthesize Student-Consistent SFT Data
Paper • 2604.14164 • Published • 35
-
Multi-Agent Computer Use
Paper • 2606.01533 • Published • 7 -
OpenSkill: Open-World Self-Evolution for LLM Agents
Paper • 2606.06741 • Published • 29 -
Socratic-SWE: Self-Evolving Coding Agents via Trace-Derived Agent Skills
Paper • 2606.07412 • Published • 12 -
Bayesian-Agent: Posterior-Guided Skill Evolution for LLM Agent Harnesses
Paper • 2606.08348 • Published • 16
-
EvoMaster: A Foundational Agent Framework for Building Evolving Autonomous Scientific Agents at Scale
Paper • 2604.17406 • Published • 6 -
Solvita: Enhancing Large Language Models for Competitive Programming via Agentic Evolution
Paper • 2605.15301 • Published • 22 -
Learning to Foresee: Unveiling the Unlocking Efficiency of On-Policy Distillation
Paper • 2605.11739 • Published • 61 -
SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution
Paper • 2605.18401 • Published • 131
-
Large Language Models Generate Harmful Content Using a Distinct, Unified Mechanism
Paper • 2604.09544 • Published • 7 -
OmniJigsaw: Enhancing Omni-Modal Reasoning via Modality-Orchestrated Reordering
Paper • 2604.08209 • Published • 27 -
Training a Student Expert via Semi-Supervised Foundation Model Distillation
Paper • 2604.03841 • Published • 11 -
How to Fine-Tune a Reasoning Model? A Teacher-Student Cooperation Framework to Synthesize Student-Consistent SFT Data
Paper • 2604.14164 • Published • 35
-
GARDO: Reinforcing Diffusion Models without Reward Hacking
Paper • 2512.24138 • Published • 30 -
DiffThinker: Towards Generative Multimodal Reasoning with Diffusion Models
Paper • 2512.24165 • Published • 53 -
Dynamic Large Concept Models: Latent Reasoning in an Adaptive Semantic Space
Paper • 2512.24617 • Published • 67 -
Coupling Experts and Routers in Mixture-of-Experts via an Auxiliary Loss
Paper • 2512.23447 • Published • 100