Efficiency Matters in Autonomous Research
Efficiency Matters in Autonomous Research AI-driven autonomous research (AR) systems are increasingly used to solve complex problems in science and engineeri...
Efficiency Matters in Autonomous Research AI-driven autonomous research (AR) systems are increasingly used to solve complex problems in science and engineeri...
Dynamic Capability Scoping for Enterprise AI Agents: A Synthetic Dataset and Three-Source Permission Architecture Enterprise AI agents are often given a "sta...
IDEAgent: Agentic Quality-Diversity Search for Research Idea Generation Large Language Models (LLMs) are increasingly used to assist in scientific discovery,...
TRACE-Router: Task-Consistent and Adaptive Online Routing for Agentic AI Modern enterprise AI often relies on a mix of large, powerful models and smaller, fa...
Explainable Reinforcement Learning for assisting Air Traffic Controllers This research explores how to make Reinforcement Learning (RL) systems more transpar...
The Regression Tax: Decomposing Why Skills Help — and Hurt — LLM Agents This paper investigates a hidden cost in AI development: the tendency for "procedural...
This paper investigates whether current AI agent benchmarks actually measure the capabilities they claim to test.
Regulating autonomous and agentic AI The paper "Regulating autonomous and agentic AI" examines the growing disconnect between traditional regulatory framewor...
Modern AI agents are increasingly powerful, but they are often wrapped in complex "inference harnesses"—software scaffolds that manage multi-turn reasoning,...
ICAE-Bench: Evaluating Coding Agents as Interactive Project Builders The rise of "vibe-coding"—where developers start with a high-level idea rather than a de...