TOUCAN: Synthesizing 1.5M Tool-Agentic Data From Real-World MCP Environments - Takara TLDR

Large Language Model (LLM) agents are rapidly emerging as powerful systems
for automating tasks across domains. Yet progress in the open-source community
is constrained by the lack of high quality permissively licensed tool-agentic
training data. Existing datasets are often limited in diversity, realism, and
complexity, particularly regarding multi-tool and multi-turn interactions. To
address this gap, we introduce Toucan, the largest publicly available
tool-agentic dataset to date, containing 1.5 million trajectories synthesized
from nearly 500 real-world Model Context Protocols (MCPs). Unlike prior work,
Toucan leverages authentic MCP environments to generate diverse, realistic, and
challenging tasks with trajectories involving real tool execution. Our pipeline
first produces a broad spectrum of tool-use queries using five distinct models,
applies model-based quality filtering, and then generates agentic trajectories
with three teacher models using two agentic frameworks. Rigorous rule-based and
model-based validation ensures high-quality outputs. We also introduce three
extension mechanisms to further diversify tasks and simulate multi-turn
conversations. Models fine-tuned on Toucan outperform larger closed-source
counterparts on the BFCL V3 benchmark and push the Pareto frontier forward on
MCP-Universe Bench.

Source link

What's Hot

Rethinking Thinking Tokens: LLMs as Improvement Operators – Takara TLDR

OpenAI and Jony Ive may be struggling to figure out their AI device

Generalized Parallel Scaling with Interdependent Generations – Takara TLDR

TOUCAN: Synthesizing 1.5M Tool-Agentic Data from Real-World MCP Environments – Takara TLDR

Rethinking Thinking Tokens: LLMs as Improvement Operators – Takara TLDR

Generalized Parallel Scaling with Interdependent Generations – Takara TLDR

Agentic Jigsaw Interaction Learning for Enhancing Visual Perception and Reasoning in Vision-Language Models – Takara TLDR

Former ARTnews Publisher Dies at 97

National Gallery of Art Closes as a Result of Government Shutdown

Almine Rech Closes London Gallery After More Than a Decade

Record Exec and Art Collector Gets Over 4 Years

Rethinking Thinking Tokens: LLMs as Improvement Operators – Takara TLDR

OpenAI and Jony Ive may be struggling to figure out their AI device

Generalized Parallel Scaling with Interdependent Generations – Takara TLDR

What's Hot

TOUCAN: Synthesizing 1.5M Tool-Agentic Data from Real-World MCP Environments – Takara TLDR

Related Posts

Subscribe to Updates