Paper Page - Discovering Hierarchical Latent Capabilities Of Language Models Via Causal Representation Learning

A causal representation learning framework identifies a concise causal structure to explain performance variations in language models across benchmarks by controlling for base model variations.

Faithful evaluation of language model capabilities is crucial for deriving
actionable insights that can inform model development. However, rigorous causal
evaluations in this domain face significant methodological challenges,
including complex confounding effects and prohibitive computational costs
associated with extensive retraining. To tackle these challenges, we propose a
causal representation learning framework wherein observed benchmark performance
is modeled as a linear transformation of a few latent capability factors.
Crucially, these latent factors are identified as causally interrelated after
appropriately controlling for the base model as a common confounder. Applying
this approach to a comprehensive dataset encompassing over 1500 models
evaluated across six benchmarks from the Open LLM Leaderboard, we identify a
concise three-node linear causal structure that reliably explains the observed
performance variations. Further interpretation of this causal structure
provides substantial scientific insights beyond simple numerical rankings:
specifically, we reveal a clear causal direction starting from general
problem-solving capabilities, advancing through instruction-following
proficiency, and culminating in mathematical reasoning ability. Our results
underscore the essential role of carefully controlling base model variations
during evaluation, a step critical to accurately uncovering the underlying
causal relationships among latent model capabilities.

Source link

What's Hot

This distributed data storage startup wants to take on Big Cloud

AgentKit Demo

The Gross Law Firm Reminds C3.ai, Inc. Investors of the Pending Class Action Lawsuit with a Lead Plaintiff Deadline of October 21, 2025

Paper page – Discovering Hierarchical Latent Capabilities of Language Models via Causal Representation Learning

OBS-Diff: Accurate Pruning For Diffusion Models in One-Shot – Takara TLDR

TTRV: Test-Time Reinforcement Learning for Vision Language Models – Takara TLDR

SHANKS: Simultaneous Hearing and Thinking for Spoken Language Models – Takara TLDR

$45 M. Basquait Painting to Headline Sotheby’s Fall Sales in New York

Guggenheim’s 2026 Shows Include Carol Bove Survey, Taryn Simon Project

Frieze London 2025 Opens in a Cautious Market

Industry Moves for October 8, 2025

This distributed data storage startup wants to take on Big Cloud

AgentKit Demo

The Gross Law Firm Reminds C3.ai, Inc. Investors of the Pending Class Action Lawsuit with a Lead Plaintiff Deadline of October 21, 2025

What's Hot

Paper page – Discovering Hierarchical Latent Capabilities of Language Models via Causal Representation Learning

Related Posts

Subscribe to Updates