Paper Page - R^2ec: Towards Large Recommender Models With Reasoning

A unified large recommender model with intrinsic reasoning capabilities is proposed, facilitating interleaved reasoning and recommendation using a reinforcement learning framework called RecPO.

Large recommender models have extended LLMs as powerful recommenders via
encoding or item generation, and recent breakthroughs in LLM reasoning
synchronously motivate the exploration of reasoning in recommendation. Current
studies usually position LLMs as external reasoning modules to yield auxiliary
thought for augmenting conventional recommendation pipelines. However, such
decoupled designs are limited in significant resource cost and suboptimal joint
optimization. To address these issues, we propose \name, a unified large
recommender model with intrinsic reasoning capabilities. Initially, we
reconceptualize the model architecture to facilitate interleaved reasoning and
recommendation in the autoregressive process. Subsequently, we propose RecPO, a
corresponding reinforcement learning framework that optimizes \name\ both the
reasoning and recommendation capabilities simultaneously in a single policy
update; RecPO introduces a fused reward scheme that solely leverages
recommendation labels to simulate the reasoning capability, eliminating
dependency on specialized reasoning annotations. Experiments on three datasets
with various baselines verify the effectiveness of \name, showing relative
improvements of 68.67\% in Hit@5 and 45.21\% in NDCG@20. Code available at
https://github.com/YRYangang/RRec.

Source link

What's Hot

YouTube’s multi-language audio feature for dubbing videos rolls out to all creators

Jus Mundi Launches Agentic Tool, Explains How It Works – Artificial Lawyer

Parallel-R1: Towards Parallel Thinking via Reinforcement Learning – Takara TLDR

Paper page – R^2ec: Towards Large Recommender Models with Reasoning

Parallel-R1: Towards Parallel Thinking via Reinforcement Learning – Takara TLDR

Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search – Takara TLDR

Visual Representation Alignment for Multimodal Large Language Models – Takara TLDR

Ralph Rugoff to Leave London’s Hayward Gallery After 20 Years

New York Foundation for the Arts Workers Move to Unionize

Patrizia Sandretto Re Rebaudengo Teams Up with New Museum

Growing Support for Parthenon Marbles’ Return to Greece, More Art News

YouTube’s multi-language audio feature for dubbing videos rolls out to all creators

Jus Mundi Launches Agentic Tool, Explains How It Works – Artificial Lawyer

Parallel-R1: Towards Parallel Thinking via Reinforcement Learning – Takara TLDR

What's Hot

Paper page – R^2ec: Towards Large Recommender Models with Reasoning

Related Posts

Subscribe to Updates