ORAL: Prompting Your Large-Scale LoRAs Via Conditional Recurrent Diffusion

arXiv:2503.24354v1 Announce Type: cross
Abstract: Parameter generation has emerged as a novel paradigm for neural network development, offering an alternative to traditional neural network training by synthesizing high-quality model weights directly. In the context of Low-Rank Adaptation (LoRA) for evolving ($\textit{i.e.}$, constantly updated) large language models (LLMs), this approach promises efficient adaptation without costly retraining. However, existing methods face critical limitations in simultaneously achieving scalability and controllability. In this paper, we introduce $\texttt{ORAL}$, a novel $\textbf{conditional recurrent diffusion}$ framework that addresses these challenges. $\texttt{ORAL}$ incorporates a novel conditioning mechanism that integrates model architecture and textual task specifications, enabling the generation of task-specific LoRA parameters that can seamlessly transfer across evolving foundation models. Our approach successfully scales to billions-of-parameter LLMs and maintains controllability. Through extensive experiments across seven language tasks, four vision tasks, and three multimodal tasks using five pre-trained LLMs, we demonstrate that $\texttt{ORAL}$ generates high-quality LoRA parameters that achieve comparable or superior performance to vanilla trained counterparts.

Source link

What's Hot

Lost Money on C3.ai, Inc. (AI)? Join Class Action Suit Seeking Recovery – Contact Levi & Korsinsky

Former Google DeepMind Core Developer Joins xAI to Assist in Grok Development_Tran_the_his

Alibaba Unveils AI Model for Character Animation and Replacement

ORAL: Prompting Your Large-Scale LoRAs via Conditional Recurrent Diffusion

LTLCrit: A Temporal Logic-based LLM Critic for Safe and Efficient Embodied Agents

From Imitation to Innovation: The Emergence of AI Unique Artistic Styles and the Challenge of Copyright Protection

VerifyLLM: LLM-Based Pre-Execution Task Plan Verification for Robots

Hidden Portrait May Be Vermeer’s Earliest Known Work

Who Are the Art World Figures on the Time 100 List?

Acquavella Signs Harumi Klossowska de Rola, Daughter of Balthus

Heirs of Jewish Collector Urge Court to Reconsider Claim to Sunflowers

Lost Money on C3.ai, Inc. (AI)? Join Class Action Suit Seeking Recovery – Contact Levi & Korsinsky

Former Google DeepMind Core Developer Joins xAI to Assist in Grok Development_Tran_the_his

Alibaba Unveils AI Model for Character Animation and Replacement

What's Hot

ORAL: Prompting Your Large-Scale LoRAs via Conditional Recurrent Diffusion

Related Posts

Subscribe to Updates