Paper page - Instruction-Guided Autoregressive Neural Network Parameter Generation

Learning to generate neural network parameters conditioned on task
descriptions and architecture specifications is pivotal for advancing model
adaptability and transfer learning. Existing methods especially those based on
diffusion models suffer from limited scalability to large architectures,
rigidity in handling varying network depths, and disjointed parameter
generation that undermines inter-layer coherence. In this work, we propose IGPG
(Instruction Guided Parameter Generation), an autoregressive framework that
unifies parameter synthesis across diverse tasks and architectures. IGPG
leverages a VQ-VAE and an autoregressive model to generate neural network
parameters, conditioned on task instructions, dataset, and architecture
details. By autoregressively generating neural network weights’ tokens, IGPG
ensures inter-layer coherence and enables efficient adaptation across models
and datasets. Operating at the token level, IGPG effectively captures complex
parameter distributions aggregated from a broad spectrum of pretrained models.
Extensive experiments on multiple vision datasets demonstrate that IGPG
consolidates diverse pretrained models into a single, flexible generative
framework. The synthesized parameters achieve competitive or superior
performance relative to state-of-the-art methods, especially in terms of
scalability and efficiency when applied to large architectures. These results
underscore ICPG potential as a powerful tool for pretrained weight retrieval,
model selection, and rapid task-specific fine-tuning.

Source link

What's Hot

Smuggled Nvidia AI Chips Worth $1 Billion Flood Chinese Black Market Despite U.S. Export Controls

Claude Code AI Automations for Community Management in 2025

Earnings Shock: Why IBM, Chipotle, and American Airlines Tumbled—and What Comes Next

Paper page – Instruction-Guided Autoregressive Neural Network Parameter Generation

Paper page – LAPO: Internalizing Reasoning Efficiency via Length-Adaptive Policy Optimization

Paper page – TTS-VAR: A Test-Time Scaling Framework for Visual Auto-Regressive Generation

Paper page – Captain Cinema: Towards Short Movie Generation

Auction House Will Sell Egyptian Artifact Despite Concern From Experts

Anish Kapoor Lists New York Apartment for $17.75 M.

Artist Loses Final Appeal in Case of Apologising for ‘Fishrot Scandal’

US Appeals Court Overturns $8.8 M. Trademark Judgement For Yuga Labs

Smuggled Nvidia AI Chips Worth $1 Billion Flood Chinese Black Market Despite U.S. Export Controls

Claude Code AI Automations for Community Management in 2025

Earnings Shock: Why IBM, Chipotle, and American Airlines Tumbled—and What Comes Next

What's Hot

Paper page – Instruction-Guided Autoregressive Neural Network Parameter Generation

Related Posts

Subscribe to Updates