Paper Page - Sel3DCraft: Interactive Visual Prompts For User-Friendly Text-to-3D Generation

Sel3DCraft enhances text-to-3D generation through a dual-branch retrieval and generation system, multi-view hybrid scoring with MLLMs, and prompt-driven visual analytics, improving designer creativity.

Text-to-3D (T23D) generation has transformed digital content creation, yet
remains bottlenecked by blind trial-and-error prompting processes that yield
unpredictable results. While visual prompt engineering has advanced in
text-to-image domains, its application to 3D generation presents unique
challenges requiring multi-view consistency evaluation and spatial
understanding. We present Sel3DCraft, a visual prompt engineering system for
T23D that transforms unstructured exploration into a guided visual process. Our
approach introduces three key innovations: a dual-branch structure combining
retrieval and generation for diverse candidate exploration; a multi-view hybrid
scoring approach that leverages MLLMs with innovative high-level metrics to
assess 3D models with human-expert consistency; and a prompt-driven visual
analytics suite that enables intuitive defect identification and refinement.
Extensive testing and user studies demonstrate that Sel3DCraft surpasses other
T23D systems in supporting creativity for designers.

Source link

What's Hot

IBM Adds Agentic AI to Network Intelligence

OpenAI unveils AgentKit that lets developers drag and drop to build AI agents

OpenAI ramps up developer push with more powerful models in its API

Paper page – Sel3DCraft: Interactive Visual Prompts for User-Friendly Text-to-3D Generation

Efficient Multi-modal Large Language Models via Progressive Consistency Distillation – Takara TLDR

Compose Your Policies! Improving Diffusion-based or Flow-based Robot Policies via Test-time Distribution-level Composition – Takara TLDR

REPAIR: Robust Editing via Progressive Adaptive Intervention and Reintegration – Takara TLDR

Tomb of Amenhotep III Reopens After Two-Decade Renovation

Limited Edition Print of Ozzy Osbourne Art Sold To Benefit Charities

Morning Links for October 6, 2025

Sotheby’s to Sell René Magritte Held in Same Collection for 100 years

IBM Adds Agentic AI to Network Intelligence

OpenAI unveils AgentKit that lets developers drag and drop to build AI agents

OpenAI ramps up developer push with more powerful models in its API

What's Hot

Paper page – Sel3DCraft: Interactive Visual Prompts for User-Friendly Text-to-3D Generation

Related Posts

Subscribe to Updates