GenTune: Toward Traceable Prompts to Improve Controllability of Image Refinement in Environment D...
GenTune: Toward Traceable Prompts to Improve Controllability of Image Refinement in Environment D... Wen-Fan Wang, Ting-Ying Lee, Chien-Ting Lu, Che-Wei Hsu, Nil Ponsa i Campanyà, Yu Chen, Mike Y. Chen, Bing-Yu Robin Chen UIST 2025: The 38th Annual ACM Symposium on User Interface Software and Technology Session: 2. Creating with Gen AI Environment designers in the entertainment industry create imaginative 2D and 3D scenes for games, films, and television, requiring both fine-grained control of specific details and consistent global coherence. Designers have increasingly integrated generative AI into their workflows, often relying on large language models (LLMs) to expand user prompts for text-to-image generation, then iteratively refining those prompts and applying inpainting. However, our formative study with 10 designers surfaced two key challenges: (1) the lengthy LLM-generated prompts make it difficult to understand and isolate the keywords that must be revised for specific visual elements; and (2) while inpainting supports localized edits, it can struggle with global consistency and correctness. Based on these insights, we present GenTune, an approach that enhances human–AI collaboration by clarifying how AI-generated prompts map to image content. Our GenTune system lets designers select any element in a generated image, trace it back to the corresponding prompt labels, and revise those labels to guide precise yet globally consistent image refinement. In a summative study with 20 designers, GenTune significantly improved prompt-image comprehension, refinement quality and efficiency, and overall satisfaction (all p .01) compared to current practice. A follow-up field study with two studios further demonstrated its effectiveness in real-world settings. DOI:: doi.org/10.1145/3746059.3747774 Web:: https://programs.sigchi.org/uist/2025... Video presentations for UIST 2025 papers

The AI Breakthrough That Will Change Everything (Google DeepMind CEO Interview)

Sculpin: Direct-Manipulation Transformation of JSON

BaroPoser: Real-time Human Motion Tracking from IMUs and Barometers in Everyday Devices

Stanford CS25: Transformers United V6 I From Language Models to Native Multimodal Intelligence

Andrej Karpathy: From Vibe Coding to Agentic Engineering w/ Stephanie Zhan

Training Sand to Think: Artificial General Intelligence & Future of Physics

Stop Prompting Claude. Use Karpathy's Method Instead.

The FULL VIDEO of Trump they didn’t want released

China Just Built What TSMC Said Was Impossible

My AI Design Workflow That Doesn't Ship Slop

I Think They Are Lying To You

Finally. Agent Loops Clearly Explained.

Anthropic is Completely F*cked.

I Made Opus 4.8 and Fable 5 Build the Same App (RAW RESULTS)

New #1 open-source AI model is here!

4K TV Art: Vintage Summer Landscape with Gold Frame | Relaxing Screensaver

We let AI buy a robot and a car, it does exactly what experts warned.

AI Does Something Horrifying To Human Thinking

Der Vater der KI: „Wir haben noch 3 Jahre!” Roboter, Singularität & die Zukunft (Jürgen Schmidhuber)

