List पर वापस जाएँ
📄Papers

Generative Interfaces for Language Models

Chen J. et al.
2025-08

परिचय

ACL 2026 Findings paper एक ऐसा paradigm प्रस्तावित करता है जिसमें LLMs केवल text reply देने के बजाय proactively task-specific interfaces बनाते हैं। यह work interface-specific representations को iterative refinement से जोड़ता है और रिपोर्ट करता है कि information-dense तथा exploratory tasks में human evaluators ने conversational interfaces की तुलना में generative interfaces को 72% तक अधिक पसंद किया।

सारांश

यह paper क्यों महत्वपूर्ण है:

1. Direct chat comparison: Generative interfaces को केवल synthetic baselines नहीं, conventional LLM conversations के सामने test करता है
2. Evaluation depth: Functional, interactive और emotional experience को समेटने वाला multidimensional assessment इस्तेमाल करता है
3. मजबूत preference signal: Complex tasks पर interface-first responses के लिए बड़े human-preference gains रिपोर्ट करता है
4. Paradigm clarity: GenUI को user goals के लिए तैयार proactive interface generation के रूप में रखता है
5. Practical relevance: Adaptive UI को plain assistant chat की जगह कब इस्तेमाल करना चाहिए, यह तय करने वाली teams के लिए उपयोगी evidence

Tags

acl-2026evaluationgenuihuman-preference
मूल लेख पढ़ें

मिलते-जुलते लेख

📄Papers

LEGOUI: Designing with UI-DSL Bricks for Transparent, Controllable Generation

अगस्त 2026 का HCI paper, जो शुरुआती design ideation के लिए staged Generative UI framework पेश करता है। LEGOUI prompt से मिले और model द्वारा अनुमानित फैसलों को provenance-aware UI DSL में दर्ज करता है, users को layout, interaction, relation और style stages में फैसले स्वीकार, अस्वीकार या जोड़ने देता है, और बदलती specification से interface render करता है।

genuiui-dsldesign-time
लेखक Zhou Y. et al. · 2026-08
और पढ़ें
📄Papers

Beyond a Single Judge: Simulating Social Persona Panels for Generative UI Evaluation

जुलाई 2026 का paper, जो ESPP नाम की evaluation method पेश करता है। यह एक LLM judge की जगह evidence-grounded और मनोवैज्ञानिक रूप से विविध personas का panel रखता है। Panelists generated-interface screenshots को स्वतंत्र रूप से rate करते हैं, bounded-confidence mechanism के जरिए राय साझा करते हैं, और उनके scores को social weighting के साथ जोड़ा जाता है।

genuiui-evaluationhuman-preferences
लेखक Wu Z. et al. · 2026-07
और पढ़ें
📄Papers

Design Theater: A Benchmark for Generative UI

जुलाई 2026 का paper, जो यह जाँचने के लिए benchmark पेश करता है कि Generative UI tools users को बताए गए design rationales सच में लागू करते हैं या नहीं। पाँच tools से बने 120 interfaces और 24 structural, styling व functional tasks में यह study बताए गए design decisions और generated result के बीच अंतर मापती है।

genuiui-evaluationdesign-rationale
लेखक Imteyaz K. et al. · 2026-07
और पढ़ें

Generative UI resources का curated संग्रह।

बनाया गया ❤️ community द्वारा