العودة إلى القائمة
📄الأبحاث

MV-Bench: Benchmarking MLLMs for Coordinated Multi-View Interface Construction

Zhao Y. et al.
2026-07

حول

July 2026 benchmark for testing whether multimodal language models can generate executable multi-view visualization interfaces from visual designs. MV-Bench uses Tableau workbooks as structured ground truth and evaluates visual fidelity, data bindings, coordinated interactions, and code executability.

الملخص

Why MV-Bench Adds New Signal:

1. Beyond single-screen appearance: Tests coordinated dashboards with multiple linked views
2. Executable benchmark: Provides 1,048 verified instances with code, rendered interfaces, datasets, and interaction annotations
3. Structured ground truth: Derives bindings and interactions from 92 Tableau workbooks
4. Semantic evaluation: Measures data correctness and interaction completeness alongside visual fidelity
5. Important capability gap: The strongest tested model reaches 75.45% layout accuracy but only 21.71% data-binding and 11.68% interaction accuracy

الوسوم

ui-generationmultimodal-modelsdata-visualizationinteractionbenchmark

مقالات ذات صلة

📄الأبحاث

Toward Frontier-Quality Declarative UI Generation at Small-Model Cost

September 2026 Harness4GenUI paper studying whether small language models can generate production-oriented A2UI from approved component catalogs. Across task-management and cloud-console domains, the authors compare training-data strategies, model sizes, catalog sizes, cost, latency, binding correctness, and rendered quality.

a2uismall-language-modelssupervised-fine-tuning
بواسطة Yang Y. et al. · 2026-09
اقرأ المزيد
📄الأبحاث

EvoGenUI-Bench: Evaluating Multi-Turn Generative UI Assistants

August 2026 benchmark testing whether LLM assistants can maintain one executable web interface as user requirements evolve. EvoGenUI-Bench contains 150 five-turn tasks across information presentation, stateful interaction, and tool-grounded external state, with browser-based evaluation over screenshots, DOM and source evidence, actor traces, and runtime logs.

generative-uimulti-turnbenchmark
بواسطة Peng Y. et al. · 2026-08
اقرأ المزيد
📄الأبحاث

Maru: Information Architecture for Aligned, Persistent Generative UI

August 2026 HCI paper introducing information architecture as shared state for Generative UI. Maru captures how users partition, order, name, and prioritize information, turns those choices into persistent rules, and applies them across successive interface generations instead of rebuilding structure from each prompt.

generative-uiinformation-architecturepersonalization
بواسطة Kim E. et al. · 2026-08
اقرأ المزيد

مجموعة منتقاة من مصادر Generative UI وGenUI وواجهات AI الديناميكية.

صُنع باستخدام ❤️ من قِبل المجتمع