Назад ΠΊ списку
πŸ“„Papers

MV-Bench: Benchmarking MLLMs for Coordinated Multi-View Interface Construction

Zhao Y. et al.
2026-07

О ΠΌΠ°Ρ‚Π΅Ρ€ΠΈΠ°Π»Π΅

July 2026 benchmark for testing whether multimodal language models can generate executable multi-view visualization interfaces from visual designs. MV-Bench uses Tableau workbooks as structured ground truth and evaluates visual fidelity, data bindings, coordinated interactions, and code executability.

ΠšΡ€Π°Ρ‚ΠΊΠΎΠ΅ содСрТаниС

Why MV-Bench Adds New Signal:

1. Beyond single-screen appearance: Tests coordinated dashboards with multiple linked views
2. Executable benchmark: Provides 1,048 verified instances with code, rendered interfaces, datasets, and interaction annotations
3. Structured ground truth: Derives bindings and interactions from 92 Tableau workbooks
4. Semantic evaluation: Measures data correctness and interaction completeness alongside visual fidelity
5. Important capability gap: The strongest tested model reaches 75.45% layout accuracy but only 21.71% data-binding and 11.68% interaction accuracy

Π’Π΅Π³ΠΈ

ui-generationmultimodal-modelsdata-visualizationinteractionbenchmark
Π§ΠΈΡ‚Π°Ρ‚ΡŒ ΠΎΡ€ΠΈΠ³ΠΈΠ½Π°Π»

ΠŸΠΎΡ…ΠΎΠΆΠΈΠ΅ ΠΌΠ°Ρ‚Π΅Ρ€ΠΈΠ°Π»Ρ‹

ΠŸΠΎΠ΄Π±ΠΎΡ€ΠΊΠ° рСсурсов ΠΏΠΎ Generative UI, GenUI ΠΈ Π³Π΅Π½Π΅Ρ€Π°Ρ‚ΠΈΠ²Π½Ρ‹ΠΌ интСрфСйсам.

Π‘Π΄Π΅Π»Π°Π½ΠΎ с ❀️ сообщСством