AI

DrawingVQA: 建設図面におけるマルチデプス画像テキスト推論のための実世界ベンチマーク

DrawingVQA: A Real-World Benchmark for Multi-Depth Visual-Textual Reasoning on Construction Drawings

arxiv2026年7月20日

日本語要約

本論文は、建設図面におけるマルチデプス画像テキスト推論のための新しいベンチマークDrawingVQAを導入する。複雑な実世界の技術文書に現在のAIモデルを適用する際の課題を浮き彫りにし、この分野でのAI能力向上のための経路を示唆している。

English Summary

This paper introduces DrawingVQA, a new benchmark for evaluating visual-textual reasoning on construction drawings, particularly in multi-depth scenarios. It highlights the challenges in applying current AI models to complex, real-world technical documents and suggests pathways for improved AI capabilities in this domain.


元記事を読む𝕏 でシェア