AI
DrawingVQA: 建設図面におけるマルチデプス画像テキスト推論のための実世界ベンチマーク
DrawingVQA: A Real-World Benchmark for Multi-Depth Visual-Textual Reasoning on Construction Drawings
arxiv2026年7月20日
日本語要約
本論文は、建設図面におけるマルチデプス画像テキスト推論のための新しいベンチマークDrawingVQAを導入する。複雑な実世界の技術文書に現在のAIモデルを適用する際の課題を浮き彫りにし、この分野でのAI能力向上のための経路を示唆している。
English Summary
This paper introduces DrawingVQA, a new benchmark for evaluating visual-textual reasoning on construction drawings, particularly in multi-depth scenarios. It highlights the challenges in applying current AI models to complex, real-world technical documents and suggests pathways for improved AI capabilities in this domain.