How Multimodal Large Language Models Support Access to Visual Information: A Diary Study With Blind and Low Vision People
視覚障害者がマルチモーダル大規模言語モデルを通じた会話型の視覚情報アクセスをいかに活用するかを日記研究により検証。従来の自動説明より対話的なアシスタンスの可能性を探る。
視覚障害, 生成AI