DALL-E 2 Fails to Reliably Capture Common Syntactic Processes

Evelina Leivada; Elliot Murphy; Gary Marcus

DALL-E 2 は一般的な構文プロセスを確実にキャプチャできない

機械知能は、感覚、言語処理、および自然言語を理解してさまざまな刺激に変換する能力に関する主張にますます結び付けられています。言語学で広く議論され、人間の言語に浸透している構成性に関連する 8 つの文法的現象を捉える DALL-E 2 の能力を体系的に分析します: 結合原理と共参照、受動態、構造的曖昧性、否定、語順、二重オブジェクト構造、文コーディネーション、省略記号、比較記号。幼児は日常的にこれらの現象を習得し、構文とセマンティクスの間の体系的なマッピングを学習しますが、DALL-E 2 は構文と一致する意味を確実に推測することができません。これらの結果は、そのようなシステムが人間の言語を理解する能力に関する最近の主張に異議を唱えるものです。将来のテストのベンチマークとして、テスト材料の完全なセットを利用できるようにします。

Machine intelligence is increasingly being linked to claims about sentience, language processing, and an ability to comprehend and transform natural language into a range of stimuli. We systematically analyze the ability of DALL-E 2 to capture 8 grammatical phenomena pertaining to compositionality that are widely discussed in linguistics and pervasive in human language: binding principles and coreference, passives, structural ambiguity, negation, word order, double object constructions, sentence coordination, ellipsis, and comparatives. Whereas young children routinely master these phenomena, learning systematic mappings between syntax and semantics, DALL-E 2 is unable to reliably infer meanings that are consistent with the syntax. These results challenge recent claims concerning the capacity of such systems to understand of human language. We make available the full set of test materials as a benchmark for future testing.

updated: Sun Oct 23 2022 23:56:54 GMT+0000 (UTC)

published: Sun Oct 23 2022 23:56:54 GMT+0000 (UTC)

arXiv

参考文献 (このサイトで利用可能なもの) / References (only if available on this site)

被参照文献 (このサイトで利用可能なものを新しい順に) / Citations (only if available on this site, in order of most recent)

Amazon.co.jpアソシエイト