LATENT REFERENCES / TAG2
Corpus
Original title: コーパス
This reference note belongs to Tag2 in Latent References, an archive curated by Keigo Yoshida. Its archive region is Media art timeline. The note preserves its source text and links so that readers can trace the material behind the 3D map.
- Collection
- Tag2
- Archive region
- Media art timeline
Archived reference note
English translation of the archived note. JP shows the original text. Source links and literal code are retained; the translation does not update or independently verify the source claims.
A database collecting natural-language texts and usage on a large scale, organized for computer searching. In Japanese it is also called a “complete collection of language.” AI requires learning from enormous amounts of data to handle natural language.
自然言語の文章や使い方を大規模に収集し、コンピュータで検索できるよう整理されたデータベースのことです。 日本語では「言語全集」などとも呼ばれます。 AIが自然言語を扱うためには、膨大な量のデータ学習が必要です。
Source updated 2023-12-31 · Snapshot 2026-10-08
Source links and calculated neighbors
Cosine values measure shared lexical features, not truth, agreement or identical meaning. Original reference links are labeled separately.
- ISMIR MIDI DatabaseComputed lexical cosine similarity 0.121 · shared title, text, tags and references
- spresenseComputed lexical cosine similarity 0.117 · shared title, text, tags and references
- World Atlas of Language StructuresComputed lexical cosine similarity 0.106 · shared title, text, tags and references