Bigram Browser

What words appear next to a given word in speech to and from children? Type a word, pick a collection, and see the words that most often occur immediately before and after it, split by corpus and speaker role. Counts are case-insensitive and come from adjacent word pairs within utterances.

Loading data — the first load can take a few seconds…

Bars show, for the selected corpora and speaker roles, the words most often adjacent to the queried word within an utterance, in descending order of total count.

Pairs are adjacent words within a single utterance in the childes-db 2026.1 token table, matched on the NFC-normalized, lowercased gloss; pairs touching an unintelligible or untranscribed gloss (xxx, yyy, www) are excluded. Counts are precomputed at the corpus level, so there is no per-child filtering here — use childesr for child-level analyses of utterance context. Word pairs occurring only once in a (collection, corpus, speaker role) cell are pruned from the precomputed tables, so the minimum observable count is 2. The first query initializes an in-browser database (~10 s); later queries fetch only the ~1 MB shard containing the queried word.