1 · Corpus text settings ▾
No corpus loaded.
2 · Collocates of a node
Statistics (per collocate c, node n): O = tokens of c inside the span windows; W = total window tokens; N = corpus tokens; E = f(c)·W/N.
MI = log₂(O/E) · t = (O−E)/√O · LL = Dunning log-likelihood on the 2×2 window table, signed (− = repelled).
Click column headers to sort; click a collocate for its concordance.
5 · N-grams (whole corpus)
N-gram MI generalises pointwise MI: log₂( f(gram)·N⁽ⁿ⁻¹⁾ / Π f(wᵢ) ), as in the Windows version. N-grams do not cross sentence boundaries.