Corpus linguistics toolkit

SemanText

Stack
Astro ยท Pages
Cost
US$ 0
Locale
ID
01

Gather articles

Search by keyword, paste URLs, or merge an existing corpus

Any keywords work. Article URLs are collected from news search, then scraped at the edge.

One URL per line. First 45 are fetched per round.

Corpora produced by SemanText: Datetime, Title, Text, URL, TextID, Publication.

02

Corpus

Everything you gathered, in one place

02No articles yet. Start with step 01: gather by keyword or URL, or upload a saved corpus.

03

Analyze

All computation happens in your browser

A

Most frequent words

B

N-gram

C

Collocations

D

Key Words in Context