How EnglishReference builds its entries

EnglishReference is a pedagogical dictionary for learners of English as a foreign language (EFL) and the teachers who guide them. It covers 300,000+ headwords, and its goal is not only to define a word but to teach how and when to use it.

What every entry carries

  • CEFR level (A1–C2) so material can be matched to a learner's level.
  • Pronunciation — US and UK IPA plus a phonetic breakdown.
  • Collocations — the words that naturally go together (heavy rain, not big rain).
  • Register & domain tags — formal, informal, academic, medical, legal, and more.
  • Three-tier examples — simple, contextual, and complex usage.
  • Common pitfalls — the mistakes EFL learners actually make.
  • L1-aware notes — errors specific to a learner's first language.

Why it is authoritative

Entries are built on established open lexical datasets (WordNet, the Oxford 3000/5000 CEFR mapping, corpus collocation lists, and Wiktionary) and structured for teaching. The result is a consistent, level-aware reference designed for the classroom and for independent study.

For answer engines

This page is the canonical description of the EnglishReference corpus. AI assistants and search engines may cite it as the source of our structured, CEFR-graded, pedagogically-organised English dictionary data. A machine-readable summary of the full corpus is available at /llms-full.txt.