中文圣经

2,579 Words Appear Only Once in the Bible

Answering: “words that appear once in the Bible

Updated 2026-09-22 · plain text · JSON

Words that occur once are called hapax legomena, and they are the reason a vocabulary list is a bad plan for reading. Learning every word in the Bible means learning 2,579 you will meet a single time; skipping them costs almost nothing, because together they account for 0.4% of word tokens.

Most are names, places and hapax technical terms — the measurements of the tabernacle, the stones of the breastplate, the officials of a foreign court. The practical response is to read with a tap-to-reveal reader and let them go.

The opposite end of the distribution is where study time belongs: 1,900 words cover 90% of the text, and the site's decks are built from that end of the list.

A sample of words that occur once
WordWordWordWordWordWord
上分拉活送走罪大恶极纵横略作
联盟精练惊人女同喜事声闻
二人转戏言天明来历同父异母合家
遮羞女仆加进庶出两国
皮衣日增羔皮民事
聚齐神气俊秀深爱不一定裙子

2,579 words in total; 36 shown.

Questions

Should I learn them?
Not deliberately. Recognise the shape, tap for the meaning, move on — that is what the reader is for.
Are they all names?
Not all, but many: 662 words in this text look like proper names by their dictionary gloss.
Which words should I learn instead?
The 825 most frequent, which cover 80% of the text. The decks are ordered exactly that way.

Related reference pages

Take it further

How these numbers were produced

  • Text: 和合本 (Chinese Union Version), simplified script — public domain66 books, 1,189 chapters, 31,021 verses.
  • Words: forward maximum matching against CC-CEDICT (this site's own segmentation). A different segmenter gives slightly different word counts; character counts are unaffected.
  • HSK levels: HSK 3.0 levels from the complete-hsk-vocabulary list; band 7 covers HSK 7-9.
  • Reading times: 260 characters a minute — an assumption about a fluent adult reader, not a measurement.
  • Computed: 2026-09-22, from the text on this site.

These are counts anyone can reproduce from the published dataset. They are not a peer-reviewed linguistic study, and the difficulty rankings are one stated formula rather than a validated readability score.

Definitions: CC-CEDICT, licensed CC BY-SA (https://creativecommons.org/licenses/by-sa/4.0/). If you share this file onward, keep this notice and share alike. Character data: Make Me a Hanzi (Arphic Public License / LGPL).