feat: 独立词库、三向认词与当日快闪复习

## 独立词库

理解词原先只能跟着课程单元走,学完 A0 十课词汇量只增加约 22 个实词,
不足以解决"记不住单词"。新增一份独立词库 assets/words/wordbank.json
(2748 词,A1–B1),挂进 receptiveWordRegistry 的合成单元 bank-A1/A2/B1,
完全复用理解词已有的状态机,不依赖课程进度,第一天就能用。

数据来源、许可与合成规则记在 tool/words/DATA-NOTE.md:CEFR-J 定等级、
公开词书提供音标、AI 重写全部释义并生成例句、OpenSubtitles 提供口语词频。
词书部分为 CC BY-NC-SA 4.0 且上游权利不明,仅供个人非商用;
若要分发或上架,须替换音标那一列。

## 背单词机制

- 间隔阶梯 1/3/7/15/30/60/120 天,连续答对上一级,答错回第一级。
  原先首次答对后要等 7 天才复习,正是"第二天就忘"的成因。
- 每日新词上限(10 分钟 8 个 / 20 分钟 15 个 / 30 分钟 20 个)。
  阶梯第一级是次日,今天引入的新词就是明天的工作量。
- 新词按口语频率发放,不再按字母序 —— A1 从 a.m./ability 变成 no/not/know/just。
- 三个方向按层级轮转:看词(英→中)→ 听词(音→中)→ 想词(中→英)。
  想词题仍是选择题,不要求产出,理解词定位不变,不进升级分母。
- 单词页独立成 tab,首页今日任务卡下方给一张认词入口卡。

## 用法对照

课程 JSON 增加 usage 字段(when/reply/swap/confuse):一个句型用在什么场合、
对方通常怎么答、还能怎么说、跟哪个学过的句型容易混。
知道 How are you? 的意思,不等于知道它不是用来问名字的。

## 复习流

- 当日快闪(recap)独立成队列,不占复习预算,也不计入积压。
- 只发放当日预算内的量,其余保持到期状态等下次,不悄悄丢弃或改期。
- 答错的项隔几题后回来,而不是立刻重问。

测试 296 通过,flutter analyze 干净。

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
This commit is contained in:
shenlei
2026-09-20 23:54:17 +09:00
co-authored by Claude Opus 5
parent 1a10ca88e0
commit 6b42b7abc3
52 changed files with 4044 additions and 82 deletions
+42
View File
@@ -0,0 +1,42 @@
#!/bin/bash
# Downloads the beginner-level word books. The repo is CC BY-NC-SA 4.0 and its
# own README calls the contents "compiled from public sources, personal use
# only" -- see the note in DATA-NOTE.md before shipping any of this.
set -u
BASE="https://github.com/lilinji/English/raw/main"
enc() { python3 -c "import urllib.parse,sys;print(urllib.parse.quote(sys.argv[1]))" "$1"; }
get() {
local path="$1" out="src/$2"
[ -s "$out" ] && { echo " 已有 $2"; return; }
curl -sL -o "$out" "$BASE/$(enc "$path")"
if [ -s "$out" ] && head -c2 "$out" | grep -q PK; then echo " 取得 $2 ($(wc -c <"$out") 字节)"
else echo " 失败 $2"; rm -f "$out"; fi
}
echo "剑桥 KET/PET:"
get "9.其他(更多)/14天攻克KET核心词汇.xlsx" ket-core.xlsx
get "9.其他(更多)/KET核心词 巧记速练.xlsx" ket-drill.xlsx
get "9.其他(更多)/21天攻克PET核心词汇.xlsx" pet-core.xlsx
get "9.其他(更多)/PET核心词 巧记速练.xlsx" pet-drill.xlsx
get "9.其他(更多)/突破英文基础词汇.xlsx" basic.xlsx
echo "人教版小学:"
for g in 一 二 三 四 五 六; do
for t in 上 下; do
get "1.全国各大教材版本中小学同步/人教版/人教版一年级起点${g}年级${t}.xlsx" "rj-p1-${g}${t}.xlsx"
done
done
echo "人教版初中(七年级):"
get "1.全国各大教材版本中小学同步/人教版/人教版初中英语七年级上册.xlsx" rj-m7a.xlsx
get "1.全国各大教材版本中小学同步/人教版/人教版初中英语七年级下册.xlsx" rj-m7b.xlsx
# Word frequency, for the order new words are handed out in. These are
# OpenSubtitles counts, i.e. spoken language: a web corpus ranks `ankle` and
# `asleep` far below where a learner of spoken English needs them.
# hermitdave/FrequencyWords is MIT, a cleaner licence than the word books above.
echo "词频表:"
if [ -s src/en_50k.txt ]; then
echo " 已有 en_50k.txt"
else
curl -sL -o src/en_50k.txt \
"https://raw.githubusercontent.com/hermitdave/FrequencyWords/master/content/2018/en/en_50k.txt"
echo " 取得 en_50k.txt ($(wc -l <src/en_50k.txt) 行)"
fi