- RTX 50 (Blackwell) needs CUDA 12.8+ but project uses cu11; use CPU mode
- GPU debugging checklist: paddlepaddle-gpu vs CPU, torch GPU build, driver
- CPU slow? check for old release (v5 server 1189ms vs v6 tiny 110ms)
The file table still listed start-backend-pipeline-gui.cmd but we removed it
from Windows packages to avoid confusion with one-click-start.cmd. Repository
keeps the file for macOS/manual use, but it shouldn't be in the user-facing
file table. Linux entry from PR #34 already present.
The Linux one-click-start.sh and setup_backend_linux.sh scripts landed in
the previous setup commits but were not surfaced in user-facing docs or
in the release zip layout, so Linux users had no on-ramp. This commit
adds the missing user-facing documentation and ships the Linux entry
point in both release zip builders.
- README.md: extend 快速开始 / 启动后端 to cover Linux alongside Windows
and macOS, update the zip table to mark online-installer as the
macOS / Linux recommendation, add Linux quick-start snippet, and
refresh 常用文件 / 常用启动方式 to list one-click-start.sh and the
new docs/linux-setup.md. Wording tightened so Windows users are not
told to follow Linux commands and vice versa.
- docs/linux-setup.md: new Chinese-language install guide covering scope,
prerequisites (Python 3.12 / uv, NVIDIA GPU optional), one-click and
manual setup paths, virtualenv layout, known limitations, port
troubleshooting, post-install verification, and a Windows/macOS/Linux
comparison table.
- one-click-start.sh: refresh the top-of-file comment so it describes
the actual entry point (start_backend.py --headless ->
captcha_server_headless) instead of the old "pipeline backend" copy
from when the Linux path was a stub.
- scripts/release/build_portable.ps1: include one-click-start.sh in the
portable zip and add a Linux section to the embedded portable README
pointing users at online-installer for Linux with a fallback chmod +
run snippet.
- scripts/release/build_release_zips.ps1: include one-click-start.sh in
the common zip items and rewrite ONLINE_INSTALLER_README.txt to give
per-platform launch instructions, document auto PyPI mirror detection,
the .venv_paddle / .venv_paddle_gpu layout, and link to
docs/linux-setup.md.
No code logic change; only docs + packaging so Linux is a first-class
release target alongside Windows and macOS.
Zhipu upgraded risk control this morning: auto-click subscribe sometimes
blocked/rejected (manual click too), shows block_message. Advisory:
- wait ~10s and retry when blocked
- if auto-click keeps failing, manually click the '特惠订阅' entry
- OCR/captcha solving still fully automatic regardless of how the entry was
triggered (auto-click and OCR are decoupled)
One-liner: 入口点不动就手动点特惠订阅,验证码自动打。
The quick-start step 4 and the rush steps both told users to double-click
start-backend-pipeline-gui.cmd, but that runs the optional backend/ FastAPI
pipeline, not what one-click-start runs (captcha_server.py). Windows users
should double-click one-click-start.cmd for both first-time install and daily
launch. pipeline-gui.cmd now only mentioned as optional alternative.
The '验证码识别说明' section described backend/ FastAPI pipeline as the main
backend and recommended start-backend-pipeline-gui.cmd, but one-click-start
actually runs scripts/tools/captcha_server.py (different Tk GUI, different
model display). Rewrote to:
- Describe captcha_server.py as the main backend one-click-start runs
- Document the Tk GUI fields (status/model/prompt/result/log with timing)
- Clarify backend/ is an optional alternative pipeline, not the default
- Update file table: one-click-start.cmd is the main entry, mention
captcha_server.py explicitly
- Mention v6 tiny+medium and the config/env override for ocr_model
- Small hidden set: add row 10 PP-OCRv6 tiny + constrained (~30ms/img)
- Stress test 379: add PP-OCRv6 tiny row (100%, ~110ms/img, ~11x faster than
v5 server), keep v5 server as comparison, note test conditions
- Strict click radius table: add v6 tiny column (100% at all radii)
- Update conclusion: v6 tiny now best on accuracy/speed/stability
- Update model journey line to mention v6
Update the pipeline backend section and 常用文件 table to surface
the new double-click launcher, alongside the existing
python backend/server.py invocation.
Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>
- change openMultipleWindows prompt default from 3 to 2 (max stays 10)
- add prominent RPM warning box in README rush section after GLM upgraded RPM limits
- strengthen language in step 5 and 重要提醒 from "may trigger" to "has caused widespread failure"
- sync root and scripts/userscripts/ copies
Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>