mrkeyoor.com_
Fri 11 Sept 06:44 UTC
Dataevaluationupdated 11 Sept 2026

shulihuazixuecongshu review

Shulihuazixuecongshu is a Chinese-language preservation and ebook-reconstruction project for a 17-volume mathematics, physics, and chemistry self-study series. Its documentation is Chinese, with no English guide present. It keeps scans, normalized Markdown, images, and scripts that build reproducible EPUB 3 files.

Verdict

Our run installed 35 packages in 14 seconds and completed the build in 7 seconds, but the repository supplied no test target. Use it for Chinese-language study or auditable ebook preservation when you can verify the source corrections and the rights for your jurisdiction. Skip it if you need English material, a small checkout, CI-enforced tests, or a license that clearly permits redistribution.

We ran it

Lab card: what happened when we ran shulihuazixuecongshuScreenshot of shulihuazixuecongshu (github.com/tradecatlabs/shulihuazixuecongshu)
Install✓ · 14s35 packages · 37 MB
Build✓ · 7s
Testsn/ano test script
Known vulns0(pip-audit)
Repo4935 files~2,066 lines of source · 154.8 MB · 0 CI workflows

Answers from our run

Does shulihuazixuecongshu build from source?

Dependencies installed in 14 seconds (35 packages), and the build succeeded in 7 seconds. We cloned commit 76e79a8 into a clean Debian container with 3 CPUs and no project-specific setup.

Does shulihuazixuecongshu have tests you can run?

Not through a standard command: the project exposes no test script or target that our harness could run.

Does shulihuazixuecongshu have known vulnerabilities in its dependencies?

pip-audit found none in the dependency tree at the time of our run.

Who should not use shulihuazixuecongshu?

Readers who need English books or English setup documentation: the catalog language is zh-CN, and the repository has no English README.

What are the alternatives to shulihuazixuecongshu?

Pandoc, Sigil, Standard Ebooks tools. Our run installed 35 packages in 14 seconds and completed the build in 7 seconds, but the repository supplied no test target.

Setup4/514-second install and 7-second build, with external PDF tools
Docs4/5Detailed Chinese provenance and maintenance docs; no English guide
Community3/5872 stars and two resolved reports, but very little history
Maturity3/5Version 2.4.0 builds cleanly; CI and a test target are absent

Who it’s for

Chinese readers and archivists who want 17 editable volumes alongside the original scans.
Ebook maintainers who value fixed source hashes, image checks, MathML, and reproducible EPUB output.
Python and Pandoc users prepared to review corrections against scanned pages.
Researchers who can determine whether their local copyright rules permit the intended use.

Who it’s NOT for

Readers who need English books or English setup documentation: the catalog language is zh-CN, and the repository has no English README.
Anyone assuming a public GitHub repository grants redistribution rights: the rights notice grants no new license for the original text, layout, or scans.
Users wanting all 17 finished books as current release downloads: releases from v2.0.0 onward contain only 3 subject-combined EPUBs.
Teams requiring CI-backed regression tests: our scan found 0 CI workflows, no tests directory, and no test script or target.

Setup reality

Our Python 3.12 sandbox installed 35 packages in 14 seconds, using 37 MB on disk, and the build succeeded in 7 seconds. There was no test script or target, so tests were skipped. Pip-audit reported 0 known vulnerabilities.

The standard build needs Python 3.10 or newer and Pandoc 3.1 or newer; GNU Make is optional. Deep PDF inspection also calls for PyPDF2 3.x, qpdf, Poppler tools, and ExifTool. No credentials are needed for a local EPUB build.

The checkout contained 4,935 files, about 2,066 source lines, and 154.8 MB of material. It had no Dockerfile, no CI workflow files, and no tests directory. The current release policy publishes 3 combined EPUBs, while per-volume output is built locally.

Seventeen Chinese books are preserved as editable sources

The repository contains 17 Chinese-language books covering algebra, geometry, trigonometry, physics, and chemistry. Its main README and maintenance documents are in Chinese, and the catalog marks the ebook language as zh-CN; no English README is present. Each volume has normalized Markdown, while the repository also keeps the original scans and image assets. This is useful for reading, correction, and ebook production, rather than as a translated course for English-speaking students.

The README inventory claims 5,927 scanned pages, 4,875 image assets, 4,878 image references, and 88,653 MathML elements in verified EPUB output. It reports 0 missing or orphaned assets and 0 empty image descriptions. Those figures describe the project's own content audit, not our sandbox. They show the sort of structural checks the scripts perform and the scale of the material a maintainer must inspect when a correction crosses formats.

What happened when we ran it

Our measurement setup used commit 76e79a8 in an unprivileged Python 3.12 container with 3 CPUs, 8 GB of RAM, and no secrets. Installation succeeded in 14 seconds, adding 35 packages and consuming 37 MB on disk. The build then succeeded in 7 seconds. For a collection that also ships large scans and thousands of images, the tooling itself was quick to prepare on our box.

The lab found no test script or test target, so it skipped tests rather than inventing a command. Pip-audit reported 0 known vulnerabilities in the installed packages. That audit is reassuring within its narrow scope, but it is not a correctness test for formulas, OCR, page order, image placement, or EPUB navigation. Our scan also found 0 CI workflow files and no tests directory, which leaves the documented Make targets as the visible quality gate.

The 154.8 MB checkout is mostly books and images

We measured 4,935 files and about 2,066 lines of source in a 154.8 MB checkout. The unusual ratio makes sense: 17 raw PDFs and 4,875 images occupy far more space than the Python scripts. Developers looking only for a reusable EPUB builder are taking on the full archive. Readers who want the books may appreciate that bundling because the scans, Markdown, figures, hashes, and catalog stay together for comparison.

The standard path requires Python 3.10 or newer and Pandoc 3.1 or newer, with GNU Make listed as optional. A deeper PDF audit adds PyPDF2 3.x, qpdf, Poppler commands, and ExifTool. The repository has no Dockerfile, so those system tools are your responsibility. No service account or API key is needed for the local build. Generated EPUBs, intermediate files, and JSON reports are excluded from version control and rebuilt through the Make targets.

Four blank scan pages were restored from another source

The known-issues note says pages 49, 50, 55, and 56 of the original solid-geometry scan remain blank. The Markdown and EPUB restore that material from another scan, using neighboring text, printed page numbers, exercise numbering, and figure sequences to map the replacement. The raw PDF is left unchanged so its bytes remain verifiable. Readers get repaired content while auditors can still see the defect in the archived source.

The provenance record gives more detail than most ebook repositories. It identifies the alternate scan, hashes both PDFs, describes OCR at 96 DPI, and says the restored pages were checked against page images. Issue 2 supplied the alternate source and closed the next day. This does not make transcription errors impossible. It does make one consequential repair traceable, including why the replacement scan's page numbers were not copied directly into the normalized text.

Release v2.4.0 ships three combined EPUBs

Release v2.4.0 was published on August 30, 2026. Its notes describe normalizing 7,433 high-confidence traditional characters, 423 Japanese or old glyph forms, and 2 context-dependent uses of 反覆, while checking 7 uncommon contexts against the PDFs. The original scans and images stayed unchanged. This is careful editorial work, but a Chinese reader should still compare suspect formulas or wording with the page image before citing the reconstructed edition.

Since v2.0.0, releases contain 3 combined EPUBs, one each for mathematics, physics, and chemistry. The 17 individual Markdown sources and local per-volume build remain in the repository. More important, the rights statement grants no new copyright license for the original writing, page design, or scans. Personal study, preservation, public hosting, and commercial redistribution can have different legal outcomes. The repository tells users to determine the status that applies in their jurisdiction.

The first five days show work, not long-term maintenance

GitHub says the repository was created on August 27, 2026, released v2.4.0 on August 30, and was last pushed on September 1. It had 872 stars, 163 forks, and 0 open issues when fetched. Both recorded issues were closed on August 30, one for Markdown compatibility and one for the missing solid-geometry pages. That is prompt early activity, but five days between creation and last push is too little history for a durability claim.

The 14-second install and 7-second build make this archive easy to try if you already read Chinese and have Pandoc available. Its better qualities are specific: fixed raw hashes, documented corrections, editable sources, and reproducible EPUB tooling. The weak points are just as specific: 154.8 MB of material, no automated test target or CI, no English guide, and no new rights grant. Keep the scans beside the rebuilt text and treat redistribution as a separate decision.

Alternatives

ProjectWhat it isPick it when
PandocA universal markup converter and the underlying document engine used by this project's EPUB build.pick this instead when you have your own source material and need a general conversion tool rather than these 17 books.
SigilA desktop EPUB editor for inspecting and changing ebook packages through a graphical interface.pick this instead when manual EPUB editing matters more than a scripted, source-controlled reconstruction.
Standard Ebooks toolsThe command-line toolset used to produce Standard Ebooks editions.pick this instead when you are preparing a new public-domain edition under Standard Ebooks conventions.

What people are saying

  1. [velocity-scout] tradecatlabs/shulihuazixuecongshu

Sources

  1. Shulihuazixuecongshu README
  2. Shulihuazixuecongshu repository facts
  3. Release v2.4.0
  4. Provenance and processing notes
  5. Known issues and rights status
  6. Maintenance workflow

More data reviews

rocketmq · free-programming-books · Ontology-Playground · awesome-osint · data-engineering-zoomcamp · zju-icicles · the whole board →