mirror of
https://github.com/blader/humanizer.git
synced 2026-09-27 18:19:51 +00:00
i18n: Simplified Chinese (zh-CN) adaptation as a separate repo #203
Labels
No labels
bug
documentation
duplicate
enhancement
good first issue
help wanted
invalid
question
wontfix
No milestone
No assignees
1 participant
Notifications
Due date
No due date set.
Dependencies
No dependencies set.
Reference
skills/blader-humanizer#203
Loading…
Add table
Add a link
Reference in a new issue
No description provided.
Delete branch "%!s()"
Deleting a branch is permanent. Although the deleted branch may continue to exist for a short time before it actually gets removed, it CANNOT be undone in most cases. Continue?
Hi @blader — following #163 (French), #138 (Spanish) and #194 (Traditional Chinese), I built a Simplified Chinese (zh-CN) localization of humanizer. Flagging it here for discoverability, not to request a merge.
Repo: https://github.com/jiji262/humanizer-chinese
You noted in #163 that you'd rather localized variants stay as separate community repositories so each language can evolve without adding duplicate runtime authorities to the core project. This one fits that shape: standalone repo, MIT, with blader/humanizer attribution kept in the LICENSE, README, and SKILL.md.
Why it is an adaptation, not a translation
Three of your rules invert or soften in Simplified Chinese, so a literal port would actively damage correct prose:
“”are the standard in Simplified Chinese (GB/T 15834-2011), not a tell. The rule is replaced with detection of full-width/half-width punctuation mixing, which is the real machine fingerprint.——is legitimate punctuation for explanation and topic shifts. Only overuse, paired parenthetical dashes, and mis-typesetting (single-width—, or-/--standing in for a dash) are flagged.被-constructions and translationese.Net: 24 patterns map one-to-one, 8 collapse into 4, 1 is dropped, and 8 Chinese-specific patterns are added, for 36 total.
Chinese-specific patterns, with sources
首先…其次…最后,值得注意的是,综上所述) — AI conjunction density measured at roughly 3× human in a 6,586-text parallel corpus study (CCL 2023, Beijing Language and Culture University).对…进行分析,作出调整) — the Europeanized-Chinese problem Yu Kwang-chung criticized in 1987, which LLM training data amplifies.Two additions that may be of general interest
Both came out of baseline testing (running "de-AI-ify this" on Chinese AI text without a skill, then reading what went wrong), so they may generalize beyond Chinese:
Distinct from #194 (zh-TW): different script, different punctuation standard (
“”vs「」), and different vocabulary; the two are siblings rather than duplicates.Happy to close this if you'd rather keep the tracker clear — it's here mainly so Simplified Chinese users can find a version that doesn't break their punctuation.