rouge-chinese
Python ROUGE Score Implementation for Chinese Language Task (official rouge score)
Decision gist · record as of 2026-08-14
Yes, if you need ROUGE metrics for Chinese text and can tolerate dormant maintenance. The package solves real problems in the original ROUGE for Chinese, has low install friction, and no known vulnerabilities. However, verify that the unclear license aligns with your use case, and confirm compatibility with your Python version since support is unspecified. Consider it stable for evaluation tasks but not actively developed.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Requires external word segmentation to tokenize Chinese text before scoring; the package itself does not include tokenization.
- Low friction: pure Python wheel with a single runtime dependency (six).
- Dormant since first release with no updates, though the repository remains active with 114 stars.
License · maintenance · safety
LICENCE.txt (unclear) — License treatment is unclear—the package references LICENCE.txt but provides no SPDX identifier. Review the license file in the repository before adopting in proprietary or restricted contexts.
last release 2022-09-18 (1426 days) · last repo commit 2024-06-27 · 114 stars
0 known vulnerabilities (OSV.dev, 2026-08-14) · 81,338 downloads/mo, #14,236 on PyPI
Alternatives
Verify before relying
pip install rouge-chinese
from rouge_chinese import Rouge
rouge = Rouge()
scores = rouge.get_scores(hypothesis, reference)- Whether the package works correctly with modern Python versions (python_support is unspecified in metadata)
- Current state of the single runtime dependency (six) and its long-term maintenance status
- Whether memory optimizations claimed in the description remain effective with contemporary hardware and text sizes
What it is and what it does
Rouge-Chinese is a Python library that calculates ROUGE evaluation scores specifically for Chinese-language NLP tasks. It addresses known limitations in the original ROUGE implementation when applied to Chinese text: incorrect sentence segmentation (missing Chinese punctuation marks), excessive memory consumption during longest-common-subsequence calculation, and inaccurate scores due to approximation methods. The library improves sentence splitting to recognize Chinese punctuation, optimizes memory usage by computing sequence lengths without generating the sequences themselves, and computes official ROUGE scores rather than approximate variants.
The package provides three main interfaces: a library API for scoring single or multiple sentence pairs with optional averaging, a file-based API for batch scoring line-by-line from text files, and a command-line tool for ad-hoc scoring. It depends only on six and requires external word segmentation to tokenize Chinese text before scoring. Output includes precision, recall, and F1 scores for ROUGE-1, ROUGE-2, and ROUGE-L metrics.
Use it for
- Evaluate machine-generated Chinese text summaries against reference summaries in NLP research or production systems.
- Batch-score multiple hypothesis-reference pairs from files to assess summarization model performance across datasets.
- Compare Chinese abstractive or extractive summarization outputs in academic papers or model benchmarking.
- Integrate into Chinese NLP pipelines where ROUGE metrics are required for quality assurance or model selection.
- Command-line scoring of individual Chinese text pairs without writing Python code.
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes, if you need ROUGE metrics for Chinese text and can tolerate dormant maintenance.
The package solves real problems in the original ROUGE for Chinese, has low install friction, and no known vulnerabilities. However, verify that the unclear license aligns with your use case, and confirm compatibility with your Python version since support is unspecified. Consider it stable for evaluation tasks but not actively developed.
Install
rouge-chinese on PyPI
Before you install
Low friction: pure Python wheel with a single runtime dependency (six). Dormant since first release with no updates, though the repository remains active with 114 stars.
Requires external word segmentation to tokenize Chinese text before scoring; the package itself does not include tokenization.
License in practice
License treatment is unclear—the package references LICENCE.txt but provides no SPDX identifier. Review the license file in the repository before adopting in proprietary or restricted contexts.
Quickstart
pip install rouge-chinese
from rouge_chinese import Rouge
rouge = Rouge()
scores = rouge.get_scores(hypothesis, reference)
Verify before relying
- Whether the package works correctly with modern Python versions (python_support is unspecified in metadata)
- Current state of the single runtime dependency (six) and its long-term maintenance status
- Whether memory optimizations claimed in the description remain effective with contemporary hardware and text sizes
Package facts
| License | LICENCE.txt unclear |
| Python support | Not specified |
| Install friction | Low. Pure-Python wheel |
| Runtime dependencies | 1 packagesix |
| Maintenance | Dormant 1,426 days since the last release |
| Last repo commit | |
| First released | |
| Downloads | 81,338 / month, #14,236 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
| Classifiers | Intended Audience :: Science/ResearchProgramming Language :: Python :: 3Topic :: Text Processing :: Linguistic |
Evidence: rouge_chinese-1.0.3-py3-none-any.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “chinese rouge metric calculation”
- rouge-chineseComputes ROUGE evaluation metrics for Chinese text summarization and…
- rougeComputes ROUGE scores (Recall-Oriented Understudy for Gisting…
- rouge-metricComputes ROUGE metrics (ROUGE-N, ROUGE-L, ROUGE-W, ROUGE-S, ROUGE-SU)…
Give your agent the search over MCP, or paste the wish link into any chat.
More Linguistic packages
Detects and normalizes text encoding from unknown or ambiguous sources, supporting all IANA character sets that Python's core library provides codecs for, with the ability to register custom codecs.
tiktoken is a fast BPE tokenizer that converts text into token sequences compatible with OpenAI models, supporting multiple encoding schemes including o200k_base and model-specific encodings.
Install it if you work with OpenAI APIs or need to understand token boundaries in GPT-family models.
Detects character encoding and language in byte sequences with high accuracy, supporting 99 encodings and returning confidence scores, language tags, and MIME types.
Install it if you need to detect character encoding or language in byte data; the rewrite makes it substantially faster and more accurate than its predecessors.
Converts Unicode text to ASCII by transliterating non-ASCII characters into their closest ASCII equivalents, with no runtime dependencies.
However, if transliteration quality or ongoing maintenance matters, consider unidecode instead despite its GPL-only license.
Lark is a parsing library that builds abstract syntax trees from context-free grammars, supporting multiple parsing algorithms (Earley, LALR(1), CYK) with automatic line and column tracking.
Python bindings to the tree-sitter parsing library, enabling incremental parsing and syntax tree analysis for source code.
See also rouge · rouge-metric · unbabel-comet · rouge-score · pycocoevalcap · jieba · sacrebleu · bert-score · jieba3k · seqeval