$npx skillfedfor your agent

rouge-chinese

Python ROUGE Score Implementation for Chinese Language Task (official rouge score)

With conditionsPyPI LinguisticReleased Sep 202281.3K downloads / moLICENCE.txtPure Python

Decision gist · record as of 2026-08-14

pure-Python wheel — rouge_chinese-1.0.3-py3-none-any.whl
v1.0.3 · released 2022-09-18 · 1 runtime deps: six

Yes, if you need ROUGE metrics for Chinese text and can tolerate dormant maintenance. The package solves real problems in the original ROUGE for Chinese, has low install friction, and no known vulnerabilities. However, verify that the unclear license aligns with your use case, and confirm compatibility with your Python version since support is unspecified. Consider it stable for evaluation tasks but not actively developed.AI-flagged interpretation of the facts on this page — verify before relying

Before you install

  • Requires external word segmentation to tokenize Chinese text before scoring; the package itself does not include tokenization.
  • Low friction: pure Python wheel with a single runtime dependency (six).
  • Dormant since first release with no updates, though the repository remains active with 114 stars.

License · maintenance · safety

LICENCE.txt (unclear) — License treatment is unclear—the package references LICENCE.txt but provides no SPDX identifier. Review the license file in the repository before adopting in proprietary or restricted contexts.

last release 2022-09-18 (1426 days) · last repo commit 2024-06-27 · 114 stars

0 known vulnerabilities (OSV.dev, 2026-08-14) · 81,338 downloads/mo, #14,236 on PyPI

Verify before relying

pip install rouge-chinese

from rouge_chinese import Rouge

rouge = Rouge()
scores = rouge.get_scores(hypothesis, reference)
  • Whether the package works correctly with modern Python versions (python_support is unspecified in metadata)
  • Current state of the single runtime dependency (six) and its long-term maintenance status
  • Whether memory optimizations claimed in the description remain effective with contemporary hardware and text sizes
Same gist for agents: .md · .json

What it is and what it does

Rouge-Chinese is a Python library that calculates ROUGE evaluation scores specifically for Chinese-language NLP tasks. It addresses known limitations in the original ROUGE implementation when applied to Chinese text: incorrect sentence segmentation (missing Chinese punctuation marks), excessive memory consumption during longest-common-subsequence calculation, and inaccurate scores due to approximation methods. The library improves sentence splitting to recognize Chinese punctuation, optimizes memory usage by computing sequence lengths without generating the sequences themselves, and computes official ROUGE scores rather than approximate variants.

The package provides three main interfaces: a library API for scoring single or multiple sentence pairs with optional averaging, a file-based API for batch scoring line-by-line from text files, and a command-line tool for ad-hoc scoring. It depends only on six and requires external word segmentation to tokenize Chinese text before scoring. Output includes precision, recall, and F1 scores for ROUGE-1, ROUGE-2, and ROUGE-L metrics.

Use it for

  • Evaluate machine-generated Chinese text summaries against reference summaries in NLP research or production systems.
  • Batch-score multiple hypothesis-reference pairs from files to assess summarization model performance across datasets.
  • Compare Chinese abstractive or extractive summarization outputs in academic papers or model benchmarking.
  • Integrate into Chinese NLP pipelines where ROUGE metrics are required for quality assurance or model selection.
  • Command-line scoring of individual Chinese text pairs without writing Python code.

Worth the install?

AI-flagged interpretation of the facts on this page. Verify before relying on it.

With conditions

Yes, if you need ROUGE metrics for Chinese text and can tolerate dormant maintenance.

The package solves real problems in the original ROUGE for Chinese, has low install friction, and no known vulnerabilities. However, verify that the unclear license aligns with your use case, and confirm compatibility with your Python version since support is unspecified. Consider it stable for evaluation tasks but not actively developed.

Install

rouge-chinese on PyPI

Before you install

Low friction: pure Python wheel with a single runtime dependency (six). Dormant since first release with no updates, though the repository remains active with 114 stars.

Requires external word segmentation to tokenize Chinese text before scoring; the package itself does not include tokenization.

License in practice

License treatment is unclear—the package references LICENCE.txt but provides no SPDX identifier. Review the license file in the repository before adopting in proprietary or restricted contexts.

Quickstart

pip install rouge-chinese

from rouge_chinese import Rouge

rouge = Rouge()
scores = rouge.get_scores(hypothesis, reference)

Verify before relying

  • Whether the package works correctly with modern Python versions (python_support is unspecified in metadata)
  • Current state of the single runtime dependency (six) and its long-term maintenance status
  • Whether memory optimizations claimed in the description remain effective with contemporary hardware and text sizes

Package facts

LicenseLICENCE.txt unclear
Python supportNot specified
Install frictionLow. Pure-Python wheel
Runtime dependencies
1 package
six
MaintenanceDormant 1,426 days since the last release
Last repo commit
First released
Downloads81,338 / month, #14,236 on PyPI 30-day window, as of 2026-08-14
Known vulnerabilitiesNone known OSV.dev, checked 2026-08-14
Classifiers
Intended Audience :: Science/ResearchProgramming Language :: Python :: 3Topic :: Text Processing :: Linguistic

Evidence: rouge_chinese-1.0.3-py3-none-any.whl

Tags

Capabilities
chinese rouge metric calculationsummarization evaluation chineserouge score chinese nlptext similarity chinesechinese nlp evaluation metricsrouge-1 rouge-2 rouge-l chinesesummarization quality assessment chinese
Topics
chinese-nlpevaluation-metricsummarization
PyPI keywords
NLCLnatural language processingcomputational linguisticssummarizationchinese

Let your AI agent find packages like this

Example. Real query, live index.

You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.

wish › “chinese rouge metric calculation”

  • rouge-chineseComputes ROUGE evaluation metrics for Chinese text summarization and…
  • rougeComputes ROUGE scores (Recall-Oriented Understudy for Gisting…
  • rouge-metricComputes ROUGE metrics (ROUGE-N, ROUGE-L, ROUGE-W, ROUGE-S, ROUGE-SU)…

Give your agent the search over MCP, or paste the wish link into any chat.

More Linguistic packages

charset-normalizer Worth it
PyPI · Utilities · released Aug 2026

Detects and normalizes text encoding from unknown or ambiguous sources, supporting all IANA character sets that Python's core library provides codecs for, with the ability to register custom codecs.

permissive licensepure Python · 3.7+
1.7Bdownloads / mo
tiktoken Worth it
PyPI · Linguistic · released May 2026

tiktoken is a fast BPE tokenizer that converts text into token sequences compatible with OpenAI models, supporting multiple encoding schemes including o200k_base and model-specific encodings.

Install it if you work with OpenAI APIs or need to understand token boundaries in GPT-family models.

permissive licensecompiled wheel · 3.9+
233.0Mdownloads / mo
chardet Worth it
PyPI · Python Modules · released Aug 2026

Detects character encoding and language in byte sequences with high accuracy, supporting 99 encodings and returning confidence scores, language tags, and MIME types.

Install it if you need to detect character encoding or language in byte data; the rewrite makes it substantially faster and more accurate than its predecessors.

0BSDpure Python · 3.10+
199.0Mdownloads / mo
text-unidecode With conditions
PyPI · Python Modules · released Aug 2019

Converts Unicode text to ASCII by transliterating non-ASCII characters into their closest ASCII equivalents, with no runtime dependencies.

However, if transliteration quality or ongoing maintenance matters, consider unidecode instead despite its GPL-only license.

GPL-2.0-or-laterpure Pythonabandoned
89.0Mdownloads / mo
lark Worth it
PyPI · Python Modules · released Oct 2025

Lark is a parsing library that builds abstract syntax trees from context-free grammars, supporting multiple parsing algorithms (Earley, LALR(1), CYK) with automatic line and column tracking.

MITpure Python · 3.8+
79.7Mdownloads / mo
tree-sitter Worth it
PyPI · Linguistic · released Jun 2026

Python bindings to the tree-sitter parsing library, enabling incremental parsing and syntax tree analysis for source code.

MITcompiled wheel · 3.10+
79.0Mdownloads / mo

See also rouge · rouge-metric · unbabel-comet · rouge-score · pycocoevalcap · jieba · sacrebleu · bert-score · jieba3k · seqeval