ROUGE automatic summarization evaluation toolkit. Support for ROUGE-[N, L, S, SU], stemming and stopwords in different languages, unicode text evaluation, CSV output.
-
Updated
Apr 9, 2020 - Java
ROUGE automatic summarization evaluation toolkit. Support for ROUGE-[N, L, S, SU], stemming and stopwords in different languages, unicode text evaluation, CSV output.
A library & tools to evaluate predictive language models.
A participatory evaluation toolkit for community-led creative placemaking. Built as a KTP prototype for The Stove Network & University of Strathclyde. Covers project setup, indicator selection, data collection, impact dashboard, and report generation.
Robustness evaluation toolkit for video classifiers: perturbations, saliency-guided tests, metrics & experiment harness
Shared offline evaluation toolkit for depth estimation checkpoints — YOLO26-Depth and Depth Anything V2, PyTorch and OpenVINO, ScanNet benchmarks and local lab-generalization tests, one consistent metric pipeline throughout.
To associate your repository with the evaluation-toolkit topic, visit your repo's landing page and select "manage topics."