macOS
brew install trec_evallocal Homebrew formula metadata
brew / rank 12175
Evaluation software used in the Text Retrieval Conference. Version 10.0 via Homebrew; verified 2026-06-17.
install
brew install trec_evallocal Homebrew formula metadata
overview
Evaluation software used in the Text Retrieval Conference
history
trec_eval is the long-running command-line evaluator associated with the Text Retrieval Conference, used to score submitted information-retrieval runs against judged relevance files.
The Text Retrieval Conference began in 1992 as a NIST and U.S. Department of Defense effort to provide infrastructure for large-scale evaluation of text retrieval systems. Within that culture, trec_eval became the shared evaluator for ad hoc retrieval experiments, pairing submitted run files with standard judged results.
The official README describes trec_eval as the standard tool used by the TREC community for evaluating an ad hoc retrieval run. Its file layout reflects that role: parsers for qrels and TREC result formats, measure implementations, tests, and a single small C command-line program that researchers can build with make.
The code has evolved with the evaluation methodology. The 2008/2009 version 9 work was a large rewrite that separated individual measure calculations, made it easier to add measures and input formats, and added measures such as nDCG and preference-evaluation support. The 2026 version 10 changelog records a breaking scoring fix, added measures, stdin support for results files, Windows/GCC fixes, and continued maintenance.
TREC itself grew from an annual research evaluation into a long-running benchmark ecosystem: NIST describes decades of datasets, measurement processes, and tools, and notes that the evaluation software is available so organizations can evaluate their own retrieval systems.
For package users, trec_eval matters because it is not just another metrics script; it is the reference implementation many IR papers, TREC tracks, and regression tests expect. Packaging it in Homebrew gives macOS researchers and search engineers the same executable name and behavior used in benchmark instructions.
The typical workflow is to run `trec_eval` with official qrels and submitted results, optionally adding `-q` for per-query output, `-c` and `-M1000` for official-style handling, or `-m` to print a selected measure.
In package-manager culture, it is often installed for reproducible evaluation notebooks, search-ranking experiments, and old TREC run comparisons where the exact evaluator is part of the experiment.
trec_eval is a compact example of a domain-standard scientific CLI: tiny build surface, old C code, public benchmark authority, and enough historical behavior that package availability helps preserve reproducible research.
security posture
narrow executable package without higher-risk signals.
green risk · low confidence · appliance
Before unattended agent use, check whether the tool reads plaintext credentials, writes remote state, publishes artifacts, or shells out to plugins.
executables
| Command | Kind | Exposure | Note |
|---|---|---|---|
trec_eval | executable | indexed executable | Discovered from the local executable index. |
freshness
These signals separate page generation age, package-manager activity, and upstream release comparison. Version lag is warned only when an evidence URL and comparable versions are present.
install metadata
| Package key | brew:trec_eval |
|---|---|
| Version | 10.0 |
| Package manager | Homebrew |
| Homepage | https://trec.nist.gov/ |
| Repository | https://github.com/usnistgov/trec_eval |
| Last updated | 2026-06-17T22:17:22Z |
| Pulse | updated |
| Bottle | not recorded |
| Service | none declared |
source trail
This page is generated by av-web from the private package SQLite artifact built by scripts/generate-pkg-sqlite.py.
View the package source record on GitHub.