Open evaluation artifact

Brali Bench

A portable 50-case suite for checking whether practical-knowledge retrieval preserves relevance, provenance, evidence boundaries, and deliberate no-answer behavior.

Boundary: this evaluates Brali retrieval and grounded packets. It is not an unpinned language-model benchmark.

Current checked report

Reproduce

git clone https://github.com/Brali-LifeOS/brali-lifeos.github.io.git
cd brali-lifeos.github.io
npm run build
npm run evaluate:check

Portable files

Manifest · Cases · Results · Methodology

Comparison layers

Reuse correctly

Pin a data-v* release for research or integration claims. Preserve the Brali dataset version, source cases, methodology, and trust boundaries when publishing derived results.