llm_eval 0.1.0 copy "llm_eval: ^0.1.0" to clipboard
llm_eval: ^0.1.0 copied to clipboard

Test harness for LLM outputs: assertion checks, LLM-as-judge scoring, response caching, and CI-friendly reports.

Changelog #

0.1.0 #

Initial release.

  • EvalCase, EvalSuite, and EvalReport with Markdown and JSON output.
  • Built-in checks: contains, notContains, matches, isValidJson, predicate, and LLM-as-judge scoring with Check.judge.
  • ResponseCache interface in the core and a file-backed FileResponseCache in package:llm_eval/io.dart (atomic writes) for deterministic reruns in CI.
  • Concurrent case execution with stable result order.
  • Repeat runs with a flakiness rate.
1
likes
0
points
924
downloads

Publisher

verified publisherdeveloperyusuf.com

Weekly Downloads

Test harness for LLM outputs: assertion checks, LLM-as-judge scoring, response caching, and CI-friendly reports.

Repository (GitHub)
View/report issues

Topics

#llm #ai #testing #evaluation #ci

License

unknown (license)

Dependencies

crypto

More

Packages that depend on llm_eval