Search 4 packages
| Packages1–4 of 4 | Version | Findings | Description | Downloads | When |
|---|---|---|---|---|---|
| vtcode-eval cargo | 0.173.2 | clean | Agent evaluation framework for VT Code: pass@k / pass^k metrics, capability and regression evals, and environment-based outcome verification. | 3.0K | 9h ago |
| skilltest-core cargo | 0.12.2 | clean | Core library for skilltest: run AI skills on harness/model platforms and score transcripts with natural-language evals. | 434 | 5h ago |
| eval-magic cargo | 0.11.1 | 1 vulnerability 1 | One-stop CLI for running skill evals — measure whether an agent skill actually shifts behavior. | 286 | 21h ago |
| artifactize cargo | 0.5.6 | clean | Declare Artifacts and the evals that review them (tests, LLM reviews, human sign-offs), and reuse every verdict while its fingerprints are unchanged. | 17 | 1d ago |
