chore(main): release trg 0.10.0 - #112
Conversation
PR SummaryLow Risk Overview The new changelog entry documents what landed since 0.9.0—eval execution (parallel runs, multi-draw cells, suite subset selection, spend caps), richer case/check declarations (skills, tools, commands, state, harness capabilities), unified transcript events and tool-call observability, OpenAI-compatible judge endpoints, and many fixes around cached/reused runs, pass-rate and cost reporting, run isolation (permissions, answer keys, timeouts), and bundle conformance vs scoring. Reviewed by Cursor Bugbot for commit 861ed26. Bugbot is set up for automated code reviews on this repo. Configure here. |
|
Important
This repository does not receive automatic reviews because it has fewer than 10 stars. ⚙️ Run configurationConfiguration used: Organization UI Review profile: CHILL Plan: Advanced Run ID: Thanks for using CodeRabbit! It's free for OSS, and your support helps us grow. If you like it, consider giving us a shout-out. Comment |
dae95fe to
33043d4
Compare
Signed-off-by: SHT Bot <61149376+sht-bot@users.noreply.github.com>
33043d4 to
861ed26
Compare
An automated release has been created for you.
0.10.0 (2026-09-13)
Features
Bug Fixes
This PR was generated with Release Please. See documentation.