SEDIMENT_ORG_ID,
SEDIMENT_DATABASE_URL, and SEDIMENT_MIRROR_PATH set in your shell.
Which model works better?
Sediment compares CI pass rates and the share of captured Inference calls linked to commits (Attribution rate). Reports include sample sizes and uncertainty. Observed differences don’t prove that a model caused an outcome. The comparison doesn’t cover agent harnesses.Did accepted work reach a commit?
The lifecycle report’saccepted_work section follows explicitly accepted work
through Attribution, observed commits, merged pull requests, and CI results.
A missing commit observation doesn’t prove abandonment.
Stock pi and Cursor accepts are implicit, as are automatic Codex approvals.
Those events don’t enter accepted_work; an empty section doesn’t mean that
their work failed to reach a commit. Observed Session commits support the
separate Session progression measurement.
How much code survived?
Sediment measures how much of an edit remains at Session end and, separately, how much of an attributed addition remains after review and merge. Low retention doesn’t identify a defect or who changed the code. Reports with incomplete pull-request history can’t support rankings. For review and merge retention:Where does work get rejected or changed?
The lifecycle report’srework section shows explicit rejects, Retry linkages,
modified or deleted edits, external line changes, and CI failures separately.
These measure different events, so Sediment doesn’t combine them into one
rework score. External changes don’t establish human authorship.
What work is linked to a CI failure?
Sediment traces a failed run to its commit, observed Sessions, and likely contributing Inference calls. The Session dossier shows a timeline of recorded events, without CI logs or call content. These links don’t establish responsibility for the failure. Aftersediment login, replace
COMMIT_SHA with the failed commit’s full hash: