AgentFlight / Changelog
What shipped, version by version.
What shipped in each AgentFlight release, mirrored from the public repository.
[0.16.1] - 2026-06-29
Fixed
- Included the real AgentFlight finish screenshot in the npm package allowlist
so the published README can render
docs/assets/agentflight-finish.png.
Verification
- Release audit:
docs/development/v0.16.1-release-audit.md.
[0.16.0] - 2026-06-29
Added
- Added
agentflight guard, a local live trust monitor for active coding-agent sessions with watch mode,--once, JSON output, exit codes, and CLI help. - Added Guard trust signals for proof gaps, Baseframe scope drift, Baseframe verification gates, stale review receipts, and ready states.
- Added finish-target hints to Guard output for Review Passport JSON, Review Passport Markdown, and Baseframe result artifacts.
- Added exported Guard package API types plus
createAgentFlightGuardSummary. - Added Guard documentation, a website update prompt, and pre-release readiness evidence.
Changed
- README workflow now places Guard before explicit verification and
finish, while keeping Review Passport as the final review packet. - Package contents now include
docs/development/guard.md.
Fixed
- Stabilized git-heavy Baseframe fixture tests under full-suite load by increasing local fixture git timeouts for Guard and Finish tests.
Security
- Guard remains local-only and source-free. It reads existing AgentFlight status and Baseframe metadata, does not run hidden verification commands, and does not upload source, send telemetry, or post PR comments.
Verification
- Release audit:
docs/development/v0.16.0-release-audit.md.
[0.15.0] - 2026-06-27
Added
- Added
agentflight finish, a local end-of-session command that writes a Review Passport and refreshes the handoff, report, replay, and resume artifacts. - Added Review Passport JSON and Markdown artifacts under
.agentflight/reports/<session-id>-review-passport.jsonand.md. - Added source-free Review Passport integrity fingerprints for session, changed-file, verification, review, Baseframe, and artifact metadata.
- Added exported
ReviewPassportV1andcreateReviewPassportfor package API consumers. - Added Review Passport documentation and README workflow updates.
Changed
- In Baseframe sessions,
agentflight finishalso finalizes.baseframe/evidence/<task-id>/agentflight-result.jsonand prints the AgentLoopKit reconciliation command. - AgentFlight-owned Baseframe output is filtered from changed-file analysis so
generated
agentflight-result.jsonand.baseframe/agent-workflow.jsondo not create scope drift.
Security
- Review Passport artifacts remain local and source-free. They store paths, commands, statuses, timestamps, counts, artifact paths, and hashes, without uploads, telemetry, PR posting, source contents, or secret reads.
Verification
- Release audit:
docs/development/v0.15.0-release-audit.md.
[0.14.1] - 2026-06-27
Added
- Added a real AgentFlight v0.14.0 Baseframe readiness terminal capture for the product README and website pipeline.
Changed
- Included the new Baseframe readiness image in the npm package allowlist so the README asset is available from the published package.
Verification
- Release audit:
docs/development/v0.14.1-release-audit.md.
[0.14.0] - 2026-06-27
Added
- Added Baseframe Suite Integration v1 so AgentFlight can start from AgentLoopKit task-contract JSON, resolve linked ProjScan assessment JSON, persist optional Baseframe integration context, and remain independent of ProjScan and AgentLoopKit internals.
- Added
agentflight finalize, which writes.baseframe/evidence/<task-id>/agentflight-result.json, refreshes local report/replay/resume evidence, updates.baseframe/agent-workflow.json, and prints the AgentLoopKit reconciliation command. - Added deterministic scope-drift detection for Baseframe sessions using
allowed and excluded path globs while preserving AgentFlight runtime-file
filters and keeping
.agentflight/config.jsonvisible. - Added required verification gate reconciliation with exact command and normalized-whitespace matching only.
- Added exported Baseframe contract/result types plus
loadBaseframeIntegrationContextandcreateAgentFlightResult. - Added Baseframe sections to status, Markdown report, HTML replay, and resume prompt: Repository Assessment, Task Contract, Scope Adherence, Verification Gates, Review Focus, Proof Gaps, Readiness, and Next Action.
- Added Baseframe contract fixtures and end-to-end tests that simulate ProjScan, AgentLoopKit, AgentFlight execution drift, verification outcomes, and final result generation without depending on sibling packages at runtime.
Changed
agentflight startnow accepts--from-task,--from-projscan, and--task-idfor Baseframe workflows while preserving the existing--task-based standalone session flow.verify,snapshot,report,replay,resume, andstatusnow refresh the Baseframe result artifact when the current session has Baseframe integration context.- Package contents now include the Baseframe Suite Integration v1 documentation.
Security
- Baseframe integration reads and writes only repository-local JSON artifacts with safe path validation, task ID validation, schema/kind checks, and atomic writes for generated result and workflow manifest files.
- Verification gate matching intentionally avoids fuzzy matching so unrelated commands cannot satisfy required AgentLoopKit gates.
Verification
- Release audit:
docs/development/v0.14.0-release-audit.md.
[0.13.0] - 2026-06-24
Added
- Added Trust Delta guidance across status, handoff, Markdown report, HTML replay, resume, and status JSON. It summarizes failed proof, stale proof, missing proof, manual review, and repo-history under-proofing from existing local metadata.
- Added a review queue that orders proof reruns, missing-proof commands, manual checks, repo-calibration guidance, and file inspection without adding a new command.
- Added local review receipts via
agentflight handoff --accept, with current or stale receipt state across status, handoff, Markdown report, HTML replay, resume, history, and status JSON. - Added role-aware review routing across status, handoff, Markdown report, HTML replay, resume, and status JSON so maintainers, verification reviewers, security reviewers, Docs/DX reviewers, and release reviewers can see their specific local review path.
Changed
agentflight startnow stores configured.agentflight/config.jsonverification commands in the session when present, so later status, handoff, report, replay, and resume guidance matches the proof commands shown at start.- Status and handoff now include a small full-command recovery block when dense review sections shorten long suggested proof commands.
- Status and resume now open existing ready artifacts only when the latest recorded review summary still matches the current ready changed-file count.
- Repo calibration now prefers accepted local review receipts when enough similar accepted sessions exist, while keeping ready handoffs as fallback history.
- Repo calibration now ignores verification runs recorded after the accepted handoff boundary, so later exploratory proof does not change what the accepted handoff taught AgentFlight.
- Repo calibration history loading now stops after enough newest ready sessions are loaded instead of parsing every local session file.
- Review routing reuses existing source-free Review Intelligence signals rather than adding a new command or hosted workflow.
- Review routing no longer sends stale non-release receipts to the Release route, and verification routing no longer describes legacy proof freshness as current.
- Verification routing now keeps the proof route clear when only manual-review files are stale after proof capture.
- Project Review Contract stale proof status now applies only to requirements whose own proof-required files changed after proof was captured.
- Project Review Contract rules that accept multiple proof kinds now choose the best current accepted proof instead of letting an older stale proof of another accepted kind win.
- Project Review Contract rules that accept multiple proof kinds no longer mark the requirement failed when another accepted proof kind is already current.
- Text review surfaces now show when Review Focus rows are capped instead of silently omitting lower-priority files.
- Markdown reports, resume prompts, and status changed-area output now cap large changed-file lists with a remaining count.
- Status, handoff, Markdown report, and resume now use one shared suggested command collector for full-command recovery blocks.
Fixed
- Accepted review receipts now become stale when an unresolved verification failure happens after local acceptance.
- History now marks accepted review receipts stale when new changed files appear after acceptance while preserving non-Git fallback behavior.
- Status JSON now uses the same open-first next action as terminal status when a ready handoff artifact already exists.
- Markdown report, resume prompt, and handoff Markdown artifacts now escape raw HTML in dynamic review text, file paths, task names, and failure excerpts while preserving raw stdout/stderr evidence files.
- Handoff failed-verification excerpts now render as fenced text blocks, so Markdown-looking failure output stays readable without changing artifact structure.
- Handoff output now promotes only unresolved failed verification excerpts when mixed resolved and unresolved failures exist.
- HTML replay now points urgent failed-run navigation at unresolved failed runs and marks resolved failures as historical in mixed failure histories.
- Clean-worktree handoff runs now preserve existing session report, replay, handoff, and resume artifacts even if the current resume prompt was removed, and restore the current resume from the session artifact when available.
- Clean-worktree resume restoration now uses a direct local artifact copy instead of routing session resume content through a read/write transform.
- Persisted session ids are now validated before session artifact paths are
built, preventing unsafe local metadata from escaping
.agentflight. - Core verification evidence writing now validates persisted session ids before reserving stdout/stderr evidence paths.
- Display excerpts now strip terminal control, OSC, and bidi control characters while preserving raw stdout/stderr evidence files.
- Passing verification commands that write more than 1 MiB no longer false-fail due to the process runner's output buffer limit.
- Trust Delta and review queue stale-proof rows now retain all stale Project Review Contract files instead of only the first stale requirement.
- Handoff now retains proof-gap related files and focus suggested proof commands when reading status JSON.
- Markdown aggregate sections in handoff, report, and resume now preserve generated list structure while escaping active Markdown syntax inside dynamic review text.
- Review Contract proof references in replay now cap dense visible links and fall back to changed-file anchors when a review-focus row is hidden by the visible-row limit.
- Escaped full-command summaries inside Markdown
<details>blocks. - Added a replay anchor for Review Readiness proof references and kept full commands visible in print/PDF output.
- Capped long proof-freshness and required-proof file lists in dense review surfaces.
Security
- Trust Delta and review queue use source-free local metadata only: paths, changed-file categories, proof gaps, proof commands, proof freshness, Project Review Contract status, and repo-calibration summaries.
- Review receipts store source-free local metadata only: receipt decision, timestamp, changed paths, proof counts, readiness state, commit/branch, and proof snapshot fingerprints. They do not add identity, signatures, upload, telemetry, or automatic PR comments.
- Review routing uses source-free local metadata only: changed-file paths and categories, proof gaps, proof freshness, Trust Delta, review queue, calibration, review receipt state, and readiness.
- Proof freshness and receipt freshness hash changed files locally for fingerprints, but AgentFlight does not store, render, upload, or analyze source contents, full diffs, or historical stdout/stderr evidence for these review signals.
Verification
- Release audit:
docs/development/v0.13.0-release-audit.md.
[0.12.0] - 2026-06-24
Added
- Added repo-calibrated proof guidance that compares current proof with similar local ready handoffs and suggests additional proof commands when current proof is weaker than local history.
- Status, handoff, Markdown report, HTML replay, resume, and status JSON now include repo calibration when local session history is available.
- Proof freshness now attributes stale proof to the files and categories that changed after verification, so docs-only changes can stay manual-review guidance while source/test changes still recommend rerunning proof.
Security
- Repo calibration reads bounded local session metadata only. It does not read historical stdout/stderr evidence files or full diffs, and it does not upload, sync, or call external services.
Verification
- Release audit:
docs/development/v0.12.0-release-audit.md.
[0.11.0] - 2026-06-24
Added
- Added a local Project Review Contract proof standard that maps changed-file categories to required proof kinds and manual review checks.
- Status, Markdown reports, HTML replays, resume prompts, handoffs, and status JSON now show required proof alongside actual proof state.
- Project Review Contract requirements now explain why each rule matched, which proof command satisfied it, and which reviewer actions remain.
- New
.agentflight/config.jsonfiles include a default local Project Review Contract baseline for auth/security/payment, database, backend/API, dependency, config/CI, frontend, source, tests, docs, and AgentFlight config changes.
Changed
- Review Contract claims now include Project Review Contract requirement claims before file-level claims when a requirement matches the changed files.
- Handoff now leads with a decision and why summary before the detailed review packet.
- Handoff, status, report, resume, and replay surfaces now show the same explainable required-proof decision path.
- Public docs now describe AgentFlight as a local-first review layer for coding agent sessions and document the Project Review Contract handoff path.
Verification
- Release audit:
docs/development/v0.11.0-release-audit.md.
[0.10.0] - 2026-06-22
Added
- Review Contract claims now include source-free proof references for changed files, proof statuses, proof gaps, readiness reasons, and suggested proof commands so reviewers can see why each claim exists.
- Review Contract output now includes a compact review path summary and next action across status, Markdown reports, HTML replays, resume prompts, and handoffs.
- HTML replays now add stable anchors for Review Contract claims, review-focus rows, and proof gaps, allowing claim-to-proof navigation in the local replay.
Changed
- README samples now show the local Review Contract path and include an ASCII local-review-flow diagram.
[0.9.0] - 2026-06-22
Added
- Added a deterministic local Review Contract claim ledger that turns task, file, proof-gap, and readiness signals into explicit supported, stale, failed, unsupported, manual-review, not-testable, or unknown claims.
- Status JSON, terminal status, Markdown reports, HTML replays, resume prompts, and handoffs now expose Review Contract claims without adding cloud, PR comments, CI JSON, telemetry, or model-based claim extraction.
Fixed
agentflight doctornow treats configured.agentflight/config.jsonverification commands as satisfying proof-command setup, avoiding misleading missing root-script warnings in monorepos that verify through subpackage commands.- Review Intelligence now recognizes passed
verifyscripts such asnpm run verifyas proof, so AgentFlight's own default verification command can satisfy source and test review gaps.
Verification
- Release-candidate verification is tracked in
docs/development/v0.9.0-release-audit.md.
Show 17 older releases
[0.8.0] - 2026-06-21
Added
- Verification runs now store source-free proof snapshots of changed files so Review Intelligence can distinguish current proof from stale proof after a file changes.
- Status, Markdown reports, HTML replays, resume prompts, snapshots, and
handoffs now surface proof freshness as
current,stale,covered,missing,failed,not required, orunknown.
Changed
- Review focus rows now include a compact proof-status line across local review surfaces, making stale proof visible without opening session JSON.
- Stale proof now creates an actionable proof gap and keeps readiness at
Needs verificationuntil the relevant proof command is rerun.
Fixed
- Verification proof capture now avoids slow git-status probing in non-git sessions with no known changed files.
[0.7.1] - 2026-06-21
Fixed
- Refreshed the README terminal hero GIF so it shows the current handoff-first review flow instead of the older replay-first flow.
- Updated the README hero caption and sample output to include
handoffandhistoryaccurately.
Changed
- Tightened README workflow copy around the local handoff packet, recent local session history, and current failure-excerpt surfaces.
[0.7.0] - 2026-06-21
Added
- Added
agentflight handoff, a local-only review handoff command that generates the report, replay, and resume artifacts and summarizes readiness, proof gaps, review focus, and failed verification excerpts. - Added
agentflight history, a read-only local command for listing recent sessions, proof counts, the current-session marker, and existing report/replay artifact paths without search indexing, export, sync, or session switching. - Added local
agentflight history --task <text>andagentflight history --state ready|blocked|needs_verification|unknown|currentfilters so engineers can find relevant local sessions without scanning long histories or adding any index, sync, export, or session switching behavior. - Added a compact Review Path section to local HTML replay artifacts so long sessions lead reviewers through proof gaps, unresolved failed runs, review focus, and verification evidence without adding scripts, exports, or hosted behavior.
- Documented the
agentflight history --limit 1latest-action workflow for reopening local handoff/report/replay/resume artifacts. - Added session-specific handoff artifacts under
.agentflight/reports/soagentflight historycan point to stable handoff packets from prior sessions. - Added session-specific resume artifacts under
.agentflight/reports/soagentflight historycan point to stable continuation prompts from prior sessions. - Added post-v0.6.0 user-research findings and a v0.6.0 website update prompt focused on the local handoff workflow.
- Added a post-v0.6.0 product direction note that keeps local handoff, first-run workspace hygiene, replay ergonomics, proof guidance, and explainable ranking as the priority order.
Fixed
- Ready
agentflight statusoutput now points at an existing local handoff, replay, or report artifact when one already exists, instead of repeating handoff-generation guidance. - Ready
agentflight resumeprompts now point at an existing local handoff, replay, or report artifact when one already exists, instead of repeating handoff-generation guidance. - Review Intelligence now ignores unfinished AgentFlight readout/artifact
commands such as
agentflight replayandnode dist/cli.js replaywhen computing incomplete proof gaps, so readiness guidance keeps pointing at meaningful verification commands. - Incomplete verification guidance now says a command may still be running and suggests waiting before rerunning, avoiding false lost-evidence alarms when status is checked during parallel verification.
- Ready
agentflight handoffterminal guidance now tells reviewers to share the local handoff packet first, with report/replay as supporting detail artifacts. - Current start-only
agentflight historysessions now guide users to run verification before handoff when no proof exists yet, while keeping handoff-only guidance once verification has been recorded. - Public positioning regression coverage now guards current public/runtime surfaces against stale assistant-style positioning.
- Clean-worktree status now tells users to open handoff/report/replay or JSON output for tucked verification details, matching the handoff-first ready review path.
- Ready-session
Open firstguidance now prefers the generated handoff packet in handoff, clean status/resume, and history surfaces, while blocked sessions still point to the report/fix path. - Full Markdown proof reports now show changed files, risk, verification, review focus, proof gaps, readiness, recommendation, and next action before the timeline, keeping long sessions faster to review.
agentflight historynow compacts non-current start-only sessions that have no proof or review artifacts, keeping recent local artifacts easier to scan.- Review Intelligence now keeps generated
.agentflight/.gitignorehelper files below real first-run review targets while leaving them visible. agentflight resumenow includes unresolved and historical failed-run context below the verification count, matching status/report trust cues.- Clean-worktree
agentflight resumeprompts now include the latest localOpen first:artifact path when current report/replay/handoff evidence exists. - Clean-worktree
agentflight statusnow shows the latest localOpen first:artifact path when current report/replay/handoff evidence exists. agentflight historynow shows the nearest previous local artifact when the current latest session has none yet, while keeping the current-session handoff next action.- First-run AgentFlight generated-file lists and local-file guidance now use
shared output helpers, keeping
initandstart --yescopy consistent. agentflight start --yesnow explains the AgentFlight files it generated during safe auto-init, including project config and local runtime evidence.agentflight initandagentflight startnow use concise ProjScan version checks instead of extra help probing, whileagentflight doctorkeeps deeper diagnostics.agentflight doctornow renders multiple detected proof-command suggestions as an indented list instead of one long semicolon-separated line.agentflight historynow prefers useful review-ready or blocked artifact metadata over later clean-worktree artifact metadata, while preserving clean-only session history.- Clean-worktree
agentflight handoffnow preserves existing session review artifacts instead of overwriting report/replay/resume evidence with an empty post-commit clean-state artifact. - Clean-worktree
agentflight handoffnow exits successfully instead of treating an informational post-commit handoff as a command failure. agentflight doctornow suggests concrete detectedagentflight verify -- ...commands when config verification commands are empty.agentflight verifynow suggests detected package proof commands when no explicit command is provided and config commands are empty.agentflight doctornow warns when package proof scripts exist but.agentflight/config.jsonhas no configured verification commands.agentflight historynow lists capped repo-relative malformed session paths instead of only reporting a skipped-file count.- Concurrent
agentflight verifyruns now reserve distinct stdout/stderr evidence paths and merge verification updates without dropping either run. agentflight history --limitnow rejects non-integer, zero, and negative values with a clear local error instead of silently falling back or returning an empty history.- Review Intelligence now treats an earlier failed verification as resolved when the same stored command later passes, so TDD red/green and format-fix loops do not leave handoffs permanently blocked.
- Status, report, replay, and handoff now distinguish unresolved failed verification from historical failed runs that later passed.
- Resume prompts now use the same unresolved-versus-resolved verification count wording as the other review surfaces.
- Clean-worktree status now points users to
agentflight history --limit 1to reopen the latest local artifacts before starting the next session. - History now shows unresolved-versus-resolved failed verification counts for prior sessions.
- History now includes stable handoff artifact paths alongside report and replay paths when those artifacts exist.
- History now includes stable resume artifact paths when those artifacts exist.
- History now suggests which existing local artifact to open first for each session.
- History now surfaces the newest session's open-first action before the full session list, reducing scan work in long local histories.
- History now tells users to run
agentflight handoffwhen the latest local session is current but no handoff/report/replay artifact exists yet. - History now shows the latest session's recorded readiness in the top-level
Latest action:block. - History now says
Open first: none yetwhen no handoff, report, or replay artifact exists. - HTML replay now reserves urgent failed-run navigation for unresolved failed verification while keeping historical failed runs visible in the ledger.
- Ready handoffs no longer inline historical failed verification excerpts once no unresolved failed-verification proof gap remains; those excerpts stay in report/replay evidence.
agentflight startnow treats AgentLoopKit'sActive task: none pinned.status as no active task instead of falsely reporting task reuse.- AgentLoopKit task-link diagnostics now use generic link-check wording instead of stale automatic task-creation copy.
agentflight initnow reports ProjScan and AgentLoopKit CLI availability using the same concise tool formatter as start/report surfaces instead of relying on repo marker files.agentflight statusnow reports a clean worktree explicitly instead of calling zero changed filesUnknownafter a completed session has been committed.agentflight statusnow compacts very long terminal verification run lists while keeping full verification runs in JSON, report/replay, and local evidence files.- Clean-worktree
agentflight statusnow tucks individual verification run details when there are no unresolved failed runs, while keeping counts and JSON evidence complete. - Review Intelligence now treats incomplete verification attempts as blocking before clean-worktree readiness, so live status cannot call a session clean while verification is still in progress.
agentflight initnow lists created and skipped files by repo-relative path instead of only showing counts.agentflight doctornow warns when.projscan-memory/memory.jsonexists without a matching project filter and reports OK once the repo filters it throughchangedFileFilters.ignore.agentflight historynow labels stored review metadata asRecorded readiness:so it is not confused with live worktree readiness.agentflight doctorno longer prints the absolute repository root in the successful repository-root check.- HTML replay now labels resolved failed verification rows as historical when no unresolved failed runs remain.
agentflight doctornow treats a missing current session as OK first-run guidance instead of warning when the rest of the local setup is healthy.- Clean-worktree status now reports
Risk: noneinstead ofRisk: unknownwhile preservingunknownfor legacy or genuinely missing metadata. - Clean-worktree risk reasons now use current-state wording instead of saying no changed files were detected "yet."
- Parallel report, replay, and resume commands now preserve each artifact event in session history instead of letting the last stale session write win.
- Review Intelligence no longer lets ProjScan risk hints boost generated
.projscan-memory/memory.jsonabove real first-run review targets, while the file remains visible with the existingchangedFileFilters.ignoreguidance. agentflight historynow includes the selected local artifact path directly on theOpen first:line, reducing lookup work in long session lists.agentflight handoffnow includes the selected report or replay path directly on theOpen first:line while preserving the full artifact list.
Changed
- Clean-worktree
agentflight resumeconstraints now tell agents to start a new AgentFlight session before unrelated work instead of applying active-task constraints to a completed clean session. agentflight startnow prefers configured no-argagentflight verifyguidance when.agentflight/config.jsonalready has verification commands, while keeping detected package-script fallback guidance for empty configs.- Current product copy now uses
coding agent sessionsand agentic engineering language instead of assistant-style positioning. - Idempotent
agentflight initnow shows a concrete detected proof command when existing config verification commands are empty. - First-run
agentflight initnow points seeded configs at no-argagentflight verifyin the primary workflow. agentflight initnow points first-run users through the handoff golden path: start a session, capture verification, then generate a local handoff, with status and doctor listed as supporting checks.agentflight initnow uses the first detected verification command in its primary workflow guidance, falling back to<proof command>when no proof script is detected.- Newly generated
.agentflight/config.jsonfiles now seed detected verification commands from package scripts while leaving profiles empty and preserving existing configs. - README and verification docs now describe the handoff-first first-run workflow and init-seeded verification commands.
- Kept ready-review report, replay, resume, examples, and demo copy aligned with
the
agentflight handoffgolden path while keeping report/replay/resume as supporting local artifacts. - Changed-file review surfaces now fail with an actionable git-status error instead of treating git-status failures as an empty diff.
- Shortened the optional ProjScan baseline budget during
agentflight startso busy local ProjScan work cannot stall session startup for too long. agentflight startnow reuses an active AgentLoopKit task instead of creating a duplicate AgentFlight placeholder task.agentflight startnow shows concise ProjScan and AgentLoopKit warning summaries when optional tooling is available but degraded.agentflight startnow inspects ProjScan availability without running the heavier optionalprojscan startbaseline on the startup path.agentflight startnow uses lightweight AgentLoopKit availability inspection on the startup path while preserving task reuse/linking.agentflight startnow reuses AgentLoopKit's local active-task state file directly instead of parsingagentloopkit statusoutput.agentflight startnow links existing AgentLoopKit task state without creating new AgentLoopKit task contracts automatically.- Start output and Markdown tooling rows now show whether AgentLoopKit has an active task linked when that local state is known.
agentflight handoffnow treats missing required proof as not ready to share: it exits non-zero, usesFix before sharing, and points users to the report first.- Start, report, replay, and handoff terminal output now display local
.agentflight/...artifact paths relative to the repo instead of absolute user-directory paths. - Handoff verification details now distinguish zero verification runs from passing runs that simply have no failed excerpts.
- Review Intelligence suggested proof now follows each proof gap's preferred
proof-kind order, so source gaps prefer
npm testwhen available and dependency gaps prefer build/install-style proof before typecheck. - Review Intelligence proof-gap rules are now centralized in one ordered table to keep future proof guidance changes easier to review.
- HTML replay verification ledgers now display long run commands compactly while keeping the full command available in the title text.
- Status and Markdown report verification evidence rows now use compact display labels for long run commands while preserving stored command evidence.
agentflight initnow writes.agentflight/.gitignorefor runtime evidence directories instead of seeding new runtime.gitkeepfiles, reducing first-run Git noise while keeping.agentflight/config.jsonand the local AgentFlight ignore file visible as project config.- AgentFlight session JSON writes now use same-directory temp files and atomic rename so concurrent report, replay, resume, and handoff commands do not read partially written session state.
- Review Intelligence now describes
.projscan-memory/memory.jsonas generated tool state instead of arbitrary unknown code while keeping the file visible. - Report and replay generation now persist a compact local readiness summary in
session events, letting
agentflight historyshow the latest recorded readiness without recalculating old sessions.
AgentFlight v0.6.0 - 2026-06-19
Local review ergonomics and automation surfaces for heavier real-world dogfood.
Added
- Replay navigation for long HTML evidence ledgers, including jump links, section anchors, failed-run anchors, and a shortcut to the first failed verification run.
- Local JSON status output with
agentflight status --format jsonfor scripts that need structured session, risk, verification, and Review Intelligence state. - Named verification profiles with
agentflight verify --profile <name>for local command groups stored in.agentflight/config.json. - Markdown report modes:
agentflight report --mode compactfor shorter local review summaries.agentflight report --mode pr-commentfor a local PR-comment draft that is never posted automatically.
- Optional typed ProjScan review hints for deterministic Review Intelligence ranking enrichment without making Review Intelligence call ProjScan directly.
- First-run guidance that explains which
.agentflight/paths are runtime evidence and which files are project config.
Changed
Clarified first-run workspace hygiene docs:
.projscan-memory/**can be added tochangedFileFilters.ignorewhen ProjScan memory is generated evidence rather than a review target.Lowered generated ProjScan memory priority in Review Intelligence so
.projscan-memory/memory.jsonremains visible but no longer outranks real first-run review targets such as.agentflight/config.jsonor docs changes.Classified first-party TypeScript/JavaScript source files under
src/as source changes so review focus gives clearer guidance and proof gaps than the previous unknown-file fallback.Aligned ready-review next actions with the handoff golden path: status, report, replay, and resume now point users toward
agentflight handoff, while the handoff artifact itself tells users to share the generated local packet.AgentFlight now describes itself as a local-first review layer for coding agent sessions across package metadata, README, and product docs.
Long suggested proof commands stay compact in high-density review surfaces while preserving the full suggested action where useful.
Local AgentLoopKit evidence paths are filtered from AgentFlight changed-file review surfaces:
.agentloop/state.json.agentloop/reports/**.agentloop/handoffs/**.agentloop/runs/**
AgentLoopKit task contracts, policies, harness files, and gates remain visible for review.
.projscan-memory/**remains suggestion-only guidance, not a built-in ignored path.
Fixed
- Reduced Review Intelligence and report noise caused by generated local workflow evidence.
- Split complex verification/profile and ProjScan-hint logic into smaller helpers without changing command behavior.
- Added regression coverage for PR-comment draft failure excerpts so stderr-preferred excerpts stay aligned with report/replay behavior.
Verification
npm run verifypassed.npm run format:checkpassed.npm pack --dry-runpassed.npm audit --audit-level=moderatefound 0 vulnerabilities.- AgentLoopKit verification passed.
- ProjScan doctor passed with score 100/A.
- ProjScan preflight/review returned a documented manual release-signoff caution for scale risk only, with no concrete blockers.
AgentFlight v0.5.1 - 2026-06-17
Focused dogfood patch for v0.5.0 review artifacts.
Fixed
- Terminal
agentflight verifynow uses the same stderr-preferred failure excerpt as report/replay artifacts. - Raw stdout/stderr evidence remains preserved.
- Reduced noisy AgentLoopKit Tooling diagnostics in Markdown reports.
Changed
- Trimmed long suggested proof commands in status, report, replay, and resume surfaces.
- Added shared output display helpers for compact proof commands and concise tool-report lines.
- Kept
.projscan-memory/**as suggestion-only guidance, not a built-in ignored path.
Verification
npm run verifypassed.npm run format:checkpassed.npm pack --dry-runpassed.npm audit --audit-level=moderatefound 0 vulnerabilities.- AgentLoopKit verification passed.
- ProjScan doctor passed with score 100/A.
- ProjScan preflight/review returned a documented manual release-signoff caution for scale risk only, with no concrete blockers.
[0.5.0] - 2026-06-16
Inline verification evidence and replay accessibility.
Added
agentflight verifycaptures a short output excerpt (stderr preferred, otherwise stdout) at run time and stores it on the verification run.- The replay shows the excerpt inline for failed runs, so a reviewer can see what broke without opening the evidence file. Passing runs keep the excerpt tucked inside the evidence details to stay calm.
- The Markdown proof report includes the failure excerpt as a fenced block for failed runs.
Changed
- Replay readout band now sticks to the top so risk and readiness stay visible while scrolling long records.
- Raised contrast on muted replay text and risk colors to meet WCAG AA for small text.
Added (replay)
- Print stylesheet for clean PDF handoffs and incident reconstruction: expands evidence, avoids mid-record page breaks, and preserves risk and readiness color.
Verification
npm run verifypassed.npm run format:checkpassed.
[0.4.2] - 2026-06-16
Replay UI redesign and refreshed product demos.
Changed
- Redesigned the HTML replay artifact as a flight-record evidence ledger: a verdict-led masthead where review readiness is the lead signal, an instrument-style readout band in place of the metric-card grid, a log-spine timeline whose event nodes encode type by shape and fill, and a verification ledger with explicit
PASS/FAILstamps. - The replay now stays calm by default and applies semantic color only when risk is elevated or proof is failing, and never relies on color alone.
- Updated ProjScan to 4.5.0 and AgentLoopKit to 0.35.2.
Added
- Reproducible demo-asset pipeline:
npm run demo:assetsregenerates the replay screenshot and scroll GIF (Playwright) and the terminal workflow GIF (VHS) fromscripts/demo/anddocs/marketing/. - Refreshed README with the new terminal workflow GIF and an animated replay walkthrough.
Verification
npm run verifypassed.npm run format:checkpassed.
[0.4.1] - 2026-06-15
Review Intelligence trust patch focused on v0.4.0 dogfood findings.
Fixed
- Detect incomplete verification attempts when a
verification_startedevent has no later completed result. - Incomplete verification now appears as a blocking Review Intelligence proof gap and prevents
Ready for review. - Report output no longer shows legacy verification-gap text that conflicts with Review Intelligence proof gaps.
.agentflight/config.jsonis classified as AgentFlight project config instead of genericunknown.
Changed
agentflight verifyemits a minimal heartbeat while long-running verification commands are still active..projscan-memory/memory.jsonnow triggers an informational suggestion to add.projscan-memory/**tochangedFileFilters.ignorewithout hardcoding that path as a built-in ignore.
Documentation
- Documented incomplete verification handling, heartbeat output, and optional ProjScan memory filters.
- Updated v0.4.0 dogfood findings with v0.4.1 patch candidate status.
Verification
npm run verifypassed.npm run format:checkpassed.npm pack --dry-runpassed.npm audit --audit-level=moderatefound0 vulnerabilities.- AgentLoopKit verification passed.
- ProjScan doctor passed with health
100/100. - ProjScan preflight returned
proceedwith no caution. - Local packed-package smoke test passed.
[0.4.0] - 2026-06-14
Added
- Added deterministic Review Intelligence for coding agent sessions.
- Added review focus ranking to highlight the files developers should inspect first.
- Added proof gap detection for missing or failed verification evidence.
- Added clearer readiness states and next-best-action guidance.
- Added config-driven generated/internal changed-file filters via
changedFileFilters.ignore. - Added documentation for generated/internal file filters.
- Added completion audit for v0.4.0.
Changed
status,report,replay,resume, andsnapshotnow include review intelligence where appropriate.- Updated replay UI screenshot to show review focus, proof gaps, readiness, and timeline.
- Hardened malformed
changedFileFilters.ignorehandling. - Kept
.agentflight/config.jsonvisible while runtime.agentflight/artifacts remain filtered.
Deferred
- PR comments.
- JSON/CI integration.
- ProjScan-enriched ranking.
- Heartbeat/progress output for long verification.
- Interrupted verification cleanup.
- Tool availability messaging alignment.
- Cloud/login/billing/Pro/Team/GitHub App.
Verification
npm run verifypassed.npm run format:checkpassed.npm pack --dry-runpassed.npm audit --audit-level=moderatefound0 vulnerabilities.- AgentLoopKit verification passed.
- ProjScan doctor passed with health
100/100. - ProjScan preflight returned a documented scale/complexity caution that was manually reviewed and accepted.
[0.3.3] - 2026-06-14
Patch candidate focused on dogfood findings from published agentflight@latest v0.3.2.
Changed
- Improved HTML replay readability with a calmer developer-review layout, clearer timeline rows, cleaner file grouping, and denser verification evidence rows.
- Moved replay stdout/stderr evidence paths into collapsed native details so proof stays available without dominating the page.
Fixed
- Excluded AgentFlight runtime artifacts from changed-file analysis by default:
.agentflight/sessions/**.agentflight/reports/**.agentflight/current/**.agentflight/evidence/**
- Kept
.agentflight/config.jsonvisible in changed-file analysis because it is user-controlled project configuration, not runtime evidence.
Documentation
- Added
PRODUCT.mdwith the confirmed product UI direction: precise, calm, trustworthy. - Added
docs/development/v0.3.2-dogfood-findings.mdwith dogfood results from AgentFlight, ProjScan, and fifa-predictor. - Updated the development log with dogfood, UI polish, and patch verification evidence.
[0.3.2] - 2026-06-13
Changed
- Added the official AgentFlight logo to the README.
- Added the Baseframe Labs AgentFlight website link near the top of the README.
- Added an animated CLI workflow preview before the 60-second workflow.
- Added a "Watch The Flow" section explaining start -> verify -> snapshot -> status -> report/replay/resume.
- Updated package homepage to the AgentFlight website page.
- Expanded packaged assets so npm README logo and animation references resolve.
Documentation
- Updated launch notes with the AgentFlight logo reference.
- Updated development log with branding and README animation verification.
Verification
npm run verifypassed.npm run format:checkpassed.npm pack --dry-runpassed.npm audit --audit-level=moderatefound0 vulnerabilities.- ProjScan preflight passed.
- AgentLoopKit verification passed.
[0.3.1] - 2026-06-13
Changed
- Reworked the README around a clearer 60-second AgentFlight workflow.
- Added sample CLI output for status, report, replay, and resume.
- Added a replay timeline screenshot to make the v0.3.0 experience easier to understand.
- Added a basic example session walkthrough and v0.3.0 launch-note drafts.
- Improved npm keywords and package metadata.
- Narrowed packaged files to include useful README-linked docs/assets without shipping marketing drafts.
Documentation
- Added launch notes draft for v0.3.0 public/demo positioning.
- Updated development log with post-v0.3.0 polish verification.
Verification
npm run verifypassed.npm run format:checkpassed.npm pack --dry-runpassed.npm audit --audit-level=moderatefound0 vulnerabilities.- ProjScan preflight passed.
- AgentLoopKit verification passed.
[0.3.0] - 2026-06-13
AgentFlight now records session events and snapshots so reports and replays show how a coding session evolved.
Added
- Added
agentflight snapshot --note "..."to record a local checkpoint for the active session. - Added session-level
eventswith backward compatibility for older sessions. - Added event recording for session start, verification attempts, snapshots, report generation, replay generation, resume prompt generation, and doctor runs.
- Added timeline sections to Markdown reports and HTML replays.
- Added latest snapshot context to status and resume prompts.
- Added changed-file groups to replay output.
- Added tests for session events, snapshot creation, missing active sessions, timeline rendering, older session compatibility, and verification events.
Changed
- Replays now feel more like flight recorder timelines instead of final-state summaries.
- Reports now include a concise event timeline before changed files and proof evidence.
- Resume prompts now include latest snapshot notes and current verification state.
Verification
npm run verifypassed.npm run format:checkpassed.npm pack --dry-runpassed foragentflight@0.3.0.npm audit --audit-level=moderatefound0 vulnerabilities.- ProjScan preflight passed with health
100/100. - AgentLoopKit verification passed.
[0.2.1] - 2026-06-13
Patch release candidate focused on friction found while dogfooding the v0.2.0 core workflow.
Changed
agentflight verifynow prints the stdout and stderr evidence paths immediately after each recorded run.- Failed verification output and downstream next actions now include the exact command to rerun.
agentflight report,agentflight replay, andagentflight resumenow avoid stale "generate a report" next actions once proof is ready.- HTML replays now include a compact summary strip for risk, changed files, proof counts, and review readiness.
- AgentLoopKit workflow files under
.agentloop/are treated as low-risk dogfooding artifacts instead of unknown code changes.
Fixed
- ProjScan and AgentLoopKit adapters now prefer repo-local
node_modules/.binbinaries before PATH-global commands, preventing stale global versions from appearing in reports. - Tool adapter version output is normalized when CLIs include decorated version text.
Verification
- Targeted v0.2.1 regression tests passed.
npm run verifypassed.npm run format:checkpassed.npm pack --dry-runpassed foragentflight@0.2.1.npm audit --audit-level=moderatefound0 vulnerabilities.- ProjScan preflight passed with health
100/100. - AgentLoopKit verification passed.
[0.2.0] - 2026-06-13
AgentFlight now captures real verification evidence and uses it across status, report, replay, and resume.
Added
- Added
agentflight verifyto run verification commands and capture local evidence. - Added verification run persistence in session records, including command, timestamps, duration, exit code, status, stdout path, and stderr path.
- Added
.agentflight/evidence/for local stdout/stderr artifacts. - Added evidence-aware status, report, replay, and resume outputs.
- Added tests for verification success, failure, persistence, evidence-aware outputs, and v0.1 session compatibility.
Changed
agentflight statusnow reports changed areas, proof gaps, review readiness, and a next action based on captured evidence.agentflight reportnow includes verification evidence and honest review recommendations.agentflight replaynow renders verification cards in the local HTML artifact.agentflight resumenow includes verification gaps, the exact next command, and stronger continuation guardrails.
Fixed
agentflight doctornow checks write permission for.agentflight/, not just path existence.agentflight --versionnow reports the package version instead of the stale0.1.0value.
Verification
npm run verifypassed.npm run format:checkpassed.npm pack --dry-runpassed.npm audit --audit-level=moderatefound 0 vulnerabilities.- ProjScan preflight passed.
- AgentLoopKit verification passed.
[0.1.1] - 2026-06-13
Fixed
- Fixed npm
.binsymlink invocation sonpx agentflight --helpruns the CLI instead of exiting silently.
Added
- Added CI verification on pushes and pull requests.
- Added tag-based npm publishing through GitHub Actions Trusted Publishing.
- Added release process documentation.
[0.1.0] - 2026-06-13
Added
- First AgentFlight MVP CLI.
- Added
init,start,status,report,replay,resume, anddoctorcommands. - Added local
.agentflight/config, session, report, replay, and resume prompt artifacts. - Added ProjScan and AgentLoopKit adapters with graceful fallbacks.
- Added TypeScript, Vitest, ESLint, Prettier, and npm package setup.
- Added README, architecture docs, dogfooding docs, verification docs, roadmap, and monetisation notes.
More from the studio
ProjScan
ProjScan tells reviewers when to bootstrap, prove, or stop. Review Gate returns one decision; bootstrap is explicit. Local proof, no code upload.
ViewAgentLoopKit
The local control plane for low-token, verifiable agent loops. It owns scope, gates, and completion decisions, with a token receipt on every step.
ViewVerisKit
veris runs the test and quality tools you already have, then returns one honest verdict, verified, failed, or partial, plus a PR-ready report.
View