Skip to content

ci(release): migrate from bumpp to tagpr - #1406

Merged
ryoppippi merged 5 commits into
mainfrom
tagpr
Jul 9, 2026
Merged

ci(release): migrate from bumpp to tagpr#1406
ryoppippi merged 5 commits into
mainfrom
tagpr

Conversation

@ryoppippi

@ryoppippi ryoppippi commented Jul 8, 2026

Copy link
Copy Markdown
Member

Summary

Migrates release management from a local just release (bumpp) flow to tagpr release PRs, while keeping changelogithub-style GitHub Release notes.

What Changed

  • Add .tagpr config: bumps all nine workspace package.json files (versionFile), syncs the Rust workspace via just sync-rust-version (postVersionCommand), and disables tagpr's own CHANGELOG.md/GitHub Release generation (changelog = false, release = false).
  • Add .github/workflows/tagpr.yaml: runs Songmu/tagpr (SHA-pinned v1.20.0) on every push to main, and hosts the whole release pipeline (native builds, npm publish, changelogithub GitHub Release) in the same workflow, gated on the tagpr job's tag output.
  • Delete release.yaml: consolidating the pipeline into tagpr.yaml removes the workflow_dispatch chaining that worked around GITHUB_TOKEN-pushed tags not triggering push: tags workflows, along with the actions: write grant it needed.
  • Replace the just release recipe with just sync-rust-version (cargo set-version --workspace from apps/ccusage/package.json), and remove bumpp and bump.config.ts.
  • Update the development skill command reference to describe the tagpr flow.

Why

The bumpp flow ran on a maintainer's machine and depended on the local environment (clean checkout, cargo-edit, push rights). With tagpr, releasing becomes merging an auto-generated release PR: the version bump diff is reviewable, unreleased changes are always visible, and tags are only created in CI. Release notes stay in the existing changelogithub style; changelogithub resolves the tag with git tag --points-at HEAD, so it works from the consolidated workflow since the release jobs check out the freshly created tag.

Release Flow After Merge

  1. Pushes to main create/update the tagpr release PR (patch bump by default; label merged PRs minor/major to raise it).
  2. Merging the release PR tags the merge commit; the same workflow run then builds native packages, publishes to npm, and creates the changelogithub GitHub Release.
  3. A failed release is retried with "Re-run failed jobs" (a full re-run finds no new tag and skips the release jobs).

Required repo setting: enable Settings > Actions > General > "Allow GitHub Actions to create and approve pull requests" so tagpr can open the release PR with GITHUB_TOKEN.

Testing

  • Verified nix develop --command just sync-rust-version syncs all four Cargo.toml files and Cargo.lock to the package.json version (temporarily bumped to 20.0.15, then reverted).
  • actionlint passes on the consolidated workflow; just fmt reports no changes.

Summary by CodeRabbit

  • New Features

    • Releases are now generated automatically from merged changes, with version bumps following patch/minor/major labels.
    • Package and native builds now publish from a tagged release flow on the main branch.
  • Bug Fixes

    • Improved consistency between release versions across JavaScript and Rust components.
  • Documentation

    • Updated release instructions to match the new automated workflow.

ryoppippi added 3 commits July 9, 2026 00:41
Introduce Songmu/tagpr to manage releases via an auto-generated
release PR: every push to main creates or updates a PR that bumps all
nine workspace package.json versions (tagpr versionFile) and syncs the
Rust workspace via the new `just sync-rust-version` recipe run as
postVersionCommand. Merging the PR tags the merge commit.

GitHub Release creation and CHANGELOG.md generation are disabled in
.tagpr because changelogithub keeps generating the release notes in
the existing style.

Tags pushed with GITHUB_TOKEN do not trigger `on: push: tags`
workflows, so tagpr.yaml dispatches release.yaml explicitly with
`gh workflow run --ref <tag>`; release.yaml gains a workflow_dispatch
trigger for that purpose.
Releases are now driven by tagpr in CI, so the local `just release`
recipe and the bumpp dependency are no longer needed. bump.config.ts
is deleted because its cargo set-version hook moved to the
`just sync-rust-version` recipe that tagpr runs as postVersionCommand.
Replace the removed `just release` recipe in the development skill
command list with a note on the tagpr release PR flow and the
minor/major bump labels.
Copilot AI review requested due to automatic review settings July 8, 2026 23:43
@cloudflare-workers-and-pages

cloudflare-workers-and-pages Bot commented Jul 8, 2026

Copy link
Copy Markdown

Deploying with  Cloudflare Workers  Cloudflare Workers

The latest updates on your project. Learn more about integrating Git with Workers.

Status Name Latest Commit Preview URL Updated (UTC)
✅ Deployment successful!
View logs
ccusage-guide cff9af6 Commit Preview URL

Branch Preview URL
Jul 09 2026, 12:10 AM

Copilot AI left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Copilot was unable to review this pull request because the user who requested the review has reached their quota limit.

@coderabbitai

coderabbitai Bot commented Jul 8, 2026

Copy link
Copy Markdown

Review Change Stack

📝 Walkthrough

Walkthrough

This PR migrates release automation from a bumpp-based tag-push pipeline to a tagpr-driven workflow triggered on main pushes. It adds .tagpr configuration, gates downstream native/npm/release jobs on a computed tag, replaces the version-bump recipe with sync-rust-version, removes bumpp dependencies, and updates release docs and comments.

Changes

Tagpr release automation migration

Layer / File(s) Summary
New tagpr workflow and configuration
.github/workflows/tagpr.yaml, .tagpr
Workflow renamed to tagpr, triggered on main pushes with concurrency: tagpr; runs Songmu/tagpr and exposes a tag output. .tagpr sets release branch, version prefix, disables changelog/release, defines versionFile manifests, and a postVersionCommand.
Downstream job gating on computed tag
.github/workflows/tagpr.yaml, nix/static-package.nix
build-native-packages, npm, and release jobs run only when tagpr produces a non-empty tag, checking out that tag explicitly; release now depends on tagpr and npm. A Nix comment now references tagpr.yaml.
Rust version sync recipe
justfile
Adds sync-rust-version recipe that reads the version from apps/ccusage/package.json and runs cargo set-version, invoked via tagpr's postVersionCommand.
Removal of bumpp dependency
package.json, pnpm-workspace.yaml
Removes bumpp from catalogs/devDependencies; adds changelogithub and pkg-pr-new from catalog:release.
Release process documentation update
.agents/skills/development/references/commands.md
Replaces just release reference with tagpr release-PR merge and label-based bump description.

Estimated code review effort: 3 (Moderate) | ~20 minutes

Sequence Diagram(s)

sequenceDiagram
  participant MainPush as main branch push
  participant TagprWorkflow as tagpr.yaml
  participant TagprAction as Songmu/tagpr
  participant NativeBuild as build-native-packages job
  participant NpmJob as npm job
  participant ReleaseJob as release job

  MainPush->>TagprWorkflow: push to main
  TagprWorkflow->>TagprAction: run tagpr
  TagprAction-->>TagprWorkflow: computed tag output
  alt tag is non-empty
    TagprWorkflow->>NativeBuild: checkout tag, build
    NativeBuild->>NpmJob: publish native packages
    NpmJob->>ReleaseJob: needs tagpr and npm, checkout tag
    ReleaseJob->>ReleaseJob: create GitHub release
  else tag is empty
    TagprWorkflow-->>TagprWorkflow: skip downstream jobs
  end
Loading

Possibly related PRs

  • ccusage/ccusage#627: Updates the same pkg-pr-new/changelogithub release command tooling introduced here.
  • ccusage/ccusage#631: Touches the same bumpp/pnpm-workspace.yaml/release workflow area this PR removes.
🚥 Pre-merge checks | ✅ 5
✅ Passed checks (5 passed)
Check name Status Explanation
Docstring Coverage ✅ Passed No functions found in the changed files to evaluate docstring coverage. Skipping docstring coverage check.
Linked Issues check ✅ Passed Check skipped because no linked issues were found for this pull request.
Out of Scope Changes check ✅ Passed Check skipped because no linked issues were found for this pull request.
Description Check ✅ Passed Check skipped - CodeRabbit’s high-level summary is enabled.
Title check ✅ Passed The title is concise and accurately summarizes the main change: migrating release automation from bumpp to tagpr.
✨ Finishing Touches
📝 Generate docstrings
  • Create stacked PR
  • Commit on current branch
🧪 Generate unit tests (beta)
  • Create PR with unit tests
  • Commit unit tests in branch tagpr

Thanks for using CodeRabbit! It's free for OSS, and your support helps us grow. If you like it, consider giving us a shout-out.

❤️ Share

Comment @coderabbitai help to get the list of available commands.

@pkg-pr-new

pkg-pr-new Bot commented Jul 8, 2026

Copy link
Copy Markdown

Open in StackBlitz

ccusage

npx https://pkg.pr.new/ccusage@1406

@ccusage/ccusage-darwin-arm64

npx https://pkg.pr.new/@ccusage/ccusage-darwin-arm64@1406

@ccusage/ccusage-darwin-x64

npx https://pkg.pr.new/@ccusage/ccusage-darwin-x64@1406

@ccusage/ccusage-linux-arm64

npx https://pkg.pr.new/@ccusage/ccusage-linux-arm64@1406

@ccusage/ccusage-linux-x64

npx https://pkg.pr.new/@ccusage/ccusage-linux-x64@1406

@ccusage/ccusage-win32-x64

npx https://pkg.pr.new/@ccusage/ccusage-win32-x64@1406

commit: cff9af6

@github-actions

github-actions Bot commented Jul 8, 2026

Copy link
Copy Markdown
Contributor

ccusage performance comparison

PR SHA: 3c674d436f30
Base SHA: 2e30ce317527

This compares the Rust PR release binary against the configured base package on the same CI runner.

Package runtime diagnostics

Compares the PR package wrapper, the installed native optional dependency binary, and the workspace release binary on the same large fixture. This identifies whether slow package results come from JavaScript wrapper overhead, the published native binary build, or the Rust core itself.

Fixtures: Claude /home/runner/_work/_temp/ccusage-large-fixture (1.01 GiB, 2597 files), Codex /home/runner/_work/_temp/ccusage-large-codex-fixture (1.01 GiB, 2597 files)
All rows run --offline --json, measured by hyperfine with 0 warmups and 1 runs. This isolates wrapper overhead from the installed native optional dependency and the workspace release binary built on the runner.

Command Runtime Input Median Throughput Samples
claude --offline --json Package wrapper 1.01 GiB 253.3ms 3.97 GiB/s 1
claude --offline --json Installed native binary 1.01 GiB 216.3ms 4.65 GiB/s 1
codex --offline --json Package wrapper 1.01 GiB 104.6ms 9.62 GiB/s 1
codex --offline --json Installed native binary 1.01 GiB 80.6ms 12.49 GiB/s 1

Committed fixture performance

Committed small fixtures for stable PR-to-PR feedback and explicit Claude/Codex command coverage.

Fixtures: Claude apps/ccusage/test/fixtures/claude (0.00 MiB, 2 files), Codex apps/ccusage/test/fixtures/codex (0.00 MiB, 1 files)
Base runs the published ccusage package from pkg.pr.new, installed before measurement; PR runs the published native ccusage binary from pkg.pr.new, installed before measurement. Both run --offline --json, measured by hyperfine with 2 warmups and 7 runs.
Peak RSS is measured separately with /usr/bin/time using 1 runs. Lower RSS ratios are better.

Command Input Base median PR median PR vs base Base peak RSS PR peak RSS PR/base RSS Base throughput PR throughput
claude daily --offline --json 0.00 MiB 26.2ms 4.3ms 6.08x 53.50 MiB 12.21 MiB 0.23x 0.06 MiB/s 0.36 MiB/s
claude session --offline --json 0.00 MiB 27.3ms 3.1ms 8.91x 53.50 MiB 10.20 MiB 0.19x 0.06 MiB/s 0.50 MiB/s
codex daily --offline --json 0.00 MiB 22.5ms 2.1ms 10.61x 53.50 MiB 8.20 MiB 0.15x 0.04 MiB/s 0.40 MiB/s
codex session --offline --json 0.00 MiB 25.2ms 2.1ms 12.17x 53.75 MiB 8.20 MiB 0.15x 0.03 MiB/s 0.41 MiB/s

Large real-world-shaped fixture performance

Generated fixtures shaped from aggregate local log statistics: thousands of JSONL files, many small sessions, and a long tail of larger sessions. No real prompts, paths, or outputs are stored in the fixtures.

Fixtures: Claude /home/runner/_work/_temp/ccusage-large-fixture (1.01 GiB, 2597 files), Codex /home/runner/_work/_temp/ccusage-large-codex-fixture (1.01 GiB, 2597 files)
Base runs the published ccusage package from pkg.pr.new, installed before measurement; PR runs the published native ccusage binary from pkg.pr.new, installed before measurement. Both run --offline --json, measured by hyperfine with 0 warmups and 1 runs.
Peak RSS is measured separately with /usr/bin/time using 1 runs. Lower RSS ratios are better.

Command Input Base median PR median PR vs base Base peak RSS PR peak RSS PR/base RSS Base throughput PR throughput
claude --offline --json 1.01 GiB 279.5ms 241.2ms 1.16x 952.34 MiB 900.32 MiB 0.95x 3.60 GiB/s 4.17 GiB/s
codex --offline --json 1.01 GiB 105.6ms 82.6ms 1.28x 405.05 MiB 417.05 MiB 1.03x 9.54 GiB/s 12.18 GiB/s

Artifact size

Artifact Base PR Delta Ratio
packed ccusage-*.tgz 18.54 KiB 18.54 KiB +0.00 KiB 1.00x
installed native package binary 4054.25 KiB 4054.25 KiB +0.00 KiB 1.00x

Lower medians and smaller artifacts are better. CI runner noise still applies; use same-run ratios as directional PR feedback, not release guarantees.

@cubic-dev-ai cubic-dev-ai Bot left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

All reported issues were addressed across 9 files

Reply with feedback, questions, or to request a fix.

Re-trigger cubic

Comment thread .github/workflows/release.yaml Outdated
Comment thread .github/workflows/tagpr.yaml Outdated
@github-actions

github-actions Bot commented Jul 8, 2026

Copy link
Copy Markdown
Contributor

ccusage performance comparison

PR SHA: 3c674d436f30
Base SHA: 2e30ce317527

This compares the PR package against the configured base package on the same CI runner.

Package runtime diagnostics

Compares the PR package wrapper, the installed native optional dependency binary, and the workspace release binary on the same large fixture. This identifies whether slow package results come from JavaScript wrapper overhead, the published native binary build, or the Rust core itself.

Fixtures: Claude /home/runner/_work/_temp/ccusage-large-fixture (1.01 GiB, 2597 files), Codex /home/runner/_work/_temp/ccusage-large-codex-fixture (1.01 GiB, 2597 files)
All rows run --offline --json, measured by hyperfine with 0 warmups and 1 runs. This isolates wrapper overhead from the installed native optional dependency and the workspace release binary built on the runner.

Command Runtime Input Median Throughput Samples
claude --offline --json Package wrapper 1.01 GiB 244.3ms 4.12 GiB/s 1
claude --offline --json Installed native binary 1.01 GiB 205.6ms 4.90 GiB/s 1
codex --offline --json Package wrapper 1.01 GiB 109.6ms 9.18 GiB/s 1
codex --offline --json Installed native binary 1.01 GiB 86.4ms 11.66 GiB/s 1

Committed fixture performance

Committed small fixtures for stable PR-to-PR feedback and explicit Claude/Codex command coverage.

Fixtures: Claude apps/ccusage/test/fixtures/claude (0.00 MiB, 2 files), Codex apps/ccusage/test/fixtures/codex (0.00 MiB, 1 files)
Base runs the published ccusage package from pkg.pr.new, installed before measurement; PR runs the published ccusage package from pkg.pr.new, installed before measurement. Both run --offline --json, measured by hyperfine with 2 warmups and 7 runs.
Peak RSS is measured separately with /usr/bin/time using 1 runs. Lower RSS ratios are better.

Command Input Base median PR median PR vs base Base peak RSS PR peak RSS PR/base RSS Base throughput PR throughput
claude daily --offline --json 0.00 MiB 30.6ms 31.2ms 0.98x 53.75 MiB 53.75 MiB 1.00x 0.05 MiB/s 0.05 MiB/s
claude session --offline --json 0.00 MiB 29.0ms 28.9ms 1.00x 54.00 MiB 53.75 MiB 1.00x 0.05 MiB/s 0.05 MiB/s
codex daily --offline --json 0.00 MiB 27.6ms 28.4ms 0.97x 53.75 MiB 53.75 MiB 1.00x 0.03 MiB/s 0.03 MiB/s
codex session --offline --json 0.00 MiB 28.0ms 28.5ms 0.98x 54.00 MiB 53.75 MiB 1.00x 0.03 MiB/s 0.03 MiB/s

Large real-world-shaped fixture performance

Generated fixtures shaped from aggregate local log statistics: thousands of JSONL files, many small sessions, and a long tail of larger sessions. No real prompts, paths, or outputs are stored in the fixtures.

Fixtures: Claude /home/runner/_work/_temp/ccusage-large-fixture (1.01 GiB, 2597 files), Codex /home/runner/_work/_temp/ccusage-large-codex-fixture (1.01 GiB, 2597 files)
Base runs the published ccusage package from pkg.pr.new, installed before measurement; PR runs the published ccusage package from pkg.pr.new, installed before measurement. Both run --offline --json, measured by hyperfine with 0 warmups and 1 runs.
Peak RSS is measured separately with /usr/bin/time using 1 runs. Lower RSS ratios are better.

Command Input Base median PR median PR vs base Base peak RSS PR peak RSS PR/base RSS Base throughput PR throughput
claude --offline --json 1.01 GiB 263.0ms 249.3ms 1.05x 952.34 MiB 932.34 MiB 0.98x 3.83 GiB/s 4.04 GiB/s
codex --offline --json 1.01 GiB 113.5ms 115.8ms 0.98x 405.05 MiB 421.04 MiB 1.04x 8.87 GiB/s 8.69 GiB/s

Artifact size

Artifact Base PR Delta Ratio
packed ccusage-*.tgz 18.54 KiB 18.54 KiB +0.00 KiB 1.00x
installed native package binary 4054.25 KiB 4054.25 KiB +0.00 KiB 1.00x

Lower medians and smaller artifacts are better. CI runner noise still applies; use same-run ratios as directional PR feedback, not release guarantees.

Gate release.yaml build/publish/release jobs behind startsWith(github.ref, 'refs/tags/') so a workflow_dispatch from a branch cannot bypass the tag-only release flow. tagpr dispatches with --ref <tag>, so the intended path is unaffected.

Move the release dispatch out of the tagpr job into a dependent dispatch-release job that alone holds actions: write, keeping tagpr on its documented least-privilege scopes (contents/pull-requests/issues).

Co-authored-by: Codesmith <[email protected]>
@github-actions

github-actions Bot commented Jul 9, 2026

Copy link
Copy Markdown
Contributor

ccusage performance comparison

PR SHA: 6e68d4f2cc40
Base SHA: 2e30ce317527

This compares the PR package against the configured base package on the same CI runner.

Package runtime diagnostics

Compares the PR package wrapper, the installed native optional dependency binary, and the workspace release binary on the same large fixture. This identifies whether slow package results come from JavaScript wrapper overhead, the published native binary build, or the Rust core itself.

Fixtures: Claude /home/runner/_work/_temp/ccusage-large-fixture (1.01 GiB, 2597 files), Codex /home/runner/_work/_temp/ccusage-large-codex-fixture (1.01 GiB, 2597 files)
All rows run --offline --json, measured by hyperfine with 0 warmups and 1 runs. This isolates wrapper overhead from the installed native optional dependency and the workspace release binary built on the runner.

Command Runtime Input Median Throughput Samples
claude --offline --json Package wrapper 1.01 GiB 272.9ms 3.69 GiB/s 1
claude --offline --json Installed native binary 1.01 GiB 212.4ms 4.74 GiB/s 1
codex --offline --json Package wrapper 1.01 GiB 113.1ms 8.90 GiB/s 1
codex --offline --json Installed native binary 1.01 GiB 86.0ms 11.71 GiB/s 1

Committed fixture performance

Committed small fixtures for stable PR-to-PR feedback and explicit Claude/Codex command coverage.

Fixtures: Claude apps/ccusage/test/fixtures/claude (0.00 MiB, 2 files), Codex apps/ccusage/test/fixtures/codex (0.00 MiB, 1 files)
Base runs the published ccusage package from pkg.pr.new, installed before measurement; PR runs the published ccusage package from pkg.pr.new, installed before measurement. Both run --offline --json, measured by hyperfine with 2 warmups and 7 runs.
Peak RSS is measured separately with /usr/bin/time using 1 runs. Lower RSS ratios are better.

Command Input Base median PR median PR vs base Base peak RSS PR peak RSS PR/base RSS Base throughput PR throughput
claude daily --offline --json 0.00 MiB 31.2ms 31.6ms 0.99x 53.75 MiB 54.00 MiB 1.00x 0.05 MiB/s 0.05 MiB/s
claude session --offline --json 0.00 MiB 29.5ms 30.1ms 0.98x 54.00 MiB 53.75 MiB 1.00x 0.05 MiB/s 0.05 MiB/s
codex daily --offline --json 0.00 MiB 27.3ms 25.0ms 1.09x 53.75 MiB 53.75 MiB 1.00x 0.03 MiB/s 0.03 MiB/s
codex session --offline --json 0.00 MiB 24.6ms 25.7ms 0.96x 53.75 MiB 53.75 MiB 1.00x 0.03 MiB/s 0.03 MiB/s

Large real-world-shaped fixture performance

Generated fixtures shaped from aggregate local log statistics: thousands of JSONL files, many small sessions, and a long tail of larger sessions. No real prompts, paths, or outputs are stored in the fixtures.

Fixtures: Claude /home/runner/_work/_temp/ccusage-large-fixture (1.01 GiB, 2597 files), Codex /home/runner/_work/_temp/ccusage-large-codex-fixture (1.01 GiB, 2597 files)
Base runs the published ccusage package from pkg.pr.new, installed before measurement; PR runs the published ccusage package from pkg.pr.new, installed before measurement. Both run --offline --json, measured by hyperfine with 0 warmups and 1 runs.
Peak RSS is measured separately with /usr/bin/time using 1 runs. Lower RSS ratios are better.

Command Input Base median PR median PR vs base Base peak RSS PR peak RSS PR/base RSS Base throughput PR throughput
claude --offline --json 1.01 GiB 270.4ms 244.0ms 1.11x 946.34 MiB 952.33 MiB 1.01x 3.72 GiB/s 4.13 GiB/s
codex --offline --json 1.01 GiB 113.0ms 109.9ms 1.03x 417.05 MiB 401.04 MiB 0.96x 8.91 GiB/s 9.16 GiB/s

Artifact size

Artifact Base PR Delta Ratio
packed ccusage-*.tgz 18.54 KiB 18.54 KiB +0.00 KiB 1.00x
installed native package binary 4054.25 KiB 4059.25 KiB +5.00 KiB 1.00x

Lower medians and smaller artifacts are better. CI runner noise still applies; use same-run ratios as directional PR feedback, not release guarantees.

@github-actions

github-actions Bot commented Jul 9, 2026

Copy link
Copy Markdown
Contributor

ccusage performance comparison

PR SHA: 6e68d4f2cc40
Base SHA: 2e30ce317527

This compares the Rust PR release binary against the configured base package on the same CI runner.

Package runtime diagnostics

Compares the PR package wrapper, the installed native optional dependency binary, and the workspace release binary on the same large fixture. This identifies whether slow package results come from JavaScript wrapper overhead, the published native binary build, or the Rust core itself.

Fixtures: Claude /home/runner/_work/_temp/ccusage-large-fixture (1.01 GiB, 2597 files), Codex /home/runner/_work/_temp/ccusage-large-codex-fixture (1.01 GiB, 2597 files)
All rows run --offline --json, measured by hyperfine with 0 warmups and 1 runs. This isolates wrapper overhead from the installed native optional dependency and the workspace release binary built on the runner.

Command Runtime Input Median Throughput Samples
claude --offline --json Package wrapper 1.01 GiB 286.7ms 3.51 GiB/s 1
claude --offline --json Installed native binary 1.01 GiB 218.3ms 4.61 GiB/s 1
codex --offline --json Package wrapper 1.01 GiB 108.1ms 9.31 GiB/s 1
codex --offline --json Installed native binary 1.01 GiB 83.5ms 12.05 GiB/s 1

Committed fixture performance

Committed small fixtures for stable PR-to-PR feedback and explicit Claude/Codex command coverage.

Fixtures: Claude apps/ccusage/test/fixtures/claude (0.00 MiB, 2 files), Codex apps/ccusage/test/fixtures/codex (0.00 MiB, 1 files)
Base runs the published ccusage package from pkg.pr.new, installed before measurement; PR runs the published native ccusage binary from pkg.pr.new, installed before measurement. Both run --offline --json, measured by hyperfine with 2 warmups and 7 runs.
Peak RSS is measured separately with /usr/bin/time using 1 runs. Lower RSS ratios are better.

Command Input Base median PR median PR vs base Base peak RSS PR peak RSS PR/base RSS Base throughput PR throughput
claude daily --offline --json 0.00 MiB 30.3ms 5.4ms 5.57x 53.75 MiB 10.21 MiB 0.19x 0.05 MiB/s 0.28 MiB/s
claude session --offline --json 0.00 MiB 31.0ms 3.7ms 8.41x 53.75 MiB 10.21 MiB 0.19x 0.05 MiB/s 0.42 MiB/s
codex daily --offline --json 0.00 MiB 33.0ms 2.5ms 13.32x 53.75 MiB 8.19 MiB 0.15x 0.03 MiB/s 0.35 MiB/s
codex session --offline --json 0.00 MiB 29.2ms 2.8ms 10.49x 53.75 MiB 8.19 MiB 0.15x 0.03 MiB/s 0.31 MiB/s

Large real-world-shaped fixture performance

Generated fixtures shaped from aggregate local log statistics: thousands of JSONL files, many small sessions, and a long tail of larger sessions. No real prompts, paths, or outputs are stored in the fixtures.

Fixtures: Claude /home/runner/_work/_temp/ccusage-large-fixture (1.01 GiB, 2597 files), Codex /home/runner/_work/_temp/ccusage-large-codex-fixture (1.01 GiB, 2597 files)
Base runs the published ccusage package from pkg.pr.new, installed before measurement; PR runs the published native ccusage binary from pkg.pr.new, installed before measurement. Both run --offline --json, measured by hyperfine with 0 warmups and 1 runs.
Peak RSS is measured separately with /usr/bin/time using 1 runs. Lower RSS ratios are better.

Command Input Base median PR median PR vs base Base peak RSS PR peak RSS PR/base RSS Base throughput PR throughput
claude --offline --json 1.01 GiB 268.6ms 237.1ms 1.13x 940.34 MiB 924.32 MiB 0.98x 3.75 GiB/s 4.25 GiB/s
codex --offline --json 1.01 GiB 120.5ms 88.9ms 1.36x 427.05 MiB 411.04 MiB 0.96x 8.35 GiB/s 11.33 GiB/s

Artifact size

Artifact Base PR Delta Ratio
packed ccusage-*.tgz 18.54 KiB 18.54 KiB +0.00 KiB 1.00x
installed native package binary 4054.25 KiB 4059.25 KiB +5.00 KiB 1.00x

Lower medians and smaller artifacts are better. CI runner noise still applies; use same-run ratios as directional PR feedback, not release guarantees.

Move the build/publish/release jobs from release.yaml into tagpr.yaml,
gated on the tagpr job's tag output, and delete release.yaml. Running
everything in one workflow removes the workflow_dispatch chaining that
worked around GITHUB_TOKEN-pushed tags not triggering `push: tags`
workflows, along with the dispatch-release job and its `actions: write`
grant.

The release jobs check out the freshly created tag explicitly.
changelogithub resolves the release tag with `git tag --points-at
HEAD`, not GITHUB_REF, so it picks the right release even though the
run's ref is refs/heads/main.

A failed release is retried with "Re-run failed jobs"; a full re-run
finds no new tag and skips the release jobs.

@coderabbitai coderabbitai Bot left a comment

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Actionable comments posted: 1

🧹 Nitpick comments (1)
.github/workflows/tagpr.yaml (1)

37-101: 🔒 Security & Privacy | 🔵 Trivial | ⚡ Quick win

Add explicit permissions to build-native-packages.

This job lacks a permissions block while tagpr, npm, and release all declare theirs. Without a workflow-level default, the job inherits the repo-default token scope, which may grant broader access than needed. The job only checks out code and builds/uploads artifacts, so contents: read is sufficient.

🔒️ Add a minimal permissions block
     timeout-minutes: 30
+    permissions:
+      contents: read
     strategy:
🤖 Prompt for AI Agents
Verify each finding against current code. Fix only still-valid issues, skip the
rest with a brief reason, keep changes minimal, and validate.

In @.github/workflows/tagpr.yaml around lines 37 - 101, The
build-native-packages job is missing an explicit permissions block, so it may
inherit broader repo-default token access than needed. Add a minimal permissions
declaration on the build-native-packages job in tagpr.yaml, using contents: read
since the job only checks out the tagged ref and builds artifacts via the
build-linux-native-package, build-macos-nix-native-package, and
build-windows-native-package actions.
🤖 Prompt for all review comments with AI agents
Verify each finding against current code. Fix only still-valid issues, skip the
rest with a brief reason, keep changes minimal, and validate.

Inline comments:
In @.github/workflows/tagpr.yaml:
- Around line 140-159: The release job in tagpr.yaml is missing an explicit
timeout, unlike the other jobs in this workflow. Add a timeout-minutes value to
the release job near the job definition so the changelogithub/Nix step cannot
hang for hours; use the same style as the existing job-level settings and keep
the change scoped to the release job block.

---

Nitpick comments:
In @.github/workflows/tagpr.yaml:
- Around line 37-101: The build-native-packages job is missing an explicit
permissions block, so it may inherit broader repo-default token access than
needed. Add a minimal permissions declaration on the build-native-packages job
in tagpr.yaml, using contents: read since the job only checks out the tagged ref
and builds artifacts via the build-linux-native-package,
build-macos-nix-native-package, and build-windows-native-package actions.
🪄 Autofix (Beta)

Fix all unresolved CodeRabbit comments on this PR:

  • Push a commit to this branch (recommended)
  • Create a new PR with the fixes

ℹ️ Review info
⚙️ Run configuration

Configuration used: defaults

Review profile: CHILL

Plan: Pro Plus

Run ID: ce39a247-1779-4b65-8e01-4856909cdbdf

📥 Commits

Reviewing files that changed from the base of the PR and between 6e68d4f and cff9af6.

📒 Files selected for processing (2)
  • .github/workflows/tagpr.yaml
  • nix/static-package.nix
✅ Files skipped from review due to trivial changes (1)
  • nix/static-package.nix

@coderabbitai coderabbitai Bot left a comment

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Caution

Inline review comments failed to post. This is likely due to GitHub's internal server error or limits when posting large numbers of comments. If you are seeing this consistently it is likely a permissions issue. Please check "Moderation" -> "Code review limits" under your organization settings.

Actionable comments posted: 1

🧹 Nitpick comments (1)
.github/workflows/tagpr.yaml (1)

37-101: 🔒 Security & Privacy | 🔵 Trivial | ⚡ Quick win

Add explicit permissions to build-native-packages.

This job lacks a permissions block while tagpr, npm, and release all declare theirs. Without a workflow-level default, the job inherits the repo-default token scope, which may grant broader access than needed. The job only checks out code and builds/uploads artifacts, so contents: read is sufficient.

🔒️ Add a minimal permissions block
     timeout-minutes: 30
+    permissions:
+      contents: read
     strategy:
🤖 Prompt for AI Agents
Verify each finding against current code. Fix only still-valid issues, skip the
rest with a brief reason, keep changes minimal, and validate.

In @.github/workflows/tagpr.yaml around lines 37 - 101, The
build-native-packages job is missing an explicit permissions block, so it may
inherit broader repo-default token access than needed. Add a minimal permissions
declaration on the build-native-packages job in tagpr.yaml, using contents: read
since the job only checks out the tagged ref and builds artifacts via the
build-linux-native-package, build-macos-nix-native-package, and
build-windows-native-package actions.
🤖 Prompt for all review comments with AI agents
Verify each finding against current code. Fix only still-valid issues, skip the
rest with a brief reason, keep changes minimal, and validate.

Inline comments:
In @.github/workflows/tagpr.yaml:
- Around line 140-159: The release job in tagpr.yaml is missing an explicit
timeout, unlike the other jobs in this workflow. Add a timeout-minutes value to
the release job near the job definition so the changelogithub/Nix step cannot
hang for hours; use the same style as the existing job-level settings and keep
the change scoped to the release job block.

---

Nitpick comments:
In @.github/workflows/tagpr.yaml:
- Around line 37-101: The build-native-packages job is missing an explicit
permissions block, so it may inherit broader repo-default token access than
needed. Add a minimal permissions declaration on the build-native-packages job
in tagpr.yaml, using contents: read since the job only checks out the tagged ref
and builds artifacts via the build-linux-native-package,
build-macos-nix-native-package, and build-windows-native-package actions.
🪄 Autofix (Beta)

Fix all unresolved CodeRabbit comments on this PR:

  • Push a commit to this branch (recommended)
  • Create a new PR with the fixes

ℹ️ Review info
⚙️ Run configuration

Configuration used: defaults

Review profile: CHILL

Plan: Pro Plus

Run ID: ce39a247-1779-4b65-8e01-4856909cdbdf

📥 Commits

Reviewing files that changed from the base of the PR and between 6e68d4f and cff9af6.

📒 Files selected for processing (2)
  • .github/workflows/tagpr.yaml
  • nix/static-package.nix
✅ Files skipped from review due to trivial changes (1)
  • nix/static-package.nix
🛑 Comments failed to post (1)
.github/workflows/tagpr.yaml (1)

140-159: 🩺 Stability & Availability | 🟡 Minor | ⚡ Quick win

Add timeout-minutes to the release job.

Every other job in this workflow sets an explicit timeout (15, 30, 10 minutes). The release job runs only changelogithub via Nix and should finish in minutes, but without a timeout it defaults to 360 minutes—leaving a hung job consuming a runner for hours if nix develop or changelogithub stalls.

⏱️ Proposed fix
     if: needs.tagpr.outputs.tag != ''
     runs-on: blacksmith-32vcpu-ubuntu-2404-arm
+    timeout-minutes: 10
     permissions:
       contents: write
📝 Committable suggestion

‼️ IMPORTANT
Carefully review the code before committing. Ensure that it accurately replaces the highlighted code, contains no missing lines, and has no issues with indentation. Thoroughly test & benchmark the code to ensure it meets the requirements.

  release:
    needs:
      - tagpr
      - npm
    if: needs.tagpr.outputs.tag != ''
    runs-on: blacksmith-32vcpu-ubuntu-2404-arm
    timeout-minutes: 10
    permissions:
      contents: write
    steps:
      # changelogithub resolves the release tag with `git tag --points-at HEAD`,
      # so checking out the tag is what makes it pick the right release
      - uses: actions/checkout@df4cb1c069e1874edd31b4311f1884172cec0e10 # v6.0.3
        with:
          ref: ${{ needs.tagpr.outputs.tag }}
          fetch-depth: 0
          persist-credentials: false
      - uses: ./.github/actions/setup-nix-cache
      - run: nix develop --command pnpm changelogithub
        env:
          GITHUB_TOKEN: ${{secrets.GITHUB_TOKEN}}
🤖 Prompt for AI Agents
Verify each finding against current code. Fix only still-valid issues, skip the
rest with a brief reason, keep changes minimal, and validate.

In @.github/workflows/tagpr.yaml around lines 140 - 159, The release job in
tagpr.yaml is missing an explicit timeout, unlike the other jobs in this
workflow. Add a timeout-minutes value to the release job near the job definition
so the changelogithub/Nix step cannot hang for hours; use the same style as the
existing job-level settings and keep the change scoped to the release job block.

@cubic-dev-ai cubic-dev-ai Bot left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

1 issue found across 3 files (changes from recent commits).

Prompt for AI agents (unresolved issues)

Check if these issues are valid — if so, understand the root cause of each and fix them. If appropriate, use sub-agents to investigate and fix each issue separately.


<file name=".github/workflows/tagpr.yaml">

<violation number="1">
P1: The npm release can fail authentication after moving the publish step into `tagpr.yaml`: npm trusted publishing validates the exact workflow filename, and the previous OIDC publish path was `release.yaml`. Consider updating each npm trusted publisher to `tagpr.yaml` as part of this migration, or keep the publish step in the configured workflow/use token auth.</violation>
</file>

Tip: Review your code locally with the cubic CLI to iterate faster.

Re-trigger cubic

@@ -1,13 +1,43 @@
name: npm publish
name: tagpr

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

P1: The npm release can fail authentication after moving the publish step into tagpr.yaml: npm trusted publishing validates the exact workflow filename, and the previous OIDC publish path was release.yaml. Consider updating each npm trusted publisher to tagpr.yaml as part of this migration, or keep the publish step in the configured workflow/use token auth.

Prompt for AI agents
Check if this issue is valid — if so, understand the root cause and fix it. At .github/workflows/tagpr.yaml, line 138:

<comment>The npm release can fail authentication after moving the publish step into `tagpr.yaml`: npm trusted publishing validates the exact workflow filename, and the previous OIDC publish path was `release.yaml`. Consider updating each npm trusted publisher to `tagpr.yaml` as part of this migration, or keep the publish step in the configured workflow/use token auth.</comment>

<file context>
@@ -31,20 +29,131 @@ jobs:
+          for archive in "$RUNNER_TEMP"/native-packages/*.tar; do
+            tar -xf "$archive"
+          done
+      - run: pnpm --filter='./apps/ccusage' --filter='./packages/ccusage-*' publish --provenance --no-git-checks --access public
+
+  release:
</file context>

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Valid concern, but the fix is an external npmjs.com change: update each package's npm trusted publisher workflow filename from release.yaml to tagpr.yaml. Publishing uses pure OIDC trusted publishing with no token fallback, so this must be done before the first release, but it is a maintainer settings step outside this diff rather than a code change.

@github-actions

github-actions Bot commented Jul 9, 2026

Copy link
Copy Markdown
Contributor

ccusage performance comparison

PR SHA: cff9af6c3241
Base SHA: 2e30ce317527

This compares the Rust PR release binary against the configured base package on the same CI runner.

Package runtime diagnostics

Compares the PR package wrapper, the installed native optional dependency binary, and the workspace release binary on the same large fixture. This identifies whether slow package results come from JavaScript wrapper overhead, the published native binary build, or the Rust core itself.

Fixtures: Claude /home/runner/_work/_temp/ccusage-large-fixture (1.01 GiB, 2597 files), Codex /home/runner/_work/_temp/ccusage-large-codex-fixture (1.01 GiB, 2597 files)
All rows run --offline --json, measured by hyperfine with 0 warmups and 1 runs. This isolates wrapper overhead from the installed native optional dependency and the workspace release binary built on the runner.

Command Runtime Input Median Throughput Samples
claude --offline --json Package wrapper 1.01 GiB 323.7ms 3.11 GiB/s 1
claude --offline --json Installed native binary 1.01 GiB 291.8ms 3.45 GiB/s 1
codex --offline --json Package wrapper 1.01 GiB 105.1ms 9.58 GiB/s 1
codex --offline --json Installed native binary 1.01 GiB 79.1ms 12.72 GiB/s 1

Committed fixture performance

Committed small fixtures for stable PR-to-PR feedback and explicit Claude/Codex command coverage.

Fixtures: Claude apps/ccusage/test/fixtures/claude (0.00 MiB, 2 files), Codex apps/ccusage/test/fixtures/codex (0.00 MiB, 1 files)
Base runs the published ccusage package from pkg.pr.new, installed before measurement; PR runs the published native ccusage binary from pkg.pr.new, installed before measurement. Both run --offline --json, measured by hyperfine with 2 warmups and 7 runs.
Peak RSS is measured separately with /usr/bin/time using 1 runs. Lower RSS ratios are better.

Command Input Base median PR median PR vs base Base peak RSS PR peak RSS PR/base RSS Base throughput PR throughput
claude daily --offline --json 0.00 MiB 25.2ms 4.0ms 6.36x 54.00 MiB 10.20 MiB 0.19x 0.06 MiB/s 0.39 MiB/s
claude session --offline --json 0.00 MiB 24.5ms 2.3ms 10.49x 53.75 MiB 10.19 MiB 0.19x 0.06 MiB/s 0.66 MiB/s
codex daily --offline --json 0.00 MiB 23.7ms 2.1ms 11.35x 53.50 MiB 8.19 MiB 0.15x 0.04 MiB/s 0.41 MiB/s
codex session --offline --json 0.00 MiB 22.3ms 2.0ms 10.96x 53.75 MiB 8.18 MiB 0.15x 0.04 MiB/s 0.42 MiB/s

Large real-world-shaped fixture performance

Generated fixtures shaped from aggregate local log statistics: thousands of JSONL files, many small sessions, and a long tail of larger sessions. No real prompts, paths, or outputs are stored in the fixtures.

Fixtures: Claude /home/runner/_work/_temp/ccusage-large-fixture (1.01 GiB, 2597 files), Codex /home/runner/_work/_temp/ccusage-large-codex-fixture (1.01 GiB, 2597 files)
Base runs the published ccusage package from pkg.pr.new, installed before measurement; PR runs the published native ccusage binary from pkg.pr.new, installed before measurement. Both run --offline --json, measured by hyperfine with 0 warmups and 1 runs.
Peak RSS is measured separately with /usr/bin/time using 1 runs. Lower RSS ratios are better.

Command Input Base median PR median PR vs base Base peak RSS PR peak RSS PR/base RSS Base throughput PR throughput
claude --offline --json 1.01 GiB 271.1ms 294.9ms 0.92x 958.34 MiB 956.33 MiB 1.00x 3.71 GiB/s 3.41 GiB/s
codex --offline --json 1.01 GiB 104.6ms 82.8ms 1.26x 461.04 MiB 417.03 MiB 0.90x 9.62 GiB/s 12.17 GiB/s

Artifact size

Artifact Base PR Delta Ratio
packed ccusage-*.tgz 18.54 KiB 18.54 KiB +0.01 KiB 1.00x
installed native package binary 4054.25 KiB 4061.50 KiB +7.25 KiB 1.00x

Lower medians and smaller artifacts are better. CI runner noise still applies; use same-run ratios as directional PR feedback, not release guarantees.

@github-actions

github-actions Bot commented Jul 9, 2026

Copy link
Copy Markdown
Contributor

ccusage performance comparison

PR SHA: cff9af6c3241
Base SHA: 2e30ce317527

This compares the PR package against the configured base package on the same CI runner.

Package runtime diagnostics

Compares the PR package wrapper, the installed native optional dependency binary, and the workspace release binary on the same large fixture. This identifies whether slow package results come from JavaScript wrapper overhead, the published native binary build, or the Rust core itself.

Fixtures: Claude /home/runner/_work/_temp/ccusage-large-fixture (1.01 GiB, 2597 files), Codex /home/runner/_work/_temp/ccusage-large-codex-fixture (1.01 GiB, 2597 files)
All rows run --offline --json, measured by hyperfine with 0 warmups and 1 runs. This isolates wrapper overhead from the installed native optional dependency and the workspace release binary built on the runner.

Command Runtime Input Median Throughput Samples
claude --offline --json Package wrapper 1.01 GiB 331.6ms 3.04 GiB/s 1
claude --offline --json Installed native binary 1.01 GiB 309.3ms 3.26 GiB/s 1
codex --offline --json Package wrapper 1.01 GiB 106.4ms 9.46 GiB/s 1
codex --offline --json Installed native binary 1.01 GiB 82.4ms 12.22 GiB/s 1

Committed fixture performance

Committed small fixtures for stable PR-to-PR feedback and explicit Claude/Codex command coverage.

Fixtures: Claude apps/ccusage/test/fixtures/claude (0.00 MiB, 2 files), Codex apps/ccusage/test/fixtures/codex (0.00 MiB, 1 files)
Base runs the published ccusage package from pkg.pr.new, installed before measurement; PR runs the published ccusage package from pkg.pr.new, installed before measurement. Both run --offline --json, measured by hyperfine with 2 warmups and 7 runs.
Peak RSS is measured separately with /usr/bin/time using 1 runs. Lower RSS ratios are better.

Command Input Base median PR median PR vs base Base peak RSS PR peak RSS PR/base RSS Base throughput PR throughput
claude daily --offline --json 0.00 MiB 26.1ms 24.8ms 1.05x 53.75 MiB 53.75 MiB 1.00x 0.06 MiB/s 0.06 MiB/s
claude session --offline --json 0.00 MiB 28.2ms 24.6ms 1.15x 53.75 MiB 53.50 MiB 1.00x 0.05 MiB/s 0.06 MiB/s
codex daily --offline --json 0.00 MiB 25.3ms 23.9ms 1.06x 53.50 MiB 53.75 MiB 1.00x 0.03 MiB/s 0.04 MiB/s
codex session --offline --json 0.00 MiB 23.1ms 22.9ms 1.01x 54.00 MiB 53.75 MiB 1.00x 0.04 MiB/s 0.04 MiB/s

Large real-world-shaped fixture performance

Generated fixtures shaped from aggregate local log statistics: thousands of JSONL files, many small sessions, and a long tail of larger sessions. No real prompts, paths, or outputs are stored in the fixtures.

Fixtures: Claude /home/runner/_work/_temp/ccusage-large-fixture (1.01 GiB, 2597 files), Codex /home/runner/_work/_temp/ccusage-large-codex-fixture (1.01 GiB, 2597 files)
Base runs the published ccusage package from pkg.pr.new, installed before measurement; PR runs the published ccusage package from pkg.pr.new, installed before measurement. Both run --offline --json, measured by hyperfine with 0 warmups and 1 runs.
Peak RSS is measured separately with /usr/bin/time using 1 runs. Lower RSS ratios are better.

Command Input Base median PR median PR vs base Base peak RSS PR peak RSS PR/base RSS Base throughput PR throughput
claude --offline --json 1.01 GiB 274.4ms 317.0ms 0.87x 952.34 MiB 942.33 MiB 0.99x 3.67 GiB/s 3.18 GiB/s
codex --offline --json 1.01 GiB 108.3ms 107.8ms 1.01x 431.04 MiB 427.04 MiB 0.99x 9.29 GiB/s 9.34 GiB/s

Artifact size

Artifact Base PR Delta Ratio
packed ccusage-*.tgz 18.54 KiB 18.54 KiB +0.01 KiB 1.00x
installed native package binary 4054.25 KiB 4061.50 KiB +7.25 KiB 1.00x

Lower medians and smaller artifacts are better. CI runner noise still applies; use same-run ratios as directional PR feedback, not release guarantees.

@ryoppippi
ryoppippi merged commit 701e1c1 into main Jul 9, 2026
36 checks passed
@ryoppippi
ryoppippi deleted the tagpr branch July 9, 2026 00:25
@github-actions github-actions Bot mentioned this pull request Jul 9, 2026
axisrow added a commit to axisrow/ccusage that referenced this pull request Aug 12, 2026
* chore(ci): remove pullfrog because they dont serve free tokens anymore

* Restore `pullfrog.yml` workflow

* feat(statusline): show reasoning effort level next to model name (ccusage#1405)

* feat(statusline): show reasoning effort level next to model name

Claude Code 2.1.119+ includes an optional top-level effort.level field
(low, medium, high, xhigh, or max) in the statusline hook JSON,
reflecting the live /effort setting. Parse it from the hook input and
append it to the model segment, e.g. '🤖 Fable 5 (high)'.

The field is absent for models without the effort parameter and on
older Claude Code versions, in which case the statusline keeps showing
just the model label as before.

* test(statusline): add Fable 5 fixture with effort level

Adds a manual statusline fixture for the latest model shape, including
the effort.level field, plus a test-statusline-fable5 recipe wired into
test-statusline-all so the effort display can be smoke-tested from the
CLI.

* docs(statusline): document effort level next to the model name

Updates the statusline guide examples to the current model display
('Fable 5 (high)') and explains that the reasoning effort level comes
from Claude Code 2.1.119+, with a fallback example for models or
versions that do not report it.

* Revert "Restore `pullfrog.yml` workflow"

This reverts commit 04f45b0.

* fix(codex): skip forked session replay history (ccusage#1369)

* feat(json): emit modelBreakdowns in per-agent JSON reports (ccusage#1395)

Per-agent subcommands (ccusage pi|opencode|amp|hermes|... daily/weekly/
monthly/session --json) compute per-model cost breakdowns — the table
view renders them with --breakdown — but the shared per-agent JSON
serializer never emitted them, forcing JSON consumers to re-derive
model costs they cannot actually reconstruct.

Add "modelBreakdowns" to agent_summary_json, mirroring the unified
serializers (summary_json / session_summary_json). Purely additive:
every pre-existing key and value is unchanged; the codex-native
serializer (models object) is deliberately untouched.

Tests: shared-shape insta snapshot now shows populated breakdowns for
all four report kinds; pi daily JSON asserts a full single-element
breakdown array with non-zero cost (the motivating case); fixture-
driven copilot (real pricing via read_otel_file) and qwen (real JSONL
fixture line) assertions pin their entire breakdown arrays.

* fix(pi): align unified session date filtering (ccusage#1394)

* fix(pi): align unified session date filtering

Filter default pi unified session entries by date before summarizing, matching
`ccusage pi session --pi-path` behavior for inclusive `--until` days.

* style: apply treefmt formatting

---------

Co-authored-by: ryoppippi <[email protected]>

* feat(unified): --sections and --by-agent for single-invocation reporting (ccusage#1396)

* feat(unified): --sections and --by-agent for single-invocation reporting

Dashboards polling ccusage today need one unified invocation per section
plus one per-agent invocation per agent — every call re-scanning all
stores. Two additive flags on the unified commands (and the bare root
invocation) collapse that to a single call:

  ccusage daily --json --sections daily,monthly,session --by-agent

--sections <csv> emits each requested grouping in one envelope from at
most TWO store scans (daily/weekly/monthly share one Daily-kind base
load; session adds one Session-kind load), one process, one pricing
load. Every section is produced by exactly the code path its standalone
command uses — load_sections delegates to the same load_rows machinery,
so section output is identical to a standalone invocation by
construction (covered by fixture equivalence tests including claude
agent-progress usage lines and codex cross-session/model-alias dedupe
cases). Envelope order is deterministic via a local ordered serializer:
invoked section first, remaining sections in canonical order, totals
last; single-section envelopes use the unchanged existing path.

--by-agent adds an "agents" array to daily/weekly/monthly rows (the
internal per-agent breakdowns, now serialized: tokens, cost, and
modelBreakdowns per agent). Session rows are already per-agent, so the
flag is a no-op there. Per-agent costs sum exactly to the combined row.

Backward compatibility: without the new flags, JSON and table output
are byte-identical to before (verified against a 20-invocation golden
matrix on real stores). Tables render requested sections sequentially;
--by-agent is JSON-only.

* refactor(unified): address review feedback on duplication and detected agents

- row_json now composes agent_json and layers on the row-level fields
  (period, metadata, agents), so the shared row shape has a single
  serialization path; output unchanged.
- Extract parse_unified_report_arg so the root, unified-command, and
  top-level-session parse sites share one --all/--sections/--by-agent
  block; the root site keeps its mark_used bookkeeping.
- Carry daily-load and session-load detected agents separately so each
  --sections table header shows the same detected list as the
  equivalent standalone invocation.

* style: apply treefmt formatting

---------

Co-authored-by: ryoppippi <[email protected]>

* feat(pi): named pi-format stores as config-declared agents (ccusage#1397)

* feat(pi): named pi-format stores as config-declared agents

Tools built on pi (oh-my-pi and other forks) keep pi-format session
stores at their own paths. ccusage could only read one pi path universe
and labeled everything it found as agent "pi". Declare named extra
stores in the config file:

  { "pi": { "stores": [ { "name": "omp", "path": "~/.omp/agent/sessions" } ] } }

Each named store loads through the existing pi parser and surfaces as
its OWN agent in the unified reports: rows tagged in metadata.agents,
sessions with projectPath/lastActivity like pi, model labels prefixed
"[<name>] ". Named stores are additive to the default pi store and use
the same path-list semantics (comma-separated, ~-expansion, dedupe) and
the same date-window filtering as `ccusage pi ... --pi-path`.

Costs are computed from the unprefixed model name — the configurable
store name never participates in pricing lookup (a store named "o3"
cannot fabricate o3 pricing; regression-tested), while prefixed
pricingOverrides keys are consulted first and keep working.

Config validation: names match ^[a-z][a-z0-9_-]{0,31}$, reject
collisions with built-in agents (single source of truth asserted
against the unified loader's registry), duplicates, empty paths, and
stores whose resolved paths overlap the default pi store or another
store (silent double-counting is never possible). Invalid stores error
through the same config-error path as other invalid config content.
Absent store paths yield clean empty results, like default pi.

Backward compatibility: without pi.stores configured, all output is
byte-identical to before (verified against a golden matrix on real
stores, including a known pre-existing until-day session-window quirk
in the default pi unified path, deliberately preserved here and fixed
in a separate patch). Committed config schema regenerated.

No CLI surface changes: per-agent subcommands remain a closed set;
named stores appear in unified reports only.

* fix(pi): reject nested/partial named-store path overlaps, dedupe path parsing

- Session files are collected recursively, so a named store rooted at an
  ancestor or descendant of the default pi store (or another named
  store) would ingest the same files twice under different dedupe
  identities. The resolver now rejects any overlap — equal, ancestor,
  or descendant — and partial collisions error instead of silently
  dropping the colliding path, matching the documented contract.
  Regression tests for a store nested inside the default pi path and a
  partial overlap across two stores.
- Extract a shared existing_paths helper in pi/paths.rs; the default
  and named-store variants now differ only in their path mapper, with
  the deliberate ~-expansion difference documented.
- Update config/pi docs for the stricter overlap wording.

* docs(pi): move trailing space out of code spans (markdownlint MD038)

---------

Co-authored-by: ryoppippi <[email protected]>

* fix(kimi): support Kimi Code new wire format (ccusage#1362)

Kimi Code (`~/.kimi-code`) emits a new `wire.jsonl` schema that the old
adapter could not parse, so its usage was silently dropped (ccusage#1261).

- Detect `~/.kimi-code` and the deeper layout
  `sessions/<ws>/<session>/agents/<agent>/wire.jsonl` (5 path components)
  alongside the legacy 3-component layout.
- Parse top-level `type == "usage.record"` lines: camelCase token fields
  (`inputOther`, `inputCacheRead`, `inputCacheCreation`), `time` in
  milliseconds, and `model` prefixed with `kimi-code/` (stripped for
  pricing lookup). Skip cumulative `usageScope == "session"` records.
- Deserialize `time` leniently so a float- or string-encoded timestamp
  degrades to the file-mtime fallback instead of dropping the whole line.
- Walk the correct number of parents in `kimi_root_from_wire_path` for the
  deeper layout so config resolution looks at the right root.
- Keep full backward compatibility with the old StatusUpdate format.
- Update the Kimi guide and data-source docs for `~/.kimi-code`.

Fixes ccusage#1261

Co-authored-by: Claude Opus 4.8 <[email protected]>

* perf: cache PricingMap::find() results and skip redundant opencode pricing checks (ccusage#1407)

* perf(pricing): cache PricingMap::find() results to avoid repeated fuzzy matching

PricingMap::find() does an exact HashMap lookup followed by expensive
fuzzy matching through all ~2,200 pricing entries when the exact model
name is not in the map. When adapters repeatedly query the same model
names, a large fraction of lookups miss the HashMap and trigger a full
scan of the pricing table for every call.

Add a OnceLock&lt;Mutex&lt;FxHashMap&gt;&gt; cache that memoizes find()
results by model name (including None for models not found in pricing).
Once a model name has been resolved, future lookups complete in O(1)
instead of O(n) over the pricing table.

Also add clear_find_cache() called from load_json_with_overrides(),
load_models_dev_models(), and apply_overrides() so the cache stays
consistent when the pricing table is mutated.

* perf(opencode): skip redundant missing-pricing check when cost is known

calculate_open_code_cost and missing_open_code_pricing independently
iterate through the same model candidates. When the cost calculation
already found a valid positive cost (either from a stored cost_usd
field or from pricing lookup), skip the missing-pricing check entirely
since pricing was already resolved.

---------

Co-authored-by: turtton <[email protected]>

* ci(release): migrate from bumpp to tagpr (ccusage#1406)

* ci(release): add tagpr release PR automation

Introduce Songmu/tagpr to manage releases via an auto-generated
release PR: every push to main creates or updates a PR that bumps all
nine workspace package.json versions (tagpr versionFile) and syncs the
Rust workspace via the new `just sync-rust-version` recipe run as
postVersionCommand. Merging the PR tags the merge commit.

GitHub Release creation and CHANGELOG.md generation are disabled in
.tagpr because changelogithub keeps generating the release notes in
the existing style.

Tags pushed with GITHUB_TOKEN do not trigger `on: push: tags`
workflows, so tagpr.yaml dispatches release.yaml explicitly with
`gh workflow run --ref <tag>`; release.yaml gains a workflow_dispatch
trigger for that purpose.

* chore(release): drop bumpp local release flow

Releases are now driven by tagpr in CI, so the local `just release`
recipe and the bumpp dependency are no longer needed. bump.config.ts
is deleted because its cargo set-version hook moved to the
`just sync-rust-version` recipe that tagpr runs as postVersionCommand.

* docs(skills): document tagpr release flow

Replace the removed `just release` recipe in the development skill
command list with a note on the tagpr release PR flow and the
minor/major bump labels.

* ci(release): gate release jobs to tag refs and isolate actions:write

Gate release.yaml build/publish/release jobs behind startsWith(github.ref, 'refs/tags/') so a workflow_dispatch from a branch cannot bypass the tag-only release flow. tagpr dispatches with --ref <tag>, so the intended path is unaffected.

Move the release dispatch out of the tagpr job into a dependent dispatch-release job that alone holds actions: write, keeping tagpr on its documented least-privilege scopes (contents/pull-requests/issues).

Co-authored-by: Codesmith <[email protected]>

* ci(release): consolidate release pipeline into tagpr workflow

Move the build/publish/release jobs from release.yaml into tagpr.yaml,
gated on the tagpr job's tag output, and delete release.yaml. Running
everything in one workflow removes the workflow_dispatch chaining that
worked around GITHUB_TOKEN-pushed tags not triggering `push: tags`
workflows, along with the dispatch-release job and its `actions: write`
grant.

The release jobs check out the freshly created tag explicitly.
changelogithub resolves the release tag with `git tag --points-at
HEAD`, not GITHUB_REF, so it picks the right release even though the
run's ref is refs/heads/main.

A failed release is retried with "Re-run failed jobs"; a full re-run
finds no new tag and skips the release jobs.

---------

Co-authored-by: Codesmith <[email protected]>

* ci(release): use Conventional Commits title for tagpr release PRs (ccusage#1409)

tagpr titles its release PRs "Release for vX.Y.Z", which fails the
check-pr-title workflow because it has no Conventional Commits type
prefix (seen on PR ccusage#1408). tagpr takes the first line of the rendered
pull request template as the PR title, so point .tagpr at a custom
template whose first line is "chore: release {{.NextVersion}}". The
rest of the template mirrors tagpr's default body, minus the unused
tag-prefix placeholder.

The template is a Go text/template, and oxfmt's markdown rewrites
break its <details> block and nested list structure, so exclude it
from treefmt.

This also restores the title style used by the previous bumpp-based
release flow ("chore: release v20.0.14").

* feat(pricing): support OpenAI two-stage pricing and add the gpt-5.6 family (ccusage#1414)

* feat(pricing): add gpt-5.6 family and OpenAI long-context tier rates

OpenAI introduced two-stage (short/long context) pricing with gpt-5.6:
requests with more than 272K input tokens are billed at higher
long-context rates. The same tier also applies to gpt-5.5, gpt-5.5-pro,
gpt-5.4, and gpt-5.4-pro on the current pricing page.

The existing tier support hardcoded the LiteLLM 200K boundary, so
Pricing gains a per-model long_context_threshold (defaulting to 200K
for LiteLLM *_above_200k_tokens data) and tiered_cost takes the
threshold as a parameter.

New built-in entries cover gpt-5.6-sol, gpt-5.6-terra, and
gpt-5.6-luna, including their cache-write rates. Long-context tier
rates live in a builtin_long_context_rates overlay that is re-applied
after every pricing load: a live LiteLLM refresh replaces whole
entries, and LiteLLM currently publishes these models with flat rates
only, so tier rates set directly on built-in entries would be silently
dropped whenever a refresh succeeds. Entries that already carry tier
rates are left untouched so upstream data wins once it exists.
Date-pinned keys such as gpt-5.5-2026-04-23 share their base model's
overlay rates.

The gpt-5.6 context limits mirror the 1,050,000-token window of the
other long-context GPT-5 flagship models until upstream data lands.

* feat(codex): bill long-context requests at OpenAI two-stage rates

OpenAI decides the pricing tier per request: once a request's input
exceeds 272K tokens, every token of that request (input, cached input,
and output) is billed at the long-context rates. Codex cost calculation
runs on per-model sums aggregated across many requests, so the tier
cannot be recovered from the totals afterwards.

CodexModelUsage now tracks the portion of tokens that came from
long-context requests. The split is recorded while token_count events
are aggregated, where each event still represents a single request,
and merged across parallel shards like the other counters.

calculate_codex_model_cost prices the aggregated usage as two
independent buckets: the short bucket at the flat rates and the long
bucket at the *_above_200k rates, falling back to the flat rates for
models without a long-context tier so their costs are unchanged. The
existing fast-speed multiplier applies to both buckets.

Report JSON and table output are unchanged; only costUSD values for
long-context requests differ.

* docs(pricing): explain all-or-nothing long-context overlay check

Codex review suggested filling missing tier fields independently when a
refreshed LiteLLM entry carries partial *_above_200k_tokens data. That
would mix rates that assume the 200K LiteLLM boundary with built-in
rates that assume the OpenAI 272K boundary under a single per-model
threshold, mispricing both tiers, so the overlay defers to upstream
entirely once any tier rate exists. Record that rationale next to the
check.

* fix(pricing): apply two-stage rates to whole request and per-model split

Co-authored-by: Codesmith <[email protected]>

---------

Co-authored-by: Codesmith <[email protected]>

* chore: use black smith more

* chore: release v20.0.15 (ccusage#1408)

[tagpr] prepare for the next release

Co-authored-by: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>

* chore(ci): rename it back to releaese.yaml

* chore: release v20.0.16 (ccusage#1416)

[tagpr] prepare for the next release

Co-authored-by: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>

* docs: update Lineman affiliate links to CCUsage landing page (ccusage#1417)

Point GitHub README and docs site sponsor links at the dedicated
LinkJolt redirect for CCUsage traffic (free tier + voucher funnel).

Co-authored-by: Cursor Agent <[email protected]>
Co-authored-by: ryoppippi <[email protected]>

* docs: update Star History chart (ccusage#1419)

* docs: update Star History chart

Switch the README and sponsorship guide to the current Star History chart endpoint, including light and dark variants. Allowlist the public read-only sealed chart token so secret scanning does not report a false positive.

Co-authored-by: ryoppippi <[email protected]>

* chore: exclude sealed token from spellcheck

Mark the exact public Star History token allowlist line as a spellcheck exclusion so its random character sequence does not fail the documentation preflight.

Co-authored-by: ryoppippi <[email protected]>

* chore: format sealed token allowlist

Use the repository's TOML formatting and bracket the random token with the supported spellchecker block directives.

Co-authored-by: ryoppippi <[email protected]>

* style: align Gitleaks TOML indentation

Match the repository formatter's tab indentation for the multiline allowlist entry.

Co-authored-by: ryoppippi <[email protected]>

* fix: match full Star History token URL

Configure the global Gitleaks allowlist to evaluate the full finding match so the narrowly scoped sealed_token pattern suppresses the six intentional chart URLs.

Co-authored-by: ryoppippi <[email protected]>

---------

Co-authored-by: Cursor Agent <[email protected]>
Co-authored-by: ryoppippi <[email protected]>

* fix(claude): count advisor model usage (ccusage#1423)

* fix(claude): count advisor model usage

Expand advisor_message iterations into distinct usage entries so their tokens and model-specific costs are included in every report path. Keep main-model iteration totals unchanged and cover both standard and daily loaders.

Co-authored-by: ryoppippi <[email protected]>

* docs(claude): clarify advisor cost modes

Co-authored-by: ryoppippi <[email protected]>

---------

Co-authored-by: Cursor Agent <[email protected]>
Co-authored-by: ryoppippi <[email protected]>

* chore: release v20.0.17 (ccusage#1418)

[tagpr] prepare for the next release

Co-authored-by: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>

* perf(nix): keep dependency cache across releases (ccusage#1424)

* build(perf): migrate benchmark harness to Babashka (ccusage#1432)

* build(perf): migrate benchmark harness to Babashka

Replace the large Nushell PR benchmark script with a Babashka implementation split by data, system, benchmark, report, and orchestration responsibilities. The new process boundary keeps argv, environment, and working-directory data explicit while preserving hyperfine, package installation, memory, size, and Markdown behavior.

Move the CI caller and profiling guidance to the executable Babashka entry point. Add focused tests behind their own Nix shebang so contributors can run the harness suite without adding Babashka to the full development shell.

* docs(agents): document implementation language choices

Route small command-oriented automation to Nushell and data-heavy, testable automation to Babashka. Keep production binaries in Rust and npm-integrated APIs in TypeScript so future tooling changes follow the same criteria used by the benchmark migration.

* test(ci): run Babashka harness tests

Execute the self-contained benchmark harness test entry point in the CI test job so changes to CLI parsing, normalization, fallback decisions, and report rendering cannot bypass pull request validation.

* fix(perf): harden platform and tarball paths

Normalize version-qualified Windows os.name values to win32 so native executable and package paths use the expected suffixes. Resolve relative pnpm pack filenames against the temporary destination while preserving the absolute paths emitted by current pnpm versions.

Add regression coverage for both platform normalization and relative or absolute tarball filenames.

* fix(perf): size local package fallbacks

Use remote tarball sizing only after the corresponding preview package was installed successfully. When either package URL times out, benchmark and size the available local checkout so fallback runs can still produce a complete report.

Cover base and head source selection and verify both unavailable URLs through a committed-fixture smoke run.

* fix(perf): bound harness child processes and skip RSS on unsupported platforms

Add a cancellable timeout to run-process and thread --package-runner-timeout-ms through the package URL probe, install, pnpm pack, and git rev-parse flows so a stalled child cannot outlive the deadline; give the curl probe and download explicit connect and read limits.

measure-memory now warns once and skips gracefully when /usr/bin/time is unavailable (unsupported platforms) instead of throwing and aborting the entire benchmark run.

Co-authored-by: Codesmith <[email protected]>

* Revert "fix(perf): bound harness child processes and skip RSS on unsupported platforms"

This reverts commit 9140a99.

---------

Co-authored-by: Codesmith <[email protected]>

* build(perf): migrate fixture generator to Bun (ccusage#1433)

* build(perf): migrate fixture generator to Bun

Replace the Nushell fixture generator with a dependency-free Bun script.\n\nKeep the generated Claude and Codex fixture layouts and command-line\ninterface while using Bun file writers and Bun Shell for file operations.

* build(perf): type Bun fixture script

Add Bun development types so the fixture generator is checked alongside the package tooling.\n\nAwait file writer operations to preserve ordered writes and satisfy the\nrepository promise lint rule.

* build(perf): avoid Bun type dependency

Keep the fixture generator dependency-free by declaring its small Bun API surface locally.\n\nRemove the Bun type package and restore the package TypeScript configuration so\npublishing the fixture generator does not expand package dependencies.

* chroe(ci): fix nix cache

* ci: add GitHub-hosted runner fallback

Keep Blacksmith runners for the upstream repository while allowing forks\nto use hosted runners by default. Forks with a Blacksmith subscription can\nopt in through HAS_BLACKSMITH=true.

* ci: skip pkg-pr previews without the GitHub App

Forks do not inherit the pkg-pr-new GitHub App installation. Skip\npreview publishing and its dependent E2E and performance jobs unless a fork\nexplicitly opts in with HAS_PKG_PR_NEW=true.

* ci: keep Windows arm release runner defined

Supply both matrix runner fields so the release workflow resolves its\nWindows ARM runner consistently with the other native package targets.

---------

Co-authored-by: ryoppippi <[email protected]>
Co-authored-by: pullfrog[bot] <226033991+pullfrog[bot]@users.noreply.github.com>
Co-authored-by: sijie-ni-0214 <[email protected]>
Co-authored-by: Ben Vargas <[email protected]>
Co-authored-by: Mint Choco <[email protected]>
Co-authored-by: Claude Opus 4.8 <[email protected]>
Co-authored-by: turtton <[email protected]>
Co-authored-by: Codesmith <[email protected]>
Co-authored-by: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
Co-authored-by: Cursor Agent <[email protected]>
Co-authored-by: ryoppippi <[email protected]>
Co-authored-by: axisrow <[email protected]>
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants