Winglet

Make your Claude Code / Codex plan go 50% further.

Results may vary by workload and workflow.

An AI usage saver that actually works. Backed by research.

Winglet seamlessly integrates with your coding agent. It cuts stale reads, repeated tool output, and wasted context before they reach the model. Everything stays on your machine.

3-day free trial · no credit card required

Product overview

Winglet savings dashboard

Winglet

agent-winglet

33% saved(14.7 MiB)Your plan goes ~49% further
33%

33% saved

Your plan goes ~49% further

BYTES SAVED

14.7 MiB

Directly measured

TOKENS SAVED

13.2M(est.)

Scaled from bytes saved

MONEY SAVED

$54.55(est.)

Uses API pricing

WHERE THE SAVINGS COME FROM

Old investigation output archived

30%
637 calls13.3 MiB

Long output trimmed

3%
245 trims1.4 MiB

Repeat output skipped

0%
24 hits5.3 KiB

SESSIONS (47)

1e59a038-618a-4b70-af6c-6a2ea4a545db
Claude34% savedYour plan goes ~52% further
8b41c19a-2bd1-4f7d-9b0a-1816f2b91241
Codex28% savedYour plan goes ~39% further
f2e7d934-2f40-4d5f-9dd4-0d1a6c2ad45a
Claude31% savedYour plan goes ~45% further
5a65dd1e-b9ce-4437-b197-5a087d97cd81
Codex19% savedYour plan goes ~23% further

Measured on August 8, 2026 while building Winglet with Winglet running.

Performance comparison

How it works

Winglet

Trajectory reduction

Drops useless command output, repeated file reads, and older observations that no longer affect the next step.

Caveman

Caveman-style replies

Rewrites prose like 'few word enough' while leaving tool calls, code, diffs, and error text unchanged.

RTK

Bash output rewriting

A PreToolUse hook rewrites eligible shell output; Read/Grep, python3, scripts, loops, and tricky shell syntax bypass it.

Headroom

Compress-then-retrieve

Stores the original output behind retrieval IDs. Claude sees compressed text first and has to ask for missing details.

Savings

Winglet

Your plan goes 26.7%-56.0% further

Based on computational cost savings of 21.1%-35.9% and input token savings of 39.9%-59.7%.

Caveman

Your plan goes 9.3% further

Based on JetBrains' evaluation measuring 8.5% cost saved (vs. 65% advertised).

RTK

Your plan goes 7.1% less far

Based on JetBrains' evaluation measuring 7.6% higher cost (vs. 60-90% advertised savings).

Headroom

Your plan goes 3.8% further

Based on a community evaluation measuring 3.7% saved, while Headroom advertises 15-20% token savings for coding agents (not cost savings).

Answer quality

Winglet

Remains the same

Parity with the original agent (ranging between -1.0% and +2.0%).

Caveman

-1.5% lower average task score

Average task score moved from 0.326 to 0.311 in JetBrains' evaluation.

RTK

No detectable quality difference

Quality remains the same despite cost increasing 7.6%.

Headroom

Remains the same on internal evals

Parity with the original agent on internal evals based on pure compression and Q&A tasks.

Evaluated on

Winglet

Industry standard benchmarks

SWE-bench Verified and Multi-SWE-bench Flash.

Caveman

Chat Q&A tasks

Despite being advertised as a coding-focused tool, its benchmark is chat Q&A.

RTK

CLI command examples

cargo test, pytest, go test, git diff, git status, git log, cat, grep, find, ls, deps, and one rtk gain screenshot.

Headroom

Generic compression suites

Benchmarks include GSM8K, TruthfulQA, SQuAD v2, BFCL, JSON logs, and search-output workloads.

Pricing

One plan covers every agent.
Every tier is included.

Single price

Your Winglet plan covers all plans across Claude Code and Codex. You do not have to switch to a different Winglet plan if you use something like Max 5x.

Available on MacOS & Ubuntu

Any issues with installation? Contact us at hi@agentwinglet.com.

Privacy-first

Your usage and savings data stays on your machine. Suitable for private repos, client work, and internal codebases.