Miles prompt compression

1,160 fewer tokens. The Happy Path still holds.

I compressed the measured Miles prompt across two passes—from 4,892 tokens to 3,732—then retested the voice agent against the same nine call behaviours. Miles Junior now passes all nine; baseline Miles passed eight.

Shorter prompts are easy. Shorter prompts that still do the job need evidence.

This Dry Dock entry shows how I compressed Miles into Miles Junior across two measured passes, protected the working call behaviours, and checked the result against a defined Happy Path.

Before

Miles passed 8 of 9 checks

The measured baseline prompt was 4,892 tokens. The reviewed call handled information, pricing, message-taking, a withheld number, number confirmation, the exact close, smooth termination, and one question per turn. It missed one operational requirement: the CRM question used the wrong lookup tool. That sample recorded one tool call.

Agent public name

Miles

The measured baseline prompt was 4,892 tokens. The reviewed call handled information, pricing, message-taking, a withheld number, number confirmation, the exact close, smooth termination, and one question per turn. It missed one operational requirement: the CRM question used the wrong lookup tool. That sample recorded one tool call.

Voice surface

Call the before state

Start a browser call with Miles to experience the original prompt before the compression passes.

Ready

Browser call

Ready for a browser call.

Process

Compress in passes, retest the same job

Each pass reduced repeated prompt structure while keeping the call safeguards, then compared one before call with one after call on the same nine-point Happy Path checklist. Pass 02 starts from the Pass 01 result, not from Miles.

  1. Pass 01 — 4,892 → 4,146 tokens (−746 / 15.25%); Happy Path 8/9 → 9/9; CRM lookup tool corrected.
  2. Pass 02 — 4,146 → 3,732 tokens (−414 / 9.99%); Happy Path stayed 9/9 on both reviewed calls.
  3. Each pass rechecked behaviour, tool use, timing, and cost against the same checklist—without treating N=1 runtime moves as proof.

Compression passes

Process walkthrough

This walkthrough shows how the compression work is done. It stands alone as process proof; it is not tied to a single pass.

Process evidence

Miles Dry Dock process walkthrough

A standalone walkthrough of the prompt-compression process used on Miles / Miles Junior—how the work is measured, reduced, and retested.

Fallback path

If the embed is blocked or skipped, use the written notes below or open the video directly on YouTube.

Open on YouTube

Transcript notes

YouTube evidence

How prompt compression moved Miles to Miles Junior

Note 1

Shows the compression method: measure the prompt, reduce repetition, keep call safeguards, and retest the Happy Path.

Note 2

Supports the Process tile as method evidence; it is not a substitute for the written token and checklist results.

Note 3

Read the Evidence note for N=1 limits and unverified deployment relationships.

Prompt diff notes

  • Pass 01: 4,892 → 4,146 tokens; CRM tool check corrected.
  • Pass 02: 4,146 → 3,732 tokens; Happy Path remained 9/9.

After

Miles Junior is 1,160 tokens leaner and passes 9/9

Against the Miles baseline, the measured prompt is now 3,732 tokens—1,160 fewer (−23.71%). The current reviewed Happy Path passes all nine checks, including the CRM lookup tool Miles missed. Miles Junior has not replaced Miles in production naming; it is the compressed working result of the two passes.

Agent public name

Miles Junior

Against the Miles baseline, the measured prompt is now 3,732 tokens—1,160 fewer (−23.71%). The current reviewed Happy Path passes all nine checks, including the CRM lookup tool Miles missed. Miles Junior has not replaced Miles in production naming; it is the compressed working result of the two passes.

Voice surface

Call the result

Start a browser call with Miles Junior to experience the current compressed result after both passes.

Ready

Browser call

Ready for a browser call.

What this proves

Prompt compression is an engineering trade-off, not a word-count exercise

Good Voice AI work protects the conversation while changing the system underneath it—and shows clearly where the evidence stops.

  1. Signal 01

    Protect what already works

    The behaviours that already passed on Miles stayed intact across both passes, including message capture, number confirmation, the exact closing phrase, and one-question-per-turn discipline.

  2. Signal 02

    Test operational details

    The review checked tool selection and call behaviour, not just whether the conversation sounded plausible. Pass 01 corrected the CRM lookup tool Miles missed; Pass 02 kept that check green while cutting more tokens.

  3. Signal 03

    Report the inconvenient results

    Individual samples still show mixed timing and cost movement. I am not presenting prompt reduction as a runtime-performance improvement.

Chronology

Dry Dock history for Miles

Dated entries record the reviewed compression comparisons behind this page.

2026-07-13

Pass 01 — Miles to Miles Junior

reviewed

Issue

The Miles baseline prompt measured 4,892 tokens and the reviewed call missed the required CRM tool-selection check (8/9).

Process applied

Patrick reduced repeated prompt instructions and compared one before call with one after call against the same nine Happy Path checks.

Outcome

The post prompt measured 4,146 tokens (−15.25%). The reviewed after call passed all nine checks, with N=1 and deployment-verification limits.

2026-07-16

Pass 02 — further Miles Junior compression

reviewed

Issue

The Pass 01 result still measured 4,146 tokens; the next reduction needed another measured pass without losing the 9/9 Happy Path.

Process applied

Patrick tightened remaining duplication in the Miles Junior prompt and compared one before call with one after call on the same checklist.

Outcome

The post prompt measured 3,732 tokens (−414 / 9.99%). Both reviewed calls passed all nine checks. Cumulative Miles → current Junior: −1,160 tokens (−23.71%).

Next step

Improve the prompt without guessing at the result

If your voice agent has become difficult to maintain, I can help you reduce the prompt while protecting the call behaviours that matter.

We start with a real call path, define what must keep working, make the prompt change, and compare the result. You see both the improvement and the limitations before deciding what ships.

Review your agent