Studio concept · fictional product. The company, product and every name and number in it are invented for this piece. About this piece

Open beta · Tuesday, November 10

See every step your agents take.

Tracewick records every prompt, tool call and handoff your AI agents make in production, flags the step that went wrong, and lets you replay the fix before you ship it.

Sample run · support-agent #48213 8steps 3tool calls 3.05 send to end 1step flagged

The problem

Agents fail quietly

Your support agent closed the ticket and replied “Done.” It had also refunded $1,240 without asking anyone. You found out from finance, three weeks later.

01 · Scattered

Three tabs, no story

Prompts in one log, tool calls in another, token counts on a provider dashboard. Nobody can read a run from start to end.

02 · Silent

Wrong, and polite about it

A run can finish, return a 200 and still do the wrong thing. Nothing alerts on a confident mistake.

03 · Unrepeatable

Every fix is a guess

You change the prompt and hope. There is no way to rerun last week's real tickets and see what it fixed, or broke.

How it works

Three steps, one clear line.

01

Wrap your agent

Python and TypeScript SDKs wrap your model and tool calls. Already emitting OpenTelemetry? Point it at Tracewick instead.

# two lines, then every run shows up
import tracewick
tracewick.init(project="support-agent")

02

Every run becomes one line

Each prompt, tool call and handoff lands on a single timeline, in order, with inputs, outputs, tokens, time and cost.

03

The broken step lights up

Write rules in plain English. Tracewick checks every run against them and points at the exact step that broke one, even when the run “succeeded”.

Rule · Refunds over $500 need an approval step.
Broken at step 6 · refunds.create · $1,240.00

Product tour

One run, start to finish.

01 · Connect

Two lines, and runs start arriving.

Install the SDK, name the project, deploy. Every run your agent makes shows up in the list within seconds.

02 · Read the run

Every step, in order, on one line.

Plan, look up, search, decide, act, reply. Open any step to see exactly what went in and what came out.

03 · Find the step

The step that broke a rule is lit.

The agent read the refund policy at step 4, then refunded $1,240 at step 6 anyway. Tracewick shows why, with the evidence.

04 · Replay the fix

Rerun the same ticket on a new prompt.

Replays use the recorded tool responses, so nothing real is refunded. Ship when the diff is green.

Use cases

For any agent that acts, not just answers.

Support agents

Refunds, returns, account changes

What gets litActions over a limit, promises the policy doesn't allow, loops that reopen the same ticket.

Back-office agents

Invoices, reconciliations, vendor email

What gets litAmounts that don't match the PO, the same invoice paid twice, a scanned PDF read upside down.

Coding agents

Pull requests, migrations, test fixes

What gets litTests edited until they pass, runaway retries, a $40 run that should have cost 40 cents.

Research & sales agents

Account research, CRM updates, drafts

What gets litSources that don't exist, stale numbers, an email drafted for the wrong account.

AI engineers

who get paged when an agent does something strange, and need the run, not a hunch.

Platform teams

running several agents on several model providers, who want one place to see them all.

Support and ops leads

who have to answer “what exactly did the agent do?” in a sentence, with proof.

Security & data

Your traces stay yours.

Traces hold your prompts, your customers' words and your tools' outputs. We treat them that way.

Redacted before it leaves

The SDK masks emails, card numbers and API keys on your servers, before anything is sent. Add your own patterns.

Your region, your retention

Choose where traces are stored and how long they are kept. When they expire, they are deleted, not archived.

Self-host when you need to

Run the collector and the storage in your own cloud account. Only the app talks to us, and it never sees trace contents.

Access you can audit

Single sign-on, roles per project, and a log of who opened which trace and when.

Never used for training

Your traces are not used to train any model, ours or anyone else's. Not in beta, not later.

Encrypted, in transit and at rest

TLS for every connection and encryption at rest for every store, with keys rotated on a schedule.

Tracewick is in beta and has not completed a third-party security audit. We'll publish reports when we have them, not badges before.

Pricing

Free while in beta. Simple after.

Builder

$0

For one agent in production.

  • 50,000 steps a month
  • 14 days of history
  • Rules and alerts
  • 1 project
Join the waitlist

Team · most teams start here

$400/ month

For teams shipping several agents.

  • 2 million steps a month
  • 90 days of history
  • Replay against new prompts and models
  • Unlimited projects, single sign-on
Join the waitlist

Scale

Custom

For regulated data and big volumes.

  • Self-hosted collector and storage
  • Retention you set
  • Security review with our team
  • A named engineer on call
Request access

Beta pricing. Every plan is free until January 2027. All prices are part of the concept.

FAQ

Questions, answered.

What does Tracewick record?

Each model call (prompt, response, tokens, time, cost), each tool call (inputs and outputs) and each handoff between agents, grouped into one run with a start and an end.

Does it work with my framework and model provider?

If your agent runs on Python or TypeScript, the SDK wraps model and tool calls directly, whatever the framework. Already sending OpenTelemetry spans? Point them at Tracewick. It works across model providers.

Will it slow my agents down?

No step waits on Tracewick. The SDK batches and sends in the background. If Tracewick can't be reached, your agent carries on and the SDK retries later.

How does a replay avoid refunding someone twice?

By default a replay uses the tool responses recorded in the original run, so no real tool is called. You choose which tools, if any, to call live.

How do rules work?

Write them in plain English, like “Refunds over $500 need an approval step”, or as code. Tracewick checks every run against them and flags the step that broke one.

Is Tracewick a real product?

No. Tracewick is a studio concept by Artrix Studio, made to show a complete product launch. The company, product and every name, number and screen are invented. The waitlist form doesn't send anything.

Open beta · November 10

Light up every step.

Join the waitlist and we'll send your invite on launch morning. One email, no drip campaign.