Cache Ttl Timer — AI skill for Claude Code
Claude Code mod: a prompt-cache TTL countdown in the prompt footer, beside the model and effort.
How to install Cache Ttl Timer
This entry records only its repository, not the path inside it, so there is no
exact command to give. Open WQGGSEY/cache-ttl-timer and copy the folder into
~/.claude/skills/, or the file into ~/.claude/agents/.
What Cache Ttl Timer does
Claude Code mod: a prompt-cache TTL countdown in the prompt footer, beside the model and effort.
Alternatives in AI
- Codex Skill — by klaudworks - Enables users to prompt codex from claude code 914 ★
- Pilotfish — Multi-model orchestration layer for Claude Code — the frontier model plans, cheaper models execute, verificati 688 ★
- STS2 AI Agent — https://github.com/user-attachments/assets/89353468-a299-4315-9516-e520bcbfbd4b English README: README.md STS2 191 ★
README
Cache TTL Timer
A Claude Code mod that shows how long the prompt cache of your conversation stays warm, right in the prompt footer beside the model and effort:
+ 🎙 ⌄ Auto ● 47m Opus 5.5 High ◔
Claude caches the conversation's prompt for a time to live (TTL) of 5 minutes or 1 hour after each request. While the cache is warm, your next message is read from it cheaply. Once it lapses, the next message writes the whole conversation to the cache again, which costs more and counts more against your usage. The timer tells you which of the two your next message will be, before you send it.
What it shows
| Footer | Meaning |
|---|---|
● 60m → ◕ → ◑ → ◔ |
The cache is warm. The ring empties as the TTL runs down, and the time counts down in minutes. |
◔ 9m in the warning color |
Less than a fifth of the TTL is left. |
○ 0:42 |
The last minute, counted in seconds. |
◌ Cold |
The cache has lapsed. Your next message re-caches the conversation. |
Nothing shows before the first request of a session. The label uses the footer's own dim color until the last fifth of the TTL, so it stays quiet until it matters.
How it works
- When the countdown starts: each time the main conversation sends a request to the model, which is when the cache is read or written. A turn with tool calls sends several requests, and each one restarts the countdown. Requests from subagents are left out, because they cache their own prompts.
- Which TTL applies: read from the session's transcript, where the API reports how many tokens each request wrote to the cache at the 5-minute and 1-hour TTL. Until a write is seen, the timer assumes 1 hour, the TTL Claude Code uses on the main conversation of a subscription.
- After a restart or a resume: the countdown picks up from the last request in the transcript, so a session you come back to shows whether its cache is still warm before you type.
The timer is an estimate from the client's side.
Related Skills
Herdr Plugin Done Timer
A prompt-cache countdown for every AI agent on herdr's agents panel, read from each agent's transcript.
Effortlane
Effortlane: open-source model and reasoning-effort routing for coding agents. Native Codex auth, Shadow evalua
Clawdy
Clawd moved in above your Claude Code prompt: a pixel pet that cooks while Claude thinks, beside live cards fo
Prompt Optimize
One-shot prompt rewrite — diagnose a rough draft, steer you to the right workflow archetype (linear, TDD, suba
Tokenslim
TokenSlim — spend fewer tokens on every LLM API call: prompt compression, 350-skill router, model cascade, sem
Claude Code Jev Smart Router
HTTP proxy for Claude Code that selects the Claude model per request to cut cost and latency. Routes on task p
Related Agents
Model Fable 5 1 High
General-purpose agent pinned to Claude Fable 5.1 (claude-fable-5-1) at high effort. Use when a task must run o
Claudehut DB Reviewer
Persistence and data-access performance review — JPA/R2DBC mappings, fetch strategy, transaction boundaries, m
Effort Grader
Blind grader for effortmining. Grades one artifact against a task prompt and a fixed rubric ONLY. The input pa