20 articles

Guides & field notes · Caching · SaveMyToken · Reviewed

Repeated document QA: does prompt order reduce cost? →

In a 24-request pilot, document-first prompts reused 12,800 input tokens. See initial versus later requests, estimated costs, answer failures and the reproducible data.

Guides & field notes · RAG evaluation · SaveMyToken · Reviewed

FAQ RAG: fixed windows or paragraph chunks? →

A 14-question pilot found equal strict answer scores, with 13.0% fewer input tokens for paragraph chunks. Inspect the failures, dataset and runnable experiment.

Guides & field notes · Personal agents · SaveMyToken · Reviewed

An agent keeps interrupting you: adjust its notifications →

Trace repetitive alerts to their tasks, separate check frequency from notification rules, and verify that a quieter agent still reports what matters.

Guides & field notes · Personal agents · SaveMyToken · Reviewed

Compare agent workflows: time, corrections, and cost →

Use a complete fictional inbox task, fixed acceptance criteria, and a blank results sheet to measure useful outcomes without inventing a leaderboard.

Guides & field notes · Agent workflows · SaveMyToken · Reviewed

Route an agent request with a decision model →

Define clear intent labels, keep uncertain cases visible and validate the selected tool before it runs.

Guides & field notes · Model routing · SaveMyToken · Reviewed

When do decision models actually save money? →

Measure complete-task cost, fallback frequency and accepted quality before claiming token savings.

Guides & field notes · RAG evaluation · SaveMyToken · Reviewed

Check RAG evidence before generating an answer →

Evaluate relevance and evidence coverage after retrieval, and measure what filtering removes.

Guides & field notes · Personal agents · SaveMyToken · Reviewed

Your agent remembers it wrong: fix stale facts and conflicting instructions →

Trace an incorrect answer to its source, make a scoped correction, and verify the result across new conversations and recurring tasks.

Guides & field notes · Personal agents · SaveMyToken · Reviewed

Move AI memories into Claude—and check what survived →

Find the correct import flow, audit a small set of memories, test behavior in a fresh conversation, and repair omissions without assuming a complete transfer.

Guides & field notes · Decision models · SaveMyToken · Reviewed

Jev vs Laya vs Kev: choose a decision model →

Compare API access, local deployment and probability outputs, then build a shortlist around your task.

Guides & field notes · Agent migration · SaveMyToken · Reviewed

Switch agents and keep the settings that matter →

An agent migration guide to preserving preferences, project rules, memory, Skills, MCP connections, and automations—with a settings worksheet and checks for what actually transferred.

Guides & field notes · Personal agents · SaveMyToken · Reviewed

Start with Muse: hand over your habits and project context →

Give Meta's personal agent a clear first task, a reviewed background brief, and explicit boundaries before adding connections or recurring work.

Guides & field notes · Personal agents · SaveMyToken · Reviewed

Before switching AI assistants, write your personal brief →

Carry your working preferences and active projects into a new assistant with a reviewable brief, a project card, and a small acceptance check.

Guides & field notes · Model evaluation · SaveMyToken · Reviewed

How to read model benchmarks →

Understand common AI benchmarks, score types, test settings, and the limits of a leaderboard rank.

Guides & field notes · Coding plans · SaveMyToken · Reviewed

Coding plans vs. API billing →

Choose a billing approach by checking compatibility, usage limits, and your actual workload.

Guides & field notes · Agent workflows · SaveMyToken · Reviewed

Set a budget for an agent task →

Control retries and tool output, then verify.

Guides & field notes · Context · SaveMyToken · Reviewed

Keep context useful, not endless →

Trim repeated history and oversized tool output.

Guides & field notes · Model selection · SaveMyToken · Reviewed

Use the right model for each step →

Route routine work and define clear upgrade rules.

Guides & field notes · Caching · SaveMyToken · Reviewed

Prompt caching, explained →

Learn what is reused and check cache usage.

Guides & field notes · Structured output · SaveMyToken · Reviewed

Ask for the output you actually need →

Reduce extra generation and parsing failures.