Stories, deep-dives and updates from the The Curator team.
Inside Test-Time Training: When an AI Model Learns From the Problem in Front of It
Most AI systems freeze their learned parameters before deployment. Test-time training breaks that boundary, briefly adapting a model to the evidence inside a live task. Here is how the mechanism works, where it differs from prompting and fine-tuning, and why controlled forgetting may be as important as learning.
Read postOn-Device AI vs. Edge AI vs. Cloud AI: Where Should Intelligence Actually Run?
A practical comparison of three deployment architectures across latency, privacy, model capability, operating cost, reliability, and product design—with a clear method for choosing the right execution boundary.
Read postHow to Build an AI Agent Honeypot That Reveals Failures Before Deployment
Create a compact adversarial environment where an AI agent can encounter deceptive instructions, unsafe tools, ambiguous permissions, and irreversible actions without putting real systems at risk.
Read postThree Myths About AI Browsers That Obscure the Real Platform Shift
AI browsers are often described as search engines with chat, autonomous agents, or inevitable replacements for websites. Each claim contains a fragment of truth—and misses the deeper contest over context, permission, and transaction control.
Read postField Notes on Agent Identity: The Missing Control Plane for Machine Workers
AI agents are beginning to acquire credentials, invoke tools, and cross system boundaries. The emerging challenge is not merely authentication, but giving each machine worker a constrained, temporary, and auditable identity.
Read postConstrained Decoding: How AI Systems Produce Outputs That Software Can Trust
A practical guide to the mechanism that restricts model generation to valid schemas, grammars, and protocols—complete with a worked example and the design choices that determine whether it succeeds.
Read postAI Guardrails: A Beginner’s Guide to Keeping Model Actions Within Bounds
Guardrails are not a single safety filter. They are a layered control system that limits what an AI can receive, decide, execute, and reveal. This guide explains the essential vocabulary, architecture, trade-offs, and first implementation steps.
Read postInside Diffusion Language Models: What Changes When Text Is Refined Instead of Written Left to Right
Most language models commit to one token at a time. Diffusion language models begin with an incomplete or corrupted sequence and repeatedly revise it. Here is how that mechanism works, where it may create new product possibilities, and why generation speed is only part of the story.
Read postPrompt Caching vs. Context Compression vs. RAG: Which Strategy Should Control Your AI Context Costs?
Three techniques can make long-context AI products more efficient, but they solve different problems. Here is how to choose among reusing prefixes, shrinking history, and retrieving evidence on demand.
Read postWhat The Curator blog is for
How it differs from the library
The library is the slow layer: ideas, methods and histories that should still read well next year. The blog is where the current argument lives — a launch worth interrogating, a strategy that keeps failing quietly, a piece of research that changes how a familiar problem looks. Posts argue; library entries explain.
Publication is paced rather than bulk-loaded, and every post is written specifically for this site, not shared across the sibling publications.
Editing and corrections
Posts are bylined and dated, and material post-publication changes are noted on the page rather than made quietly. Any reader can file a correction from an article page; the ones that check out are applied and credited.
External claims link to their source so readers can verify rather than trust. Our methodology page describes precisely where AI assists in drafting and where human editorial judgement is required before anything is published.
How the notebook differs from the archive
Two different jobs
The archive holds the durable essays. The notebook is for shorter, more provisional writing — what a new release changed about a convention, a myth worth puncturing, the reasoning behind a product decision.
Keeping them apart lets a field note stay tentative without diluting the archive, and lets an essay take months without missing the conversation entirely.
Cadence and authorship
Posts are paced rather than bulk-published, and each is written for this site specifically — nothing is syndicated from the sibling publications. Every post is bylined and dated, with an author page collecting that writer's work.
Comments, ratings and correction requests are open on every post and moderated for spam and abuse only. Disagreement about a reading is the entire point.
Corrections here too
A post that turns out to be wrong is corrected on the page, with the change noted where it alters meaning. Nothing is quietly removed to tidy the record.
External claims link out so you can check them, and a rotted link is a correction like any other — report it and it gets replaced.