FROM THE ORIGINAL NOTEBOOK / UPDATED MAY 29, 2026

Agent trust boundaries.

The latest published /now entry, preserved as the journal evolves.

CURRENT WORK

Questions on
the workbench.

01

Writing

Turning agent security notes into practical posts with examples that avoid publishing working abuse payloads.

02

Testing

Sketching eval harness patterns that can catch model regressions without leaking the private adversarial corpus.

03

Defending

Designing small trust gates around tool output, retrieval results, and generated actions.

04

Reading

Prompt injection research, browser-agent failure modes, and practical isolation designs for tool-using models.

NEXT

Follow the boundary.

Upcoming notes: quarantined LLM architecture, schema rejection ergonomics, and how to review tool calls without turning every workflow into a modal storm.

Explore agent security