All the research, one tab.
Every research item behind the daily briefs, newest first, in fast pages. Filter by topic, open any card's source, share any page — the URL is the state. The live map stays on the Brief page.
- Research
SABER benchmark: leading coding agents violate safety in over half of tasks
Coding agents doing real repo work is exactly AI Uni's build model — a reminder that autonomous edits need guardrails measured on outcomes, not refusals.
- Research
Anthropic publishes a Zero Trust framework for enterprise AI agents
A useful audit checklist even outside Anthropic's own stack — the seven control domains it names are a reasonable starting list for anyone standing up agents with real write access.
- Research✓ verified
A survey and paper list mapping agent-memory architectures, so you don't have to design one blind
Agent memory is the least-settled piece of most production agent stacks — a maintained survey and paper list is the fastest way to see which memory architecture actually matches your agent's failure mode before building a bespoke one from scratch.
- Research✓ verified
Agentic Context Engineering (evolving contexts for self-improving LMs)
The academic framing of what the substrate-batch + landscape-brief + KB-principle loop does informally. Indexed via a curated list — read the underlying papers before teaching specifics.