Research · every section, one archive

All the research, one tab.

Every research item behind the daily briefs, newest first, in fast pages. Filter by topic, open any card's source, share any page — the URL is the state. The live map stays on the Brief page.

  1. Research

    SABER benchmark: leading coding agents violate safety in over half of tasks

    Coding agents doing real repo work is exactly AI Uni's build model — a reminder that autonomous edits need guardrails measured on outcomes, not refusals.

    arXiv (SABER)2026-05-31
  2. Research

    Anthropic publishes a Zero Trust framework for enterprise AI agents

    A useful audit checklist even outside Anthropic's own stack — the seven control domains it names are a reasonable starting list for anyone standing up agents with real write access.

    Anthropic2026-05-27
  3. Research✓ verified

    A survey and paper list mapping agent-memory architectures, so you don't have to design one blind

    Agent memory is the least-settled piece of most production agent stacks — a maintained survey and paper list is the fastest way to see which memory architecture actually matches your agent's failure mode before building a bespoke one from scratch.

    Agent-Memory-Paper-List2026-04-01
  4. Research✓ verified

    Agentic Context Engineering (evolving contexts for self-improving LMs)

    The academic framing of what the substrate-batch + landscape-brief + KB-principle loop does informally. Indexed via a curated list — read the underlying papers before teaching specifics.

    awesome-ai-agent-papers2026-03-01