Research · every section, one archive

All the research, one tab.

Every research item behind the daily briefs, newest first, in fast pages. Filter by topic, open any card's source, share any page — the URL is the state. The live map stays on the Brief page.

  1. AI in Education✓ verified

    Google brings Gemini into Classroom + free ACT/GRE practice at ISTE 2026

    Adaptive, personalized learning is going free-and-mainstream from a platform giant — the competitive backdrop for any AI-tutoring product.

    Google2026-06-25
  2. AI in Education✓ verified

    Microsoft's 2026 AI-in-Education report: adoption is mainstream, support lags

    The gap between 'schools use AI' and 'schools use AI well' is exactly the gap a structured learning product is built to close.

    Microsoft2026-06-24
  3. Dev Tooling & Infra✓ verified

    GitHub secret scanning adds a Supabase-credential detector that blocks the commit

    One of the most common AI-built-app failures is the database key shipped to the browser; free push protection on a public repo catches a class of that at commit time — but know which tier you're on.

    GitHub2026-06-17
  4. Open Source & Self-Hostable

    Z.ai's GLM-5.2 ships with permissive MIT open weights

    An MIT-licensed model you can run and modify on your own hardware is a real self-hosting option — no per-token bill, no access gate.

    Z.ai2026-06-13
  5. Research

    Google DeepMind launches a Robotics Accelerator, putting Gemini robotics models in startups' hands

    When a frontier lab puts its vision-language-action models directly in startups' hands, embodied AI stops being a lab demo and starts becoming an ecosystem — the same pattern that scaled language-model apps.

    Google DeepMind2026-06-09
  6. Open Source & Self-Hostable✓ verified

    NVIDIA releases Nemotron 3 Ultra — a 550B open-weights model

    A frontier-adjacent model you can run yourself narrows the gap between hosted APIs and self-hosted stacks for serious agent work.

    NVIDIA2026-06-09
  7. Market & Business✓ verified

    Anthropic Partner Network: Services Track + Partner Hub

    A distribution channel candidate for AI-Uni-built products (Agent Engine, Classroom) — worth a strategic look.

    Anthropic2026-06-03
  8. Dev Tooling & Infra✓ verified

    Anthropic Claude Security / codebase scanning (Project Glasswing)

    Security tooling from the platform AI Uni builds on — relevant to the three-skill security-review discipline and to the Anthropic Security Plugin install this session.

    Anthropic2026-06-02
  9. Agent Frameworks & Orchestration✓ verified

    Coding-agent market consolidates around parallel orchestration

    The "stack 2-3 agents" workflow is now the senior-dev default — validates the multi-terminal model as industry direction, not idiosyncrasy.

    Multiple2026-06-02
  10. Dev Tooling & Infra✓ verified

    Researcher shows one malicious GitHub issue could hijack repos running Claude Code's GitHub Action

    CI/CD-embedded coding agents inherit the write access of the workflow they run in — treat any agent-triggering input (issue titles, PR bodies, comments) from an untrusted user as untrusted, patched or not.

    GMO Flatt Security (independent research)2026-06-01
  11. Market & Business✓ verified

    Anthropic passes OpenAI at ~$965B; confidential IPO filing

    The platform AI Uni's agents run on is scaling fast and heading public — relevant to platform-dependency risk and pricing-stability planning.

    Multiple (CNBC, Fortune)2026-06-01
  12. Dev Tooling & Infra✓ verified

    GitHub Copilot moves every plan to token-metered 'AI Credits'

    Usage-metered agent tooling makes cost track how hard your agents actually work — the same budgeting shift teams hit running their own multi-agent lanes.

    GitHub2026-06-01
  13. Research

    SABER benchmark: leading coding agents violate safety in over half of tasks

    Coding agents doing real repo work is exactly AI Uni's build model — a reminder that autonomous edits need guardrails measured on outcomes, not refusals.

    arXiv (SABER)2026-05-31
  14. Market & Business✓ verified

    Claude pricing: base stable, fast mode down 66%, effort control

    Direct input to agent-run token budgets and AI Uni's own unit economics — more throughput per dollar favors the multi-terminal model.

    Finout2026-05-28
  15. MCP & Interop✓ verified

    MCP goes stateless + 2026-07-28 spec release candidate

    The substrate Intel's own agent-facing MCP layer will sit on. Stateless + Apps + Tasks materially simplify hosting a daily-accumulating, agent-queryable KB.

    MCP project2026-05-28
  16. Frontier Models✓ verified

    Claude Opus 4.8 + effort control + 3x cheaper fast mode

    A per-call cost-vs-depth lever — dial low-effort for mechanical lanes, max-effort for hard reasoning. Materially changes how agent runs are budgeted.

    Anthropic2026-05-28
  17. Frontier Models✓ verified

    Claude Opus 4.8 ships sharper judgment and longer independent runs — same price as before

    Same price, real upgrade — the added honesty about its own progress and longer independent runs are exactly the traits agentic workflows need most, with no new pricing tax to get them.

    Anthropic2026-05-28
  18. Agent Frameworks & Orchestration✓ verified

    Bun port (Zig to Rust) via dynamic workflows

    A concrete existence-proof of large-scale autonomous multi-agent work shipping real code — a teachable case study for AI-economy curriculum.

    MarkTechPost2026-05-28
  19. Agent Frameworks & Orchestration✓ verified

    Claude Code dynamic workflows (research preview)

    The productized version of the multi-terminal + subagent pattern an agent-orchestrated org hand-rolls today — direct input to agent-engine design.

    Anthropic2026-05-28
  20. Research

    Anthropic publishes a Zero Trust framework for enterprise AI agents

    A useful audit checklist even outside Anthropic's own stack — the seven control domains it names are a reasonable starting list for anyone standing up agents with real write access.

    Anthropic2026-05-27
  21. MCP & Interop✓ verified

    NSA issues formal security guidance for Model Context Protocol deployments

    The first government-issued checklist specifically for MCP deployments — worth a direct read before your next MCP server goes into production, not just a headline.

    National Security Agency (AI Security Center)2026-05-20
  22. Frontier Models✓ verified

    Gemini 3 / 3.5 ship; "agentic + vibe coding" framing

    A credible second frontier source for any model-portability or fallback strategy.

    Google2026-05-19
  23. Frontier Models✓ verified

    DeepSeek V4 Pro / V4 Flash (open weights, MIT, 1M ctx)

    Open-weights frontier-adjacent models are a hedge against platform/pricing risk; MIT licensing matters for productization.

    llm-stats (aggregator)2026-04-24
  24. Frontier Models✓ verified

    OpenAI ships GPT-5.5 — and makes the low-hallucination Instant variant ChatGPT's new default

    GPT-5.5 Instant becoming ChatGPT's default with lower hallucination in law, medicine, and finance raises the reliability bar for the highest-stakes consumer use cases, while 5.5-Codex signals a dedicated agentic-coding line, not just a general-purpose model.

    OpenAI2026-04-23