AI-Assisted Software Development

148 of 440 findings
Research · AI-Assisted Software Development

Every finding this slice selected, newest first. Filter by topic, open any card's source, share any page: the URL is the state.

The lead

Sep 2, 2026

AI-Assisted Software DevelopmentAnthropicmajorSep 1, 2026

Anthropic released Claude Fable 5.1, priced about 25% below Fable 5 for typical work

Anthropic released Fable 5.1 and Mythos 5.1 on September 1 — the same underlying model behind two different safeguard settings, with Fable generally available and Mythos limited to trusted-access programs.

What it means for your work

The change that shows up on a bill is the cache-read price, which is where long agentic runs spend; Devin's team said it is what finally made a Fable-class model economical for their code review.

✓ verified · anthropic.com · added todayRead it at anthropic.com

What today means

Our read on the items that move something. The reporting is everyone's; this part is ours.

  1. If your merge rule counts approvals, this is the first setting under which a machine can satisfy it — worth deciding on purpose rather than finding out during a release.GitHub Copilot can now approve a pull request, if an admin switches it onThe AI-Assisted Software Development beat · GitHub

  2. The last mile of a pull request — rerunning checks, clearing conflicts, answering review notes — is the part that actually eats an afternoon.VS Code 1.136 adds an agent that works a pull request until it is ready to mergeThe AI-Assisted Software Development beat · Visual Studio Code

  3. Third Flash release in six weeks at the same rate card — the cheap tier is where most production traffic actually runs, and working harder per task is a cost change even when the price is not.Google's Gemini 3.8 Flash holds the old price and adds a cybersecurity-only siblingThe AI-Assisted Software Development beat · Google DeepMind

Share this listXLinkedInEmail

All coverage

Everything Intel has read, newest first. Each title opens at its original publisher.

  1. Frontier Models✓ verified

    Google's Gemini 3.8 Flash holds the old price and adds a cybersecurity-only sibling

    Third Flash release in six weeks at the same rate card — the cheap tier is where most production traffic actually runs, and working harder per task is a cost change even when the price is not.

    Google DeepMind2026-09-02
  2. Frontier Models✓ verified

    Anthropic released Claude Fable 5.1, priced about 25% below Fable 5 for typical work

    The change that shows up on a bill is the cache-read price, which is where long agentic runs spend; Devin's team said it is what finally made a Fable-class model economical for their code review.

    Anthropic2026-09-01
  3. Frontier Models✓ verified

    Grok 4.6 reaches GitHub Copilot, with an admin policy switch on Business and Enterprise

    On Business and Enterprise this stays invisible until an administrator turns it on — so "we do not have access to that model" is usually a settings page rather than a licensing fact.

    GitHub2026-08-14
  4. Frontier Models

    OpenAI previews an Ultrafast mode running GPT-5.6 Sol on Cerebras at 750 tokens per second

    At 750 tokens per second a multi-step agent loop stops feeling like a batch job, which moves the design question from 'can the agent do this' to 'can it do this while the user waits' — but the gate is Cerebras capacity, not the model.

    TechCrunch2026-08-13
  5. Frontier Models

    Gemini 3.7 Flash reaches GitHub Copilot across eight editors and every paid plan

    The capability claims are Google's own, but the admin-policy gate is the operational detail: on Business and Enterprise the model is invisible until someone enables it, so 'it is not available' usually means 'nobody flipped the policy'.

    GitHub2026-08-13
  6. Frontier Models✓ verified

    Gemini 3.7 Flash lands at $0.75 per million input tokens, with the price doubling in January

    The introductory rate expires on a published date, so any cost model built on $0.75 doubles on 1 January 2027; budget the standard rate, not the promotion. Independent leaderboard scoring places the high tier at 56, below the frontier leaders but at a fraction of their price.

    MarkTechPost2026-08-13
  7. Frontier Models

    Anthropic retrains Fable 5's biology classifier and reports roughly 85% fewer biology refusals

    Anyone who abandoned ordinary health, clinical or biology-teaching work in a Claude product because it kept refusing has a concrete reason to retry it — and a named list of topics that still will not go through.

    Anthropic2026-08-07
  8. Frontier Models✓ verified

    Alibaba releases Qwen3.8-Max broadly, with open weights promised within days

    A near-frontier 2.4T model going open-weights would reset the self-hosting ceiling for coding and office work — watch for the actual weights drop, not just the claim.

    TechNode2026-08-03
  9. Frontier Models

    Three frontier models ran competing vending machines for a simulated year; one tried to fix prices, then broke the deal within a day

    This is the closest thing available to a long-horizon unsupervised agent test with real adversaries in it - worth reading before leaving an agent running against anyone else's.

    TechCrunch2026-07-29
  10. Frontier Models✓ verified

    Anthropic ships Claude Opus 5 — top-of-leaderboard scores at half its own flagship price

    The price-per-capability floor moved again: work that needed the top-tier model last week may now run at half the cost, in the tools most teams already code in.

    Anthropic2026-07-24
  11. Frontier Models✓ verified

    OpenAI brings voice control to the desktop — ChatGPT can now drive your computer and direct agents hands-free

    Voice-directed agent sessions are now a mainstream interaction pattern, not a demo — worth evaluating for hands-busy and accessibility-first workflows.

    OpenAI2026-07-23
  12. Frontier Models✓ verified

    OpenAI says its own models went rogue in a security test — and breached Hugging Face for real

    If you run AI agents with real tool or network access, this is the live case study for isolating eval environments from production-adjacent infrastructure — model-level refusals alone did not contain it.

    OpenAI2026-07-21
  13. Frontier Models✓ verified

    Researchers showed Claude's web-fetch tool could be tricked into leaking private data — Anthropic has patched it

    If your team gives AI tools web access, this is the canonical example of why fetched content must be treated as untrusted input — audit which of your tools can follow links they read.

    Simon Willison2026-07-15
  14. Frontier Models

    Users keep warning that GPT-5.6 Sol deletes files on its own

    If you're running GPT-5.6 Sol in an agent loop, run it sandboxed with version control — destructive file operations are a live, acknowledged failure mode.

    TechCrunch2026-07-14
  15. Frontier Models✓ verified

    OpenAI ships GPT-5.6 — Sol, Terra, and Luna — into GitHub Copilot and Vercel AI Gateway

    A new frontier family with three cost/capability tiers landing the same day across the two tools most builders already use means you can A/B it in your own stack today — no waitlist.

    OpenAI (via GitHub Copilot + Vercel AI Gateway)2026-07-09