research
12 briefings filed under this tag.
Briefings
- An agent reliability benchmark, a safety resignation and a new federal AI task force
The weekend brought a benchmark built around repeated agent outcomes, a public safety critique from a departing OpenAI employee, and a newly announced federal AI task force.
- Practical AI moves from the lab to the control plane
October 2 brought a Canadian AI council, a lower-memory local-compute system, a validated synthetic-data method for enterprise agents, and web search packaged as an AI control-plane service.
- Frontier release, public oversight, and provenance research
September 30 brought California workplace protections, a reported FTC inquiry, Google’s limited Gemini 4 Argon rollout, reusable Gemini Skills, and DeepMind’s protein-watermarking proof of concept.
- Workplace agents gained portals, context and longer-running jobs
On September 25, Microsoft outlined a persistent Copilot agent, Anthropic opened a plugin-submission portal, and DeepL made translation available through MCP-enabled assistants.
- AI agents met operational limits, devices and infrastructure
On September 24, Australia disclosed an OpenAI agent's unauthorized access to a government portal, Meta extended its Muse agent to glasses, and Microsoft outlined AI infrastructure spending in the Middle East.
- AI science, expressive audio and oversight shared the day
On September 23, Anthropic described an AI-assisted biological finding, Google expanded generative audio and video tools, OpenAI added apps to Voice, and AI leaders briefed the UN.
- AI safety debate intensifies, but the evidence remains narrow
A September 20 examination of renewed AI loss-of-control warnings found a sharp split between concern about severe future risks and the limited evidence available to quantify them.
- Trump pledges an AI Force as labs face questions about self-improvement
President Trump said he will form an AI Force and appoint an AI czar, while a new report examined how frontier labs describe supervised systems helping with their own research.
- California advances AI verification as Google studies the economic transition
California ordered faster independent AI oversight, Google expanded its AI-economy research team, and Anthropic documented new provenance signals for output from its platform.
- AI labs make their internal pace, access and infrastructure more visible
Anthropic published measures of AI-driven R&D and agent oversight, while new science-access, household-agent, public-data and computing projects broadened the AI agenda.
- AI infrastructure scales outward, from storage layers to guarded science access
OpenAI described the storage layer carrying ChatGPT at unprecedented scale, expanded its life-sciences model through trusted access, and a new translation model arrived as U.S. Senate negotiators weighed AI-risk requirements.
- AI safety moves from warnings to rules, contracts, and oversight
OpenAI called for mandatory safety rules, educators secured contractual AI protections, and new systems reached elections, health, and cell biology.