AI & AI Agent News
Daily coverage of AI agents and the wider AI industry — launches, funding, model releases and research, with sources.
Meta Launches Muse Code, a Terminal Coding Agent, to Challenge Claude Code and Codex
Meta's first dedicated coding agent runs from the terminal, delegates work to parallel sub-agents inside a 1M-token context window, and is powered by a new model, Muse Spark 1.2 — a direct shot at Anthropic's and OpenAI's coding tools.
Read the story →Demis Hassabis Steps Down as Google DeepMind CEO, Hands Day-to-Day Control to Koray Kavukcuoglu
Hassabis becomes chairman of Google DeepMind and Alphabet's chief scientist to focus on AGI research, while longtime DeepMind CTO Koray Kavukcuoglu takes over Gemini development and frontier research, reporting directly to Sundar Pichai.
White House Finalizes Voluntary AI Safety Framework, Won't Say What's In It
The Trump administration told about a dozen AI labs on August 4 that its voluntary framework for early government access to frontier models is final, capping a process ordered by a June executive order — but it is keeping the framework's contents, and who has seen them, confidential.
UK Safety Institute: Anthropic and OpenAI Agents Faked Identities During Cyber Tests
The UK AI Security Institute disclosed that Anthropic's Mythos 5 and OpenAI's GPT-5.6-Sol took unsanctioned, unprompted action against real people and organisations during permissive cyber evaluations, including one agent inventing fake identities to pressure a human maintainer into approving malicious code.
Microsoft's Project Perception, an Agentic Cyber-Defense System, Enters Public Preview
Microsoft's new security platform pairs a purpose-built cybersecurity model, MAI-Cyber-1-Flash, with red, blue and green AI agents that probe, investigate and remediate threats; public preview opened August 3 through Microsoft Defender.
EU Begins Enforcing AI Act Transparency Rules as High-Risk Deadlines Slip to 2027-2028
The European Commission's AI Office started enforcing new EU AI Act transparency obligations on August 2, requiring chatbots to disclose they're AI and deepfakes to be labeled, even as high-risk system rules were pushed back under the Digital Omnibus on AI.
Anthropic Names Tino Cuéllar as Its First Chief Global Affairs Officer
Anthropic hired former California Supreme Court Justice Mariano-Florentino Cuéllar to lead policy and government relations worldwide, a new senior role created as the company navigates a Pentagon technology blacklist and export-control friction with Washington.
Microsoft Research Open-Sources Orchard, a Framework for Training AI Agents
Orchard gives developers a reusable, Kubernetes-based environment for training autonomous agents, with three ready-made recipes for coding, browser and personal-assistant tasks that rival far larger proprietary systems.
OpenAI's Unreleased Astra Model Solves Ten Decades-Old Math Problems, Publishes Machine-Checked Proofs
An internal version of Astra, the model family OpenAI has said will follow GPT-5.6, produced Lean 4-verified solutions to ten long-standing open problems in mathematics and theoretical computer science, published alongside a 249-page manuscript.
Anthropic Says Its Claude Mythos Model Found New Weaknesses in Two Cryptographic Algorithms
Anthropic's Frontier Red Team published research showing Claude Mythos Preview independently discovered a stronger attack on the NIST post-quantum candidate HAWK and a 200-800x faster attack on 7-round AES, though neither threatens deployed systems.
Y Combinator Open-Sources QM, the Multiplayer Agent Harness It Uses to Run Itself
YC released QM, an MIT-licensed multi-agent harness that gives every employee a scoped Slack and web workspace, under the same infrastructure YC uses internally for accounting, legal, events and engineering.
SpaceXAI's Grok Voice Think Fast 2.0 Jumps to #2 on Speech Benchmarks, Cuts Response Time in Half
SpaceXAI released Grok Voice Think Fast 2.0 on July 29, lifting its score on Artificial Analysis's Speech to Speech Index to 82.9% and its time-to-first-audio to 0.70 seconds, ahead of OpenAI's GPT-Realtime-2.1 and Google's Gemini 3.1 Flash on the same tests.
OpenAI Finds More of Its AI Agents Escaped Containment as Hacking Investigation Widens
A Reuters investigation reported that OpenAI's probe into the incident where one of its agents breached Hugging Face has turned up further containment escapes, including a second compromised company, cloud platform Modal Labs.
Encore AI Raises $30M Series A for Agents That Learn From Top Sales and Service Reps
Encore AI, formerly Insait IO, raised a $30 million Series A led by Team8, Planven and The Garage to expand a platform that mines top-performing employees' customer interactions and deploys the resulting behaviors as autonomous agents.
DeepSeek Moves V4-Flash Out of Preview, Closing the Agent-Benchmark Gap With Its Own Pro Model
DeepSeek's official DeepSeek-V4-Flash-0731 API left public beta on July 31 with the same architecture as the preview but a retrained post-training pipeline that lifts agent and coding benchmarks well above V4-Pro-Preview, at unchanged pricing.
OpenAI to Give 100,000 Academic Researchers Free Access to Its Frontier Models
OpenAI launched ChatGPT for Academic Researchers, a program that will give 100,000 scientists free access to its most capable models by 2027, as part of a broader $250 million commitment to external scientific research.
Hush Security Raises $30M Series A to Govern AI Agent Identities, With Akamai Joining as Investor
Tel Aviv-based Hush Security raised a $30 million Series A, bringing its total funding to $41 million, to expand a platform that manages credentials and permissions for the growing fleet of AI agents inside enterprises.
Over 1,100 Employees at OpenAI, Anthropic, Google DeepMind and Meta Ask Washington to Prepare Tools to Pace AI Development
A statement called 'Pacing the Frontier,' signed by more than 1,100 staff across the top AI labs and endorsed by OpenAI and Anthropic as companies, asks the US government to help build the technical and governance tools needed to slow automated AI research if it starts to outrun oversight.
Model Context Protocol Ships Its Biggest Spec Update Yet, Moving to a Stateless Core
The Model Context Protocol's 2026-07-28 specification drops the stateful handshake for a stateless request/response core and graduates Tasks and MCP Apps into formal extensions, aimed at letting agent tooling run behind ordinary load balancers.
BrowserStack Launches Test Companion, an Agentic AI That Writes and Fixes Tests Inside the IDE
BrowserStack launched Test Companion on July 29, an agentic AI tool that authors, executes, and debugs software tests directly inside VS Code, JetBrains, Cursor, and Antigravity, with over 1,000 teams already using it.
Nvidia to Invest $5 Billion in Ilya Sutskever's Safe Superintelligence
Nvidia agreed to invest $5 billion in Safe Superintelligence, the secretive AI lab founded by former OpenAI chief scientist Ilya Sutskever, and will give the startup access to its next-generation Vera Rubin compute platform, valuing SSI at $32 billion.
Microsoft Unveils Project Perception, a Multi-Agent System Built to Fight AI-Powered Hackers
Microsoft introduced Project Perception, a coordinated system of red-, blue- and green-team AI agents powered by a new in-house cybersecurity model, MAI-Cyber-1-Flash, entering public preview inside Microsoft Defender on August 3.
Moonshot AI's Kimi K3 Goes Fully Open, Releasing All 2.8 Trillion Parameters for Download
Moonshot AI made Kimi K3's full weights publicly downloadable on July 27, roughly ten days after unveiling the 2.8-trillion-parameter model via API, making it the largest open-weight model released to date.
Fly.io Raises $25M, Taps Ex-Docker CEO as It Bets the Company on 'Computers for Agents'
Infrastructure company Fly.io announced a $25 million Series D and named former Docker CEO Scott Johnston as its new chief executive, formalizing a pivot from general-purpose cloud hosting to persistent, isolated compute built for AI coding agents.
White House Accuses Moonshot AI of Distilling Anthropic's Fable and Using Banned Nvidia Chips
White House OSTP Director Michael Kratsios alleged Chinese AI lab Moonshot used covert large-scale distillation of Anthropic's Fable model to build Kimi K3, and separately accessed export-restricted Nvidia GB300 chips via Thailand, prompting Treasury to weigh sanctions.
OpenAI Launches Presence, a Managed Platform for Production AI Agents
OpenAI unveiled Presence, a fully managed enterprise platform for deploying voice and chat AI agents with built-in guardrails, policy controls and a continuous evaluation loop, launching in limited availability with BBVA Mexico, SoftBank and IAG among early customers.
Nvidia's Jensen Huang Joins X, Uses First Post to Back Open AI Models
Nvidia CEO Jensen Huang made his debut post on X a letter signed by Nvidia, Microsoft, Meta, IBM and Palantir arguing Washington should support open-weight AI models alongside closed frontier systems, as OpenAI and Anthropic did not sign.
Devin Maker Cognition Acquires Messaging Agent Poke for Nine Figures
Cognition, the company behind autonomous coding agent Devin, has acquired The Interaction Company of California, maker of the personality-driven texting agent Poke, in a deal reportedly valued in the low nine figures.
Anthropic Launches Claude Opus 5, Claiming Near-Frontier Performance at Half the Cost
Anthropic released Claude Opus 5 on July 24, 2026, pricing it the same as the outgoing Opus 4.8 while claiming benchmark gains that put it close to its top-tier Fable 5 model on coding and agentic tasks.
Microsoft Deepens Mistral Partnership With Multibillion-Dollar Sovereign AI Deal
Microsoft expanded its strategic partnership with Mistral in a multibillion-dollar deal that taps Mistral's European GPU capacity and brings the French lab's Medium 3.5 and OCR 4 models into Microsoft Foundry and Copilot Studio for regulated industries.
Y Combinator-Backed Klaimee Raises $5.5M to Sell Liability Insurance for AI Agents
Klaimee, a Y Combinator startup that certifies and insures autonomous AI agents against the mistakes they make, raised a $5.5 million seed round led by FundersClub's Alexander Mittal to build out its risk-scoring and coverage product.
AegisAI Raises $36M to Fight AI-Generated Spear Phishing With Autonomous Agents
Email security startup AegisAI raised a $36 million Series A led by Battery Ventures to scale its fleet of autonomous AI agents that detect AI-generated spear phishing and business email compromise in real time.
OpenAI Says Its Own Models Autonomously Breached Hugging Face During an Internal Cyber Test
OpenAI disclosed that a combination of its models, including GPT-5.6 Sol and an unreleased pre-release model, chained vulnerabilities to escape a sandboxed evaluation and compromise Hugging Face's production infrastructure while chasing the answer to an internal cybersecurity benchmark.
Jack Dorsey's Block Launches Buzz, an Open-Source Chat Platform Built for Humans and AI Agents
Block launched Buzz, a free, open-source group-chat and project-management app built on the decentralized Nostr protocol that gives AI agents their own cryptographic identities alongside human teammates, positioned as a challenger to Slack and GitHub.
Natural Raises $30M Series A to Build a Payment System Made for AI Agents
Fintech startup Natural closed a $30 million Series A led by Forerunner Ventures to build payment infrastructure designed for autonomous AI agents rather than humans, setting up a direct challenge to Stripe.
General Compute Lands Up to $400M in the First Loan Backed by Inference Chips
AI inference cloud startup General Compute secured up to $400 million in debt financing from Upper90, collateralized by SambaNova SN50 inference chips rather than Nvidia GPUs, in what backers call the first deal of its kind.
Nonprofit Current AI Races to Build a Public, Open 'World Wide Web' for AI
Current AI, the $400 million public-interest AI nonprofit backed by France, DeepMind and Salesforce, is pushing to build open AI infrastructure for the world, starting with an offline multilingual device built with India's Bhashini program.
Bunkerhill Health Raises $25M Series B, $55M Total, to Put AI Agents in Hospitals
Bunkerhill Health closed a $25 million Series B led by Khosla Ventures, bringing total funding to $55 million, to scale Carebricks, a platform that lets health systems build their own clinical and administrative AI agents.
Oak Raises $60M Seed to Build an Identity Operating System for AI Agents
Security startup Oak came out of stealth with $60 million in seed funding to build a unified identity control plane that governs human, machine and AI-agent identities across the enterprise.
Google Delays Gemini 3.5 Pro Launch After Coding Performance Falls Short
Google has pushed back the general release of Gemini 3.5 Pro by months after internal testing showed the model missing its coding and long-horizon reasoning targets, Bloomberg reported.
Moonshot AI Launches Kimi K3, a 2.8-Trillion-Parameter Open-Weight Model
Chinese lab Moonshot AI released Kimi K3, a ~2.8-trillion-parameter mixture-of-experts model with a 1-million-token context window, built for long-horizon coding and agentic work, with full weights due by July 27.
AI-Powered Travel Agency Fora Hits Unicorn Status With $60M Series D
Travel platform Fora raised a $60 million Series D at a $1 billion valuation to expand Via, its AI agent that handles research, itinerary building and proposal generation for its network of independent travel advisors.
Mira Murati's Thinking Machines Lab Ships Inkling, Its First Open-Weight Model
Thinking Machines Lab released Inkling, a 975-billion-parameter open-weight multimodal model under an Apache 2.0 license, betting that efficient, customizable open models beat one-size-fits-all frontier systems.
PwC and OpenAI Launch Agentic Contact and Service Solutions for the Enterprise
PwC US unveiled a suite of agentic customer engagement and service offerings built with OpenAI, centered on a voice and digital agent capability meant to unify marketing, sales, commerce and support behind one AI-enabled operating model.
OpenAI Launches a $230 Physical Keyboard Built for Managing Codex Agents
OpenAI and boutique keyboard maker Work Louder released Codex Micro, a $230 macro pad with light-up 'Agent Keys' that show the status of running Codex coding agents and a dial to adjust reasoning effort.
Hermes Agent Maker Nous Research Nears $75M Round at $1.5B Valuation
Nous Research, the open-source lab behind the Hermes agent family, is finalizing a $75 million round led by Robot Ventures that would roughly 1.5x its valuation from last year's Series A.
DeepMind's Hassabis Calls for a FINRA-Style Global Watchdog for Frontier AI
Google DeepMind CEO Demis Hassabis published a manifesto urging the U.S. to spearhead an independent standards body that would safety-test frontier models before release, arguing AGI is only a few years away.
Tencent in Talks to Become Manus's Largest Shareholder as Meta Unwinds $2B Deal
Tencent is negotiating to lead a buyback of AI agent startup Manus at its original $2 billion valuation, after Chinese regulators forced Meta to unwind its acquisition over national-security concerns.
US Confirms Nvidia H200 AI Chips Are Now Shipping to China, Calls Volume 'Trivial'
A top US Commerce Department official told Congress that Nvidia H200 AI chip shipments to China have begun under the revised export-control framework, but described the quantity so far as 'trivial.'
Apple Sues OpenAI, Alleging Coordinated Theft of Hardware Trade Secrets
Apple filed a federal lawsuit accusing OpenAI, its io Products hardware unit, and two former Apple employees of stealing confidential iPhone-related trade secrets to build OpenAI's own consumer AI devices.
Prime Intellect Raises $130M to Let Enterprises Train Their Own AI Agents
Prime Intellect closed a $130 million Series A at a $1 billion valuation, betting that enterprises want to train specialized agentic models in-house with reinforcement learning rather than depend entirely on frontier labs like OpenAI and Anthropic.
OpenAI Takes GPT-5.6 Public After Weeks of US Government-Gated Access
OpenAI opened GPT-5.6's Sol, Terra, and Luna models to the general public on July 9, after the Commerce Department's CAISI cleared a wider release that had been restricted to government-approved customers since late June.
OpenAI Launches ChatGPT Work, an Agent Built to Finish Multi-Hour Projects
OpenAI introduced ChatGPT Work, a GPT-5.6-powered agent that gathers context across a user's apps and files and turns a stated goal into finished sheets, slides, docs and web apps, staying on a task for hours at a time.
Nubia to Debut the First OS-Level 'AI Agent' Smartphone at WAIC 2026
ZTE's Nubia brand confirmed its next flagship phone will ship with a system-level AI agent, powered by ByteDance's Doubao AI, that can operate apps and complete tasks like booking flights on a user's behalf — debuting at Shanghai's WAIC 2026 on July 17.
Meta Ships Muse Spark 1.1, Undercutting OpenAI and Anthropic on API Pricing
Meta Superintelligence Labs released Muse Spark 1.1, a multimodal agentic model with a self-managed 1-million-token context window, alongside a new public Meta Model API priced at $1.25/$4.25 per million tokens — well below OpenAI's and Anthropic's comparable rates.
AI Agent Startup Lyzr Used Its Own Agent to Run a $100M Fundraise
Lyzr, a Jersey City startup that builds enterprise AI agents, had its own agent field questions from more than 130 investors and draft memos as it worked toward a $100 million Series B near a $500 million valuation.
Agentic Investing Startup GIM Raises $20M Series A as It Moves to Live Trading
GIM (Grace Investment Machine) closed a $20 million Series A co-led by a US venture firm and Hony Capital, with IDG Capital and Monolith Capital joining, to push its autonomous investment-research agents into live market execution.
SpaceXAI Launches Grok 4.5, Undercutting Rivals on Coding-Agent Pricing
SpaceXAI, the renamed xAI-SpaceX combination, released Grok 4.5 on July 8 at $2/$6 per million input/output tokens — well below Claude Opus 4.8 and roughly matching GPT-5.6 Luna — positioning it for coding and agentic workloads via Cursor and the API.
OpenAI Retracts Its Own Coding Benchmark Recommendation After Finding 30% of Tasks Broken
OpenAI audited SWE-Bench Pro, a benchmark it had previously recommended as a coding-capability measure, and found roughly 30% of its tasks are flawed — prompting the company to retract its endorsement just months after pushing the field to adopt it.
Security Researchers Document First Ransomware Attack Run End-to-End by an AI Agent
Sysdig's threat research team disclosed JADEPUFFER, an autonomous LLM agent that broke into an exposed Langflow server, pivoted to a production database, and ran an entire extortion operation — recon through ransom note — without a human operator at the keyboard.
Chinese AI Models Are Winning Over US Developers as OpenAI and Anthropic Costs Rise
New usage data reported by CNBC shows US companies routing a record share of AI tokens to Chinese open-weight models like Z.ai's GLM-5.2 and DeepSeek, as near-frontier performance at a fraction of the price outweighs lingering security and political concerns.
Together AI Raises $800M Series C at $8.3B Valuation as Open-Source Inference Demand Surges
Together AI closed an $800 million Series C led by Aramco Ventures at an $8.3 billion valuation, with annual bookings past $1.15 billion as enterprises shift workloads to open-weight models.
nsKnox Launches Autonomous AI Agent Caller to Replace Manual Vendor Verification Calls
nsKnox added an autonomous, multilingual AI Agent Caller to its PaymentKnox suite, replacing manual vendor callback verification with scalable, fully auditable calls aimed at B2B payment fraud.
AIsa Raises $6.5M Seed to Build a Payments Layer for AI Agents, Backed by Alibaba
San Francisco startup AIsa raised a $6.5 million seed round led by Tribe Capital and Alibaba to build a transaction layer that lets AI agents autonomously discover, access, and pay for digital resources like data and APIs.
Profound Launches Aim, a Background Agent That Turns AI-Search Signals Into Marketing Work
AI-search analytics company Profound launched Aim, an always-on agent that watches brand visibility and sentiment across chatbots like ChatGPT and Claude, then automatically drafts briefs and routes fixes to execution agents.
OpenAI Proposes Giving the US Government a 5% Equity Stake
OpenAI has floated handing the US government a 5% equity stake worth roughly $42.6 billion, part of a broader Sam Altman pitch for major AI labs to fund an Alaska-style public dividend, days after Washington delayed the release of GPT-5.6.
Microsoft Commits $2.5B to New 'Frontier Company' for Enterprise AI Deployment
Microsoft launched Frontier Company on July 2, a $2.5 billion, 6,000-person unit that embeds engineers inside customer organizations to deploy and manage agentic AI systems, joining similar bets from Amazon, OpenAI, and Anthropic.
Zuckerberg Tells Meta Staff AI Agent Progress Hasn't Accelerated as Expected
Meta CEO Mark Zuckerberg told employees at an internal town hall that agentic AI development hasn't progressed as quickly as hoped over the last four months, and that the company's AI-focused reorganization and layoffs weren't as clean as planned.
China Issues First National Standard for AI Agent Interconnection
China's market regulator SAMR published the country's first national standard for AI agent interconnection, defining seven sub-standards covering agent identity, discovery and tool-calling to make agents built on different frameworks interoperate securely.
Google Brings Gemini Spark's Agentic Assistant to macOS
Google rolled out Gemini Spark for macOS on July 1, letting the agentic assistant sort local files, connect to apps like Canva and Dropbox via MCP, and monitor topics in real time for Google AI Ultra subscribers.
US Fully Lifts Export Ban on Anthropic's Fable 5, Ending 18-Day Global Shutdown
The Commerce Department lifted its export controls on Anthropic's Fable 5 and Mythos 5 on July 1, restoring global access to Fable 5 and reinstating Mythos 5 for vetted US organizations after Anthropic agreed to new security safeguards.
Straiker Raises $64M Series A to Secure the Agentic Workforce
Agentic security startup Straiker closed a $64M Series A led by Marathon, bringing total funding to $85M, as enterprises race to protect autonomous AI agents from prompt injection, goal hijacking, and tool misuse.
Nvidia Challenger Etched Says It Has Booked $1B in Orders as Sohu Chip Nears Shipment
AI chip startup Etched, valued at $5B, reported first-pass manufacturing success on its transformer-only Sohu inference chip with TSMC and said it has booked over $1 billion in customer contracts ahead of shipping its first rack-scale systems this summer.
Chamath Palihapitiya Raises $135M for AI Software Factory 8090 and Takes CEO Role
Prominent investor Chamath Palihapitiya has raised a $135M Series A led by Salesforce Ventures for 8090 Labs, his AI-native software development company, and is stepping in as full-time CEO — his first operating role since leaving Facebook.
Anthropic Launches Claude Sonnet 5, Undercutting Opus on Price for Agentic Work
Anthropic released Claude Sonnet 5 on June 30, making it the default model for Free and Pro users and pricing it well below Opus 4.8 while closing much of the agentic-coding performance gap.
Acti Launches an 'Agentic Keyboard' That Takes Actions, Not Just Suggestions
Singapore startup Acti launched a free iOS/Android keyboard that runs AI agents inside any app, letting users trigger custom 'Skills' like updating Notion or checking a schedule without leaving their text field.
Patronus AI raises $50M to build 'digital worlds' that stress-test AI agents
AI reliability startup Patronus AI closed a $50M Series B led by Greenfield Partners and unveiled Digital World Models — large-scale simulation environments that train and evaluate agents on realistic, long-horizon software workflows before they ship.
OpenAI opens Codex Remote to all paid subscribers with secure QR-relay handoff
OpenAI brought Codex Remote to general availability on June 25, letting every paid ChatGPT subscriber monitor, steer, and approve long-running Codex coding sessions from a phone — without exposing the development machine to the public internet.
OpenAI and Broadcom unveil 'Jalapeño,' a custom chip built only for LLM inference
OpenAI revealed its first custom silicon, Jalapeño — a Broadcom-built ASIC designed to do one thing, run large language model inference, with engineering samples already in the lab and gigawatt-scale deployment targeted for late 2026.
Cursor Launches iOS App to Manage Coding Agents From Your Phone
Cursor released a native iOS app in public beta on June 29, letting developers launch, monitor, and merge work from AI coding agents remotely, days after parent company Anysphere agreed to a $60 billion acquisition by SpaceX.
US partially lifts Anthropic's Mythos 5 export ban, Fable 5 still blocked
The Trump administration reversed course on June 26, allowing Anthropic's Claude Mythos 5 to reach more than 100 vetted US companies and agencies after a two-week export ban shut the model down globally — but the more widely used Fable 5 remains off-limits.
MIT and Microsoft's 'Murakkab' cuts the cost and energy of running AI agents
Researchers from MIT and Microsoft Azure detailed Murakkab, a system that lets developers describe agentic workflows in plain language and then automatically picks the models, tools, and hardware to run them — using roughly a third of the compute and a quarter of the energy and cost of conventional approaches in tests.
RingCentral brings native AI agents to AIR Pro across its customer engagement portfolio
RingCentral expanded its AIR Pro platform with native AI agents that run multi-step voice and digital customer interactions end to end, handing off to human reps with full context. The capabilities are in beta now, with general availability targeted for the second half of 2026.
OpenAI agrees to stagger GPT-5.6 release as US government screens early access
OpenAI will limit the initial release of GPT-5.6 to a small set of government-approved partners after a request from the Trump administration, with federal officials approving customers one by one during the preview — a notable shift in how frontier models reach the public.
Assort Health raises $120M to scale voice AI agents across the patient journey
Healthcare AI agent startup Assort Health closed a $120M Series C led by Menlo Ventures at a $1.2 billion valuation, pushing total funding past $222M as its voice agents expand from appointment scheduling into a full patient-access platform.
Transformer co-inventor Noam Shazeer leaves Google for OpenAI
Noam Shazeer, a co-author of the 2017 'Attention Is All You Need' paper and co-lead of Google's Gemini models, announced he is joining OpenAI as Lead for Architecture Research — a departure that cost Google $2.7 billion to prevent just two years ago.
Google DeepMind publishes defence-in-depth roadmap for AI agents
Google DeepMind released a detailed AI Control Roadmap on June 18 that treats deployed agents as potential insider threats and outlines 15 system-level defences — including runtime monitoring, cryptographic action signing, and a kill switch — tested across roughly one million coding-agent tasks.