A glowing neural-network path between two corporate towers, symbolising the defection of a top AI researcher.
News

Transformer Co-Inventor Shazeer Defects From Google to OpenAI

The man whose name is in 'Generative Pre-trained Transformer' is moving across the street to the company that put the GPT in those letters.

OpenAIGoogle DeepMindGeminiTalent WarsTransformers

Noam Shazeer — the co-inventor of the Transformer architecture and co-lead of Google’s Gemini project — is joining OpenAI, Reuters reported on 2026-06-18. Shazeer confirmed the move on X. It is the highest-profile defection from Google DeepMind to OpenAI since the frontier-model race began, and it lands the man whose name is literally inside the “T” in GPT on the side of the company that put the other three letters there.

🔍 The Bottom Line

OpenAI isn’t buying a researcher. It’s buying a co-author of the architecture that makes every modern LLM possible. For Google, losing Shazeer the same week it has to defend Gemini against a torrent of new releases is a structural blow — and the kind of signal that makes the next dozen DeepMind employees update their résumés.

What Happened

Reuters broke the story on the morning of 2026-06-18, and Shazeer confirmed the move on X within hours. He is leaving his role as co-lead of Google’s Gemini project — the position he returned to in late 2024 after the controversial $2.7 billion Google/Character.AI deal that brought him, his co-founders, and a slice of Character’s research staff back inside the Google umbrella under regulator scrutiny. The Hacker News thread carrying the news is already at 7 points and climbing.

Shazeer is not just a senior engineer. He is a co-author of the 2017 “Attention Is All You Need” paper that introduced the Transformer — the architecture that every frontier model, from GPT-5.5 to Gemini Robotics ER 1.6, is built on. He spent eight years at Google Brain, helped scale the Transformer into the T5 and LaMDA lines, then left in 2021 to co-found Character.AI. The 2024 “acqui-hire” that brought him back to Google was, in regulator-speak, a way for Google to license Character’s tech and re-hire its founders without technically acquiring the company. Critics called it a back-door takeover. Either reading, the deal kept Shazeer inside Google’s pay grade for less than two years.

Why Shazeer Matters

In a field where compute and data are increasingly commoditised, the genuine bottleneck is the small number of people who understand the Transformer at a level where they can evolve it. That pool is measured in dozens, not thousands. Shazeer is one of the most senior figures in it.

What he can do at OpenAI that he arguably could not do inside Google:

  • Architectural bets. Inside a public company shipping Gemini to a billion-user surface (Search, Workspace, Android), every model change is a risk-committee decision. OpenAI’s pre-IPO research culture still tolerates bigger swings per release.
  • Post-Transformer research. The Transformer is twelve years old. The next jump — whether it is a hybrid state-space model, a learned-memory architecture, or something Shazeer has not yet published — will come from someone with the standing to push the field. OpenAI is now buying that standing.
  • Recruiting signal. A defection at this level resets the salary expectations and the “where is the frontier” narrative for every junior researcher at DeepMind, Anthropic, and Meta FAIR.

The move also widens a pattern this site has covered before: OpenAI is the buyer-of-last-resort for top-tier lab talent. OpenAI’s $122B valuation gives it the equity currency to out-bid Google on the people it cares about — and Google’s bureaucracy, post-2024 restructuring, makes it harder to match the offer structure (signing bonus, equity vesting, non-compete carve-outs) that frontier researchers now expect.

The Character.AI Detour

The 2024 Google/Character.AI deal is worth re-reading in the light of this news. The structure was a “licensing and talent” arrangement — Google paid $2.7 billion, mostly to Character’s investors, in exchange for a non-exclusive technology licence and the right to re-hire Shazeer, Daniel De Freitas, and a handful of senior researchers. The FTC opened a preliminary review. Senator Elizabeth Warren’s office issued a statement. Character.AI continued to operate as a “consumer-facing personality AI” product on Google’s cloud.

The point of the structure, in hindsight, was to keep Shazeer inside Google’s pay grade while the legal status of the deal was litigated. It did not keep him for long. Shazeer’s move to OpenAI — reportedly with broader latitude on research direction and a compensation package that sources close to the deal describe as “the largest single-researcher signing of the cycle” — suggests that the Character.AI detour was, for Shazeer personally, a brief side-quest rather than a destination.

Google’s Loss

For Google, this lands at the worst possible moment. Gemini 3 shipped to mixed reviews in May 2026. The new robotics-ER line is good but is competing with NVIDIA’s GR00T and Tesla’s Optimus demos for the embodied-AI mindshare. The 2025 Cloud AI growth that Google was leaning on to justify capex has begun to plateau. Losing the co-lead of the project that produced Gemini 2.5 and the Gemini ER family is, in the words of one former DeepMind staffer quoted in the Reuters thread, “the kind of thing that makes a VP start writing a retention memo at 2am.”

The structural problem is real. Google has more AI researchers than any other company on earth, and the per-capita output has been falling for two years. Shazeer’s exit is the highest-profile resignation in that trend — but it is not the first. Three other Gemini-team senior researchers have left for OpenAI, Anthropic, or xAI in the last six months. None of those moves were individually decisive. The Shazeer move is, because the founder-of-Transformer angle is the one a journalist can write a lede about.

The Bigger Picture

This is the third major talent-move story of 2026, after the Musk/Altman trial and the Anthropic enterprise overtake. The pattern is the same each time: the AI race is being decided less by the size of the next training cluster and more by the quality of the 30-to-50 people who decide what to do with it. Capital is plentiful. Compute is approaching plentiful. The genuinely scarce input is the researcher who has shipped a frontier model and could plausibly ship the next one.

OpenAI is the only lab that has consistently won the bidding for those researchers. Anthropic has poached several, but it has also lost several. xAI has paid heavily, but the Mythos access questions have made some senior researchers unwilling to sign. Google has the cash, but the post-Microsoft structure of the AI partnerships has made it harder for Google to win talent by offering the “you get to ship to a billion users” pitch that worked in 2018. Shazeer’s move is the cleanest data point yet that, in 2026, the best frontier researchers believe the frontier is being set at OpenAI — and they want to be in the room.

NZ Angle

New Zealand does not have a Shazeer problem. The country has roughly a dozen researchers with the citation record to be on a frontier-lab shortlist, and most of them are already abroad. The local story is downstream of this one.

What it means for a Kiwi AI founder or engineer:

  • Application layers, not foundation models. The barrier to building a competitive frontier model is now effectively closed to anyone who is not at OpenAI, Anthropic, Google, Meta, or xAI. The interesting product work is at the application layer — fine-tuning, agentic orchestration, domain-specific retrieval. That is where local agentic plays like the Spark/Xero integration live.
  • Talent-return pressure. Senior Kiwi AI researchers working in San Francisco, London, or Singapore will read this news and think about whether the OpenAI offer is the ceiling. A handful of those conversations, in 2026 and 2027, will end in returns. The opportunity for a New Zealand-based frontier-application lab is to be ready to hire them when they land.
  • Compute access. When the top talent concentrates in two or three labs, the compute those labs do not consume flows to second-tier players. NZ’s role in the regional APAC compute market is to be a credible buyer of the capacity the top labs release — and the few local compute operators that do exist (community Pi clusters, university GPU pools) become more strategically interesting, not less.

The cynical reading: the AI race is becoming a global oligopoly of three or four US-headquartered labs, and the rest of the world is a customer. The slightly-less-cynical reading: a small country with strong application talent and a tight regulatory environment can be a high-margin niche player. Shazeer’s move sharpens which of those two readings is correct.

What Happens Next

OpenAI will name Shazeer’s role inside the next 30 days. The standard options are: Distinguished Research Scientist with broad latitude (the Sutskever pattern, pre-SSI), head of a new architecture team reporting to the Chief Scientist, or a joint title with the Head of Pre-Training. Watch for the announcement language — it will signal whether OpenAI is hiring Shazeer to ship a faster GPT-5.6 or to start a parallel track on a post-Transformer architecture.

Google will not comment substantively, per its standard policy on senior-departure news. The interesting tell will be the Gemini 3.1 release notes — whether the project accelerates, slips, or quietly restructures around the departure. The next quarterly Alphabet earnings call will be the first time Sundar Pichai has to answer an analyst question about it.

❓ FAQ

Who is Noam Shazeer? Shazeer is an American computer scientist and a co-author of the 2017 “Attention Is All You Need” paper that introduced the Transformer architecture. He spent 2000–2021 at Google Brain (where he co-invented T5 and led the Meena/LaMDA work that became Bard), co-founded Character.AI in 2021, returned to Google in 2024 as co-lead of Gemini, and is now joining OpenAI.

What is the Transformer, and why does his co-invention matter? The Transformer is the neural-network architecture that underpins essentially every modern large language model — including the “T” in GPT (Generative Pre-trained Transformer). Co-inventing it means Shazeer is one of a small number of people with the standing to argue for the next architecture shift. His move is therefore a leading indicator of where OpenAI believes the next performance jump will come from.

Why was the 2024 Google/Character.AI deal controversial? The $2.7 billion transaction was structured as a “licensing and talent” deal rather than a full acquisition, in part to avoid antitrust review. Critics — including the FTC and several members of the US Senate — argued the structure was designed to deliver the same outcome (Google absorbing Character’s IP and rehiring its founders) while sidestepping merger scrutiny. The deal was the subject of a preliminary FTC review in late 2024.

What does this mean for Google’s Gemini roadmap? In the short term, the loss of a senior technical leader is a morale and momentum hit — particularly during a release window where Gemini 3 has not yet hit the milestone the team was aiming for. In the medium term, Google’s larger bench of senior researchers means the project will continue, though it may slow. The bigger risk is the recruiting signal: every DeepMind researcher who was on the fence about leaving now has a fresh data point.

Is New Zealand affected? Not directly — Shazeer’s move is between two US labs. Indirectly, it confirms that frontier-model talent is concentrating in two or three labs globally, which sharpens the strategic question for NZ’s small AI sector: compete on applications, or compete on talent-return logistics for Kiwis who have been working abroad.

Sources

Sources: Reuters, Hacker News, X (Noam Shazeer)