Business Insider reports that Google staff are testing an internal Gemini 4 variant called Carbon on its Jetski coding platform, with some employees saying it feels comparable to Anthropic’s Claude Opus 5.5 on coding tasks. The scoop, published October 9, 2026, comes as Google prepares to roll out Gemini 4 Argon to customers.
This article aggregates reporting from 1 news source. The TL;DR is AI-generated from original reporting. Race to AGI's analysis provides editorial context on implications for AGI development.
This leak matters because it shows how quickly the frontier labs are iterating beyond what they are willing to ship publicly. Argon is not even rolled out yet, and Google employees are already hands-on with a follow-on checkpoint, Carbon, that some compare to Anthropic’s latest Claude Opus 5.5 on coding work. That is the classic signature of a tight internal loop between research and deployment, where public releases always lag what the lab itself is using.
For the race to AGI, the signal is that Google is still pushing hard on highly capable coding and agentic models, even if public benchmarks briefly suggest it is behind Anthropic and OpenAI. Internal-only models like Carbon are often where labs experiment with more aggressive architectures, larger context windows and riskier training regimes before they bake that into customer-facing products. If Carbon really closes the coding gap with Anthropic, it will intensify the competition for enterprise agents and developer workflows, where marginal gains in reliability and tooling support drive real revenue.
It also reinforces a broader trend: the true performance frontier is drifting further from what is visible on leaderboards. Investors and policymakers should assume that when a lab announces “Model X,” there are already X+.1 and X+.2 variants being tested inside, and those are the ones shaping real-world risk and capability.



