+ Gemini 4 Argon: tops 13 of 19 rows
- nobody outside can use it
+ Netlify Edge Functions: ~5x fasterYou'd think a launch means you get the model. Google just launched Gemini four Argon, and unless you're a trusted cyber defender, you can't touch it. You've probably seen the hype videos already. I read the benchmark table instead, and two rows in it never made it into the post's text. I'll get to them at the end. Two stories today, and the big one is Gemini four Argon, which Google calls its next era of frontier intelligence. One answer can now run to a million tokens, up from sixty-four thousand, which is several novels in a single reply.
The launch price is two dollars per million tokens in and ten out. That's exactly what OpenAI charges for GPT six point one Sol, the model I covered yesterday, and Sol is one you can actually buy. Inside Google, it's already at work. Google says Argon agents are porting C and C++ to Rust across the company, up to eight hundred thousand lines for the Fuchsia kernel. On Hacker News, one engineer remembered when Google's C++ team wouldn't even consider Rust and looked at Carbon instead. Now a model is doing the migration they argued about. Now the table. Google puts Argon next to GPT six Astra, Claude Fable and Claude Opus on nineteen rows, and Argon comes out on top in thirteen of them. The headline is DeepSWE, a long coding benchmark, at seventy-eight percent, about four points clear.
The strangest win is legal drafting, where Argon scores twenty percent and the others stay in single digits. It's the best grade on a test that everyone fails, and on the slide-deck benchmarks, which are undefeated, that counts. The real pitch is security. Google says Argon can find, validate and patch critical vulnerabilities on its own, and that trusted defenders get it without cyber guardrails. Wiz used it to find a critical hole in hospital software that earlier models had missed. One small detail. Wiz belongs to Google, since a thirty-two billion dollar deal closed in March. So the black-box hacking test in the post is a Google company grading a Google model. And on the bug-fixing leaderboard, three models share sixty-eight percent, and the bar on the far left belongs to Grok.
So what did outside testers measure? The independent leaderboard I could find with Argon on it is Artificial Analysis. It scores Argon fifty-three on its intelligence index, on the high setting, which ties it with Astra and Fable. Claude Opus five point five sits five points ahead. A week ago I told you Opus took the top spot there, and it's still holding it. The fair part for Google is the bill. An Argon run costs about two dollars per task on that index, roughly a third of what Opus costs. And when do you get it? Google won't give a date. The post promises access for developers, enterprises and consumers as soon as possible, with paid API customers first. Sundar Pichai's post on X says, so hold tight.
Until then, access runs through Fairwind, the program Google started a month ago with Gemini three point eight Flash Cyber. It has over six hundred fifty partners, and they may only hand Argon to their security teams and must track who uses it. Hacker News took it well. One commenter wrote, Gemini not beating the can't release a model allegations. Another predicted that models turn into vaporware, a bunch of numbers on a table. Meanwhile, Netlify rebuilt Edge Functions, which run about a billion times a day. They moved from V8 isolates in a hosted service to Firecracker microVMs inside Netlify's own network, and a warm call dropped from up to forty milliseconds to about six.
Hacker News pointed out that Cloudflare Workers are V8 isolates too and run much faster, so most of the win looks like the request no longer leaving the building. Another commenter worried about the snapshots, because cloned microVMs can share random number state, which is how you get two identical UUIDs. A cold start, when a region has never seen your function, hits about one percent of calls and takes around nine milliseconds. And the microVM itself is Firecracker, which Amazon built for Lambda, so one commenter suggests you remember that the next time you curse AWS. Now, those two rows. On FrontierSWE and Terminal-bench, two coding tests the post's text never mentions, Argon comes last of four, on Google's own table. Terminal-bench puts an agent in a real shell, which is how most of us would use it, and Opus leads it by nine points.
Verdict: NEEDS REVIEW — a press release I can't call yet
Sources
https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-4-argon/
https://news.ycombinator.com/item?id=49913571
https://artificialanalysis.ai/models/gemini-4-argon
https://cloud.google.com/blog/products/identity-security/google-completes-acquisition-of-wiz
https://techcrunch.com/2026/03/11/google-completes-32b-acquisition-of-wiz/
https://deepmind.google/fairwind-program/
https://x.com/sundarpichai/status/2105387952478277979
https://x.com/sundarpichai/status/2105387954474746264
https://x.com/demishassabis/status/2105417239432200636
https://x.com/GoogleDeepMind/status/2105388084154056939
https://x.com/koraykv/status/2105392843120611660
https://www.netlify.com/blog/edge-functions-firecracker-microvms/
https://news.ycombinator.com/item?id=49912444
And that's the diff for today. I'm Niko from Axrisi. Merge responsibly.
YouTube · thedailydiff.dev · forward this to the intern who deployed on Friday.

