Gemini 4 Argon

201 points | by bradleyg223 20 minutes ago

74 comments

  • taylorfinley 8 minutes ago

    Ten days ago I had an experience with Gemini 3.8 flash that made me wonder if I was being routed to a different model under test. I was trying to use rocm with llama.cpp on my 128gb Strix Halo but could only get it to run Vulkan. I pasted the error message into agy and it proceeded to attach GDB to my GPU driver, reverse-engineer the kernel queue ioctl interface, and author an LD_PRELOAD C shim to get ROCm llama.cpp working on my Strix Halo. My jaw was hanging open the whole time.

      IndeanCondor a minute ago

      Can confirm, I was doing a routine internet search thing for a curiosity 3 days ago (about the only thing I used Gemini for) and was surprised by how suddenly thorough and quality the response seemed, almost overnight.

      spankalee 2 minutes ago

      3.8 Flash is just quite good, and so is the Antigravity harness.

      I use a mix of Fable 5.1, Opus 5.5, and Gemini 3.8 Flash and Gemini holds it's own. Especially in writing, frontend, and sysadmin work. agy for configuring a NixOS system has been truly incredible.

  • nickysielicki 5 minutes ago

    The important take away here: the leapfrogging we’ve seen this year doesn’t seem to be a temporary thing. The famous theory of Dario Amodei was that AI was this winner-takes-all field where the first team to get a head start would never cede ground back. The term he liked to use was, “concentrating”. This is yet another datapoint that he was wrong about that. AI seems more distributed amongst neoclouds and traditional hyperscalers, FAANG and startups, GPUs and ASICs than it did this time a year ago.

    Nobody has a moat.

  • babelfish 18 minutes ago

    > We’ll continue to gather feedback from early testers as we iterate on guardrails before making Argon available to developers, enterprises, and consumers as soon as possible.

    Gemini not beating the "can't release a model" allegations

      modeless 15 minutes ago

      When I said I was tired of Google launching waitlists I didn't think they would respond by simply not having a waitlist.

        ionwake a minute ago

        i know this is like "hey guys we got such a cool thing at home ,its rad and uhm we playing with it with our friends"

        ok bro thx

  • SwellJoe 16 minutes ago

    My girlfriend, you wouldn't have met her, she lives in Canada, has seen it and she thinks Gemini 4 Argon is amazing.

      hn_acc1 a few seconds ago

      I know someone who works for Google Canada with AI. Her parents and mine were friends and some thought something might happen there at one point in time..

      jastanton 9 minutes ago

      HA, this might be my favorite HN comment. Well done

      blueaquilae 3 minutes ago

      My grandma saw it too, it's really secure more than Astra 6.1 but she asked me to not talk about it.

      greenchair 9 minutes ago

      my uncle who works at nintendo said the same thing!

  • tazjin 14 minutes ago

    > Argon agents are working on migrating C/C++ codebases to Rust across Google

    Man, I remember back in the days when the cppnext team was refusing to even consider Rust, instead looking at absurd stuff like Carbon and Swift (!), even though half of the engineering staff already knew where this was headed. I hope they got a few good promos out of the delays at least.

      baq 12 minutes ago

      Not many people can hold grudges as strong as principal engineers

      timmg 12 minutes ago

      I wonder if this means Carbon is DOA.

      I was excited to see what it would be. But I don't think I can argue that it makes as much sense anymore.

        qalmakka 7 minutes ago

        Carbon was clearly DOA the moment it was announced, IMHO. It looked cool but it served none but Google, and now with LLMs you have a massive incentive not to use a niche or new language due to how better LLMs get the bigger the corpus is

        The only somewhat realistic proposal in this space is Herb Sutter's cpp2, which is arguably a massive improvement and I'm puzzled why nobody in the standard thought to give it a spin, there's just to much cruft they'll never be able to get rid of unless they make an alternate yet backward compatible syntax with C++ that changes the defaults from "random 80s nonsense" to something better

      vovavili 11 minutes ago

      What exactly makes Carbon absurd?

        boshalfoshal 5 minutes ago

        There is 0 practicality in inventing an entirely new coding language that only one company uses, and you have to teach it to thousands of new engineers. Rust exists and fits the job totally fine and is used in more places and has actual support outside of a single entity (i.e you can actually hire people that feasibly know the language).

        It was clearly done because some PL guys at google really wanted to make a new cool language and Google was the perfect place to incubate it without it getting axed. Probably got a couple of promos out of it too. This is clearly not the best use of time or money, but I guess if you're google you have so much of both it probably doesn't really make a dent, and you can keep a few very smart people happy with shiny new projects.

        Also, LLMs being used for a large portion of coding nowadays sort of remove the need for these types of languages, IMO. They make less "silly" mistakes (both logical and structural), and they are much better at languages that are better represented in the training corpus. This somewhat obviates the need for very niche "type/dummy-safe" languages like carbon (and even rust/zig, imo). So even if you did want to use Carbon, you'd likely have to bootstrap a decent amount of your own "good" carbon code to post train an LLM, and even then, it likely won't have that big of a gain vs just having an LLM write C++ or even Rust. If you are a company that still reviews code, you should just have an LLM code in a language most people can understand anyway to make verifiability tractable.

        bvinc 4 minutes ago

        I’m not op. But I think it’s not Carbon itself that is absurd.

        It’s absurd to think that Carbon is the solution to memory safety when rust exists and Carbon’s memory safety story is basically “TBD”.

        fg137 5 minutes ago

        I wouldn't call it absurd, but very questionable at least. Most companies are not going to even consider throwing money at this adventure.

        Maxatar 6 minutes ago

        The fact that it will never exist.

        gorbot 5 minutes ago

        rust's existence?

      ChickeNES 12 minutes ago

      Heh, I use my clankers to rewrite Rust in C

  • gopalv 14 minutes ago

    > taking careful precautions against feeding the findings back into training so as to not risk shaping Argon’s reasoning to evade our monitoring. We strongly encourage the rest of the industry to preserve reasoning transparency in these pivotal moments of increased capabilities while navigating alignment risks, so that model thoughts remain helpful in identifying and diagnosing misalignment.

    This is good, but they're the slow mover due to this exact thing.

    Google is getting punished for not letting the models enter an echo chamber and go faster than humanly possible.

      polotics 9 minutes ago

      Mmh ok. How much theoretical speed or 'intelligence' gain is realized by allowing reasoning to occur in some inscrutable intermediate representation? Has this been actually tested, how much is it slowing them down, and compared to whom exactly?

      janustimes 3 minutes ago

      OpenAI is the company that originally proposed and popularized chain-of-thought monitoring: https://openai.com/index/chain-of-thought-monitoring/

      So no, Google is not being punished, nor are they the people behind this technique.

  • elAhmo 4 minutes ago

    > Quantum algorithmic optimization: Argon is helping our quantum computing researchers optimize the spacetime resources (qubits × gates) of subroutines that bottleneck important applications. In one example, it beat the published baseline by 40% in a matter of minutes.

    Amazing breakthrough! So useful in day to day life, glad they put this as the first bullet of how it is making changes at Google.

  • arjunchint 4 minutes ago

    I dont get it, why even make this announcement, nothing's available and only one real benchmark for comparison?

    Only theory is team wanted this out before perf/promo reviews to kick it over the line and then its not their problem

  • iamronaldo 17 minutes ago

    Argon will launch at an introductory price of $2 per million input tokens and $10 per million output tokens, with cached input tokens priced at 95% off input token price. Wow

      LucasBrandt 15 minutes ago

      5x cheaper than Astra for input and output, 10x cheaper for cached input.

        h14h 4 minutes ago

        watch it somehow use 20x more tokens tho

  • darksaints 3 minutes ago

    > Argon agents are working on migrating C/C++ codebases to Rust across Google

    If anybody at google is reading this, please please pretty please prioritize or-tools. I absolutely love the project and use it all the time, but for the entire life of the project they've never had a repeatable working build system, and the whole SWIG framework is a nightmare to deal with. There's so much potential as an open source project, and a lot of external researchers would love to contribute, but the codebase is an example of everything wrong with the C++ ecosystem.

  • jjcm 14 minutes ago

    Big number results, and impressive pricing. That said it really feels like benchmarks have been hyper saturated these days. I’ll wait for hands on before getting too hyped that Google is back. It would be nice having more than just OAI / A\ in the running for SOTA top tier intelligence.

      nurettin 4 minutes ago

      With these numbers, I'm holding my breath for the pelicanbench.

  • helsinkiandrew 5 minutes ago

    > Google Grapples With Employee Skepticism About New Gemini Model

    https://www.bloomberg.com/news/articles/2026-09-30/google-gr...

  • bottlepalm 16 minutes ago

    Gemini is the model that is routinely borderline psychotic. It scares me. If we get paperclipped I won't be surprised if it's Gemini.

      eamsen 5 minutes ago

      Anecdote: Gemini 3.5 casually added a DROP TABLE for an actual production table in a system test.

      It had previously attempted to create that table as part of the test setup, so it apparently concluded that it was a test table.

      During human review, it explained that it had simply chosen a table name inspired by the codebase.

      rsstack 11 minutes ago

      If there's a company that culturally doesn't understand alignment, on a human or systemic or AI-research level, it's going to be Google. (or Oracle, but they're not in this race)

      colordrops 14 minutes ago

      Examples? What makes you say thatm?

        Scrapemist 7 minutes ago

        Experience? Ask it to write a prompt to generate an image and it generates an image instead.

      polotics 7 minutes ago

      traces or it didn't happen!

  • dom96 7 minutes ago

    Why announce this if it’s not available yet? Why not at least announce when it will be released to the public?

    None of the other AI labs do this. Really frustrating.

  • scirob 8 minutes ago

    "Rolling out soon" don't let them hype without any release

  • netdur 7 minutes ago

    I started my antigravity ide and I do not see gemini 4 there, does it mean google need government approval?

      tom1337 4 minutes ago

      Are you enrolled in Fairwind?

      > Today, we’re announcing our new frontier model, Gemini 4 Argon, which is rolling out to a set of trusted cyber defenders through our Fairwind Program.

  • osiris970 7 minutes ago

    Hopefully their harnesses aren't unusable when they release this

  • bananaflag 7 minutes ago

    I wonder how it will be at solving open math problems.

  • lanthissa 10 minutes ago

    deepswe vs frontierswe spread is huge.

    I think that should be a really bad sign, but hope its great.

  • wewewedxfgdf 13 minutes ago

    Gemini is so far behind that it is effectively useless compared to Claude.

    It's a surprise that Google has let themselves lose the game given their infinite cash, massive computing resource, gargantuan information store/training data, and vast number of programmers.

    The truckloads of ads revenue mean they don't have the single focus drive needed to win.

      jjice 10 minutes ago

      We're like 3.5 years into this new era - I'm not counting winners or losers yet.

      dhdjcjcjnd a minute ago

      Google's strategy is to let their competitors bankrupt themselves while they continue to offer good-enough models near breakeven.

      mattlondon 5 minutes ago

      How is it far behind? The benchmarks published in the blog post show it is superior to Opus 5.5 and Astra 6?

      Behind how?

        wewewedxfgdf 2 minutes ago

        Within one question of their web interface, it has lost context and asks you to clarify what you are talking about.

      bel8 8 minutes ago

      I wonder if Google bans internal use of Claude/Codex.

      And I wonder if Google's main monorepo is already in Anthropic/OpenAI training data because of some stubborn dev.

      VirusNewbie 11 minutes ago

      I use it and claude back and forth and Argon is better imo.

        handfuloflight 9 minutes ago

        You have access to Argon?

          matthewfcarlson 4 minutes ago

          Their profile says: > Currently at Google as a Sr. SWE SRE on the cloud.

          osti 7 minutes ago

          Google employees do.

      LoganDark 8 minutes ago

      I've tasted Gemini through an intermediary and it feels far better at attention to detail than other models I've tested (Claude Opus/Sonnet, GPT whatever it's called nowadays). But it's less likely to get one-shots right.

  • linksbro 15 minutes ago

    Personally, I'm waiting for Gemini Krypton, Xenon, and Radon.

    Jokes aside, looks like an impressive model!

  • kccqzy 15 minutes ago

    Unfortunately it’s not actually released yet to mere mortals.

  • LoganDark 9 minutes ago

    Is there a way to use Gemini models without linking your usage to your personal Google account yet?

      alehlopeh 3 minutes ago

      Use your work google account

  • Razengan 5 minutes ago

    Oh we're down to gas names now?

    Goshdarnit they didn't see my suggestion: https://news.ycombinator.com/item?id=49899171

  • jasonjmcghee 13 minutes ago

    > 1M output token limit

    what about input?

    (Maybe I missed it)

      murkt a minute ago

      Input token limit is 1M for Gemini models for a long time. Haven’t they been the first with 1M input?

  • retropragma 6 minutes ago

    no Pareto frontier graph?

  • TacticalCoder 13 minutes ago

    > Large Scale Codebase Migrations and Optimizations: Argon agents are working on migrating C/C++ codebases to Rust across Google

    So Google is migrating codebases from C to Rust? That is interesting...

  • nikope 12 minutes ago

    Looks like an impressive model

  • FranzFerdiNaN 5 minutes ago

    Can’t wait to get my hands on yet another model that’s only good coding, because clearly that’s what the world needs.

    I still miss the days of Sonnet 4.5 and 4o, those models were actually good at creating stories and writing text that was actually readable by a human being.

  • tamimio 10 minutes ago

    Now AI models will turn into vaporware, a bunch of numbers on a table without even releasing the model, because it’s toooo scary to release!

  • VirusNewbie 12 minutes ago

    It's fucking insanely good.

  • gravisultra 7 minutes ago

    Google has the audacity to "protect us from ourselves" and talk about "safety" and in the very same blog post highlight the Israeli "security" company Wiz, that they acquired for a very exaggerated sum of money.

    This is why I will never take any of these leading model houses seriously when they talk about alignment. They are literally complicit in genocide and the worst crimes against humanity imaginable.

  • pliiight 15 minutes ago

    Hate to say i will never be touching this model for anything except for youtube video understanding