43 comments

  • revolvingthrow 13 minutes ago

    I find the Sunday roast comparison of 5.6 vs 6 very interesting. I have no doubt most people will prefer 6, yet I am almost repulsed by all the images, so much needless whitespace, checklist and so on. Feels like I'm being condescended to and treated like a child.

    Given that OpenAI is making noises about merging work with chat (a horrible idea imo), and Work is very similar to Codex... I dearly hope things like these won't have any meaningful cross-polination into the actual work tools.

    Seeing that the chat is based on 6.0 and not 6.1 is disappointing. The "Visual and interactive explanations" seems genuinely useful, but 6.1 is just so much better. I wouldn't truly trust the 6.1 with the explanations, but I'd trust them a fair bit more than 6.0. I understand that compute isn't infinite, but tons of people only interact with the chat and having your "things-explainer" be as good as it can be is important when people increasingly treat AI models as the source of truth, or even use them for academic learning and whatnot.

    Still, the models will improve, so the visual explainer seems pretty good as an idea / mvp.

      361994752 4 minutes ago

      I really hope they don't merge chat and work. That's kinda the only edge they have over Anthropic at this time point...

  • mortenjorck 18 minutes ago

    Of everything that could be automated, Bartosz Ciechanowski really was the last on my list.

    In all seriousness, his lovingly and expertly crafted explainers are still going to age like a handcrafted heirloom clock in a world of plastic-clad quartz movements. But it’s absolutely incredible that we are now in an age where a computer can manufacture a serviceable interactive explainer on whatever niche topic you desire.

      FLeXMurphy 7 minutes ago

      The reason why B.C. became a thing is because the art of drafting died from CAD. The attention to detail, the minutae of walking the reader through a highly sophisticated thing was replaced by short-form video explanations. His work is very much a callback to the days of old. So too, is his turn to be relegated to a relic of his time.

        bayindirh 3 minutes ago

        Also, it's hand coded WebGL. It's smooth, faultless and has no peers.

        > So too, is his turn to be relegated to a relic of his time.

        No, it'll be tasteful artifact, not a relic. Records, fountain pens and automatic watches did not die. They are used by people who discern things, and no, none of these things have to be expensive (i.e. Neither Seiko 5, nor Lamy Safari are expensive, yet they are as dependable as their 100x expensive brethren).

        Human touch still has that finesse and warmth.

  • xpct 23 minutes ago

    I've had the most success with GPT explaining things to me by making it take a few sentences at a time back and forth, instead of reading full write-ups of whatever I asked. It also often poisons the conversation if it misunderstood some part of the question, and I can lead it better by continuously questioning its statements. It's also more engaging that way.

    I've been learning music lately and it kept re-pasting the same one chord visualization throughout many conversations, almost randomly and often barely related to the question. So I at least hope this won't be as aggressive so I can prompt it away!

  • jjcm 9 minutes ago

    I've been calling this "disposable UI" or "paper plate UI", eg something meant to be used once. One thing I'll be curious about is overzealousness to produce this, when sometimes what you want is just a simple response. Overall though I'm a big fan of it, if it can be provided fast enough. I'd be curious on how much impact it has on latency of a response.

  • franze a few seconds ago

    "If you call something intelligent you know what it isn't."

  • julesrms 5 minutes ago

    They seem to be pitching something that pretty much all the decent models can already do..?

    Obviously if your model is stuck inside a CLI terminal, then not so much. But in a GUI harness (shameless plug for my own one: https://juggler.studio, but I assume others can do this too), you just ask them to answer in HTML and they'll happily draw pretty pictures inline in the conversation. I've been doing this for ages with claude, GPT, Deepseek and others.

  • Tiberium 35 minutes ago

    I find it extremely strange that they're adding GPT-6 Sol to Chat over GPT-6.1 Sol which is significantly more capable.

      MikhailTal 6 minutes ago

      6.1 is Astra minor. Way more capable but also way heavier+slower. Its really 2 different models, they just shipped it as sol to recover from the gpt6 disaster lunch, where they tried to pass terra 6(or a cheaper model) as sol but it was worse than expected

      xpct 29 minutes ago

      I really don't like having to open Work sessions for one-off questions, only because they're limiting what models they put in the chat. The older models are just too dumb for some things.

      exitb 8 minutes ago

      Given the extreme short time between 6 Sol and 6.1 Sol, I suspect they don’t actually have much in common and 6.1 is a heavier model rebranded as Sol in a panic response to poor agentic capabilities of 6.

      wincy 31 minutes ago

      I’ve hit model not available limits for GPT 6.1 Sol multiple times over the last week. It’s never for very long, and I can switch to Astra, but it seems like OpenAI is struggling with capacity. This has happened with my personal $200 Pro connection and my Codex enterprise connection.

  • yread 9 minutes ago

    1.2B weekly active users? WAU!

  • kingstnap 21 minutes ago

    GPT-6's design sense is kind of ridiculous imo.

    I have explicit instructions to tone it down. Less taglines, eyebrow text, subheadings, decorative spacing, pills, cards.

    Hopefully this doesn't bleed into the chat...

      briga 13 minutes ago

      More UI elements == more tokens == more money for OpenAI

      It's a pretty clever way to sell more tokens, I have to admit

        wvenable 10 minutes ago

        Except that chat is a fixed monthly fee. More tokens = more cost for them.

  • cameronh90 5 minutes ago

    Is there any relationship between this and AG-UI or A2UI?

  • solarkraft 37 minutes ago

    AFAIK, they already had a simpler form of this. It was kind of an obvious next step. Now connect it up to tools so that we can again comfortably do the things that are more precise by hand! The “pick a color” or “select the width on a slider” use case is coming closer.

  • haute_cuisine 28 minutes ago

    In the video, they showed examples of ChatGPT making interactive tutorials on how to fold origami, how to arrange colours/interior and how to assemble a bike.

    Supposedly, people were struggling to follow written manuals and they needed an interactive explanations.

    I'm not a mathematician and I would certainly love having a tool that would do ELI5 on some complex stuff, but I'm really worrying about using this too often and outsourcing my ability to do stuff to some mega corp.

  • blakeashleyjr 23 minutes ago

    This seems like a natural progression of models becoming better at frontend coding in general.

    "Here is a library of [svelte/react/whatever] components, use them to construct a helpful visual to demonstrate your point."

    The deconstructed bike at the beginning was in a class of its own, however.

  • throwaway7783 40 minutes ago

    Is this OpenAI catching up with Anthropic artifacts? At the same time they say "We’ve trained GPT‑6 to compose responses using text, visuals and interactive elements...", rather than a harness.

  • wincy 28 minutes ago

    It makes sense they’re doing this - I’ve noticed lately when asking a more complex question involving a lot of nonlinear data using Codex work mode, Astra and Sol will write a fully html document to better display the info with a Cliff’s Notes version in chat.

  • utilize1808 19 minutes ago

    The irony really is that LLMs are partly responsible for the walls and walls of text as seen in the video in the first place.

    And now we are asking LLMs to solve it.

  • abroszka33 27 minutes ago

    The intro video is just lame. Those things never happen in real life.

      twoodfin 18 minutes ago

      They're making the point—apparently too subtly—that text alone is a highly limiting "UI" for many tasks. So it's great that GPT-6 can now communicate in a a richer interactive medium.

      bogdiyan 22 minutes ago

      So you never have to assemble a bike or paint a room. How those are not real world examples? I find their video pretty cool. For people with ADHD or people learning by watching or children this is spot on.

  • robertlagrant 21 minutes ago

    I wish things like bikes came with mostly-written manuals. Maybe a few diagrams. If anything, things are too pictorial these days.

  • hollowturtle 25 minutes ago

    > The compiler allows the interface to appear progressively as the model generates it, without waiting for the entire response to be complete.

    why not just stream html?

      Aarostotle 19 minutes ago

      The element can’t render until it’s closed, presumably.

        hollowturtle 18 minutes ago

        Html partial streaming is a real thing, presumably

  • fang2hou 32 minutes ago

    I'm also building something similar for an internal project, based on the vercel's json-render design. It hasn't been too challenging, especially since the release of faster models like 5.6 Luna.

  • maherbeg 22 minutes ago

    Love it. I've been using $visualize a lot in the codex desktop app, and having even richer experiences will be sweet.

  • 2sk21 22 minutes ago

    How would a user know whether a generated visualization to illustrate some process is accurate or not?

      hollowturtle 20 minutes ago

      It wont they're just trying making the chat look like more a session iron man would have with jarvis

  • BeetleB 25 minutes ago

    So I guess it's going to nail the bicycle the Pelican rides on, right?

  • bogdiyan 15 minutes ago

    I read the comments here and I am really surprised by many people. One picture is worth a thousand words so with better UI you can make much better user experiences. This unlocks personal assistants to be better adopted by elderly or disabled people. I see so many benefits of it and having models which can do this (if they can do it constantly with good quality) is amazing. Much better products - I am really tired of dumb chatbot - if I can do something with one button or view the whole information in one diagram/image, this is amazing.

  • tamimio 32 minutes ago

    This looks bad, I want the output to be as much as text based so I can easily export it and further processing it, plus, this might make the resource-eating app even worse.

  • fsniper 24 minutes ago

    Is it me or GPT-6 Instant food answer gives the vibe to check the hell out immediately? It's like the it will start to explain a sunny Sunday afternoon from 20 years ago for 5 full pages.

  • simianwords 27 minutes ago

    I made a prediction last year that there'd be a new UI protocol (like HTML) but for agents. I believe that custom or personalised UI's are going to be the new browser interface. More radically, I think browsers can be completely replaced. My news feed can be personalised to me, based on what AI thinks might be important from all sources like Reddit, X, HN.

  • topsykreet 24 minutes ago

    Written manuals/guides come with pictures. The ad is dishonest.

      password54321 20 minutes ago

      Plot twist: The guides were written by ChatGPT. But now you can use ChatGPT to solve problems by ChatGPT.