GPT-6 Astra

(openai.com)

520 points | by kibae 1 hour ago

5 comments

  • dang 2 minutes ago
    Argh! I hit a wrong keyboard shortcut and moved the entire thread.

    Please stand by... it will all come back shortly

    • the_duke 1 minute ago
      500 upvotes with 2 comments would have been a new record. ;)
  • abixb 2 minutes ago
    I want to take a step back: So, this is GPT-6 -- the natural number version release comparable to GPT-4 and GPT-5 from the past few years. The ARC-AGI-3 score is obviously impressive at 99.9% (we'll need to wait for more details on how they used the response API harness on GPT-6 Astra, wrt reasoning retention and compaction), but every other benchmarks seems to be a relatively modest improvement, comparable with any of the 'point' updates from AI labs.

    If this is truly AGI (subject to one's definition of AGI still), then this is a very boring release of an AGI model. No video announcement, no presser, just a blog post (with some Twitter promo vids)?

    As others mentioned, I'm starting to think OpenAI was under immense pressure to deliver an 'AGI' model for certain contractual reasons, but I never expected GPT-6 release to be this mundane and banal.

  • herpdyderp 1 minute ago
    The FrontierCode 1.1 Extended benchmark is the only benchmark that aligns with my actual LLM experiences and Astra isn't significantly better or cheaper. All this celebration, and yet it's only on-par with an already existing model? I don't get it.
  • damsta 1 minute ago
    Why release it now instead waiting those few days until it is available for everybody?
  • cesarvarela 44 minutes ago
    So, on the one hand, we have AGI; on the other, the release page is returning 500s.
    • aldanor 35 minutes ago
      Artificial general incompetence?
      • nazgulsenpai 22 minutes ago
        I believe this is naturally occurring
      • baq 14 minutes ago
        Nobody can escape the dilbert phb law
      • ReptileMan 15 minutes ago
        Artificial intelligence will be no match for natural stupidity.
    • glenstein 21 minutes ago
      I think something about its benefiting from a harness is significantly contributing to its performance, though I honestly am not smart enough to know whether that makes it "count" more or less.
    • ignoramous 18 minutes ago
    • kingleopold 28 minutes ago
      just because they have smart AI, it does not mean they are great at software and cdn.
      • nemomarx 18 minutes ago
        if it's really smart AI you might expect them to have it create the software and cdn

        don't they use their coding agents internally?

      • stigz 24 minutes ago
        If they can't get the blocking and tackling right, it doesn't lend confidence they are doing the more complicated stuff well.
      • ForHackernews 8 minutes ago
        Their entire sales pitch is based on the idea that no one will need to be any good at anything, because the AI will do it for them.

        If they can't ask the AI to configure their CDN correctly, it undermines the validity of their claims. Maybe it's trivial, but any Omnipotent Machine-God deserving of the name wouldn't forget to zip up his fly.

    • jonplackett 39 minutes ago
      Welcome to the future