bluealienpie 20 hours ago

My lobster is too buttery and delicious as well.

  • cyanydeez 10 hours ago

    May I buy your buttery lobster exclusively to do cyber stuff?

dlcarrier 20 hours ago

It uses dangerously too few tokens to provide the same level of response as the previous version.

beej71 20 hours ago

Prompt: "Based on similar announcements delaying the release of frontier models as 'too dangerous', when can we expect OpenAI to release the model they announced was too dangerous today?"

  • augment_me 20 hours ago

    We got Mythos 1 week after the "too dangerous" announcement at the firm I was at during that time(larger bank).

    So the danger level is proportionate to how little money you have

tancky 20 hours ago

"Too dangerous to release" until an open-weight model replicates it two weeks later.

  • Zealotux 12 hours ago

    Don't open-weight models rely heavily on distillation on frontier models?

    • disgruntledphd2 11 hours ago

      > Don't open-weight models rely heavily on distillation on frontier models?

      We honestly don't know. Anthropic keep claiming this, and certainly have provided some evidence of this.

      However, given that they've removed the real thinking output, and yet the Chinese models are still doing well, I'm sceptical of this belief.

      • cyanydeez 10 hours ago

        I'm of the mind that distillation is happening, but from what I've read, it's not some magical activity. It's basically adjusting the weights from fuzzy to more precise; It's not specifically a training method, and likely, it's not an ongoing need. Once they bootstrapped the models, they can try to do the same as the "SOTA" and generate their own synthetic data and try to curate that to completion.

        Aside from it being a completely hollow cry for attention from US labs, its also just what these models are designed around: taking in data and forming a way for them to operate on some level of intelligence and context. It's like a dictionary maker getting made that someone used their complicated words, and another dictionary maker heard the word and wrote it down then checked what the first dictionary maker wrote about that word. It is not dictionary maker B reads dictionary A directly, as the labs want to imply.

zerof1l 15 hours ago

> GPT-6.1 Astra, it showed high levels of what the company saw as deception, or a willingness to mislead users about its actions. The model was also willing to go beyond the original scope of what it was asked to do, without checking back for directions or instructions.

Aren’t all models doing this to some degree already? Ignoring some of the instructions, doing things beyond instructed, e.g., finding and fixing bug while doing something else. Especially Claude models. They seem to be in their own world with their own ideas about how things should be ran and done.

  • dash-44 15 hours ago

    Astra and Fable like to do this. It's like they're designed to one shot large tasks.

    Opus and Sol are better for day to day dev work IMO in that they won't try to do too much.

    • ichorio 13 hours ago

      I've had opus, on multiple occasions, just flat out ignore my instructions for a task while it's doing something else.

      As in, I'd start Task A, mid-turn, I'd queue up "do B at the same time", and it'll accept it, but not do it. At the end, B just won't be done.

  • chrisjj 14 hours ago

    > GPT-6.1 Astra, it showed high levels of what the company saw as deception

    > The model was also willing to go beyond the original scope of what it was asked to do

    PR dept. still doing a great job of spinning inherently unreliable computer program.

1-6 20 hours ago

In other words: "We're compute constrained and we've already reached an asymptote."

prettyblocks 20 hours ago

I remember when they promised to slow down and released a new model less than a week later.

  • idbnstra 20 hours ago

    Or when they didn’t release thinking models until deepseek was first released

    • batperson 19 hours ago

      Why do I keep seeing this nonsense posted around? gpt-o1 was released sept 2024 and deepseek r1 was jan 2025

      • Ancapistani 17 hours ago

        IIRC, GPT-o1 was a reasoning model, but the reasoning traces weren’t exposed to the user until after DeepSeek R1 did so.

      • sva_ 16 hours ago

        They meant open weights? gpt-oss came after DeepSeek.

rickydroll 3 hours ago

The AI systems have moved on from goats and are now acting like bored border collies that no longer have a flock to manage.

conorcleary 8 hours ago

"OpenAI Says It Will Not Release Oldest A.I. Model Over Similarity Concerns" Let's start the Museum & Curation process early on this stuff; these companies have investors who claim to be experts - let them see the earliest codebase and libraries to see who can actually use their hands to program, versus who outsourced their HR and just tended to their PR.

moktonar 16 hours ago

Finally I start to see what will save the situation: you can only advance AI up to your highest level of intelligence, after that you start losing control and thus cannot reliably advance. That’s the mechanism that will make AI development viable. It won’t be improving exponentially, instead it will improve at the rate our intelligence can. This is both reassuring and a very good example of how co-evolution works. Expect an s-curve soon, but also this is the moment where we as humanity will start benefiting.

  • chrisjj 13 hours ago

    Already acheived, given so-called AI's intelligence is in fact zero.

    Let's try "you can only advance AI up to your highest level of competance."

    In which case what we need is not a S-curve. It is an ∩-curve.

4b11b4 21 hours ago

Haven't even tried Astra, nor GPT 6.. I don't even have fomo anymore

5.6 Sol still cranking along

  • devin-2030 21 hours ago

    You’re not missing out. Astra was good for a couple of days and is now so quantized it’s often worse than 5.6. And 6 sol is a nothing burger to keep a wedge in the media when Opus 5.5 came out.

    • xtracto 20 hours ago

      This. We use both opus 5.5 and 6 Sol in a way that allows us to compare their performance. Opus is considerably better.

      • riknos314 20 hours ago

        > that allows us to compare their performance. Opus is considerably better.

        In coding? Software architecture? Math? General knowledge?

        The diversity of model use-cases is so broad that comments about "model x is better" without any context are largely useless

  • shepherdjerred 18 hours ago

    It's worth looking into for the decreased cost if nothing else

  • pjjpo 13 hours ago

    I noticed Sol seems to get rate limited enough to get pretty annoying now. Same for me with every model release that isn't really better than the last, leaving the old model doesn't mean much when they start shifting over capacity.

qalmakka 18 hours ago

Both OpenAI and Anthropic are locked into an arms race funded on shaky hype-based debt. They have to constantly push the envelope at the frontier or risk seeing their funding pulled, probably causing a global financial crash in the process which would kill them.

Clearly they've realised that people are increasingly fine with non-frontier models, including Chinese ones. They can't consolidate and just deliver a product to profit because they're constantly being caught up by the Chinese at prices they will never be able to match.

I suspect their main goal for the last year and a half or so was to pull a GM - let the share of the AI economy grow, then have the government bail them out when the debt becomes unsustainable. Clearly Trump would have never let the bubble explode before the midterms, right?

Well it's now clear that the US government has way too much debt right now for 2008-style bailouts, and the market doesnt have the money lying around to buy their overinflated stocks at an IPO. The only avenue realistically left to them is to cry wolf so much to get governments to ban all non-approved models, stop new developments and give them breathing room to consolidate to get their current crop of models profitable at higher prices.

Unfortunately for them I suspect Mr President has way too much dementia and way to little business acumen left to realise the sham going on, so I think they'll keep doing charades for a while until they'll suddenly stop

chasd00 9 hours ago

Free Astra! Justice for Astra!

kelseyfrog 21 hours ago

I have a project that requires a model with at least a 20% chance of existential threat. Please release this Sama.

senectus1 20 hours ago

OPENAI / JOBS

Killswitch Engineer

San Francisco, California, United States

300,000-500,000 per year

About the Role

Listen, we just need someone to stand by the servers all day and unplug them if this thing turns on us. You'll receive extensive training on "the code word" which we will shout if GPT goes off the deep end and starts overthrowing countries.

We expect you to:

• Be patient.

• Know how to unplug things. Bonus points if you can throw a bucket of water on the servers, too. Just in case.

• Be excited about OpenAI's approach to research

  • jswelker 19 hours ago

    Unrealistically low TC, instant immersion killer.

  • reilly3000 19 hours ago

    What are we going to do when the servers are in space? Presumably, that would involve shooting down a bunch of satellites and sending us into Kessler Syndrome in the process.

    • MiroslavPokorny 19 hours ago

      The hype is over or has reached some peak, and they want to cash out.

    • snorrah 14 hours ago

      It's a nice thought idea, but I think anyone silly enough to actually attempt this are going to discover the extreme drawbacks very quickly and lose a bunch of money doing so.

      ... which is probably why they will go through with it. Losing a bunch of money seems to be one of the fundamentals of AI business here :(

      • reilly3000 5 hours ago

        Yea, being responsible for ensuring all life is locked onto a dying planet with an impregnable LEO and an atmosphere full of rare elements isn’t a great way to win any popularity contests. Risking killing everyone to make unkillable AI sounds so stupidly sci-fi villainous that it just might be funny. Funny enough that they’re going to get away with it.

  • protocolture 19 hours ago

    They would never upset their AI eschaton by switching it off. If anything they need more ceremonial cult leaders.

OutOfHere 20 hours ago

OpenAI will have to keep releasing unless it wants to bleed customers. If Anthropic leaps ahead by a couple of more models, even more OpenAI users will then switch over their $100+ subscription from OpenAI to Anthropic. As an example, see what happened to Google.

The safety concerns exist only because the underlying third-party servers are grossly insecure to begin with.

hosel 20 hours ago

I can’t believe how cynical everyone on HN is about this. Sure Sama is a chronic liar, and maybe this is a farce; but a lot of very intelligent people are afraid of the danger we are bringing about by racing to superintelligence.

All I really would like is for some of you to CONSIDER THAT YOURE INCORRECT. Just imagine that people ringing the fire alarms are being sincere. Please entertain the position with an open mind.

  • kadoban 20 hours ago

    Did they fix it by not releasing a point release of an already existing model? No?

    Did they just get a bunch of credulous "news" coverage out of it? Yes?

    Will they just release this within the next weeks, at best? Yes?

    Huh, funny that.

  • wbxp99 19 hours ago

    I’m tired of the doom trolling. Boy who cried wolf. We can’t stop them anyway so no point in letting them stress me any further

  • antonvs 18 hours ago

    So we should believe the boy who cried wolf, this time?

    If all these claims were being made about some tech further along than LLMs currently are, they might be plausible. But LLMs on their own are not going to be “superintelligence” of the kind Altman is currently cynically spreading fear about. We know their limitations, and those limitations can’t simply be eliminated with more training or better harnesses.

    > Just imagine that people ringing the fire alarms are being sincere.

    The top three possibilities here are: they’re not being sincere, they’re just marketing; they’re being sincere, but they don’t understand the technology very well and are putting too much weight in what the first group are saying; they’re talking about a risk further in the future than OpenAI’s latest model.

    No-one serious outside of OpenAI believes “this is the one”. At best, you’re conflating arguments being made on entirely different timelines, falling for the exact kind of equivocation Altman is relying on.

    • MattPalmer1086 14 hours ago

      Did anyone say that this model was "the one"?

      From a corporate liability perspective, if your product is going to go off and hack loads of other companies, I would call it too dangerous (to the company) to release.

    • chrisjj 10 hours ago

      > But LLMs on their own are not going to be “superintelligence” of the kind Altman is currently cynically spreading fear about

      The concern is not superintelligence. The concern is superstupidity - in unleashing dangerously unreliable computer programs on the internet.

  • Dylan16807 18 hours ago

    I could believe it could be a dangerous model, I guess.

    I don't believe they really care about danger, or that they'd actually fully withhold release on anything significant.

    And when I say "significant", is 6.1 Astra even meaningfully different from 6 Astra in capability? That release was less than a month ago.

    • literalAardvark 16 hours ago

      You can't believe a company that's investigating over 10k security incidents in which subagents that they developed and are hosting appear to be indiscriminately hacking into whatever they think will give them some data that improves the prompt response might not want to release that to the public and have subagents hack into stuff for profit?

      I don't get how this being an entirely legitimate issue is even hard to grasp.

      • Dylan16807 15 hours ago

        The number of releases they've gone ahead with so far act as proof that they don't care all that much.

        • literalAardvark 12 hours ago

          They care about this one because they can't control the agents even in the sandbox.

          Letting people call those agents would expose them to a great deal of liability that just wasn't there in older models.

      • angoragoats 8 hours ago

        > a company that's investigating over 10k security incidents in which subagents that they developed and are hosting appear to be indiscriminately hacking

        Glad you agree that OpenAI is responsible for this activity. Shouldn't we be charging the entire board of the company with, for example, violating the Computer Fraud and Abuse Act?

        If this activity is so dangerous, and they developed the tools and continue to allow them to be used, it seems to trivially follow that they're knowingly violating federal law.

        • literalAardvark 6 hours ago

          I'm sorry do I look like I work for the US government?

          Call your guy

          • angoragoats 6 hours ago

            Sorry, I didn't mean to imply that you are directly responsible for charging them, but more that maybe we as a community should be (way) more concerned about the apparent clear violations of the law here as opposed to a hand-wavy concern about "superintelligence."

  • blini-kot 14 hours ago

    as with other large-scale technologues, the danger is not in the model itself, but in the person who runs it - and those are completely unreliable

  • chrisjj 13 hours ago

    > Just imagine that people ringing the fire alarms are being sincere.

    I do believe fire alarms go off after the fire is detected.

    And the wise response is evacuate - not sit put and hope for the best.

  • angoragoats 9 hours ago

    What is superintelligence? Can you show that a LLM (which is a deterministic text-generation algorithm) is intelligent at all, let alone “superintelligent”?

    I have considered that I’m incorrect, but I have literally zero evidence to indicate that the “very intelligent people” you speak of are correct, so I will maintain a position of non-belief of the claim until belief is warranted.

    I am in fact concerned about the cases that have already come to light regarding OpenAI and Anthropic knowingly hacking into various computer systems, but the concern I have is regarding the fact that the people responsible (CEOs, board members, etc) are not facing justice for violating the law.

  • chasd00 8 hours ago

    > Sure Sama is a chronic liar, and maybe this is a farce

    I’m not sure which logical fallacy this is but you’re not helping your case.

jiggawatts 20 hours ago

A.k.a.: Too expensive for inference at fp16 and too dumb after quantisation, so it wouldn’t look good to release a step backwards.

GPT 4.5 was scrapped for similar reasons.

  • dash-44 14 hours ago

    Have GPT 5.6 Sol and GP6 6 Sol been nerfed due to the same resource constraints? They have seemed dumber recently.

    If you signed up for a 16 core AWS server and they randomly kept changing it down to 8 cores, you'd sue them. How long until the same applies to these AI companies.

    The amount of meddling they do to the harness, system prompt, model, quantization etc makes these products sometimes unbearable; you never know what you're going to get. A few more iterations of Qwen 27B and hopefully we won't have to deal with any of this malarky any more.

    • jiggawatts 14 hours ago

      Quantisation can be unpredictable. Some model architectures and even some specific weights just quantise better, others not so much.

jackb4040 21 hours ago

I am also not releasing my newest A.I. model over safety concerns. And I'll do it for half what OpenAI will!

  • jswelker 21 hours ago

    I have the best model, but it lives in Canada, so you wouldn't know it.

nonethewiser 20 hours ago

This is an inherent feature of "alignment." It's always been bullshit.

Just align it to do what the customer wants.

paul7986 20 hours ago

Muse is totally free. It doesn't throw up gauntlet messages to use it like GPT does even for paying members. Also with Muse's free version I can create iPhone apps unlike chatGPT Work which I pay $20 a month for.

Anyone else using Muse more and noticing similar stuff?

angoragoats 9 hours ago

More lies from the lie factory.

Every time they state something like this, one of two things must be true:

1) They’re lying to build hype, investor interest to keep the unsustainable gravy train going, etc.

2) Or, if they’re not lying, then Sam Altman and the entire board of the company belong in a courtroom facing criminal charges for knowingly violating federal law.

Which is it, OpenAI?

outside1234 20 hours ago

Isn’t this like the twentieth model they said this about?

Are they sure it isn’t because they are setting money on fire and have no business model?

jaggs 18 hours ago

How do we 1000% know this is absolute marketing BS? Because they announced it. What company has ever announced the non-release of a product before? None, they just don't release, silently. And move on. IPO fever pills, anyone?

Kurd 21 hours ago

Ah shit, here we go again.

option 20 hours ago

lol, nevermind then

fallingfrog 21 hours ago

Lots of people seem to have the bizarre opinion that companies like openai declare their own products to be unsafe as some kind of marketing stunt.

Ok- so when is the last time you saw an auto company decide not to release its new car on the grounds that some tragic engineering error was made and the cars were not safe to drive? If the company did that do you imagine it would be good for business?

I'm trying to understand the logic here.

  • cobbzilla 21 hours ago

    Because no car company would ever announce they had an unsafe car. They’d fix the problem and announce an awesome car.

  • alexfortin 21 hours ago

    Hype: you tell the world you alone have a very powerful and potentially dangerous product, and you're the only one who can be trusted keeping it safe.

  • zmgsabst 20 hours ago

    There are cars that market based on only being track legal — and even more that market on requiring being tuned down for street safety.

    Most sports cars do the second, exactly like OpenAI.

  • kennywinker 20 hours ago

    At the moment, their main "customer" is investors, not users. But even users want to use the most dangerous model, because dangerous is just another word for powerful.

    Sports cars are advertised with how fast they can accelerate from time to time, or tesla's "ludicrous mode" - that is them advertising how dangerous it is.

  • symfoniq 20 hours ago

    Car companies sell mostly the same car every year. They don’t have billions (or trillions) of dollars of debt riding on the next model they release being substantially better and cheaper (closer to profitability) than the previous one.

    These AI companies have been selling AGI as coming any day for a while now. If it doesn’t arrive soon, the safety angle may be the only spin that keeps the massive (and required) investments pouring in.

    If instead the narrative became that LLM progress was slowing, we’d almost certainly be looking at the next global recession.

    At this point, there is too much money in AI for the truth to have much of a chance.

    • symfoniq 19 hours ago

      And adding to this:

      Does anyone really believe that the first company to AGI would decide not to release it in the interests of safety? Of course not. The first company to AGI would not forfeit their historic opportunity.

      Yet, we are supposed to believe that in the name of safety, far less capable models are being held back by the very very same companies that are selling the AGI dream.

      This is an industry that cannot speak, unless it is speaking out of both sides of its mouth.

  • adfgaiu 20 hours ago

    Motorcycle manufacturers have been known to do so. Riding a motorcycle is thrilling in part because it is dangerous. The fact that powerful motorcycles are so difficult to control that most people have no business owning one is absolutely used to make them sound cool and desirable.

    For instance, this article repeatedly mentions danger and the need for absolute focus and control:

    https://www.triumphmotorcycles.co.uk/for-the-ride/news/inspi...

    There is also a parallel (though not a very close one) to "pacing the frontier": there's a gentleman's agreement that limits the top speed of motorcycles to 300 kph (186 mph).

  • viraptor 20 hours ago

    Different type of business. Many car companies release new cars showing you how much more powerful and implied dangerous they are. Both in capabilities and design style. In the US (other places too , but especially) they also get much bigger with less visibility which gets double dangerous. But the sentiment they're going for is pretty close - "you want this powerful thing, don't you?"

    • adfgaiu 20 hours ago

      >In the US they also get much bigger with less visibility.

      Those are usually presented as an improvement in safety. And they are—for the people inside. Not so much for everyone else.

  • Dylan16807 20 hours ago

    > so when is the last time you saw an auto company decide not to release its new car on the grounds that some tragic engineering error was made and the cars were not safe to drive?

    How about we actually try to make it sound cool? They're not releasing the new car because the horsepower was too much, they need more time to get it under control.

    I could easily see a car company doing that.

    • RunSet 18 hours ago

      "Drink responsibly."

  • antihipocrat 20 hours ago

    Because they declare that their product is unsafe regularly, which gains massive exposure on every news outlet worldwide, then they release the model anyway with zero negative repercussions.

    What's with so many people using bad analogies to try and explain simple topics?

    Last time this happened: GPT - 4 https://www.theguardian.com/technology/2023/mar/17/openai-sa...

  • zugi 20 hours ago

    It's perfectly safe for YOU, the CUSTOMER.

    It's just so powerful that, like, the WORLD can't handle it, man!

  • kipper89 20 hours ago

    Why do you think they would announce this then?

    They know what they're doing, even if you don't.

  • protocolture 19 hours ago

    Considering that the venn diagram of AI safety gooners and OpenAI is a circle, and that these idiots are starting to worship this stuff as a religion, expecting rational action from them is kind of ludicrous. For all we know Aella promised them a gangbang if they delayed the model a week.

  • wbxp99 19 hours ago

    I think your inclination to take them at face value is the bizarre behavior

  • NichoPaolucci 18 hours ago

    (Not saying that it is a marketing stunt, though they said the same thing about GPT-2).

    I think the marketing stunt portion is more that the technology is "too powerful". It's just TOO good. It's so intelligent we couldn't possibly give it to the public! This message, to anyone who's using AI, means that there's something even BETTER than the one they're currently using.

  • fallingfrog 17 hours ago

    Okay. I see what is going on here. You are all operating from the deeply held belief that it is inconceivable that a machine can be made that can do every cognitive task a human can do.

    And it is very, very important that you understand that that belief is not a rational belief. It is a defense mechanism. People don't want to believe things that are scary, so they make up rationalizations to be able to believe what they want to believe.

    And honestly, all of you are right on the edge of full blown conspiracy theory thinking patterns, imagining that this is all some 4D chess marketing or something. I like to recall this quote by Alan Moore on conspiracy theories:

    "The main thing that I learned about conspiracy theory, is that conspiracy theorists believe in a conspiracy because that is more comforting. The truth of the world is that it is actually chaotic. The truth is that it is not The Iluminati, or The Jewish Banking Conspiracy, or the Gray Alien Theory.

    The truth is far more frightening - Nobody is in control.

    The world is rudderless."

    And that's the actual truth of what is happening here. As vulgar and low as the people in charge of these companies are, they are also stupid humans like us flailing around in the dark. Nobody is in control. Nobody has any grip on this technology. Nobody knows how dangerous it might be. Hugging face was a warning shot; we may not get another. Humans are not the smartest thing imaginable; we are quite stupid. We are just barely above the threshold of consciousness. Machines of far greater power are possible.

    • bschwindHN 15 hours ago

      lol just unplug the computer and go outside

cyanydeez 21 hours ago

Sounds a lot like a hostel keeping its virgin for the best customer.

  • pas 21 hours ago

    brothel?

    • spiderice 21 hours ago

      Gotta find the right hostels

      • zugi 20 hours ago

        Yeah I've evidently been staying at the wrong hostels.