Balooga 8 hours ago

Jevons Paradox [1]

> when technological improvements that increase the efficiency of a resource's use lead to a rise, rather than a fall, in total consumption of that resource.

[1] - https://en.wikipedia.org/wiki/Jevons_paradox

Las Vegas replaced the expensive incandescent lighting on the strip with cheaper to run LED equivalents. But the costs didn't come down because they were able to add more lights and larger displays.

I think the same will happen with tokens. As the cost of tokens comes down, these models will just consume more tokens.

  • perching_aix 7 hours ago

    Sounds like a variation on the induced demand principle: https://en.wikipedia.org/wiki/Induced_demand

    Related:

    - Parkinson's law: "Work expands to fill the available time." https://en.wikipedia.org/w/index.php?title=Parkinson%27s_Law

    - Lewis–Mogridge position: "Traffic expands to meet the available road space." https://en.wikipedia.org/wiki/Lewis%E2%80%93Mogridge_positio...

    And I pretty much just plain agree, this is exactly what will happen.

    I don't think there's anything wrong with it (in isolation) either, though I do already find myself pointing out that we're misusing LLMs at work sometimes (most notably, a recent mini project could have been a jinja template - and it did become one thanks to me pushing back on this). Abundance is one thing, waste and misuse is another.

    • bena 7 hours ago

      Also Jevon's Paradox: the more efficient a system becomes, the more it gets used, making it use more resources rather than less.

  • AndrewKemendo 7 hours ago

    Turns out Grey Goo was just thermal paste

  • bee_rider 7 hours ago

    Definitely not going to argue against the Jevons paradox in general, it is observed in various cases.

    The lightbulb thing seems different though? Or at least it is a specific subset. Lights in Las Vegas are sort of an advertisement, right? In the sense that having the brightest or most interesting (or whatever) lights draw attention to your show, casino, hotel, whatever. It’s kind of a zero sum game in that the different shops are competing for the finite attention of a more-or-less set number of tourist. I think part of the Jevons paradox is that society generally finds more useful applications of the newly cheap thing. If the thing’s only purpose is to compete better in a competition with a set prize (all of the tourists’ money), that’s constrained in some way.

    Intelligence is weird though. I guess we could eventually hit the point where, I dunno, maybe there’s some information theoretic bound where we process all of our signals as cleverly as possible and aren’t bound by intelligence anymore. Obviously we’re nowhere near that. It would be a very alien environment.

  • Frieren 7 hours ago

    > I think the same will happen with tokens. As the cost of tokens comes down, these models will just consume more tokens.

    That sounds a race to the bottom for AI companies profits.

    I do not think that LEDs are high profit margin items.

  • everdrive 7 hours ago

    > But the costs didn't come down because they were able to add more lights and larger displays.

    People are a gas; they expand the fill the space they're in. If you give someone a big house, they'll fill it with crap. If you make food cheap, they'll eat too much, even when it harms their health.

    Making things cheaper usually just makes making them more prevalent. It's why computers are not faster than 20 years ago. They're merely more capable -- developers quickly and aggressively fill up (and overflow) all that added capability until you're back in the same place you used to be.

    • vanuatu 7 hours ago

      You're right, computers are definitely the same speed as they were 20 years ago.

      • wongarsu 6 hours ago

        20 years ago is about when SSDs came to consumer PCs. Which is the last time I remember thinking that my computer became faster than the previous one was when it was new (if by computer we mean the hardware combined with a mainstream up-to-date software stack)

        • behringer 6 hours ago

          Your computer may not feel faster than a 20 year old PC, but CPUs have gotten way, way faster.

          Compare the AMD Ryzen 7 9800X3D to the AMD A12-9800, you go from 4 cores to 8, and you double the power consumption. But it's not twice as fast. It's 10 times faster. Depending on the exact metric, it could be as little as twice as fast or as much as 1000 times faster, depending on the exact operations.

          On top of that memory throughput between DDR4 and DDR6 is about 2.5 times faster as well.

          Oh and this isn't 20 years apart, this is 7 years apart. 20 years will see almost an exponentially larger gap even still.

          • everdrive 6 hours ago

            I kind of hoped my point would come across more clearly. Obviously my current computer is faster than one from 20 years ago. But, it doesn't really react more quickly. It just does more: higher resolution video, advanced graphics, etc. The moment we add capability we rush to use it all up. This is not _always_ bad when it comes to computing, but I wouldn't mind if people in general had more restraint.

            • behringer 5 hours ago

              I suppose that's true superficially, but how many of us utilize 100 percent of our CPU and GPU power.

              AI datacenters, absolutely. People? Nah we probably use 10 percent for 90 percent of the time.

              But when I need to compile MAME, oh boy is it nice having a fast CPU.

              • everdrive 5 hours ago

                Right -- in the details this is heavily task dependent. If we're talking about word processing, then computing has mostly gotten worse in the past 20 years.

                If we're talking about video game graphics it's a bit of a mixed bag.

                If we're talking about local LLM models then you need all the power you can get.

                Etc, etc.

          • alphakappa 4 hours ago

            That is OP's point. The underlying technology has gotten faster but the computer you use feels the same as 20 years ago because the software that utilizes it does a lot more too.

        • nemomarx 6 hours ago

          You had SSDs in 2006? Was I missing out?

          I don't think I had any in my builds until 2014ish

          • missedthecue 2 hours ago

            Samsung commercialized the first high-volume consumer notebook PC utilizing an SSD in 2006. The Intel X25-M was released in sept 2008 and cut SSD hard drive prices per gigabyte by 60% pretty much instantly

        • wpm 4 hours ago

          I can still feel performance increases, most recently was when I got an M4 Mac Mini. Compared to my M2 Air, which is no slouch by any means, there was a very perceptible improvement in latency, app opening, general task completion, and so on. But very slight, and some coming from placebo effect/marketing.

          These days, I feel the most performance uplift when I switch from Windows/macOS over to a Linux or BSD installation, if only because there is so little wasted effort in running a bajillion little background tasks for crap I didn't ask for. Windows 10 on my PC is highly debloated and optimized, very little in my startup services and so on, but CachyOS on the same hardware feels leagues better.

        • qeternity 1 hour ago

          I presume you're a Windows user then.

          When the M1 was released, I couldn't believe how fast it was.

    • jayd16 7 hours ago

      Anyone not living paycheck to paycheck is just constantly gorging on fast food?

      • everdrive 7 hours ago

        I don't mean to be glib, but for the most part, yes. https://www.nytimes.com/2025/09/11/well/child-obesity-un-rep...

        Obesity is nearly everywhere. In the old days, "prosperous" might have been a euphemism for "fat," as only the rich could afford to be fat. Now even people who are "food insecure" are often quite fat. Some of this of course is due to food quality, but ultimately there are no starving fat people.

        • jayd16 6 hours ago

          How does "people will uncontrollably eat if given the chance" turn into "there are no starving fat people".

          You're right that obesity is not necessarily linked to overeating . How is this relevant to your original point? Isn't it a complete contradiction to what you were saying before?

          The original post was hyperbolic to the point of nonsense and now it seems we're on a completely different topic.

          • everdrive 6 hours ago

            >You're right that obesity is not necessarily linked to overeating

            This can be true, but I think some people are confused. The vast majority of the time, it is linked to overeating. There are some edge cases where there is something else going on metabolically, but most people do over eat.

        • ButlerianJihad 5 hours ago

          > ultimately there are no starving fat people

          It has been credibly hypothesized that fat people are indeed malnourished. What would you do if all your food sources were appallingly lacking in essential nutrients and the sorts of things your body craved in order to get fed? You would begin to eat more, and more and more, seeking to satisfy those cravings for nutrition. But if nutritional cravings aren't satisfied, you're still reaching for those foods, because they are faking, very accurately, the tastes and feels that you would get from something that's nutritious and good for you.

          So, it stands to reason: the more fake nutrition available to someone, the more they will try and eat, and get fatter. On the other hand, if we're eating real food, perhaps those cravings will be easier to satisfy without overindulging?

          • everdrive 5 hours ago

            >What would you do if all your food sources were appallingly lacking in essential nutrients and the sorts of things your body craved in order to get fed?

            When I hit 300 lbs I might suspect there was a flaw in my reasoning.

            • ButlerianJihad 5 hours ago

              Sure, you may be a rational actor and come to your senses, but you are also poor, you have no vehicle, and you live in a food desert. So will you realistically be up for major modifications to your lifestyle and habits?

              • everdrive 4 hours ago

                I hear what you're saying -- I think there are a few points here:

                - Obesity is not isolated to food deserts. The people in food deserts may be in a pitiable situation, but most of us are obese; not just the people in food deserts. If obesity were isolated to food deserts, or even concentrated in food deserts, I might feel differently.

                - Even the people in food deserts have access to nutrition labels. One thing that makes a lot of sense is ensuring you get your vitamins, and then keeping the rest of your food cheap. In other words, the most cost-effective food regarding calorie-per-dollar might be zebra cakes, but obviously you don't actually to just buy Zebra cakes. Living on black beans and rice might not sound appealing, but it is not the path to obesity. When you live on black beans and rice, other foods are your garnish. Sparing, small servings, but present from time to time. Most breakfast cereal is also effectively just grass seed + vitamin powder, as well.

                [edit]

                There was a time in my life where I had very little money, and I lived on black beans and rice for quite a while. I biked and walked around the city and I was thinner than I've ever been since.

                • ButlerianJihad 1 hour ago

                  People are often not aware of how their ability to function far exceeds some other people who are not so fortunate.

                  To describe a lifestyle of "surviving on black beans and rice" and riding your bike around, implies that you maintained a functioning household and kitchen with paid, active utilities; you were able to clean the room, vessels, appliances and utensils on a regular basis; you did pest control as necessary, and you were able to accomplish basic cooking tasks 2-3 times a day, every day.

                  Accomplishing all these tasks well enough to survive, feed yourself without hazard, and not become overrun with cockroaches is a sure sign of baseline mental capacity and cognitive function. Again, there are lots of poor people who were homeless, or grew up in dysfunctional households, or are on the brink of homelessness, and struggle with mental illness that militates against them "surviving on black beans and rice" day by day.

                  So I congratulate anyone who can accomplish these seemingly-trivial tasks. You deserve a pat on the back for sure!

    • css_apologist 6 hours ago

      these things you are talking about are only true due to the unbelievably LARGE marketing ~propaganda~ spending

      what you're saying reflects corporation tendencies

      take away the propaganda, let people live, see how they react

      • everdrive 6 hours ago

        I'm not convinced. Propaganda only works when it relies on something that people are dispositionally inclined towards. eg: hating foreigners, calls towards group unity, etc.

        You can't use propaganda to trick people into becoming experts at calculus, for instance.

        • simonra 4 hours ago

          > You can't use propaganda to trick people into becoming experts at calculus, for instance.

          Perhaps not an individual instantly. But how about settling for making a generation of parents fear that there will be no future for their kids unless they get fantastic grades, and that they have to use them to go into stem?

      • jostylr 4 hours ago

        How would you distinguish corporation influencing consumers from corporations following consumers? One thing would be to see if there are items that are heavily advertised by corporations but that they do not succeed. You can also take items that everyone knows about, no one wants, and then launch a propaganda campaign and see what happens. Different propaganda works in different ways so one would need to try varieties. I imagine a randomly controlled trial of having various ubiquitous goods that are not heavily advertised nor heavily purchased and then randomly assigning different advertising strategies (including none). Do that in various regions, countries, ethnicities, etc, and then see what can be seen.

        My guess is that advertising raises desire by a few percent in most cases. There would be some instances of a run away craze, probably with a massive rise followed by a collapse, but that most products would largely be unaffected. I assume some advertising strategies would backfire and reduce demand.

        For me, I feel like most of the time it is been exposed to the existence of an item and my ability to see how it could be useful. I don't think random advertisements I come across impact me that much since most of the advertisements I see don't seem to correspond to anything I want or do or spend money on.

        Humans use their imagination to figure out what choice of action leads to a future of less uncomfortability. Often times, these are choices between short term future and long term future. Advertising is a way of helping shape that argument as to what choice is to be taken, but the underlying drive is already there.

        You could also look for cases where propaganda is used against consumerism and see how impactful that tends to be.

  • Flamkuchlo 7 hours ago

    Tokens for sure because we still waiting for agents talking to agents as a default. Like your agent team and very long or 24/7 running agents.

    But we never scaled intelligence like this. The industrieal revolution created for the people at that time quite a huge issue / it was disruptive.

    What will hapen to us though?

  • conjecTech 7 hours ago

    Jevon's Paradox says nothing about the size of the incremental demand though. Lighting is a great example. The US went from spending 15% of our electricity on lighting to ~4%. LEDs are maybe 5x as efficient, so that would be 3% without any change in usage. So we did see maybe a 30% increase in consumption as a result of lower prices, but that was nowhere near large enough to offset the efficiency gain in terms of overall consumption. Las Vegas may have, but that is a very small part of overall usage.

    I anticipate we will see something similar with intelligence. There is probably headroom to consume 100x as much intelligence in R&D. But that isn't most of the economy. Will run of the mill service jobs increase their use of intelligence by enough to offset the effect of cheaper prices? I think that's the real question.

    • TheOtherHobbes 6 hours ago

      Intelligence is a lever, not a fluid. Power depends on how much force it can apply to specific points, not how much exists in general.

      True high-intelligence outputs (Maxwell's equations, the Fourier Transform, quantum theory) are fundamentally transformative in ways that mid-high-competence (starting a generic B2B SaaS, making another CRUD app) aren't.

      Which is why we've assumed we're already pretty far down the path to AI, but we really aren't. Solving random Erdős problems isn't the same as opening up a completely new kind of math/science with game changing practical applications.

      I don't think you can get to that level with more compute and more tokens. I think it's going to take new higher level knowledge representations and new kinds of training to get there.

      And the token count and compute may turn out to be lower than what we're using now.

    • HDThoreaun 2 hours ago

      Is that because non lighting electricity use increased though? 30% increase in the percentage of a growing pie isnt the same as a straight 30% increase

      • conjecTech 59 minutes ago

        No, per capita electricity usage dropped slightly over the last 25 years, driven in part by LED adoption.

andai 8 hours ago

> Reading everything becomes the default. At a cent per document, a model can read every paper

I love how in our day "reading everything" means "the computer reads it for me".

I expect soon the computer will be able to go on bicycle rides, and spend time with my wife.

  • bellowsgulch 7 hours ago

    Futurama did it first!

    • gs17 6 hours ago

      "Robot, experience this tragic irony for me!"

  • sssilver 7 hours ago

    Hasn't the computer already been spending time with your wife?

    • smugtrain 7 hours ago

      Asymmetric multiprocessing with your mom

  • rjsw 7 hours ago

    Douglas Adams captured this in "Dirk Gently's Holistic Detective Agency", a VCR watches TV for you and an Electric Monk believes things for you to save you the effort of doing it yourself.

  • nradov 6 hours ago

    The next generation of bike computers will probably have links to an LLM cloud subscription service for real-time coaching and route planning.

jbotdev 7 hours ago

I think speed is actually going to be a bigger factor than cost. Even projects where “money is no object” often hit a wall with LLM response times.

Sure you can speed things up with parallel work under subagents, but as with parallelizing traditional computational tasks, there are diminishing gains.

I keep hearing people saying just change the way you work to trust long-running agents and multi-task more, because they’re too slow to work with interactively for many use cases. I think that’s painful in a world where we expect humans to still heavily guide and interact with agents for their day-to-day work.

  • perching_aix 7 hours ago

    Given the 750 tok/sec GPT 5.6 Sol Ultrafast (via Cerebras), the many-1000 tok/sec Chinese models, and the 15000 tok/sec Taalas HC1, I think we're well on the way towards seeing that solved too. Combine the two, and yeah, wild ride incoming.

    What's especially bewildering to me is that translated back to raw bandwidth, even 15000 tok/sec is just like what, 75 KB/s? Extremely meager amounts of data, moving mountains.

    It's already kinda funny seeing LLMs throw out effort estimates in wall time terms. It's always some "hours, days, weeks" tier thing, when in reality, it's gone and done in minutes.

    • cactusplant7374 6 hours ago

      One of my projects has an estimate of 3,000+ hours. It seems accurate. It has spent months working on it.

      • perching_aix 6 hours ago

        Around the clock?

        • cactusplant7374 6 hours ago

          Yes, with a few exceptions. It's much easier now that the five hour limit has been removed from codex.

      • wyre 6 hours ago

        Can I ask what you are working on?

    • newAccount2025 42 minutes ago

      Haha. Yeah. I somehow find Claude Code’s utter lack of time reasoning to be very fun. “The work we did two weeks ago…” oh friend that was this morning…

  • jbstack 7 hours ago

    > they’re too slow to work with interactively for many use cases

    This just demonstrates how much we already take for granted the LLMs that we have now. If you compare it to what we had before (hand the task off to a junior dev and wait for them to complete the work) then it doesn't seem slow at all.

AnotherGoodName 8 hours ago

I think a big one is robotics. A robot can today fold your laundry. It takes ~10mins per item. Seriously. It takes a long time to process the image find the corner move the claw to the corner of the shirt and attempt to straighten before folding.

Robots right now generally move at glacial speeds. You might have seen robots doing flips in semi controlled environments but watch how slowly they open doors etc. processing time is a major bottleneck.

  • andai 8 hours ago

    Wait til the robots get on Cerebras, it'll set your pants on fire.

    • LoganDark 8 hours ago

      Only if the fire manages to escape your wallet!

  • segmondy 8 hours ago

    You must not have been paying attention to development with robots, there are many videos of robots moving really fast in "non controlled environments"

    • airstrike 8 hours ago

      Yes, I think I saw one in a video titled "Robocop"

    • TheAceOfHearts 7 hours ago

      Honestly, most of the videos I've seen of robots moving around quickly aren't actually doing anything useful. We've had really impressive tech demos for the past 15 years of robots dancing and jumping around. But I don't want a dancing robot, I want a robot to make me a BLT, wash my dishes, take out the trash, and fold my laundry.

      The most recent video which actually impressed me was a demonstration from Gemini Robotics 2, where a robot was shown autonomously removing the bag from a trash can and folding the loops closed in real time.

      I don't follow robotics advances closely so it's possible I'm just ignorant, do you know any autonomous robotics demonstrations of useful activities that you would suggest checking out?

    • infecto 6 hours ago

      Most of these videos are still controlled in some manner. We are not yet at the age of throwing a robot in a home and doing n different dynamic tasks.

  • CodingJeebus 8 hours ago

    Have LLMs improved at being able to process physics-based problems and environments? I remember that issue being discussed around generative gaming a while ago but I hadn't heard much about it recently.

    • AnotherGoodName 7 hours ago

      Transformers in general are huge but if you say LLM you’re specifically saying the language model transformer. Robots use vision transformers

  • Alien1Being 8 hours ago

    Xiaomi robots do this in double digit seconds.

    Still slow compared to humans, but Chinese robots will be as successful as Chinese EVs, phones and solar panels.

    • kaashif 7 hours ago

      Which is to say, they'll be banned in the US.

      • warkdarrior 5 hours ago

        The 1950s in US didn't have vaccines, robots, or data centers.

      • qzw 5 hours ago

        How long until we’re like Cuba, still tooling around in 50s cars. But maybe that’ll actually be seen as some kind of Luddite utopia in the future.

    • infecto 6 hours ago

      ~1min per shirt in a controlled environment. Better but still far away.

      • andriy_koval 3 hours ago

        Also, 24h 365d, which is superior to humans.

        • infecto 2 hours ago

          And? Still in a controlled environment. I cannot pick one up today from a store.

          • andriy_koval 2 hours ago

            laundry folding is some superficial benchmark, real impact will be made in controlled environments in factories, packing centers, etc.

  • wongarsu 6 hours ago

    Sensors are a huge challenge for robotics. We have very precise force-feedback on our joints, pressure and heat (temperature gradient) sensors all over our body, and our hands have a sensor density that allows us to count needle heads and detect the exact grip strength needed by feeling the micro-slippage of objects in our hands. Robots don't have that.

    You can do backflips with pretty much just visual sensors for your environment, a good IMU for your spatial orientation, and some feedback on the position of a small number of really beefy joints and the force exerted on them. Folding laundry and opening doors is much more difficult, and trying to compensate with mostly vision requires going slow enough that things have time to move over appreciable distances before you take the next adjustment

    • andriy_koval 3 hours ago

      > We have very precise force-feedback on our joints, pressure and heat (temperature gradient) sensors all over our body

      You can do all the same with robots. Current advantage of human that on top of imperfect sensor data we have hyper-efficient brain connecting all dots and making calculations and approximations, and this gap looks very closeable today.

newAccount2025 7 hours ago

I’m loving small models. The gemma4 26/31b models have been deeply impressive on weird prose analysis tasks that I am working on. Nova-micro is really stupid but is extremely fast when it’s smart enough to do something. I’m trying to be disciplined about able to evaluate quality vs cost everywhere for real systems built on this stuff. I probably need to get off Bedrock because it’s missing a lot of other little models that might be good competitors.

bdhdhduuyd 8 hours ago

Personally I still see LLMs as very advanced search engines which lack intelligence. To me it seems that the cost of getting data is reduced by LLMs, not the cost of intelligence. I mean: we tell the model what we want to achieve, and the model responds with the right data in de form of code in seconds.

That's why 'stackoverflow programmers' will have a hard time competing with LLMs but engineers are still needed for their intelligence.

Well that's just my 2 cents.

  • dboreham 7 hours ago

    I think this viewpoint fails to understand what "intelligence" is. The idea must be that intelligence is some special thing that only humans have. So when machines couldn't do jack s... we said "it's the Turing test". When machines blew through the Turing test we said "that was just prediction..not really intelligence, that's different".

    It's not different. The delusion humans have is that intelligence is special and magical. It's not. It's just nature's prediction machine. A very fancy version to be sure. But not qualitatively different .

    All statements that "oh but it'll never be able to do that" will prove false.

    • perching_aix 7 hours ago

      I notice that a lot of these terms are squarely humanist for certain people, so any kind of allegation that machines are exhibiting them as traits will be a complete showstopper for them. You'll either be strung along on an infinite goalpost moving exercise, or be accused of either anthropomorphization [0] in the kinder cases, or straight up mental illness in the less so kind cases. Never will they stop to consider that maybe you're simply working with a post-humanist understanding of these words, as that basically doesn't make sense to them, and as they are usually quite vested for it to stay that way.

      I remember in one of the Hugging Face incident threads here, simply acknowledging that the agents were operating autonomously was super controversial. Thousands of years old concept [1], still inherently human for a lot of people.

      To be clear, I'm not trying to be judgemental with this, I more consider it to be a communications breakdown, and find that to be frustrating instead. I'm not really sure how to meaningfully help it either, cause no matter how one slices it, you will in the end ask these people do desecrate these terminologies in favor of some more twisted-seeming ones. Same the other way around, the humanist understanding of these terms is basically non-workable.

      [0] as opposed to personification, which is what people are actually doing almost always: https://en.wikipedia.org/wiki/Personification

      [1] https://en.wikipedia.org/wiki/Automaton

    • goatlover 6 hours ago

      It's not about magic, it's that humans have an evolved embodied intelligence for surviving in the world as living organisms that reproduce, care for young and live in communities. That's quite a bit different from language models trained on human data.

      Artificial intelligence is artificial. It can still be called intelligent, just not the same kind as biological since it's not remotely biological. It's human-like but also alien. To have something artificial be human you'd need something like replicants from Blade Runner which are synthetic biological robots.

    • 2ferr 5 hours ago

      Until you contribute something substantial to humanity that is novel, you ought to reduce the tone of confidence.

  • Flamkuchlo 7 hours ago

    Sooooo what does it matter if we do 99% of things a LLM can just solve as a 'advanced search engine'?

    Btw. an advanced search engine is probably the worst comparision i have read so far.

    A LLM is a latent space which is capable of a lot of things a search engine can't do. It can apply different type of patterns and flows onto data, it can combine these etc.

    My 'advanced search engine' was just able to create a working PR for exactly what i wanted it to solve (fixing a bug) by analysing the bug, finding a valid solution then commiting the solution itself.

  • iririririr 6 hours ago

    You're correct on the technology assessment.

    But you forget all development today is just searching for a template, copy pasting, changing some small things. And a smart-ish search engine can do all that.

MichaelNolan 8 hours ago

100x seems like an underestimate. Even with no model improvements, we should see that sort of reduction. Looking at TSMC’s margins, Nvidia’s margins, and OAI/Anth (alleged) margins on inference, there is a room for a 100x reduction.

Right now all three of those are at abnormally high levels. Competition will come for all three.

vanuatu 7 hours ago

I think what a lot of people miss about jevon's paradox is the elasticity of demand of the underlying resource

textiles had jevons paradox, and many more textile workers were employed even when textile machines were being created, until we saturated the demand for cheap clothing in the world and then textile workers were kaput (same for farming, and horses)

software is currently undergoing jevons paradox, but it's very unknown how high the ceiling of demand for software is. web dev might be doomed, but software in general i think is probably limitless

Intelligence is also probably unbounded (atm software and intelligence are very closely tied together). its very possible token spend rides up the curve forever.

  • quantified 5 hours ago

    Brings up the question of what the intelligence is used for. Humans exploited intelligence for competition. With each other to wipe out other Homo species, mate more and collect resources, with other animals to limit predator impact and gain food. Intelligence will be used offensively by corporations and their people to extract more from consumers (make pricing opaque, terms of service more complicated, etc.) and scams far more sophisticated. The "consumer" will need extra intelligence to fight all that off. There are only so many meals you can expertly produce, shirts to fold, itineraries to fun places you can execute, but there's a practical infinity of traps to set and avoid.

nchmy 8 hours ago

I've been working with the chinese open models for 4 months. They are more than capable for a tiny fraction of the cost of the frontier ones. And yet they also continue to get significantly better and (Deepseek's recent price increase aside) cheaper. Its hard to fathom how the truly frontier stuff will be able to compete long-term.

  • jostmey 7 hours ago

    And why won’t the frontier models continue to become better? The open models are getting better but so are the frontier models. The frontier models might remain in a constant race to remain ahead

    • ForHackernews 7 hours ago

      They are running out of novel, clean training data and compute. There is probably a limit to how much improvement can be squeezed out of LLMs. Recent improvements have been more about orchestration and "reasoning" loops (i.e. iteratively feeding context back through the model).

      • svachalek 5 hours ago

        For base models they really must be running out of new training materials. It's more about size and architecture. But they still seem to be getting big strides out of improving the post training. They keep dropping point upgrades in under 8 weeks lately, which is an insane pace for product release. At some point it's going to slow down but we're not close yet imo.

    • HappyPanacea 7 hours ago

      Diminishing returns on both intelligence and training, mostly

    • sweetjuly 7 hours ago

      I don't think it's a matter of "staying ahead"; the proprietary frontier models are better, but the trouble for them is that open weight models are good enough in increasingly many cases. This raises the floor on the frontier companies and cuts their total addressable market by commodifying the easier LLM tasks. This is really the central argument of tfa :)

      • cman1444 6 hours ago

        The big question is if there are larger economic gains to be made from ever greater intelligence. It seems to me like that may be the case for only a select few hard problems, while the vast majority of tasks approach their economic ceiling asymptotically with intelligence.

    • hadlock 4 hours ago

      I suspect due to model distillation, their need to stay on top for IPO value is more important than absolutely crushing their competition, who will simply harvest their output and distill it for training data on a ~3 month delay. Also they are probably hitting a wall on increase in intelligence vs training time.

brotchie 8 hours ago

There’s still 50-500x cost reduction in “this is only an engineering problem” low hanging fruit from specialized chips to run the models + improved distillation.

Entirely feasible that by 2031, Fable 5 (or greater) intelligence level models will run cool on smart phones, if not sooner.

  • andai 8 hours ago

    I saw a 1B model yesterday that was fine tuned on Fable output. I found that hilarious, but it did actually make all the scores go up.

    (Actually talking to it, it was about as coherent as you'd expect, i.e. 3/10)

    The floor for "actually usable model" keeps dropping though. (Seems to be about 27B right now?)

  • tgv 7 hours ago

    You're betting on getting getting ridiculously powerful chips to run on batteries in a tiny housing without cooling, while we can't even get enough RAM? It would be a terrible waste of resources. Now we already have TFLOPs wasting in our pockets and backpacks, then we'll have PFLOPs idling, because there's so much time between prompts. Much more efficient to batch it on a server.

    • kaashif 7 hours ago

      At some point we'll have enough RAM, surely. The incentives to produce more are huge and all the fabs are booked out.

      Maybe it'll take 10 years or 20 years. <5 years is not long enough for manufacturing to catch up.

      Not much of a comment on the phone stuff but I'd caution against suggesting technology will never be good enough to do X. Maybe it'll be horrendously wasteful but it might happen.

  • missedthecue 1 hour ago

    If fable 5 could eventually run on a smartphone, I wonder what we'll get out of datacenters. Musk and others are trying to build 100gw of compute by 2030. Will we get something 1-10 million times better than Fable 5, or will the parity gap between local and datacenter capability pair down enormously?

andai 8 hours ago

A year ago I had an aha moment, when I realized that for my purposes, Gemini Flash was not only 9x cheaper, but 3x faster than Gemini Pro, while producing identical output. Who's the best model now!

For a lot of tasks, even small models have saturated them a while ago, and then going cheaper and faster is just pure gains.

For coding I also prefer to do it interactive/realtime, micro-prompting, surgical edits, which the small models can handle just fine.

And then at the top, the real question is consistency. Not "can they do it" but "reliably enough that you don't need to constantly double check everything." (In my experience, not quite there yet, although it's getting way better.)

LordHumungous 8 hours ago

> Reading everything becomes the default. At a cent per document, a model can read every paper in a field, every record in an archive, every email, or every message in a support queue as a matter of routine

Pretty much already happening

Multiplayer 8 hours ago

This means a great de-risking is happening for the costs of deploying somewhat autonomous agents. This has profound implications on the timeline of deployable personal agents. Cost was a significant factor for many people during the OpenClaw frenzy, specifically when they let their agents run somewhat wild. It will become much more palatable, or already has, to install whatever the next generation of token consuming autonomous systems will be.

maxglute 5 hours ago

I imagine some form of token consensus maxing. Repeat tasks X times concurrently to find agreement or outliers. Millions of token burnt deciding what sock to wear.

cush 7 hours ago

It’s way too early to tell the true cost of intelligence. Inference is still heavily subsidized, and training is apparently being funded by a mountain of free money

  • infecto 6 hours ago

    How do we know inference is heavily subsidized? Are all these third party Chinese model providers subsidizing the true cost?

    • cush 4 hours ago

      > How do we know inference is heavily subsidized?

      Balance sheets

      > Are all these third party Chinese model providers subsidizing the true cost?

      It could be argued the subidies are even heavier than American model providers as the price war among Chinese providers is so fierce. https://www.scmp.com/tech/big-tech/article/3358868/after-tri...

      • infecto 4 hours ago

        You did not answer the question at all.

        Third party inference providers for Chinese models are in the US and there are so many players it would be shocking they are all heavily subsidizing it.

        • cush 4 hours ago

          It would be shocking to you if in a market with many competing sellers would there be price competition such that those sellers are operating at a loss?

          If you’re trying to get me to argue that every seller is operating at a loss, I don’t know. But overall the market absolutely is, again because balance sheets

          • infecto 2 hours ago

            Balance sheets is not an answer. Sorry I am just asking for material proof. It would indeed be pretty shocking for every market participant, large or small, funded or unfunded to be operating inference at a loss.

            Plans are definitely subsidized. Inference (token consumption) has never been proven to be operating at a loss and now that you can run many of the large SOTA models from China on sparks or other clustered servers you can figure out some of the math.

            Not trying to argue but just saying “balance sheets” makes zero sense.

            Yes lots of capex spend, hard to say if anyone has spent too much. At the same time demand is increasing for compute.

  • unknownfuture 2 hours ago

    I don't think we have evidence that inference itself is subsidized, but I think we can assume the margins on inference are so low that the foundation model providers have absolutely no chance of recouping training costs and the whole thing is riding on a pile of VC cash and debt.

    My bet is the assumption was one company would ultimately monopolize the space and then they could jack up rates. Alas, the opposite seems to be happening.

imnotr0b0t 7 hours ago

The article is solid. But there is a nuance what he skipped — quality vs price. Sure, GPT-5.6 Luna for pennies can do the same thing what Claude 4.5 Sonnet did for a dollar a year ago. Except Sonnet back then actually carried the codebase, while Luna... eh, not so much. And another thing, speed. You can make it cheaper as much as you want, but if a model thinks for half a minute you save cents but lose time.

  • Tade0 6 hours ago

    The other day I saw a benchmark of coding models that fit in 8GB VRAM - for reference some version of Mistral was added, normally requiring 32GB, but moving along at 4-5tok/sec when partially offloaded to CPU.

    Surprisingly, some of the small models would not only give worse results, but also took longer than Mistral, because they were thinking so much.

    That is an important detail which I was previously overlooking.

qsera 8 hours ago

Idiocy becomes rampant!

  • andai 7 hours ago

    AGI ≈ Ralph × Infinite persistence

    Just like real life!

  • perching_aix 7 hours ago

    becomes rampant? Where have you been the past decades? Hell, centuries and millenia?

    If we're going cynical, may as well go full throttle.

  • Flamkuchlo 7 hours ago

    Whats your problem?

    Even if its not intelligence, a LLM found a bug due to one error message, fixed it, created a PR and it solved it.

    If an LLM is only able to do all of this after training on it and never achieving AGI, we already at the point were it is cheaper to teach one LLM one problem than teaching humans to do so.

iririririr 6 hours ago

all the optimistic pundits ignore the blatant second order consequence of this: the entire economy is pegged on this NOT being cheap!

all the US economy is tied to video cards being used in lieu of gold. cost dropping 100x means the economy bottom falls out.

bkd9 8 hours ago

Author here. I made these plots because I had been searching for them for months and never found quite what I wanted: how the cheapest way to reach a fixed capability level has moved over time. Artificial Analysis publishes enough data to reconstruct it. If someone knows of a source that already tracks this, with historical prices, please share.

  • andai 7 hours ago

    Thanks for the effort you put into this.

    I wonder if it might drive the point even further if the graph scales were linear? Or maybe the progress has been so great that this would make the graphs unreadable?

    https://xkcd.com/1162

  • brabel 7 hours ago

    It blows my mind how fast models are getting better and this is the first article I’ve seen that shows just that and leaves almost no room for disagreement. Well done.

  • jnwatson 5 hours ago

    The animated Pareto front graph by time is truly inspired. Well done!

fabsalvadori 6 hours ago

There is a second-order effect I see, here: when generating work becomes dramatically cheaper, failed work also becomes dramatically more common.

With coding agents, cheap intelligence doesn't just mean producing the same software for less, but rather trying five implementations, letting agents run longer, touching larger scopes, and supervising fewer intermediate steps.

That changes which infrastructure matters. When attempts are expensive, you optimize for success, while when attempts are cheap, you start optimizing for how cheaply you can inspect, reject and recover from failure.

Git was an enormous enabler of cheap human experimentation, and I suspect we'll end up building analogous primitives around autonomous work.

dominotw 6 hours ago

nothing because this shit is not "intelligence"

ppl are doing all sorts of gymnastics to tell claude to slow its roll with verbosity.