OpenAI Jalapeño: Better than Nvidia Blackwell

(newsletter.semianalysis.com)

79 points | by bmulholland 5 hours ago

6 comments

  • jimmySixDOF 14 minutes ago
    I love how now you have to consider the possible s** posting motivation behind analysis of a trillion dollar industry being conducted at a world-class level by a bunch of ex Reddit and 4Chan adjacent mods -- it's one of the best stories in AI that SemiAnalysis is not cut from the same cloth as Gartner McKinsey et al
    • tmp10423288442 1 minute ago
      SemiAnalysis’ founder was roommates with Anthropic people, not OpenAI, so he may be slightly (very slightly) more objective here.
    • xyzsparetimexyz 0 minutes ago
      s* posting? sex posting?
    • verall 5 minutes ago
      semianalysis is pretty good
  • anthonypasq 28 minutes ago
    Continued hardware improvements really make it hard for me to believe token prices will not continue to plummet.
    • dgellow 7 minutes ago
      There is just so much downward pressure on token price, from every direction. We would need a completely new understanding of economics to explain why the price shouldn’t go down. Or market collusion/regulatory manipulation.
      • simianwords 2 minutes ago
        The price has been going down for ages, its not clear what you are pointing at
    • datakan 21 minutes ago
      Token prices coming down means nothing if the models keep wasting them
    • mathisfun123 3 minutes ago
      this is a story about a proprietary accelerator being built/designed by a token provider. and you think they're going to return the efficiency gains to the customer instead of capture the value for themselves? interesting take.
    • gwerbin 12 minutes ago
      Hopefully this also means billionaires can stop trying to drop data centers into residential neighborhoods with zero noise control and polluting on-site generators, signing local politicians on with NDAs, calling for eminent domain to seize homes to build power lines to data centers, etc. etc. etc. Not to mention the water use controversy.

      Token prices plummeting is probably a good thing, but not without the regulatory backstops that prevent these effectively industrial facilities from being operated with no regard for the externalities they impose on people who live near them.

  • ChoosesBarbecue 1 hour ago
    This is most impressive. The interesting question to me, is outside of the LLM accelerator space: will generalized chips have massive leaps in performance once LLM technology is used to create the next generation? In general, will we see rapid advances while we extract the value of these models in creating architectures? I'm so far removed from the space that this is a very naive interpretation of all this, but I'm curious.
  • empath75 2 minutes ago
    When people talk about the commodification of inferencing, they imagine a future where everyone has access to frontier models and can run them at the same cost, and what will actually happen is closer to the commodification of _oil_, where only a few companies have the scale to produce it at a competitive price, and advances like this are _why_.

    Once models are more or less interchangeable, the price of LLMs will drop to essentially the price of energy required to run them, and the big labs will be able to run them cheaper than anyone else.

  • epistasis 19 minutes ago
    It's so funny to see FP4.... I remember 20 years ago being asked what sort of HPC we needed in genomics, and the answer was basically, "lower precision, faster" for the stuff I was working on. But FP4 is, well, almost comical.

    One thing not on that comparison table: die size. If I'm understanding that correctly, it's about the same as the Rubin, but at 1/3 the number of NVFP4 PFLOPs. (The text disagrees with the table, I'm taking the table as truth, perhaps that's wrong...)

    • nxtfari 10 minutes ago
      Agree, I remember when even half precision made its way into C# sometime around 2020 (I didn’t know much about ML then) and I thought, well I guess that’s a worthwhile tradeoff but I can’t imagine going lower. Lo and behold (1-bit Bonsai) how much lower you could go.
  • varispeed 24 minutes ago
    Why they don't research how to make their own RAM and they have to buy it from the common market?

    They should GTFO with this crap.

    Create barriers to computing for ordinary people while milking businesses for tokens.

    • petcat 19 minutes ago
      Building a custom-designed ASIC is much easier than producing state of the art memory chips.

      There's a reason why Micron and Nvidia are the crown jewels of American technology right now and for the foreseeable future.

      • JV00 15 minutes ago
        Nvidia does not make RAM
      • brcmthrowaway 17 minutes ago
        NVIDIA produces memory?
        • fc417fc802 10 minutes ago
          Fabless AFAIK. And that's the actual problem - drawing up CAD diagrams doesn't help if the factories are fully booked out.
    • datakan 18 minutes ago
      People keep saying stuff like this without understanding what it takes to make RAM. It's one of, if not the most, heavily patented things in the world. The second you dip your toes into those waters the lawsuits begin.

      If somehow you get around the patent issues, you're now faced with huge research and development costs, fabs to build, processes to sort out and all of that has very high failure rates.

      Last time I checked Micron was the largest patent holder in the world and even for them this is a hard area where they are number 3 in the market.