• CompactFlax@discuss.tchncs.de
    link
    fedilink
    English
    arrow-up
    0
    ·
    2 days ago

    “Look how low our cost of inference is”

    “Pay no attention to the marketing budget that exceeds Coca-Cola’s for a small fraction of their revenues”

    What technical, fundamental reason is there for the crash in price? The article just accepts the MSRP as fact. The cost of training can indeed be spread over time but it’s not spread across enough time. The inference cost doesn’t actually drop in reality.

    • Grimy@lemmy.world
      link
      fedilink
      English
      arrow-up
      0
      ·
      2 days ago

      There are constantly new techniques being developed to speed up inference or reduce model size post training. There’s about a hundred different levers to pull that play on inference cost, and some of them don’t have much of an impact on quality.

      • Jiral@lemmy.world
        link
        fedilink
        English
        arrow-up
        0
        ·
        2 days ago

        Yet all major players are still incapable of covering their costs and have to resort to all sorts of nasty accounting tricks to make it apear otherwise. How so?

        • Grimy@lemmy.world
          link
          fedilink
          English
          arrow-up
          0
          ·
          2 days ago

          Because of the massive investments into datacenters. Why is everyone so quick to drink the kool-aid. They are spending stupid amounts of money to capture the market and force everyone into a subscription service, not because it cost money to actually run the models. They are all switching their models to fine tuned quants after the first two weeks as well.

          The open source community has found dozens of ways to lower costs, do you really think these big companies aren’t using the same tricks and haven’t developed even more advanced techniques?

          • Jiral@lemmy.world
            link
            fedilink
            English
            arrow-up
            0
            ·
            2 days ago

            It doesn’t matter what tricks they supposedly use afterwards, the data centets that are built need to run successfully or the bubble pops and then those companies will fall like lead. In other words for those data centers to be successfull the entire real economy has to spend crazy amounts of money on stuff that needs crazy amounts of compute.

            Open AI and Anthropic would really need some good business numbers, right now. The claim that they are just deliberately making their business case much worse than it is, is insane.

            • Grimy@lemmy.world
              link
              fedilink
              English
              arrow-up
              0
              ·
              2 days ago

              The inference cost doesn’t actually drop in reality.

              My comment was about this which is just completely wrong.

              OpenAI is the scummiest company on earth but everyone is ready to believe them when they say shit like "our 20$ plan lets you run up the equivalent of a billion dollars in API fees. Giggles. It’s a good plan for you but not for us ;) "

              The main thing bringing up costs is Sam Altman spending all the “inference” money on raw wafers just to strangle consumer GPU sales.