• eicker@lemmy.world
    link
    fedilink
    English
    arrow-up
    73
    ·
    13 days ago

    We are repeating an old pattern in computing: throw more hardware at the problem until efficiency becomes impossible to ignore. Bigger models have delivered remarkable gains, but they’re increasingly expensive. The next breakthroughs may come less from adding parameters and more from smarter architectures, better algorithms and more efficient inference.

    • Vlyn@lemmy.zip
      link
      fedilink
      English
      arrow-up
      45
      ·
      13 days ago

      DeepSeek has really led the way here, especially as they are a bit more hardware constrained. Plus they openly publish their findings and release open source models, so high hopes there.

      It’s probably China’s play to pop the AI bubble, but I’m all for it (:

    • Rothe@piefed.social
      link
      fedilink
      English
      arrow-up
      20
      ·
      13 days ago

      Except there likely won’t be a lot of further breakthroughs if we burn down our planet faster than we already do.

      This is all an expenditure of vast amounts of energy for literally no gain for anybody except a handful of billionaires and their corporations.

    • JustDorky@lemmy.world
      link
      fedilink
      English
      arrow-up
      3
      ·
      12 days ago

      That’s literally exactly what Chinese researchers are doing at DeepSeek and they’ve built frontier models with that philosophy