• mysteryhumpf@feddit.org
    link
    fedilink
    English
    arrow-up
    6
    arrow-down
    10
    ·
    1 month ago

    Local LLMs are cool but also pretty slow compared to cloud. If you have to wait half an hour for your Feature while coding you might still opt for the cloud agent.

      • mysteryhumpf@feddit.org
        link
        fedilink
        English
        arrow-up
        3
        ·
        1 month ago

        Yes ofc I ran Gemma 4 for example, but compare that to the speed of Gemini in the cloud the difference is massive.

        • irate944@piefed.social
          link
          fedilink
          English
          arrow-up
          3
          ·
          1 month ago

          How much RAM do you have and which version of the model did you run?

          Local LLMs can be just as fast as long your device clears the requirements. If you noticed a huge difference, there’s a really good chance that you tried to use a model that requires more RAM than you have

    • f314@lemmy.world
      link
      fedilink
      English
      arrow-up
      1
      ·
      1 month ago

      Yes, they are slower. However, I think that the pricing we’re going to see from the cloud providers might be enough to deter quite a lot of people. At least I hope so:

      The fact that we’re already used to blazing speed generation kinda sucks. Local models are a much more sustainable way of unlocking the benefits of LLMs than giant ecosystem- and community-destroying data centers.