• Azazel@lemmy.ml
    link
    fedilink
    English
    arrow-up
    2
    arrow-down
    2
    ·
    17 hours ago

    I run a 122B model locally. I love local models. If you think they’re anywhere near as capable as the data center models you’ve either not tried them or are lying to yourself

    • humanspiral@lemmy.ca
      link
      fedilink
      English
      arrow-up
      1
      ·
      44 minutes ago

      qwen flash is great. Privacy/sovereignty and protection from others stealing your data is very valuable and equivalence to frontier 3-4 months ago is very impressive. Your point is correct, and larger models have the capacity to solve problems with more general knowledge baked in, where smaller models simply can’t find critical information from search, or develop a strategy as well if they are clueless, and just repeatedly randomly guessing.

    • realitista@lemmus.org
      link
      fedilink
      English
      arrow-up
      5
      ·
      16 hours ago

      I know that is the case today. But hardware and software will improve and the task sets they will be capable of will grow.

      Are you using Qwen? Where do you find it useful/not useful?

      • Azazel@lemmy.ml
        link
        fedilink
        English
        arrow-up
        5
        arrow-down
        1
        ·
        16 hours ago

        Yeah Qwen 3.5 122B-A10B. It’s the most capable model I’ve found that I can run on my machine. It’s useful for simple tasks but honestly it’s around the border if correcting/checking its works is comparable effort to doing it myself. So I don’t really use it in reality it’s more a novelty. I’ve also messed around with Kimi K3 since the benchmarks said it was awesome. It’s definitely a league above Qwen but must be benchmaxxed to hell cuz it’s not even close to the same league as opus (I can’t speak to OpenAI models as my team has a Claude account so those are the paid models I know)