i have been following Chinese models for about two years now because they are open-weight and qwen is fun to run on my kubernetes cluster, the news about this Apache licensed model complete with a recipe to make it again with potentially different ingredients is making we want to abandon them for something more ideologically sound and way more interesting

this new model, couldn’t universities rebuild it with different contexts and study it in ways you can’t reproduce in other models? like, is k2 horizons a good scientific foundation on which to study machine learning?

or am i just lacking way too much context and falling for hype?

  • Natanox@discuss.tchncs.de
    link
    fedilink
    English
    arrow-up
    0
    ·
    19 days ago

    My question would be if it uses stuff like CommonCrawl as training material. Given their size I assume they used anything incl. unethical training data, but at least admit it?

    The only models I ever found that even tried to only resort to ethically obtained data, being FOSS etc. were the tiny ones from PleIAs. And as expected they’re completely useless.

    So far I concluded that an “ideologically sound” LLM is impossible due to lack of training data. Unless your ideology allows to steal stuff.

    • melfie@lemmy.zip
      link
      fedilink
      English
      arrow-up
      1
      ·
      19 days ago

      I’m assuming a lot of its training data is synthetic and distilled from Chinese models that were themselves trained from pirated data and distilled from American models trained on pirated data. It would be quite remarkable if the training data involved no piracy whatsoever. Then again, it’s open source, so I suppose it would be essentially reversing a reverse Robin Hood.