K2 Horizon: Frontier Performance, Radically Open

(ifm.ai)

71 points | by karimf 1 hour ago

8 comments

  • jjordan 42 minutes ago
    Fully open models really need to be a big part of the AI future. That includes all source code, open training data, how it's organized, fed to the model, processed, etc. Until that becomes a thing you're always going to be left wondering what exactly lies underneath the closed model you are using, leaving open the possibility for societal manipulation.
    • kibae 15 minutes ago
      The training data would need to have a permissive license for this to be possible.
      • ux266478 5 minutes ago
        You could sidestep it by running non-permissibly licensed training data that you purchased through an LLM. Legal attitude so far seems to be that this is transformative as long as it's not 1:1. The question on whether or not the end result is copyrightable of course remains controversial and inconsistent, but that question is also fairly irrelevent. You don't get more libre than public domain.

        That's a fair amount of computational and labor overhead mind you, as you'll need to verify and prune the quality of your mountain of synthetic data, but certainly possible.

        Though this assumes the legal system is a rational actor playing by the set of rules it claims to. In fact, I highly suspect you could get very unlucky and get an unfavorable ruling against you, because you stepped on a big pile of money's toes in the process of doing this.

    • trvz 20 minutes ago
      Why? Sure, I’d prefer it, too, but this is just another GNU/Linux vs. macOS situation: most of us would prefer the first, but actually get shit done on the latter.
      • zufallsheld 5 minutes ago
        Without open-source, there'd be no macOS.. So good thing, it exists.
      • homarp 18 minutes ago
        which is why everyone runs docker on mac, to get shit done.
  • mmastrac 9 minutes ago
    The comparisons with other models here are odd.. the other models change depending on the task. It would be far more useful to at least compare against the more recent open models (DS4Flash/GLM53Flash/Qwen38).
  • jon9544hn 41 minutes ago
    Here’s the link (K2)[https://ifm.ai/k2/] as the originally linked link is a login url.
  • piinbinary 44 minutes ago
    A bit off topic, but I think I'm starting to get model fatigue. These come out 10x faster than new Javascript frameworks were coming out 10 years ago (at least new models are far easier to adopt).
    • wuhhh 33 minutes ago
      At least this one can claim being fully open to differentiate it
    • kelseyfrog 33 minutes ago
      Just wait until RSI gains enough traction. We'll be compute-limited rather than labor-limited.
  • sottol 39 minutes ago
  • afzalive 19 minutes ago
    Not to be confused with Kimi K2. Out of all the names they could've used, they picked one that would be confusing.
  • kamranjon 44 minutes ago
    it's funny that the tagline is Radically Open, but you're immediately hit with http login - maybe this was the wrong link?
    • sottol 39 minutes ago
      It's not the blog post, but there's some info here:

      https://ifm.ai/k2/

      375 A23B, 36 A4B, 32B, 7B, 3.7B, 0.9B variants.

      > 32B: Ranking among the top models in its class, 32B is our most powerful dense model, balancing capability, adaptability, and local deployability.

      > 7B: The industry’s best-performing model under 10B combines strong software engineering and expert knowledge in a package small enough to run on a phone.

    • gs17 42 minutes ago
      https://ifm.ai/k2/ seems to work for me.
  • luckydata 35 minutes ago
    both repositories for pre-training and post-training are actually empty... someone might have jumped the gun on the release.