K2 Horizon: A connected fleet of six open models (ifm.ai)
jjordan 13 hours ago
trvz 12 hours ago
homarp 12 hours ago
zufallsheld 12 hours ago
didibus 12 hours ago
verdverm 11 hours ago
I personally find the analogy unconvincing, the UX dimension is completely different as I can use the same harness with any model; and the year of the linux desktop is coming soon (tm)
eikenberry 9 hours ago
kibae 12 hours ago
ux266478 12 hours ago
That's a fair amount of computational and labor overhead mind you, as you'll need to verify and prune the quality of your mountain of synthetic data, but certainly possible.
Though this assumes the legal system is a rational actor playing by the set of rules it claims to. In fact, I highly suspect you could get very unlucky and get an unfavorable ruling against you, because you stepped on a big pile of money's toes in the process of doing this.
alightsoul 11 hours ago
Are LLMs what we need to make all data public domain? This way it could be used for that purpose
echelon 12 hours ago
The first broadly useful fully open source models will do this.
We already have open data / open code / open weights for some domain-specific cases, such as audio models trained on large open datasets, eg. Tacotron / LJSpeech from waaay back in the day, though that is certainly not SOTA anymore.
Distillation could possibly be considered an early case of this as raw AI outputs are themselves not copyrightable unless humans enrich, filter, or transform them. Granted, that does not handle the cases where the outputs are sufficiently similar to copyrighted original works.
waffleiron 11 hours ago
chaosharmonic 11 hours ago
That said, I don't necessarily disagree with you. Talkie[1] presents an interesting case for it being at least possible to do this entirely on public domain material.
But even that used Claude somewhere in the course of its training pipeline (it's listed as a contributor on their GitHub), so again, how granular you want to get with that is still a question.
jjordan 11 hours ago
Decentralized unstoppable storage, combined with decentralized unstoppable training, sorta like SETI for AI training. The seed of this tech already exists with IPFS and others like it.
We know (some? all?) of the big labs have skirted copyright laws at one point or another. Truly open models would just build on what is publicly available.
alightsoul 11 hours ago
embedding-shape 10 hours ago
idiotsecant 5 hours ago
embedding-shape 11 hours ago
chme 10 hours ago
embedding-shape 8 hours ago
dotancohen 9 hours ago
idiotsecant 5 hours ago
dotancohen 5 hours ago
ceejayoz 3 hours ago
rmunn 2 hours ago
Personally, I'd prefer a fixed term. I know enough independent authors making a living from selling their books that I'm willing to allow the fixed term to be large, like 50 years from date of completion of the work. (With a good definition of "completion" so someone can't cheat by editing a couple lines per year to keep something copyrighted indefinitely). The simpler the rule is, the easier it is to understand, and the harder it is to cheat it. The more complicated you make a rule, the more loopholes get found.
idiotsecant an hour ago
jrm4 8 hours ago
tshaddox 8 hours ago
jchw 2 hours ago
I think it is serious. In which case, I gotta say, it really seems like you didn't spend much time thinking about this. "A 1 year grave period for everyone to pull stuff off they don't want to be a part of" - How does that work when the Internet is already full of unauthorized reproductions, most of which people aren't even aware of? Even ignoring practical considerations, when literally everyone is basically stuck using the Internet for everything, this seems a bit unfair to anyone who isn't onboard, akin to The Onion's Google Opt-out Village. But there are so many practical issues with this, it would be easier to list the number of problems this doesn't have. You accidentally leak something to the Internet and it becomes commons? What happens when other people leak things to the Internet? How about revenge porn?
Not minor stuff that can easily be papered over, this literally reintroduces the problem of needing to care about the provenance of data again, in a way that can't be automated, which makes the whole thing entirely moot. All just to make training data for AI models easier to distribute?
I'm all for intellectual property reform, maybe even fairly radical. But this just seems like it wasn't thought out.
If this was satire, well, I took the bait. Oddly convincing despite being hard to believe.
ignoramous 11 hours ago
cute_boi 11 hours ago
__MatrixMan__ 11 hours ago
Sure there are all kinds of problems with that situation. But it still demonstrates that they can be coerced: play nice or don't play at all.
verdverm 11 hours ago
Open models can be used/changed for social manipulation too, by anyone, which scares a bunch of people, as opposed to the dark pattern manipulation from Big Ai/Tech
culi 7 hours ago
theplumber 6 hours ago
a11r 12 hours ago
All that said, the headline claims do not match the self-reported performance. For example, the dense 32B model is significantly behind Qwen3.8 27B (chart towards the bottom of https://ifm.ai/blog/k2). Gemma4 31B is not in the comparison set. This is the most important sweet spot for self hosted open-weight models today and real competition here will be very welcome.
xienze 12 hours ago
The 7B does look very, very good however.
WithinReason 12 hours ago
bluejay2387 11 hours ago
baron3dl 11 hours ago
cogman10 10 hours ago
It failed my basic test I like to ask models and generated incorrect code. When prompted about the bug, it preceded to start hallucinating non-existent APIs. After doing that it got caught in a loop trying to desk check the solution that didn't work.
dotancohen 9 hours ago
cogman10 9 hours ago
The reason I personally like my question is because it's pretty close to some of the real world work we do. It's mostly mundane and easy to bang out, but really easy for someone to do a n log n solution where an n solution exists.
A good example (but not my question) would be something like
"I have a list of People objects with a `first` and `last` name. Write a function which groups together all the People with the same last name in `your language of choice`"
dotancohen 8 hours ago
cogman10 8 hours ago
But much earlier they did and, apparently, these really small models still do. At this point it serves as more of a smoke test for me. Success means little, failure means a lot.
dotancohen 5 hours ago
xienze 9 hours ago
cogman10 9 hours ago
I wouldn't have dreamed to use this as an agent model.
7B models of the past have been able to pass this question. I've not tested it on a 4B model until now.
cogman10 7 hours ago
The first attempt with 7B the model got stuck in an infinite loop.
RandyOrion an hour ago
piinbinary 13 hours ago
kelseyfrog 13 hours ago
JSR_FDED 10 hours ago
wuhhh 13 hours ago
hungryhobbit 11 hours ago
But over time, more and more people got into the chip-making business, and the big players started releasing more and more chips. Now only the die-hard CPU trackers worry about every new CPU and exactly how it's better ... while everyone else just worries about "which CPU will be good enough at this moment".
I think models are on that same arc.
pantelisk 9 hours ago
Or who remembers the dancing disease of 1518, were people would stop what they are doing and start randomly doing the same dance. The lords? Out of their minds. The priests? Terrified the devil had taken hold of the flock! I have come to believe that it was probably some tik-tok like hype trend of doing a fortnite dance while waiting in line for bread and communion. And the energy back then, like now, was off the charts.
Hype and memetic trend seeking encoded deep in human psyche.
dgellow 11 hours ago
cesarvarela 11 hours ago
mzmzmzm 9 hours ago
culi 7 hours ago
uniclaude 12 hours ago
jon9544hn 13 hours ago
justin_ 10 hours ago
Some other open models I'm aware of:
- OLMo
- Apertus
- Soofi
- OpenEuroLLM
- llm-jp
OLMo is perhaps the most famous, and their Dolma training corpus has been reused in other projects. It looks like the K2 training materials haven't been released yet, but I'm interested to see what they did for training "long-horizon agentic tasks". I'm aware of SWE-smith + SWE-gym but I'm guessing there's a lot more out there now.I'm no expert, which is part of why these projects excite me. I'm hoping they can be good projects to learn from as well.
mmastrac 12 hours ago
cogman10 11 hours ago
There is, for example, no Qwen3.8 7B.
It is odd to me, though, that they didn't run the same benchmark suite for the various quants.
kzrdude 8 hours ago
kamranjon 13 hours ago
gs17 13 hours ago
esafak 13 hours ago
wmedrano 13 hours ago
sottol 13 hours ago
verdverm 11 hours ago
sottol 13 hours ago
375 A23B, 36 A4B, 32B, 7B, 3.7B, 0.9B variants.
> 32B: Ranking among the top models in its class, 32B is our most powerful dense model, balancing capability, adaptability, and local deployability.
> 7B: The industry’s best-performing model under 10B combines strong software engineering and expert knowledge in a package small enough to run on a phone.
throwawayffffas 6 hours ago
afzalive 12 hours ago
bee_rider 12 hours ago
Topfi 9 hours ago
TechSquidTV 10 hours ago
edit: Tried signing up and using the internal playground. Holy shit thats fast.
sottol 13 hours ago
reasonableklout 10 hours ago
luckydata 13 hours ago
prometheus1992 12 hours ago
villish 11 hours ago
luciana1u 11 hours ago
dakolli 11 hours ago
adrian_b 11 hours ago
For example, 3.3 Tbyte for code reasoning, 4.5 Tbyte for mathematical reasoning, 8.4 Tbyte of pre-train behaviors, and so on.
I did not compute the sum of the dataset sizes, but it appears to be some tens of Tbyte. Nonetheless, I assume that this amount of training data is more than an order of magnitude less than what OpenAI, Anthropic and the like have used, which must have been at least many hundreds of Tbyte, but more likely several thousands of Tbyte of data.
luciana1u 7 hours ago
lambda 3 hours ago
* https://github.com/ifm-ai/xllm * https://github.com/ifm-ai/horizon-post-train
Their previous model, K2 Think V2, was release with fully open training data and recipe, so I would imagine that they are committed to that, but yeah, the repos for this new model are still just placeholders.
* https://mbzuai.ac.ae/news/k2-think-v2-a-fully-sovereign-reas... * https://github.com/LLM360/Reasoning360