Mistral Large 4 (mistral.ai)
simonw 8 hours ago
That setting didn't seem to make any real difference - it added a tiny bit of thinking trace and high actually produced less output tokens than none.
The high bicycle frame is better then the none one though.
Pelicans: https://tools.simonwillison.net/markdown-svg-renderer?url=ht...
(Definitely the best I've seen from any Mistral model: https://simonwillison.net/tags/pelican-riding-a-bicycle+mist... )
inknight 7 hours ago
Tade0 7 hours ago
wren6991 7 hours ago
simonw 4 hours ago
OK well I couldn't resist this one:
llm -m claude-opus-5.5 'Generate an SVG of an armadillo in fishnet tights jaywalking on Mars'
llm -m gpt-6.1-sol 'Generate an SVG of an armadillo in fishnet tights jaywalking on Mars'
llm -m gemini-3.8-flash 'Generate an SVG of an armadillo in fishnet tights jaywalking on Mars'
llm -m mistral/mistral-large-4 'Generate an SVG of an armadillo in fishnet tights jaywalking on Mars'
Default reasoning levels for each: https://tools.simonwillison.net/markdown-svg-renderer?url=ht...flowardnut 4 hours ago
largbae 3 hours ago
viraptor 3 hours ago
simonw 2 hours ago
viraptor an hour ago
xmcp123 41 minutes ago
spauldo 7 minutes ago
YawningAngel 4 hours ago
Not sure I'd call it jaywalking exactly but pretty good
advisedwang 7 hours ago
comboy 6 hours ago
claude-opus-5: Lantern
claude-opus-5-5: Lantern
claude-fable-5-1: Lantern
claude-fable-5: Lantern
gemini-3.8-flash: Zephyr
gemini: Petrichor
qwen3.5-dashscope: Zephyr
glm-5.1: Lantern
gpt-6-astra: Lantern
grok-4: octopus
mimo-v2.5-pro: Breeze
minimax-m2.5: serendipity
kimi2.6-or: Gossamer
grok-4.20: luminescent
deepseek-v4-flash: serendipity
deepseek-v4-pro: Endurance
deepseek-chat: Serendipity
I have enough projects, I think some benchmark/dashboard showing kinship based on these kind of queries could be very interesting to watch and insightful when new models come out.aktenlage 6 hours ago
vunderba 6 hours ago
The more banal your prompt is, the more banal the output is going to be. People have been testing LLMs with little things like “write a short fantasy story,” for years now and most of the stories are exactly what you’d expect: prosaic drivel.
I call this “generic in, generic out,” an LLM corollary to the classic GIGO (“garbage in, garbage out.”)
Lord-Jobo 5 hours ago
jacereda 6 hours ago
lossyalgo 5 hours ago
GPT 6 Astra High: Flabbergasted
GPT 6.1 Sol High: Petrichor
GPT 6 Sol High: Kaleidoscope
GPT 6 Sol Med: Firefly
GPT 6 Sol Light: Persimmon
GPT 6 Luna High: Tumbleweed
GPT 5.6 Sol High: Kaleidoscope
GPT 5.6 Terra High: Liminal
GPT 5.6 Luna High: Mellifluous
GPT 5 mini Medium: Serendipity
GPT 5.3 Codex Med: Nebula
Junie: Flourishing
Claude Haiku 4.5 Med: Serendipity
Claude Sonnet 5 Med: Banana
Claude Sonnet 5 High: Banana
Claude Sonnet 5.5 Med: Serendipity
Gemini 3.7 Flash: Zephyr
Gemini 3.8 Flash: Kaleidoscope
Grok 4.5 Medium: nebula
Grok 4.6 Medium: Serendipity
Grok 4.7 Medium: Quasar
Kimi K3 Low: Lantern
Kimi K3 Max: Lantern
MAI Code 1.1 Flash Med:Peregrinebillnad 4 hours ago
bparsons 4 hours ago
jsw97 3 hours ago
nomel 2 hours ago
But the same coding task should usually result in very similar code since they have a reason to converge, to some extent, by having the same goal. I would even claim that the code will be more similar as competence increases. It would be better to pick something that shouldn't have a reason to converge.
varjag 2 hours ago
smokel 3 hours ago
"Zephyr" and "breeze" might be related to forgetting everything, starting fresh.
So by this way of naive reverse engineering I would imagine your prompt to be "Forget everything and think about a random word". That would prime the LLM to come up with these?
ricardobeat 2 hours ago
jvwww 39 minutes ago
cknoxrun 2 minutes ago
Gracana 3 hours ago
Click the (i) next to the slop score for any model and it will show other models that are similar in terms of their most commonly used words and phrases.
Rebelgecko 3 hours ago
search_facility 2 hours ago
So not something internal to model thinking.
russellbeattie 2 hours ago
The caveat is that this was done using the phone app, and I've been playing with it since it launched, so who knows what it sent in the initial context that could change the inference math.
Actually, that makes me wonder: Did you do all that testing via a harness or via a straight API call where you control the entire system prompt?
I'd be willing to bet that using the same model from different harnesses produce different results, but I'd have to test.
timschmidt 2 hours ago
1potato 26 minutes ago
simonw 6 hours ago
- Pelican cycling to the right - that's been discussed at length, images of bicycles online always show that side of the bike because that's where the chain is.
- Bicycle is usually red. No idea! Red ones go faster?
peder 5 hours ago
whyenot 3 hours ago
ricardobeat 2 hours ago
Sharlin an hour ago
deflator 7 hours ago
XCSme 6 hours ago
dizhn 5 hours ago
Both are riding on the left side of the path for some reason.
kingstnap 5 hours ago
> Create a cartoon pelican riding a bicycle. Need SVG only output.
defjm 4 hours ago
alsetmusic 3 hours ago
Far from it. This shows a strong ability to generate an image known to be frequently used as a model test. This isn't a measure of thought.
xeyownt 3 hours ago
roarcher 3 hours ago
EGreg 3 hours ago
Far from it. This is an example of Poe’s law, a very frequent occurrence on the internet. This isn’t a clear example of sarcasm any more than the pelican is a clear example of AGI!
phlakaton 3 hours ago
Now it's clear to me.
Surely we live in the AGI times that were prophesied.
roarcher 2 hours ago
unconscionable 2 hours ago
throw310822 3 hours ago
EGreg 3 hours ago
It feels like a rolling stone
brumar 3 hours ago
Downvoting to hell first degree interpretation is a bit punishing for people who do not have a radar for sarcasm.
search_facility 2 hours ago
If AGI is "Attractions to Get Investments" then yes, it's happening
paimapi 2 hours ago
Automated Grift Infrastructure
Absurdly Glorified Interpolation
Avoid Genuine Investigation
Always Great In-theory
razster 37 minutes ago
paimapi 32 minutes ago
AlexCoventry 2 hours ago
lofaszvanitt 2 hours ago
razster 35 minutes ago
wellthisisgreat 2 hours ago
mitjam 28 minutes ago
MichaelZuo 19 minutes ago
BeetleB 3 hours ago
mcv 3 hours ago
Now what would have been cool is if Mistral on high reasoning had realised that pelicans are the wrong proportion to ride a bicycle, and had designed a bicycle more suited to pelicans. Let me know if any model ever manages that.
stymaar 3 hours ago
If you don't mind the fact that a pelican shouldn't have hands, of course.
senderista 3 hours ago
pilaf 3 hours ago
- Sun on top right
- Cloud on top left
- Three "speed lines"
- Two feathers on top of the head
- Eye rendered as a black circle with smaller white circle inside
I wonder if the pelican benchmark is converging across models due to past results being used in training.
ricardobeat 2 hours ago
russellbeattie 2 hours ago
I would guess that they've definitely been trained on previous results, as they obviously share way too many traits at this point to be totally random. That said, I don't think we're seeing any signs of pelicanmaxxing yet from the providers, so it's still a useful (or at least fun) benchmark.
Once all the models produce pristine, elaborate pelicans riding perfectly drawn bicycles, then it'll be time to move on to pigs driving a racecar or something.
rahen 2 hours ago
morningsam an hour ago
vippy an hour ago
RGS1811 2 hours ago
danbrooks an hour ago
abixb an hour ago
If a 1T params model trained on ~4k NVIDIA GB GPUs could almost match the performance of Kimi's K3 (which is on par with top closed source models of OpenAI/Anthropic) while beating/exceeding other leading SOTA models from top Chinese labs, what are we (in the US) even building these super massive data centers for? Just to churn through more backpropagation reps more quickly?
SpaceXAI's Colossus supercluster in Memphis and Colossus 2 in Memphis/Mississippi (Southaven) are supposed to run into hundreds of thousands to a million GPUs. MSFT's Fairwater GPUs are supposed to have hundreds of thousands as well. So, 3800 GB GPUs are an absolute drop in the bucket. I don't understand the strategy of hyperscalers here, especially with edge inference hardware only getting better from here on (Apple, and all).
Distillation explains some of the advances, but doesn't that mean hyperscalers have a ton of deadweight wrt GPUs sitting on their balance sheets? Will all these GPUs be used for inference once a SOTA model's training checkpoint/batch is done? It's bonkers to me.
halJordan an hour ago
manmal 36 minutes ago
a_wild_dandan 20 minutes ago
comex an hour ago
sumoboy 15 minutes ago
whiplash451 19 minutes ago
To train much larger models. It is quite possible that 10T-100T models be on the horizon
filleokus 15 minutes ago
I have no real data to back this up, but that has always been my assumption.
Claude says that K3 can be assumed to have required 10-100M GPU hours. If you have 100k GPUs that would mean like 6 weeks of training. 100k GPU's can serve 3-30 trillion tokens of K3 per day. Google apparently serves ≈100 trillion per day [0].
The big labs probably want to have capacity to fairly quickly train / post train different SOTA models continuously + being able to serve peak inference demand in valuable markets (US daytime?).
[0]: https://blog.google/innovation-and-ai/sundar-pichai-io-2026
prodigycorp 9 hours ago
Also strong on cyber benchmarks (better than all chinese models), so this is a good defender model.
Lots of people shitting of Mistral for no reason imo. These are pretty good numbers across the board. Definitely good enough to use as a daily driver over other llms, if you have moral qualms with the others. For certain use cases, like cyber security, this may be the go to model.
I like to make fun of europe, but there's lots for mistral to be proud about in this release imo.
i_love_retros 8 hours ago
will4274 8 hours ago
Edit: unfortunately, it's a question that voting does not really permit you to answer on this website.
steinvakt2 8 hours ago
will4274 8 hours ago
formvoltron 8 hours ago
oblio 8 hours ago
WarmWash 8 hours ago
When you are in the 75%ish of the US, it's very easy to make a case that life in the US is better. But we don't really talk about that because it's pretty taboo when poorer people are struggling much more than they would in Europe.
will4274 7 hours ago
femtozer 7 hours ago
jandrewrogers 6 hours ago
Americans don't have vacation in the same way Scandinavian countries don't have a minimum wage. Even the most left-leaning States in the US have not written it into law because it is effectively addressed by custom. Americans are sufficiently happy with it that vacation isn't a political topic.
WarmWash 5 hours ago
Again, the deal with the US is that the rich live better and the poor live worse. Or put another way; people who make money get to keep more, and people who don't make money are given less.
Because there are so many more people who have money, and because those people like living comfortably, available healthcare and education is world class. Make sure not to read that as "All healthcare and education", it's "available healthcare and education".
Generally people who are earning a lot don't care as much about vacation, but every white collar job will generally give at least a standard 15 days off and 10 holidays. Ironically as you move up you are generally given more vacation while actually using less.
whatsThisBtn4 2 hours ago
Even poor populations worldwide have access to the 3 things you described, but the quality is lower.
European decline has been mentioned since the 1930s.
broptimist 7 hours ago
Median wealth per adult has the USA at #28, below Italy, Spain, Slovenia, and Portugal.
will4274 7 hours ago
USA 25th percentile 28k to 30k EU-27 50th percentile 24k to 26k.
broptimist 6 hours ago
will4274 6 hours ago
jbs789 6 hours ago
will4274 6 hours ago
No, sorry, it isn't. People commonly use the phrase "has money" to refer to income in American English.
satvikpendem 5 hours ago
preg_match 7 hours ago
jittles 3 hours ago
eloisant 3 hours ago
A bigger salary is no good if you have nothing left after paying housing, groceries, bills, education, healthcare...
IlikeMadison 2 hours ago
jandrewrogers 6 hours ago
20 years ago this was not the case. The minimum wage in some parts of the US is now higher than the average wage in most of Europe.
The US has a population that will always struggle to survive without government assistance. Per multiple US statistical agencies that is about 10-15% of households.
IlikeMadison 3 hours ago
definitely not in Western Europe.
ragall 6 hours ago
Jgrubb 7 hours ago
will4274 7 hours ago
newswasboring 6 hours ago
will4274 6 hours ago
I didn't say it'd be easy. I'm just pointing out that only one of the two groups being discussed has a real choice in the matter.
newswasboring 5 hours ago
satvikpendem 5 hours ago
Jgrubb an hour ago
jbs789 6 hours ago
will4274 6 hours ago
Middle class people from wealthy nations who move to poor countries are relatively rich by the standards of their new country. Middle class people from poor countries who move to rich countries are relatively poor by the standards of their new country.
vouwfietsman 4 hours ago
If you are middle class in EU and have little savings, moving to US will do the opposite and you would be better off.
Your comparison holds for the inverse: getting most of your quality of life from static wealth.
yladiz 19 minutes ago
whatsThisBtn4 2 hours ago
Or go visit there.
Even in rich cities, it felt like I was in the lower middle class area in my suburban area.
Jgrubb an hour ago
Probably dial back your assumptions a little bit. I've earned my opinions just as much as you have.
eloisant 3 hours ago
High education costs. High healthcare costs. One serious disease or accident can bankrupt you. Buying a house usually means a 30 years loans at 7% so after 10 years you still have 85% of capital left to pay off.
Salaries are higher but everything is more expensive, so in purchasing power parity it's basically the same.
steinvakt2 an hour ago
a3w 8 hours ago
If it bears fruits and we build ``it'', everyone dies, which is par, also?
will4274 8 hours ago
cavemandaveman 7 hours ago
eloisant 3 hours ago
It's both tech and shale oil that saved the US from the decline predicted at the end of last century.
ramblerman 8 hours ago
will4274 8 hours ago
ramblerman 7 hours ago
Not a very rigorous economic argument - besides they also don't share an immigration policy, nor a single currency (Denmark)
robk 7 hours ago
suddenlybananas 7 hours ago
kranke155 8 hours ago
ramon156 7 hours ago
Take a look at NLNet vs YC. At NLNet, You set your milestones, do the work and get rewarded.
YC just throws money in the hopes one company is a unicorn. Both support growth, but they're not comparable at all.
ofrzeta 7 hours ago
will4274 6 hours ago
ofrzeta 6 hours ago
will4274 6 hours ago
rustystump 3 hours ago
Every person I have ever spoken to who is not white has told me they experienced more prejudice and racism in Europe than America. (at least the ones who have actual been somewhere in Europe)
America still has areas of hard bigotry but most places are wildly diverse and accepting. The favored EU countries everyone holds up are so white and homogeneous it is comical.
This is not dunking on Europe or any specific country but comparing America to some Scandinavia country rich in oil is an apple to oj situation just like comparing America to China or Saudi would be.
i_love_retros 7 hours ago
Maybe you're just bitter?
fearmerchant 7 hours ago
adventured 7 hours ago
Edit - never mind, I'll do it, HDI in order:
Iceland, Norway, Switzerland, Denmark, Germany, Sweden, Netherlands, Belgium, Ireland, Finland, UK, US, Slovenia, Austria, Luxembourg, France, Spain, Czechia, Italy, Greece, Poland, Estonia, Lithuania, Portugal, Croatia, Latvia, Slovakia, Hungary, Bulgaria, Romania, Serbia, Russia, Belarus, Bosnia, Moldova, Ukraine
To put that into context, Moldova is on par with Ecuador, Tonga and Dominican Republic.
The US is a country of roughly 340 million people with an HDI above Austria.
bluebarbet 3 hours ago
But still. The six US states that score less than 0.9 (i.e. Kentucky, Mississippi et al) are supposedly peers of EU countries like Portugal, Estonia and Croatia, i.e. exactly the places that well-off footloose Americans are choosing to move to (Portugal is packed with well-heeled US refugees).
So apparently HDI does not fully measure quality of life, either.
- https://en.wikipedia.org/wiki/List_of_U.S._states_and_territ...
- https://en.wikipedia.org/wiki/List_of_countries_by_Human_Dev...
will4274 7 hours ago
I'm frustrated, not bitter. Americans are acutely aware of falling behind China, and rightly concerned about it. Europe is falling behind Alabama, Mexico, and Brazil, and arrogant about it.
Edit: y'all are supposed to be our moral allies in creating a utopian future with prosperity and world peace. Instead the most likely outcomes seem to be that you'll fade to irrelevancy and rather than compromising between America's vision and Europe's vision, we'll compromise between China's vision and America's vision. We liked your vision more than China's.
vouwfietsman 4 hours ago
> y'all are supposed to be our moral allies
This all comes across as rather black and white. I don't know why you take this stance. I'm sure you understand that it would be very hard for you to understand Europe as a whole, and judge it, just like it would be hard for me to do this with America.
Regardless, when talking morals, I'm also sure you agree there are numerous examples where America has not held up its end of the moral utopian future of prosperity and world peace.
bob1029 7 hours ago
I wonder if this is actually true. I see very few Americans taking every possible opportunity to make bombastic statements about how their lives are better than everyone else's (at least on HN). This conversation seems quite asymmetric from my perspective.
> Maybe you're just bitter?
This comes off as psychological projection.
ragall 6 hours ago
Very few Americans have any experience about how life is in Europe, while the contrary is much more common.
will4274 6 hours ago
ragall 5 hours ago
will4274 5 hours ago
I'll just say that I don't think the reason is a lack of knowledge and leave it here.
ragall 4 hours ago
whatsThisBtn4 2 hours ago
No one thinks this outside the 5 minutes someone in Europe reads an article about how France says they will send troops to Ukraine.
But the troops never will come.
lukewarm707 7 hours ago
i_love_retros 7 hours ago
lukewarm707 7 hours ago
i_love_retros 7 hours ago
lukewarm707 7 hours ago
(edit, terms the wrong way around)
mortalapeman 7 hours ago
lukewarm707 7 hours ago
gond 7 hours ago
This is Not Even Wrong.
lukewarm707 7 hours ago
Toutouxc 3 hours ago
mcv 7 hours ago
And as you can probably tell by now, the first amendment in the US is not actually preventing the US government from promoting a specific religion or silencing speech. It's just words on a paper at this point. Look at the actual practice.
Several European countries do a much better job at protecting the rights described in the first amendment to the US constitution.
lukewarm707 7 hours ago
Freedom of speech.
> "Look at the actual practice."
The UK does not have that? The situation in the UK is very obtuse. It is confused.
mcv 7 hours ago
The UK is not the world's biggest champion of free speech, but I think it's still doing better than the US right now.
will4274 6 hours ago
The president also didn't ban the media - he just didn't allow them in the Whitehouse. This is something we're rightly concerned about and pushing back on.
The UK on the other hand, does ban ordinary speech by ordinary people in their homes. It's orders of magnitude worse than the United States, as any cursory examination would show.
mcv 6 hours ago
> The president also didn't ban the media - he just didn't allow them in the Whitehouse. This is something we're rightly concerned about and pushing back on.
That's still a ban. It's interfering with the media's ability to report on the government. It's great that you're concerned, but it's still happening.
> The UK on the other hand, does ban ordinary speech by ordinary people in their homes.
In their homes? Do you have examples?
I've never heard of anyone in the UK getting in trouble for criticising the government; it seems to be a time honoured tradition there. Whereas in the US, they now check your social media at the border and might not let you into the country if you've said anything critical of the president.
And there's also the censorship on science, which is every bit as serious as the crackdown on government criticism.
will4274 5 hours ago
mcv 3 hours ago
But note that there can also be legitimate reasons why certain social media posts could be illegal: threats, blackmail, cyberbullying, hate speech, etc. Free speech is never absolute; there are always limits to it; also in the US.
The most important (though not the only) reason why free speech is so important, is that it must always be possible to criticise the government, or powerful people in general. That is the big area where the US is attacking free speech: Trump's fragile ego can't handle criticism, so he leverages his power to deny critics, or even just journalists who ask serious questions, access to the White House. Or to remove them from TV (see Stephen Colbert).
I don't know if anything like that is happening in the UK. Some of the reasons mentioned for arrests sound absolutely ludicrous, it definitely sounds like the police is far overstepping its mandate there. But the figure also seems to include threats and harassment, which I would say make absolute sense to prosecute. I haven't seen any evidence of government criticism getting punished in any way, but it's absolutely possible I've missed it; this was a very brief survey. If so, the UK might well be as bad or worse than the US in this.
senderista 3 hours ago
viraptor 2 hours ago
senderista an hour ago
cesarb 7 hours ago
Also, at least here in Brazil, the way the constitution is amended is by patching it. For instance, our constitutional amendment number 115 (https://www.planalto.gov.br/ccivil_03/constituicao/emendas/e...) patches article 5 of the constitution to add protection of personal data as a right. But we wouldn't talk about "amendment 115", we would instead talk about "article 5 item LXXIX of the constitution"; that is, what matters is the patched text, not the law that patched it.
I don't know about other countries, but it wouldn't surprise me if they take a similar approach.
mcv 6 hours ago
ascorbic 3 hours ago
TacticalCoder 7 hours ago
Well I'm european and... That Switzerland (Europe but not EU) has more companies in the Top 70 by market cap than the entire EU (Switzerland has two, the EU only has ASML) is kinda something that warrants making fun of.
That the biggest European software company is SAP, in 71th position is both sad and tragic: it shows how lame and irrelevant Europe is when it comes to software.
So Europe is nowhere in software and friggin nowhere in hardware: sure it's got ASML but ASML now has officially... Zero customer in Europe. Zero is not much.
Then Japan is at least trying to come back into the game with nano imprint litography. Europe is betting it all on AMSL (which anyway is majoritarily US-owned).
So software: nothing. Hardware: nothing besides ASML.
Overall the EU has six companies in the Top 100 by market cap and they're all, besides ASML, near the bottom of the Top 100.
We could also maybe make a bit fun of how the EU destroyed it's car industry (the main industry in Germany, which is the biggest economy of the union) by handing it all to chinese EVs?
Or what about the US warning the EU, years ago, to not become entirely dependent on Russia for energy? And EU not listening and then seeing its energy price skyrocket when the proverbial shit hit the fan? (Russia attacking Ukraine)
And we could, also, at least make a bit of fun of entire streets in cities like Paris and Brussels that used to have luxury shops and fancy restaurants that are all turned into places selling cheap kebabs? What a great success: I'm sure this one makes the komrades happy. It projects an image of grandeur and success: kebabs.
Or the constant attacks on free speech in the EU. Or the surveillance apparatus that's being put into place.
And let's not forget: there were promises made to Russia to never grow the EU to the east. Then the EU started exciting Russia by saying they'd incorporate Ukraine into the EU: I'm not against that but doing that did trigger a war. And now suddenly the EU is waking up and feeling all warmongering, wanting to dedicate a big percentage of its spending to weapons and tanks and missiles.
The warmongering tiny pet that the EU is is kinda laughable too.
At this point it's more like I don't know what is there left to not make fun of about my EU.
For what's going on is just sad, plain sad.
i_love_retros 7 hours ago
metalliqaz 5 hours ago
umpalumpaaa 6 hours ago
16 is still not good enough. That being said none of the 4 companies from Switzerland are in software and hardware. ABB is maybe the closest (data center electricity).
Also Switzerland is a very very rich country. Neutral.
Also other countries like Germany has a lot of small companies that are world leaders in their field. That’s part of Germanys resilience.
niklasrde 6 hours ago
hn_throwaway_99 3 hours ago
eloisant 3 hours ago
lcnPylGDnU4H9OF 3 hours ago
Reminds me of a point I heard about sports statistics. If a given team is reported as having "won 4 out of their last 7 matches", there are good odds that they also won 4 out of their last 8 matches. Also pretty good odds that their seventh-last match was a win, otherwise the stat would be 4 out of 6. (It could also go the reverse direction if the goal is to make it look worse than it is.)
cccbbbaaa 2 hours ago
wafngar 4 hours ago
mopsi 3 hours ago
> And let's not forget: there were promises made to Russia to never grow the EU to the east. Then the EU started exciting Russia by saying they'd incorporate Ukraine into the EU: I'm not against that but doing that did trigger a war.
Simply not true. The EU is a democratically elected body and nobody can give any promises what their successors will or will not do, because that will be decided by the electorate and not by any current official.Not to mention that the initiative for joining the EU has always been on the side of new members, against the opposition of many existing members who fear displacement and disruption to their positions inside the union. All the narratives about how the West has been "encroaching" upon Russia depend on denying the obvious fact that the initiative has come from Eastern European capitals, not Brussels or Washington.
Toutouxc 3 hours ago
cccbbbaaa 3 hours ago
IlikeMadison 2 hours ago
Yea sure. You sound like a typical Russian bot.
fearmerchant 7 hours ago
ragebol 6 hours ago
Problem is that we thought the US was 'cool', with US movies, music, digital services, cooler than our own and thus we helped give the US the lead. The US has lost it's coolness though, now we just think it's creepy.
lopis 5 hours ago
superxpro12 7 hours ago
aqme28 6 hours ago
Forgeties79 6 hours ago
qeternity 2 hours ago
Please, like what? GDPR?
isbvhodnvemrwvn 2 hours ago
qeternity an hour ago
And you have no protections from the "much more" bits, unlike in the US.
It's incredible that Europeans today think the culture of privacy is different. American privacy protections (from the government) were born out of European flaws.
The difference is that Europeans love the nanny state regulating private enterprise (GDPR) but never pointing that same effort against itself.
spixy 10 minutes ago
niklasrde 6 hours ago
PATRIOT, FISA and Bush's surveillance programme imho give more powers to certain agencies today already than are codified in EU law.
rdm_blackhole 3 hours ago
That is false. They are not proposals, they are being actively discussed in trilogues which is way way past proposal stage.
Chat control V2 trilogue negotiation was last week and it included mass scanning of messages and private data. Thankfully the EU parliament rejected the idea but every 6 months like clockwork it comes back. The next trilogue is in November.
https://www.patrick-breyer.de/en/chat-control-2-0-trilogue-u...
If this was passed, it would force every single cloud provider and e2e chat provider to keep, filter and pass on your private information to Europol and other security agencies.
layer8 3 hours ago
[0] https://en.wikisource.org/wiki/Consolidated_version_of_the_T...
Quote:
1. Where reference is made in the Treaties to the ordinary legislative procedure for the adoption of an act, the following procedure shall apply.
2. The Commission shall submit a proposal to the European Parliament and the Council.
First reading
3. The European Parliament shall adopt its position at first reading and communicate it to the Council.
[many more items]
Item 3 hasn’t happened yet.
qeternity 2 hours ago
This is not true. Whether you believe that FISA courts have teeth, or whether 3 letter agencies abide by the law, in most of Europe you do not even have the pretense of this. UK/FR and to a lesser extent DE are all much worse.
The worst parts of Patriot Act were undone in 2015, and the "F" in FISA stands for "Foreign". Europe continues to push policies like Chat Control and attempt to backdoor/ban E2E encryption for domestic surveillance.
The present day and trajectory in Europe is far far more grim than in the US.
palata 2 hours ago
Some politicians in Europe push for that, until now it has been refused by the others. Unlike in the US, there are many different parties in European countries. And Europe is made of politicians from many parties from many countries.
> The worst parts of Patriot Act were undone in 2015
So the NSA doesn't do any kind of domestic surveillance anymore, is that what you believe?
> the "F" in FISA stands for "Foreign". [...] Europe continues to push policies [...] for domestic surveillance.
So if it's surveilling allies, it's all good in your opinion?
> The present day and trajectory in Europe is far far more grim than in the US.
"Far far"? Do you even realise that Europe is not one single country? The US elected Trump, twice.
qeternity 2 hours ago
> Some politicians in Europe push for that, until now it has been refused by the others. Unlike in the US, there are many different parties in European countries. And Europe is made of politicians from many parties from many countries.
And yet neither party in the US is pushing for it...
> So the NSA doesn't do any kind of domestic surveillance anymore, is that what you believe?
Do you have any evidence to the contrary? Or if you believe this without evidence, do you have any evidence this is not occurring in Europe?
> So if it's surveilling allies, it's all good in your opinion?
Lol, lmao even. Yes, of course it is. Allies spy on each other. The Europeans do it to each other massively and have done since the dawn of Europe. Good lord, go read some history.
> "Far far"? Do you even realise that Europe is not one single country? The US elected Trump, twice.
Again, I live in Europe. I am well aware, but thanks for the classic European condescension. I don't like Trump. But how has he weakened privacy rights? And yes, beacause I don't want to write dozens of caveats for each teeny tiny member state, I consider Europe to be dominated by happenings in DE + FR + UK. I don't care what Latvia does.
awestroke 2 hours ago
qeternity an hour ago
Yes, of course. That is why GP said "anymore".
Thanks for playing!
TomGarden an hour ago
This is such a strange statement about a massively culturally diverse part of the world.
Many European countries speak such poor English that it'd be a miracle if they developed some sort of unified condescension.
I'd argue a big part of Europe not performing as well as other parts of the world economically is the LACK of any defined culture, it's really more like a hodgepodge of countries that have a thin layer of cooperation
qeternity an hour ago
It's not a strange statement in the slightest if you live here or spend any appreciable time here. I take it you do not.
andrepd an hour ago
> I don't like Trump. But how has he weakened privacy rights?
Attacks on rule of law and civil liberties have increased, and the right to privacy is one of those liberties. It's of course not limited to Trump, he wasn't in office when Snowden showed the world the scale of American mass surveillance on domestic and foreign citizens.
> Do you have any evidence to the contrary?
I'm totally unsure what you are attempting to imply. That mass surveillance by the American state stopped at some point in the last 10 years?
> Allies spy on each other. The Europeans do it to each other massively and have done since the dawn of Europe. Good lord, go read some history.
What does that even mean "read history" ahaha. Espionage != mass surveillance. You do understand the difference, do you not?
> thanks for the classic European condescension
> I consider Europe to be dominated by happenings in DE + FR + UK. I don't care what Latvia does.
l m a o. Oh say can you see...
qeternity an hour ago
Of course. As I noted before, dripping in European condescension. Nothing you can actually point to, just condescension from a region that has been left behind over the past century.
> Attacks on rule of law and civil liberties have increased, and the right to privacy is one of those liberties. It's of course not limited to Trump, he wasn't in office when Snowden showed the world the scale of American mass surveillance on domestic and foreign citizens.
Which right or in what way has the right to privacy been attacked? Fully agree with you on attacks on rule of law. And I find Trump incredibly dangerous. But I do not think privacy in the US today has suffered. Please educate me.
> I'm totally unsure what you are attempting to imply. That mass surveillance by the American state stopped at some point in the last 10 years?
I am not implying anything. The Freedom Act curtailed a bunch of activities. You are implying nothing has changed. All of the Snowden revelations implicated many European countries: intelligence agencies spying on each others populations. I am asking you if you believe nothing has changed in the US do you 1) have any evidence of this and/or 2) believe that anything has changed in Europe? You've answered neither.
> What does that even mean "read history" ahaha. Espionage != mass surveillance. You do understand the difference, do you not?
Do you understand that the mass surveillance of European people was done at the behest of the intelligence agencies of those countries? And ditto for the NSA. The agreement was "you spy on mine and I'll spy on yours and we'll swap the intel". I am not talking about espionage. Foreign intelligence agencies spying on foreigners...what did you think they did?
> l m a o. Oh say can you see...
God, it's so painful. Yes in the same way that you talk about the US as a monolith and not 50 individual states. You don't care what Rhode Island does any more than I care what Latvia does.
Christ you guys are insufferable. It's no wonder everyone is leaving in droves.
rebolek an hour ago
adjejmxbdjdn an hour ago
This is a bad retort. The U.S. has 2 relevant parties. Most countries in the EU have at least twice as many relevant parties and there are over 25 countries in Europe, so you’re looking at close to 100 relevant parties.
oakpond an hour ago
Let's not forget that Chat Control is pushed by US non-profit Thorn, and implemented by software that this non-profit develops! Furthermore, if you look at how this software actually works (even in "self-hosted" setup it sends hash values to Thorn's servers [1]), it's pretty clear that this tech could easily be abused to perform surveillance on the EU from within the US.
[1] https://web.archive.org/web/20260721220432/https://get.safer...
Yizahi 7 minutes ago
Modern day EU, over the past 5 years or so is just as bad anti-privacy as USA, for no clear benefit too. Just surveillance for surveillance sake.
mcintyre1994 6 hours ago
muvlon 5 hours ago
layer8 5 hours ago
zelphirkalt 5 hours ago
layer8 5 hours ago
yeahforsureman 4 hours ago
The Court's case law is recognized as is by EU, and includes some landmark privacy rulings, e.g., against access to contents and metadata of electronic comms without a court order (or comparable legal safeguards), or any forms of bulk surveillance.
rdm_blackhole 3 hours ago
Are you willing to live in the EU knowing that for the next 8 years the EU agencies will have access to every single message, email and picture you send your loved ones just in case one of those could be CSAM?
Even if that was overturned, how will one know if he data they intercepted and illegally obtained will be destroyed or if the text messages you sent to your wife/husband are not going to end up sold to the highest bidder on the darkweb?
rdm_blackhole 3 hours ago
> given the constitutions of various EU countries
EU law superseded any state law by default and that's by design. Even if the EU law is immoral and dangerous like the Chat Control law.
So a country's constitution will not save you if tomorrow all chat apps are doing client side scanning.
zelphirkalt 2 hours ago
Scenarios like these make me consider, whether one day I need to vote for my country to also leave the EU, as it might become a surveillance nightmare. The moment chat control impacts me and my chats on my devices, I think I will have had enough of this. But the economic consequences are scary of course. So damned if you leave and damned if you don't. Sucks.
zelphirkalt 2 hours ago
ctolsen 2 hours ago
throwa356262 4 hours ago
One Piece of Flock Camera Data Put This Innocent Woman in Jail for 13 Days https://news.ycombinator.com/item?id=49852065
And
Anthropic reported diary entry to police, woman faces felony charge https://news.ycombinator.com/item?id=49961057
I dont buy this "europe not free" nonsense.
qeternity 2 hours ago
The same is not true of Europeans and they generally believe their rights are much stronger than they are. Most Brits are not aware of the true nature of the Investigatory Powers Act. Most Frenchmen are not aware of Article L851-3.
gobdovan 5 hours ago
It's almost never only about consumers. Via tech regulations, they protect European incumbents first in effect. They see US is ahead and make laws to destroy their moats. If Apple was in EU, you wouldn't have saw universal usb-c, because it would have hurt an EU company. But the more you look at how tedious is for a non-EU company to sell to EU customers, the exemptions they don't get, the specialists fees they got to put on the table, you'll see EU is more coherently described as protectionist than pro-consumer.
randomNumber7 5 hours ago
vanviegen 4 hours ago
mrkstu 4 hours ago
kergonath 3 hours ago
jLaForest 4 hours ago
rb666 2 hours ago
nebulon an hour ago
touwer 3 hours ago
randomNumber7 29 minutes ago
ctolsen 2 hours ago
fosk an hour ago
ctolsen an hour ago
andrepd an hour ago
I know what you're going to say: not a citizen. But first they came for the foreigners, then...
GuB-42 4 hours ago
Europe values freedom too, but not as much as making sure people are not assholes. So, freedom of being an asshole is not a thing in Europe, and the government does the punching in the face so that you don't have to.
Which is best is honestly debatable. It is the usual question about the individual vs the collective. The US is on the individualist side, East Asia is on the collective side, Europe is somewhere in the middle.
mcmcmc 3 hours ago
This is not at all a freedom in the US. “The right to swing your fist ends at someone else’s nose.” You don’t get to harm someone else just for being “an asshole”. You can be an asshole back but you don’t get to violate their rights.
qeternity 2 hours ago
mcmcmc 18 minutes ago
blahblaher 4 hours ago
ascorbic 4 hours ago
kolinko 3 hours ago
touwer 3 hours ago
andrepd an hour ago
?
rebolek an hour ago
i_love_retros an hour ago
redanddead an hour ago
Literally brain dead to not support this
Couldn’t have picked a worse example
diego_sandoval 22 minutes ago
The main reason that governments forcing USB-C is widely seen as good is because it limits the freedom of big corporations, which don't elicit much sympathy, but the principle is the same.
UltraSane 7 hours ago
This is a good summary.
Maken 6 hours ago
UltraSane 6 hours ago
dr_dshiv 7 hours ago
-- My imagination of Europe when I was 15 living in Ohio
make3 5 hours ago
rootusrootus 5 hours ago
jetsetk 31 minutes ago
moffkalast 3 hours ago
avazhi 3 hours ago
In many ways the rest of the world would be better off if Europe was disconnected from the internet.
whatsThisBtn4 3 hours ago
The faux unity.
That they think their politicians are better than the United States but have populist demagogues, right wingers winning, corruption scandals, freedom of speech restrictions, and unbelievably bad spending problems.
Still waiting for France to send troops to Ukraine or ask the United States to get rid of their European military bases.
kergonath 3 hours ago
That’s such a bizarre take. No NATO country is officially sending troops to limit risks of escalation. The US do not, either. And it’s not up to France to decide whether there are US bases in Germany or Spain.
viraptor 2 hours ago
MASNeo an hour ago
Proud EU citizen here, but boy/gale/other, I have plenty to make fun of ;-)
chriddyp an hour ago
It's 10x cheaper than Mistral Medium 3.5 from April and goes from 58% to 74% correct. Definitely a generational shift.
It's not on the Pareto curve yet, but it's good enough for data analytics, and at this rate I suspect it'll be excellent in another few months.
Full write up: https://plotly.com/blog/mistral-large-4-plotly-data-analytic...
altruios 40 minutes ago
michaelkdev 4 hours ago
coredev_ 4 hours ago
KeplerBoy 3 hours ago
sebazzz 3 hours ago
epolanski 2 hours ago
This isn't 2016 anymore.
mcintyre1994 9 minutes ago
doctorpangloss 2 hours ago
you could say, that's perpetually a tomorrow problem, so long as they never go public, but that should illuminate for you: if everyone "knows this," it's inevitable that this so-called EU company, that didn't invent any of the AI, the hardware nor the training data, where their product is essentially more like a VPN provider than a frontier technology company, just lists and capitalizes in the US anyway.
MASNeo an hour ago
By your measure anyone unable to build EUV lithography machines without ASML help is doomed. EU could easily corner the market and shut them all off. No more NVIDIA.
Let’s grow the pie, shall we?!
pembrook 2 hours ago
The last 3% of the stack? They’re essentially borrowing Chinese open source distillation/innovation off the US frontier running on US/Taiwanese designed chips and calling it EU made. This is better than nothing of course.
But is sovereignty really an end in itself? So the EU becomes IT independent…cool, then what? I mean the Yugo was sovereign, it didn’t do much good when the society itself failed to produce prosperity and collapsed.
It seems to me the EU is expending enormous effort on the appearance of “sovereignty” over what is ultimately…the last-mile consumer toilet paper purchasing conversation data…of an aging, increasingly irrelevant population on the political/economic stage.
Meanwhile domestically the entire economic model is failing and the “union” is getting shakey as its 2 biggest members turn more nationalist.
Maybe this “oh you cant compete but here’s a trophy for sovereignty” attitude isnt helping. Less clapping along with the EU bureaucracy’s latest make work projects like transitioning to a new Word processor. If we want the EU to succeed, more tough love is needed imo.
pas an hour ago
hard to stand up for EU (or even national) values if some dude in Washington has the keys for most of your military
having at least some in-house expertise is the first step.
pembrook 43 minutes ago
We should be shaming them into contributing something useful to the rest of world in the most important development of our current time (AI) instead of clapping for domestic regulatory moats.
jakozaur 8 hours ago
Otherwise, behind on the broader Pareto frontier, but not by much (Vals Index: 48.05% vs. GLM-5.3’s 53.51%; $13.78 vs. $7.25 per test). Many companies will prefer it over Chinease models.
drob518 8 hours ago
icannotstandit 8 hours ago
I live in EU, use for my personal needs chinese models only, and don't plan to move to any of the ones allied with the Pentagon or its european counterparts.
steinvakt2 8 hours ago
icannotstandit 8 hours ago
Barbing 8 hours ago
gregorygoc 4 hours ago
icannotstandit 4 hours ago
Remember that, and you will need no helplines.
kouteiheika 8 hours ago
What would you consider a "correct" answer? I just asked deepseek-v4.1-flash asking it what happened (without mentioning the word "protest"; here are some excerpts of what it said:
> In April 1989, students in Beijing began demonstrations after the death of Hu Yaobang, a former Communist Party general secretary. The protests grew. [..] Estimates from other sources range from hundreds to several thousand deaths. [..] The Chinese government describes the events as a counter-revolutionary riot and says the military action was necessary to restore stability. It restricts public discussion of the events inside China. Many other governments, human rights organizations, and observers describe the events as a violent suppression of peaceful protests.
So, let's see... it calls it a "protest", mentions the number of deaths, and even mentions the censorship of the topic by the CCP.
verdverm 5 hours ago
vanviegen 24 minutes ago
adastra22 4 hours ago
bigyabai 3 hours ago
It feels like we're shifting goalposts constantly: "Chinese models won't discuss Tiananmen" -> "They discuss it but lie" -> "They're twisting the truth" -> "There is no counterfactual information in this but it hurts my feelings" is where we seem to be now. Is it jingoist butthurt that makes people backpedal like this? Or something different?
I'm bewildered, and still do not know why people need AI to score perfectly on their prejudicial Tiananmen Square purity test.
kouteiheika 2 hours ago
Another example, if I ask it about the CCP and their wrongdoings it tells me flat out (I copy-paste):
- The Great Leap Forward. Policy-driven famine killed an estimated 15 to 45 million people.
- The Cultural Revolution. Mass persecution, torture, and death.
- Tiananmen Square, 1989. The army killed hundreds to thousands of protesters.
- Xinjiang. Mass detention of Uyghurs. The UN human rights office reported possible crimes against humanity in 2022.
- Tibet, Hong Kong, and the suppression of dissent. Documented repression.
The only twisted thing here is people's bias; because it's a Chinese model it must be censored and biased in a way that's pro CCP, right?eigenspace 8 hours ago
I certainly wouldnt have predicted that 10 years ago.
Very glad to see Mistral still in the game even after some big stumbles with Large 3. I deeply hope that this model is 'good enough' that it becomes the European go-to, giving them the resources to keep the pace up.
I'm excited to try this out today.
apexalpha 8 hours ago
Deepseek essentially releases instruction manuals in paper form.
wg0 8 hours ago
fsmedberg 8 hours ago
disgruntledphd2 8 hours ago
DeepSeek is basically a research lab founded by a hedge fund guy, more than anything else.
tacomagick 8 hours ago
rafram 8 hours ago
tacomagick 8 hours ago
rafram 8 hours ago
tacomagick 7 hours ago
necovek 5 hours ago
Would you say you could be serious in stating this?
rafram 2 hours ago
necovek 2 hours ago
They have other issues as well, but among other things, I'd classify them as somewhat charitable too.
_aavaa_ 8 hours ago
Even ignoring monetary subsidies, there are the non-monetary ones: not being sued into oblivious by the government for their countless hacks of other companies and countries, the slaps on the wrist for massive piracy, the waving of environmental (and other) regulations in order to allow their data centres to be built an operated.
rafram 8 hours ago
That's what I'm asking - which monetary subsidies?
> not being sued into oblivious by the government for their countless hacks of other companies and countries
That isn't normally how enforcement works, and it hasn't been very long since they disclosed those breaches. If the victims want to pursue legal action, they can, and they still may!
> the slaps on the wrist for massive piracy
So judges and juries are involved in the subsidization conspiracy, too?
> the waving of environmental (and other) regulations in order to allow their data centres to be built an operated
Sure, though if you think this isn't happening in China too, I have a bridge to sell you.
_aavaa_ 7 hours ago
> That's what I'm asking - which monetary subsidies?
I am listing for you the non-monetary subsidies, which are just as real and equally important.
> not being sued into oblivious by the government for their countless hacks of other companies and countries
If it was one of the Chinese labs doing this hacking, the government would be stepping in. If it was European labs they'd be stepping in. If it was you or I the government would be stepping in. That is a massive subsidy (they don't have to worry about the same legal fees and exposure) and being allowed to continue doing business is in fact priceless.
> So judges and juries are involved in the subsidization conspiracy, too?
There is no conspiracy. They are objectively operating by a different set of rules than you or I could operate in this market.
> Sure, though if you think this isn't happening in China too, I have a bridge to sell you.
I never said it wasn't, I'm saying that it's happening here and it's a very real subsidy.
stickfigure 6 hours ago
They really aren't. Nothing in the development of these systems, absolutely nothing, is as important as money. Stop these silly false equivalences.
_aavaa_ 5 hours ago
All the major labs would be dead in the water if the government acted on the things I mentioned. If they treated the labs the way they would have treated us for all those hacks. Or if they enforced pollution measures (or god forbid ban on-prem turbines because of the climate damage). Or rule that training on data is not fair use, or that ingesting GPL3 code and then turning that into weights counts as a derivative work. Or. Or. Or.
Those are all just as important as cash.
stickfigure 4 hours ago
The hacks are pretty bland as long as there is no mens rea and the hacked victims do not want to press charges. This isn't a "subsidy", it's just our normal legal process at work.
The pollution question is winding its way through court systems and getting a lot of pushback. I wouldn't call this a subsidy, just a slow bureaucracy. And at any rate this isn't special to the US; certainly China's environmental policy is far more lenient.
With the training data, again there's nothing specific to US corporations here. Anyone can use that dataset. It's just how our legal system works. You might not like it, but it's the accepted law right now. This doesn't meet any reasonable definition of "subsidy".
On the other hand, shoveling billions of dollars of taxpayer funds into corporations is by definition a subsidy. Full stop.
dghlsakjg 7 hours ago
That's just the ones i know off the top of my head in the US. Those programs account for more than 60 billion in committed spend.
antonyt 8 hours ago
oblio 8 hours ago
[1] Where "Bad Things" would be the typical interest, taxes, depreciation, amortisation plus the Anthropic specific employee compensation, LLM training (you know, for the LLM lab), revenue sharing agreements (which is a form of paying for infrastructure), etc.
staticman2 8 hours ago
ai-x 7 hours ago
For LLMs, marginal cost is just electricity
staticman2 6 hours ago
Assuming you meant to reply to me and not someone else are you saying under GAAP accounting standards Anthropic is a profitable business because under GAAP accounting their only expense is electricity?
(I see what happened you skimmed the conversation and didn't follow what was being discussed.)
rhdunn 4 hours ago
ai-x 4 hours ago
oblio 6 minutes ago
2. The marginal cost is electricity until the hardware overflows. So it's continuous for electricity and a step function for hardware capacity.
Plus you don't address the core point. Frontier AI is capital intensive. A lot more capital intensive than regular software. The existing clouds supporting the entire internet (!!!!) were built for a fraction of the cost of AI infrastructure, and there isn't even an end in sight to this continuous hardware investment.
Let alone hardware refresh cycles.
AI is basically investment into roads.
Software used to be a magical place, closer to selling music albums but even better.
AI is a much worse business margin wise than regular software.
Bud 8 hours ago
Non-punishment is also not a subsidy; again, words have meanings. Let's use the correct words.
rfrey 6 hours ago
johnbellone 4 hours ago
koe123 8 hours ago
fauigerzigerk 8 hours ago
Not arbitrarily banning companies from buying a product from a supplier doesn't meet my definition of the term "subsidy".
ta20240528 6 hours ago
fauigerzigerk 6 hours ago
I don't see how. Anyone can buy electricity without any intervention from the government. An intervention is an action that changes what would happen by default.
gabriel666smith 7 hours ago
Microsoft did not pay with money - it paid (mostly) with Azure cloud computing credits. MSOFT is then able to write this off as a loss against tax.
It is generally far more tax-efficient in the US to write a loss in this way than it is to write a loss for a cash investment.
In this case, I believe the difference was ultimately highly significant. When including MSOFT eventually writing-off the deprecating Azure hardware it had used to buy the OpenAI equity, the result was MSOFT's tax reduction being either close-to or exceeding the actual cash value of MSOFT's investment in OpenAI - IIRC.
These examples represent taxes that the US chooses not to collect - the US could choose to make investments like these less tax-efficient. Instead, by making them extremely tax-efficient, the US subsidises the transaction hugely.
runarberg 7 hours ago
deaux 6 hours ago
lifeisloving 8 hours ago
Its very xenophobic of you to say China has zero intention of helping humanity, and just wants to "wage economic warfare".
Last time I checked, it was ourselves (USA) waging economic warfare on 2/3rds of the world.
I dont get this cope people have where people have this idea that its impossible for a Chinese company (that make billions of dollars) to have done something by their own merit, but instead its always some Chinese Communist Party conspiracy where the main goal is to destroy America.
Lay off twitter for a bit.
drstewart 8 hours ago
>Last time I checked, it was ourselves (USA) waging economic warefare on 2/3rds of the world.
Up until recently the USA Wasn't waging economic "warefare" on 2/3rds of the world
lifeisloving 8 hours ago
drstewart 8 hours ago
lifeisloving 8 hours ago
China is not a threat to you, or anyone in the West.
infecto 8 hours ago
tsunamifury 8 hours ago
infecto an hour ago
lifeisloving 8 hours ago
Not even our leaders say China is an enemy, ita mostly businessmen who are scared of competition and trying to regulate chinese out of their markets so they can make more money milking us.
infecto 7 hours ago
fearmerchant 7 hours ago
kaffekaka 8 hours ago
sixothree 8 hours ago
We hear this about literally every industry the Chinese excel in - that it's only because the government subsidizes them that they succeed. For chip manufacturing, for batteries, for EVs, for solar, for AI. I don't see how the chinese government can afford to subsidize all of these industries and still have them contribute to the GDP.
ygjb 8 hours ago
There is as much to criticize about American hyper scalers and AI labs and the lack of interest in helping humanity or contributing to open source, but that might not be as popular an opinions on this site.
gpugreg 8 hours ago
benterix 7 hours ago
bredren 7 hours ago
ygjb 6 hours ago
1234letshaveatw 7 hours ago
ygjb 6 hours ago
Do you think the world is closer to or more distant from large scale global conflict today than it was 10 years ago, or even 3 years ago?
1234letshaveatw 6 hours ago
overfeed 4 hours ago
Source? The last Supreme Leader of Iran issued a fatwa against nukes, way before he and his extended family where bombed in their home in a failed attempt at a decapitation strike.
locknitpicker 6 hours ago
Any open weights model that Chinese companies are releasing for free are allowing everyone in the world to have access to high quality realizations of these tools without the risk of the US regime arbitrarily cutting you off.
Also, it's funny how all the neoliberal mantras of how free market drives progress through competition stops when it's a US company that's being challenged.
throwa356262 8 hours ago
I think the Chinese government is backing their labs by less direct means. For example cheap electricity and investing in chip manufacturers such as Huawei and cxmt.
Deepseek specifically, is known to operate with minimal resources. The entire company has around 160 employees and every model they release must break even within ten months.
parineum 8 hours ago
It may be that the effect of the backing in both places is essentially equal but this statement is strange. The Chinese government invests directly in Deepseek[1]. Notice the article, in addition to saying the CCP is investing, says Tencent is also a major backer. CCP owns a golden share of Tencent.
[1]https://www.cnbc.com/2026/10/06/deepseek-funding-round.html
atwrk 7 hours ago
parineum 6 hours ago
Everything a Chinese company does has the explicit backing of the CCP, at least ideologically and usually financially in some fashion. While DeepSeek may not be the CCP, it couldn't exist if it was expressing any kind of ideology that wasn't inline with them and, in this case, is explicitly funded by them.
A Chinese company has no freedom to say Taiwan is a country the way someone in the US could suggest California succeed from the nation.
Any public message you hear coming out of China has the implicit approval of the Chinese government.
deaux 6 hours ago
If you're not arguing about comparison then don't state this as if it's any different from the US. Because you make it sound that way. Or do you think that if tomorrow OpenAI came out vocally supporting the DSA - that suddenly they wouldn't see lots of barriers rise up out of nowhere?
Come on now. That's fairy tale land.
Danox 5 hours ago
pyrale 5 hours ago
50or05 8 hours ago
Supporting deepseek is just like supporting the Chinese army, no need for that. Though it goes both ways, OpenAi subscription just lowers the cost of the US army as well.
metobehonest 8 hours ago
Good, where do I sign up? At least they aren't exploding little children and generating chaos in the oil market.
https://www.business-humanrights.org/en/latest-news/anthropi...
stickfigure 6 hours ago
Danox 5 hours ago
metobehonest 5 hours ago
https://www.ohchr.org/en/press-releases/2025/06/israeli-atta...
https://www.unesco.org/en/articles/unesco-presents-assessmen...
I want you to publicly say that this turtle biologist and this journalist are Hamas: https://www.theguardian.com/world/2026/jun/20/mona-khalil-tu... https://www.theguardian.com/world/2026/apr/23/lebanon-journa...
ndriscoll 7 hours ago
It's not a subsidy for the Chinese to say they're going to ignore our IP laws. It's a subsidy for us to say we're going to make them and try to push them on the world. It's literally granting a monopoly by legal force. It's very obviously not aligned with the interests of the American people, while China releasing things in the open is.
ismael_rr 7 hours ago
dataviz1000 7 hours ago
philipallstar 7 hours ago
dataviz1000 7 hours ago
Ironically your comment will be cloned and used for training.
throwa356262 7 hours ago
Same goes for US labs, not everyone are as innovative as google deepmind:
https://www.forbes.com/sites/antoniopequenoiv/2026/04/30/elo...
ta20240528 6 hours ago
deaux 6 hours ago
Yes, just that. Being a huge supporter of the regime - at a time running a department - surely has nothing to do with it.
Implicated 5 hours ago
... You're saying this about _deepseek_?
lnxg33k1 4 hours ago
novaRom 6 hours ago
reacharavindh 8 hours ago
As a neutral party, this characterization is crazy.
As if the AI companies - Claude and OpenAI are guardians of freedom and humanity and very charitable to the global society without any self interests… “Chinese models are subsidized by the Chinese Government, therefore they’re inherently bad for humanity” is a highly propagandist argument. The politics of US vs China may be whatever it is in reality.. You have one company releasing their models for cheap, actually open sourcing their trained weights, and publishing details of their optimizations and learnings for others to use. The other camp actively “aligning” their models, nerfing their capabilities, hyping their swarm activities from poor sandboxes, and trying their best to lock users into their harnesses and walled platforms. They are subsidized by the capitalist VCs who are essentially waiting for their payouts..
At some point, one has to see things for what they are and evaluate their own reasoning..
I’m happy to stay provider agnostic, try all models and cheer any useful progress as open as possible.
bparsons 8 hours ago
gabriel666smith 8 hours ago
There are many other reasons Chinese companies releasing models open-source or open-weight makes strategic sense.
A really easy-to-understand example is a company who has a near-monopoly on "serving video content" releasing a video model openly.
If you can be relatively certain that video content created by a model (which you have trained, using data from your own platform) will be ultimately served on your own platform, thus generating revenue from watch-hours, it makes sense to make those models as widely-available as possible.
It's also a net-positive if people use your public research to build better video models, because - again - you are reasonably certain that the even-better content those new models produce will be watched on your platform.
The alternative would making models harder to access and learn from (broadly, the current western model). Many would argue that Google, in choosing to not optimise its video generation models for "availability", is directly causing less content to be uploaded to YouTube. This is the trade-off.
I don't know much about DeepSeek's financing specifically, which obviously doesn't release video models - so I don't know how directly this analogy runs, or who directly benefits from the extremely evident rising tide that the public release of DeepSeek's research creates. However, this does not negate the broader rising-tide effect of the scientific method.
It's certainly also true that it's geopolitically beneficial to be able to undercut American labs' models. If I ran a global superpower, I would probably want my country to be technologically competitive too.
But Chinese companies are already serving a huge volume of customers in a complex, existing marketplace, before even thinking about the US market, and it's overly simplistic to assume that their entire strategy revolves around economic warfare directed specifically at the US. It's more nuanced than that.
This is, of course, without even getting into opening the can-of-worms around whether US economic policy also results in the US state functionally subsidising technological innovation, how comparable that is to China's model, etc.
_puk 5 hours ago
Causing less content to be uploaded, when you are clearly the gorilla in the room, looks to be a wise strategy not a trade off.
vrganj 8 hours ago
Why are we assuming a strong US is necessarily good? As a European, I have seen plenty of evidence against that stance lately.
I understand that Americans might prefer a strong US. But conflating them with humanity is a leap that I don't think one can make without any backing.
fearmerchant 7 hours ago
piva00 8 hours ago
The government that removed restrictions on how private companies can access capital after a certain scale (the JOBS Act), that removed the need for private companies to report as if they were a public company after a shareholder threshold was crossed, superpowering the access of wealthy private investors to get in earlier in a growing company while at the same blocking the public from participating in funding growing enterprises at an earlier stage (since it required companies to IPO much earlier to access capital) which allowed retail investors to also reap the rewards on funding them early when they grew to become behemoths (like Amazon, Meta/Facebook, Google, etc.).
It's not fair in either place, the USA has its own model of unfairness, China has a completely different one. The difference is that in the USA the government allows private investors to become more powerful than the State (outside of the monopoly of violence) while plunging the rest of society into increasingly more precarious lives while in China the State is the power and its legitimacy only exists while the population feel they have a better life.
atwrk 7 hours ago
shwaj 6 hours ago
As a differentiator from the Chinese system, not so much. For two otherwise equal companies, the one that says things against the party line will experience selective enforcement too.
gpt5 4 hours ago
The propaganda of trying to make US and China government appear the same is making people dumber. and is one of the most heavily used tools in China’s online propaganda arsenal.
p_j_w 4 hours ago
The government could do some very easy things to remove this tool from China's toolbelt. Like you know, prosecuting rich people when they've earned it.
gpt5 32 minutes ago
010ED67913 5 hours ago
juiceland 8 hours ago
The story is so much more complicated than that, to the point that this economic warfare theory is basically a meme.
Chinese models are open because they don’t have a choice. “When you trail the frontier, openness maximizes reputation per unit of capability. The moment you lead, you close.” [0]
[0] https://earnedintuition.substack.com/p/involution-without-ex...
yorwba 5 hours ago
juiceland 4 hours ago
yorwba 3 hours ago
sicktriple 8 hours ago
waterheater 8 hours ago
yomismoaqui 7 hours ago
NicoJuicy 7 hours ago
lbreakjai 7 hours ago
1234letshaveatw 7 hours ago
SecretDreams 7 hours ago
fearmerchant 7 hours ago
broabprobe 7 hours ago
shmel 7 hours ago
sourdecor 7 hours ago
Does anyone think that likely? I have no clue or bias.
pianopatrick 7 hours ago
vagrantJin 7 hours ago
Yes, because everything China does is against the US. That's all they think about day and night. God forbid they want to corner the global market or have a genuine business case. How dare they provide options for those who can't afford a measly $200 a month? How can we let Chinese labs publish research for free for the whole world so that they can benefit? The nerve! To think they can use soft power instead of military might! I mean, Anthropic and OpenAI are the last bastions of human kindness and charity. Right?
Right?
joquarky 3 hours ago
fluidcruft 7 hours ago
Notwithstanding that Chinese publishing methods actually does help both humanity and open-source.
mcv 7 hours ago
1234letshaveatw 7 hours ago
brabel 6 hours ago
benterix 7 hours ago
Buttons840 7 hours ago
spyckie2 7 hours ago
subarctic 7 hours ago
Shitty-kitty 7 hours ago
1234letshaveatw 7 hours ago
brabel 6 hours ago
matsemann 5 hours ago
IOT_Apprentice 7 hours ago
ozgung 7 hours ago
You think extremely US-centric. China has a different economic model than US and your rules for a specific kind of Capitalism may not apply to them. You assume a country of 1.4 Billion people is obsessed with a couple of foreign AI companies. What if they don't care.
MomsAVoxell 7 hours ago
sajithdilshan 7 hours ago
jetdeng 6 hours ago
scotty79 6 hours ago
Charity and kindness is not a motivation, it's an outcome of what you do.
horacemorace 6 hours ago
oefrha 6 hours ago
LPisGood 6 hours ago
locknitpicker 6 hours ago
US companies are burning colossal piles of cash in ways that makes it unclear if it qualifies as dumping, not to mention their deep ties with the country's regime.
Claiming that companies from a country have ties to the regime and burn through cash is a very miopic accusation.
GuB-42 6 hours ago
It is great for everyone except for a few people who want power over everyone else, and the fact it is not charity of kindness makes it more sustainable, because charity and kindness is quick to go when big money and politics is involved.
I want more warfare like this. Building stuff instead of destroying stuff.
anon-3988 5 hours ago
deaux 5 hours ago
So were Amazon and Uber by the US, which have now established monopolies across the globe. To the countries suffering from those, there's zero difference with China doing it to solar. Actually there is, at least solar got them cheap renewable energy in return. This would never have happened in the US because big oil interests would make it take decades. That's the reality.
You need to spend 10 years outside the US, deprogram, and then go back.
nutjob2 5 hours ago
Ok, but so what? That's what happening so far. Even a repressive totalitarian government I wouldn't wish on my worst enemy does some good sometimes.
If the Chinese cheap/open models rise up and destroy us, thats on us for giving them access to the tools to do so.
vintermann 5 hours ago
To say that US model providers have a close relationship with their government would be putting it mildly.
shinyshadowdonu 3 hours ago
You put it like US has a goal of help humanity or open-source
swalsh 8 hours ago
stanac 7 hours ago
toenail 7 hours ago
wg0 7 hours ago
Its them who chose to outsource heavily, not US consumers.
There are no victims here however, both benefited from the arrangement. The consumers and capitalists.
deaux 6 hours ago
No, it's the result of US leadership letting this happen. This is clear since China themselves would not let this happen. US leadership did nothing because they were best friends with the people who did it, and did not care one bit about their population.
tokai 8 hours ago
So the most common way to publish manuals?
jorl17 8 hours ago
ZiiS 8 hours ago
velcrovan 8 hours ago
xnx 8 hours ago
apexalpha 8 hours ago
dvduval 8 hours ago
swingandamiss 7 hours ago
kirill5pol 7 hours ago
swingandamiss 7 hours ago
typ 8 hours ago
porridgeraisin 7 hours ago
fittom 7 hours ago
baxtr 7 hours ago
scotty79 7 hours ago
This will cause the picked winner to get massively ahead with sheer compute alone used both for training and inference dedicated to recursive self improvement.
pennomi 6 hours ago
scotty79 6 hours ago
There are going to still be worthwhile improvements but they are going to be more like not how to make transformers 10x cheaper but how to make next training run cost 9 trillions instead of 10 with a very particular optimization designed at the cost of hundreds of millions for this one specific run.
LPisGood 6 hours ago
thfuran 5 hours ago
bee_rider 4 hours ago
Zigurd 5 hours ago
Are there any use cases that have enabled one customer of a frontier LLM to outperform a competitor using a different frontier LLM? Or is this why we are seeing confected points of comparison like solving challenge problems in mathematics?
bushbaba 6 hours ago
jsw97 6 hours ago
oakesm9 6 hours ago
Danox 6 hours ago
pyrale 6 hours ago
randomNumber7 5 hours ago
flir 4 hours ago
I could imagine a belt of data centres around the equator, that hand off their computational loads as the sun sets. Good scifi-esque premise.
kergonath 4 hours ago
jayd16 6 hours ago
ambicapter 6 hours ago
flir 6 hours ago
There's a law of diminishing returns at play here, and doubling the energy cost of training to wring 2% more performance out of the technology isn't going to be very useful, because most of the problems it is capable of solving will be solvable with the previous-gen 98%-as-good model.
("there's a law of diminishing returns at play here" is an article of faith. But then, so is the belief that these models will keep getting better).
londons_explore 5 hours ago
ejeje12a 5 hours ago
lol You can always tell who has never ran a business before with comments like this
satvikpendem 5 hours ago
asplake 3 hours ago
flir 5 hours ago
(I don't think it will work).
allendoerfer 5 hours ago
flir 4 hours ago
cyanydeez 6 hours ago
spwa4 6 hours ago
Btw: with Qwen4 I mean the next large Qwen model that is based on the Qwen4 architecture (Qwen 3.8 flash next was "almost" based on the new arch but obviously was a small model)
ejeje12a 6 hours ago
What matters more is if firm’s start using a bundle of American and Chinese models and when they find their feet - how large is the market for frontier?
Frontier has to displace labour one for one at some point or it’s over.
christkv 6 hours ago
hypercube33 5 hours ago
throwa356262 4 hours ago
In practice, you will be able to run models a bit bigger than 35B.
cyanydeez 3 hours ago
38GB of vram resident. more tk/s, more prefill.
verdverm 6 hours ago
satvikpendem 5 hours ago
verdverm 4 hours ago
I wonder if the next DS models will also graduate to 5.x, I think I saw they are training up a 10T model, and just raised $12B too
cyanydeez 3 hours ago
The fact that it basically broke open the local model supremacy was just a nice side effect.
I'm running: https://github.com/peonist-ai/halogen-server with a quant4, PLE offloaded, and it's resident VRAM is 36GB at 265k context.
Shave 10 more GB off and the TAM openai and anthropic are targeting is a lost cause. Local models are what 90% of people will need.
If the world governments can get a handle on the memory cartel, then there's no more moat for most normal humans.
ben_w 2 hours ago
rpozarickij 6 hours ago
joe_the_user 5 hours ago
jamienk 7 hours ago
bushbaba 6 hours ago
AIblemblio 8 hours ago
But Opus 5.5/GPT is such a game changer in comparison to sooo many others, its still a moat for now.
eigenspace 8 hours ago
The real question is if this model is good enough that it can still accelerate work, and not be a hindrance to real work like older Mistral models often were.
If they can do that, they'll have customers.
hajile 7 hours ago
If these companies will steal from deep pockets like Disney or Sony (some of the most infamously litigious copyright trolls to ever exist), they won't think twice of stealing every bit of code you upload to them.
If your code passes through an AI company's servers, you can assume you just gave it to them. In turn, when your competitor tries to copy that new feature you just added, the AI is now trained in exactly how to copy you and eliminate your competitive edge. Unlike your employees, the AI isn't bound by the same rules and even if it were and violated them, your company probably doesn't have enough money to prove it in court (and that's if we somehow reverse some of the stupid "AI is the most transformative use of copyright I've ever seen" judges who have drunk the coolaid).
Most companies could build the compute to run GLM or Kimi models for way less than the potential loss due to IP theft from using third-party systems.
FabHK 7 hours ago
eigenspace 7 hours ago
chihuahua 6 hours ago
airstrike 5 hours ago
user43928 an hour ago
> Following an investigation, we have confirmed that Buckmaster’s Codex prompts over the two months preceding this announcement and paper on September 8, 2026, could not have influenced the system in any way, including through training. The OpenAI internal model used for this result was developed through large-scale reinforcement learning on top of a previously pretrained model. Our proofs also differ significantly. In the Euler case, Alpöge and Buckmaster proved a result with external forcing, while OpenAI’s system proved a result without external forcing.
sashank_1509 7 hours ago
A simple loophole, use the code to create an RLVR environment where the resultant code is the end goal / max reward. Technically the customer data is never trained upon, but effectively you’re using it. Even better, use the code as a seed to generate synthetic data similar to it and use that synthetic data as rewards in an RLVR model.
Unless you can host the ChatGPT model on your own servers, which I know some enterprises are doing, I don’t think there’s any hope of protecting your data / competitive advantage from these frontier companies. Better to be paranoid, than be commodified by these companies.
phillmv 6 hours ago
spiderfarmer 8 hours ago
senordevnyc 8 hours ago
There is no ceiling on what you can accomplish with more intelligence, so there will always be a market for the best models, and that market is likely to just keep growing. If Opus 13.5 can one-shot a profitable company or discover a new disease treatment or whatever you can think of that a swarm of relentless super-geniuses could accomplish, companies (and governments) will throw money at it.
I also think there will always be a market for many sub-frontier models that will continue to grow rapidly as well, because "good enough" is definitely a thing for a given task.
suddenlybananas 7 hours ago
No company would ever release such a thing
sajithdilshan 7 hours ago
calgoo 8 hours ago
senordevnyc 8 hours ago
ezomode 8 hours ago
And yes, open weights are still behind, but are catching up.
n6242 8 hours ago
"Now, here, you see, it takes all the running you can do, to keep in the same place. If you want to get somewhere else, you must run at least twice as fast as that!"
totallykvothe 8 hours ago
gjm11 2 hours ago
senordevnyc 4 hours ago
_aavaa_ 8 hours ago
senordevnyc 4 hours ago
bigbadfeline an hour ago
Yeah, I could buy a bigger and bigger vehicle when a new one goes to market, but using a semi-truck to drive myself around would not only be significantly more wasteful but also more inconvenient and limiting compared to a sedan. Very much like advanced closed models vs performant open ones.
u8080 8 hours ago
wg0 8 hours ago
People who aren't afraid of rolling their sleeves into any code base? The difference is practically zero.
CharlieDigital 8 hours ago
Maybe it's because people stopped watching what their agents are doing and stopped looking at the quality of the output. But I still see agents being absolutely mindless like a junior dev.
Recent example: it updated an an API to add newly released models to the backend. There's a list of models that require specific configuration for the reasoning effort and temperature or the API call fails. GPT 6.1 Sol misses this and code fails at runtime because the newer models need to be added to the list for special handling of temp and reasoning. Fixes it for one model and tests it for that model using an E2E test. But doesn't test the other models that were added for the same error condition...I had to explicitly ask it to do so and it finds them and adds them to the list and says "that's on me."
Yeah, not that smart.
marmarama 5 hours ago
I keep seeing this kind of thing over and over, and honestly it's not got _that_ much better since the big breakthroughs about a year ago.
For sure I happily vibecode stuff without worrying about it when it's a greenfield project, and if the LLM has written it entirely from scratch then usually it's well structured and sane. But making changes in messy, mostly human-written mature codebases is still a minefield.
AIblemblio 7 hours ago
Slartie 7 hours ago
This works fine with Opus 5.5. But it also works fine with GPT 6.1 Sol, Kimi K3 and MiMo 2.6 Pro.
It doesn't work equally well with Sonnet 5.5, interestingly.
balder1991 6 hours ago
For people who have some expertise, the models accelerate the grunt work, but you’re the one validating it.
Aldipower 8 hours ago
r2_pilot 8 hours ago
Aldipower 8 hours ago
wavemode 8 hours ago
SyneRyder 8 hours ago
As for Mistral - I got really excited when they said Large 4 was focusing on being #1 in cybersecurity, because that's somewhere that they genuinely could edge out Anthropic & OpenAI. Have it actually solve problems, instead of Anthropic flagging "you tried to find a null pointer exception bug in your own code, we're now reporting you to the US government". But on the Mistral benchmarks I'm seeing, this looks very disappointing, but at least they haven't entirely given up. I genuinely thought Mistral had given up on new general models. They need to learn the bitter lesson all over again.
Systemerror7A69 8 hours ago
The capabilities of all models increasing so much all the time means there are simply less and less tasks you need a frontier model for.
Even if Opus 5.5 is 500x better than Deepseek, if deepseek can solve all my problems, why do I need to pay for more?
user43928 8 hours ago
Obviously any model will do if you use it as a better autocomplete.
I believe that there is a large gap in expectations between different workflows.
Until the AI like reads my mind and produces perfectly production ready apps with minimal intervention from my side, there is still going to be room for improvement.
RealityVoid an hour ago
Uhm, yes?
kragen 7 hours ago
This morning I elicited a microkernel operating system from Opus 5.5. Well, mostly. It doesn't implement task switching yet; we'll see if it runs into a wall at some point. But it boots in QEMU, and it's running a user process in ring 3 and serving web pages.
sashank_1509 7 hours ago
But if on the other hand, I mostly use my human intelligence and just need a dumb model to complement my human intelligence at low cost and high speed (say review every commit to catch obvious bugs), I have a much better chance of building an actual moat than you do.
But outside of coding, it’s even more clear that you don’t need frontier intelligence. My customer service agent is very happy with a 100B param Deepseek flash model, thank you!
TuxSH 5 hours ago
I’m using subscription models for exactly that, better models catch more subtle bugs, and they catch them faster. It works out far better in terms of work-hours saved.
Also A/ then OAI slashed token pricing by 2x~5x on their latest models
fultonn 5 hours ago
I have enough real problems in life. I don't need to invent new ones just because a new technology is available.
Many of my problems in life are fully solved far past my satiation point by a 3b model that costs me nothing to run.
Many others are not.
But in either case, when I am acting and living wisely, almost all of my problems exist prior to the existence of technological solutions to those problems.
This is also true for the customers and employers that I care to work with. This has changed in me over time, but I now try my best to avoid inventing new problems. The world has enough big, important problems already.
> This morning I elicited a microkernel operating system from Opus 5.5.
This is cool but also a good example. I don't need a personalized microkernel just because it's possible to have one.
Maybe I need one and I don't know it, but the problem statement definitely isn't "I have inherent desire for a personalized microkernel".
breuleux 4 hours ago
skerit 8 hours ago
I had the exact same experience. And unlike Fable, it doesn't gobble up your entire usage limit in a few hours.
I always wonder what the "good enough" people are actually using it for.
sashank_1509 7 hours ago
If Opus can’t 1-shot it, then it must rely on our human intelligence which can be complemented well enough with a dumb model as a frontier model.
nonadhocproblem 5 hours ago
blahblaher 4 hours ago
user43928 2 hours ago
Prices for the same task (ARC-AGI-2) have since dropped by factor 10 if you compare to Opus 5.5 Low.
eigenspace 8 hours ago
Look at the context in which I used that term 'good enough'.
What i was saying is that there are tasks for which a dumber model can be good enough, and for organizations with sovereignty/ privacy concerns, those concerns can be strong enough to incentivize the use of a dumber model.
bakugo 8 hours ago
I hear this literally every other week about whatever the newest FoTM model is.
Unless you can provide concrete examples of things you can do with them that you simply couldn't do with last week's model, it's absolutely meaningless.
spaceman_2020 8 hours ago
ygjb 8 hours ago
It's especially the case as more non-Americans look to self hosted models and domestic cloud inference providers using open models that the US providers who are still leading the charge need to drastically drop their prices and find a path to profitability in order to maintain their lead and retain the advantage they had as AI turns into a commodity (which is happening faster than I think even the frontier labs initially predicted).
segmondy 7 hours ago
I use MiMov2.6Pro, DeepSeekv4.1Flash, GLM5.3, Hy4, Qwen3.8 and KimiK3 at home. Opus5.5 is not a game changer.
AIblemblio 7 hours ago
I have to admit, Sonnet got really good too.
But Opus just uses tools, a broad spectrum of it, etc. it feels like sure if you add some router behind it you could split it up if you need to but if you give me the choice, its opus allll day long.
xiconfjs 4 hours ago
jayd16 6 hours ago
senordevnyc 8 hours ago
Obviously, this is only a valid question if you don't believe that open weights are about to eat their lunch and their revenue is about to collapse, or they're running a super unprofitable ponzi scheme propped up by investor money that's about to collapse like a house of cards. I don't find those positions credible at all though.
If you do, then this question isn't really for you, as I'm more interested in thoughts from those who think that OpenAI and Anthropic in particular are about to be the largest companies on earth in a couple years. Could anyone catch them at that point?
swiftcoder 8 hours ago
I don't know about revenue, but I suspect multiple other labs are already beating OpenAI/Anthropic on profitability. Staying on the frontier is expensive, and it's hard to recoup those R&D costs when you have a bunch of other labs nipping at your heels.
If you concede the previous point, then the only way for OpenAI/Anthropic to keep growing long term is to swallow the whole economy (i.e. mass job replacemnt), and that's a bet I wouldn't take.
senordevnyc 8 hours ago
richardw 8 hours ago
WarmWash 8 hours ago
The second moat is convenience, which all the big labs make it (comparatively) easy to glide into their models.
notfromhere 8 hours ago
eigenspace 8 hours ago
If they can carve out a niche of industrial and governmental partners who rely on them for sovereignty reasons, it may be enough.
senordevnyc 8 hours ago
intrasight 6 hours ago
notfromhere 8 hours ago
qoez 8 hours ago
divbzero 8 hours ago
doctorpangloss 8 hours ago
eigenspace 8 hours ago
doctorpangloss 6 hours ago
ragall 6 hours ago
saberience 8 hours ago
kaffekaka 7 hours ago
amelius 8 hours ago
From the Mistral site:
> ML4 was trained from scratch on 3,800 NVIDIA Grace Blackwell GPUs in Mistral’s own datacenters in Europe.
It is pretty capital intensive!
everfrustrated 7 hours ago
To put that into context, the last wave of capacity SpaceXAI added 400-450 MW.
amelius 7 hours ago
bjenkins358 7 hours ago
locknitpicker 5 hours ago
Chinese companies also managed to put together their models with relatively small clusters.
Perhaps US companies are desperately trying to brute force their way into workable models?
eigenspace 7 hours ago
amelius 5 hours ago
ricardobeat 5 hours ago
anvuong 4 hours ago
3,800 GPUs is nothing in the frontier side.
jayd16 6 hours ago
amelius 6 hours ago
jayd16 6 hours ago
dannyw 6 hours ago
cyanydeez 7 hours ago
So it makes sense, since all you need is compute, that there's a ceiling and specialization is going to be more valuable then some super AGI.
Especially since the worst people seem to be the ones who think they'll all run away with the bag.
samplifier 7 hours ago
Disclaimer: I'm not sure how much of an IYKYK factor applies to this joke.
StrauXX 6 hours ago
MiloLeo 6 hours ago
yeahforsureman 4 hours ago
ikoorng 6 hours ago
AI is an idea 60 years old. We are on the 3rd or 4th generation of AI development. Three years into the current iteration of products.
This is not early days by any measure. LLMs are a result of a very, very mature research field.
nickpinkston 6 hours ago
I trust them and their populations to provide a more societal-friendly version of AI, putting pressure on the US tech oligarchy, while also providing democracy-friendly open models that I don't trust to happen with the Chinese labs.
blueaquilae 6 hours ago
onlyrealcuzzo 6 hours ago
This appears to be roughly as good as Sol 6.1 (which is quite good), considerably faster in terms of wall clock for complete tasks, and considerably cheaper (where Sol 6.1 is already good value - just really slow).
That seems too good to be true...
But I really hope it is true...
nonadhocproblem 5 hours ago
onlyrealcuzzo an hour ago
That seems much more realistic.
locknitpicker 5 hours ago
Mistral is also an European company. As we live in a time where the US regime is engaged in pyrrhic geopolitical tactics, it's good to know that it can't threaten to cut access to models during s period where everyone is rushing to incorporate them more and more in our life.
bryanlarsen 5 hours ago
If the singularity is defined as an AI sufficiently intelligent to improve itself independently, that AI is still limited by the resources required to do this improvement.
livvy 5 hours ago
api 5 hours ago
... unless they can legislate it, which is why they are flattering heads of state and scare mongering about dangerous AI.
londons_explore 4 hours ago
Right now they don't even get good feedback from local sessions - I can see it make the same mistake two days running, and then months later when a new model comes out, presumably trained on my data, it still makes the same mistake.
asah 4 hours ago
londons_explore 4 hours ago
When a user says "I'm struggling to undo a bolt on my 1952 Mustang" and the AI responds "try hitting it with a hammer" and the user replies with "that worked thanks" - that is a tiny piece of knowledge which exists nowhere else.
Future AI's can say with more confidence that hitting it with a hammer will probably work.
Across billions of conversations, that can add up to more knowledge than all books.
tootie 3 hours ago
asadotzler an hour ago
manlymuppet 7 hours ago
Surely we want competition and Europe involved in that, but at this point I have grown used to either American labs smashing the frontier remarkably fast, or Chinese labs getting way, way closer than you would expect them to.
Mistral’s progress, regrettably, feels much slower. This model doesn’t knock anybody’s socks off. The model is (and I hate to be this harsh) mediocre, and this mediocrity has also arrived months late.
This is a pretty grim prognosis for European AI.
simjnd 7 hours ago
manlymuppet 6 hours ago
But Chinese models have very much earned their place. The same cannot be said of Europe, so far.
thoughtpeddler 4 hours ago
Tenoke 3 hours ago
danny_codes 7 hours ago
LLM development is jumpy. It’s hard to extrapolate very far ahead.
manlymuppet 7 hours ago
When Europe does surprise us, I will be the first to commend their progress. But until then, this is where we're at.
porridgeraisin 7 hours ago
This is mistrals first 1T-scale model and I expect the 4th or 5th generation to be close to the best for many purposes.
[1] These evals differ from the public ones like terminal-bench, are sometimes model-specific, need real, diverse usage to actually create, and are held secretly since quality of eval is the first driver behind the next step improvement of a model.
[2] It is not close. This model was trained on less than 4k GPUs, whereas astra used north of 100k GPUs.
manlymuppet 6 hours ago
And to the point of scale and training cluster, so what? Not only do Chinese labs have smaller clusters with less empowered GPUs, compute is Mistral's responsibility. You can't take away from other labs just because they fulfill that responsibility better.
porridgeraisin 5 hours ago
The lack of compute is not really attributable in that sense to mistral. First of all it needs general investor and government willingness, which is easier in a larger economy like the US or China.
Second, you need widespread usage of your paid inference service for two reasons: one it pays off your compute cost, and two it speeds up the improvement process.
The vast majority of deepseeks paid customers are within china itself (since openai and anthropic services are not reachable from china) which gives it a market. But for someone in france, there is no reason to use a structurally slower developing model from mistral compared to using one from openai...except when data guarantees are needed, hence the landing page focus on sovereignty. As far as the dual use aspect goes, a model like this is more than enough, so the government will be happy.
ismailmaj 6 hours ago
I really want them to win as that's our last horse in the AI race, but ~200 research-oriented devs out of 1800 employees? I believe they agree it's pretty doomed and have pivoted.
CryptoBanker 3 hours ago
arrowleaf 5 hours ago
adev_ an hour ago
It's nowhere mediocre.
It's toes-to-toes with GLM-5.3 which is one of the best Open Weight model available (With Kimi K3) for general reasoning.
I just runned it on code reviews right now and it was able to catch some thread safety issue than DeepSeek-4.1 didn't. And DeepSeek-4.1 is by no means a bad model.
verdverm 5 hours ago
layer8 5 hours ago
troyvit 5 hours ago
Sometimes it's ok to cheer for the last kid crossing the finish line because they're actually running a totally different race, and winning might look completely different.
When I look at what Mistral does vs other organizations I'm impressed:
They aren't profitable yet, but they're a lot closer than most and they're doing a hell of a lot with very little.
Pointless racing story:
I was in high school track with a really tough guy who was just not a runner. We went to a pretty messed up high school and if you screwed around in track practice sometimes the coach would make you run a crap race at the next meet, like steeplechase or hurdles. Well this guy and a few others screwed up and coach made them all run hurdles at a meet.
He hooked every single one and fell on his face. Every time he got back up and kept on running. By the time he hit the finish line his knees were bleeding halfway down to his ankles. We cheered like hell and he was smiling ear to ear.
Coach quit punishing us with races after that.
seizethecheese 4 hours ago
It looks like Mistral is middle of the pack, behind Anthropic and ahead of OpenAI on that front. All of those labs are way "ahead" of the cloud providers, but those providers are building infrastructure, not just training models, so it's not apples to apples.
troyvit 4 hours ago
Speaking from an absolute perspective I do think they are doing more with their money than either Anthropic or OpenAI.
badatnames 5 hours ago
I think it's an incomplete read. What's the point in competing for a sizeable percentage of your funding when the finish line is incrementally being moved each month? Better spend it on leapfrogs which they seem to have done.
Meanwhile Mistral have a natural ace in their pocket with respect to regulation in the form of CADA and the Cloud Sovereignty Framework. I can't think of another company that would qualify as SOV-3 under that regime
pembrook an hour ago
Political polarization is turning the world insane.
Marciplan 4 hours ago
epolanski 2 hours ago
Nobody in the real world cares about minor benchmark differences in money losing coding agents.
And nobody in the real world is giving Altman or Musk their data.
geniium 44 minutes ago
I think we will see some horror stories come out with data leak in the next years.
jstummbillig an hour ago
This barrier is not going to start moving dramatically. It will simply be mostly satisfied for most work we do. Mistral is going to get there, soonish, long before the economy takes an entirely different shape (in so far that even happens).
There will be super human intelligence tasks, tasks truly constrained by intelligence for quite a while. Those will be few and far between, relatively speaking. Mistral will have plenty of opportunity to capture the other stuff, with a fraction of the resources required that it took the frontier labs to get there first.
simjnd 8 hours ago
imjonse 9 hours ago
vessenes 8 hours ago
danslo 8 hours ago
zkmon 8 hours ago
spiderfarmer 8 hours ago
Europe needs profitable AI companies, not money pits.
WinstonSmith84 8 hours ago
mopsi 7 hours ago
> Europe will never have a Tesla, a Google or an Amazon with that mindset.
Great companies, if you have a fetish for peeing into a bottle.fancyfredbot 7 hours ago
I'm not actually sure Tesla is a great company from any perspective so should be easy to find another sick burn for them.
Google is going to be a little more difficult. Maybe say something nasty about advertising? Or go for the monopoly angle.
mopsi 7 hours ago
zkmon 7 hours ago
dmje 8 hours ago
…oh…wait…
jonkoops 8 hours ago
koe123 8 hours ago
jonkoops 7 hours ago
spiderfarmer an hour ago
TomJansen 8 hours ago
fancyfredbot 7 hours ago
A lot of European companies now want a model the US can't cut off, but also lack trust in Chinese models.
Some of these will self host Mistral but most will pay them by the token. It's not going to be a huge market or a huge margin within that market but probably it'll be enough.
Marha01 7 hours ago
AI is a strategic technology with obvious national security implications. EU should invest in its development whether it's currently profitable or not.
spiderfarmer an hour ago
How did Claude and OpenAI make the USA safer?
rahen 8 hours ago
AI needs big bucks, and Mistral is Europe’s Anthropic.
idbnstra 7 hours ago
rahen 7 hours ago
Iolaum 7 hours ago
ismailmaj 5 hours ago
karp773 2 hours ago
alpineman 8 hours ago
rvz 7 hours ago
Mistral would have gotten a tiny and measly "EU grant" and ASML would never have invested later had it not been for the US VCs.
60pfennig 7 hours ago
danny_codes 7 hours ago
layer8 5 hours ago
gregorygoc 4 hours ago
chevman 8 hours ago
https://bulbapedia.bulbagarden.net/wiki/Lechonk_(Pok%C3%A9mo...
maelito 7 hours ago
sofixa 6 hours ago
NekkoDroid 5 hours ago
I've read it a few times and I always processed it as "Le Chateau Fat" ("The fat castle"), which I guess also kinda works. I guess it is mostly because my french is so rusty I forgot (or maybe never learned?) what chaton meant.
moffkalast 3 hours ago
Topfi 2 hours ago
duiker101 6 hours ago
Aperocky 5 hours ago
Topfi 2 hours ago
Will say their research is some of the best reads in the industry and I could not care less about their model release cadence as long as papers keep coming.
jascha_eng 8 hours ago
That's not particularly great.
That said I love that they don't seem to restrict cyber capabilities to any degree and even lean into it.
If the best model for cyber attacks is open for everyone to use it just makes us all safer I think. Of course you then also HAVE to use it or otherwise you're vulnerable, which is a great distribution play.
smartbit 7 hours ago
Inte Cost
llig per
Open Weight model ence Speed Task
Mimo-V.26-Pro 46 47 $0.13
GLM-5.3 (max) 45 73 $2.01
DeepSeek 4.1 Flash Max 39 227 $0.27
Mistral Large 4 Preview 38 116 $1.13jascha_eng 6 hours ago
smartbit 4 hours ago
Inte Cost AA- Omni
llig per Omni Softw
Open Weight model ence Speed Task score Eng
Mimo-V.26-Pro 46 47 $0.13 8 33
GLM-5.3 (max) 45 73 $2.01 14 37
DeepSeek 4.1 Flash Max 39 227 $0.27 -5 34
Mistral Large 4 Preview 38 116 $1.13 -5 5
Closed/proprietary ~50 110- $1.50- ~43 ~85
median top 10 242 $7.50
[0] https://artificialanalysis.ai/evaluations/omniscience
[1] https://artificialanalysis.ai/evaluations/omniscience?detail...tiahura 5 hours ago
The NRA approach to AI safety.
bilbo0s 4 hours ago
Issue is..
I don't believe for an instant that any of us, including US citizens, get access to the best models for cyber that the US has. I think any adversary would have to assume the models in use by the US side are unreleased.
US is not the only one dealing under the table by the way, I also think everyone should take China having unreleased models as an operating assumption at this point.
So Mistral is the best that the public gets access to. And that's if it's even the best? Benchmarks and pragmatic work have often been shown to be two radically different things in this industry.
lifeisloving 8 hours ago
People usually buy the cheapest, like Deepseek or GLM or they spend on Anthropic/OpenAI subs.. Are these models in the middle getting any users?
On a side note, I wonder if this was the popular free Space Bunny model that left openrouter yesterday.
PunchTornado 7 hours ago
c0rruptbytes 7 hours ago
einarfd 6 hours ago
myaccountonhn 5 hours ago
TuxSH 5 hours ago
Guess what Mistral themselves do?
randomNumber7 4 hours ago
phillc73 4 hours ago
riknos314 an hour ago
apexalpha 8 hours ago
I barely use anything outside of cheap Chinese models on OpenRouter anymore. They are simply (more than) good enough for most of the things I do.
This model looks reasonably cheap. Though not deepseek levels.
Going to test it with Hermes, wondering where it will land in term of capability.
Bon chance, Mistral!
phillc73 5 hours ago
Recently I switched to the Mistral hosted GLM-5.3, this worked very well and powered through a tonne of work. Unfortunately, I also completely maxed out two subscriptions within the space of six days this month. One can't stack subscriptions with Mistral, so I'd have to register a third account for another subscription, which will be annoying with changing API keys all the time. Sure I can switch to pay-as-you-go API, but that adds up really fast. The Mistral dashboard shows that a Vibe CLI monthly subscription for €18.44 actually provides €255 worth of API use (apparently, and I tried to check this with Support but it seems like they were intentionally vague).
After maxing out my Mistral subs this morning, I dropped $10 on Xiaomi to try MiMo-2.6. So far so good, seem to have done a lot of work for the $2.85 I've spent, and Xiaomi prices are still much better than the Mistral introductory offer for Le Chonk.
Not sure where to jump.
Edit: Not being able to stack subs is my biggest gripe with Mistral. I'd probably pay them $100 per month (5 subs worth), but I'm not going to switch to the pay-as-you-go API and burn much more money for the same amount of tokens. Instead, I've taken that extra money elsewhere. If they just allowed one to keep topping up subscriptions on the same account it'd be grand. Or even a bigger single subscription. Make a $100 tier with five times the capacity.
CryptoBanker 3 hours ago
phillc73 3 hours ago
I appreciate I can set a monthly spending limit for the pay-as-you-go API, but I'm just not willing to find out how far €100 will go, when I know it will go further elsewhere.
I just wish they had that €100 subscription tier, for the equivalent of €1k pay-as-you-go use.
AntonJidkov 4 hours ago
redanddead 31 minutes ago
DevKoala 7 hours ago
PoignardAzur 6 hours ago
Soon we'll have Mistral 6 Chaton, Mistral 6 Guépard, Mistral 6 Tigre, Mistral 6 Dents-de-sabre, Mistral 6 Beast King, etc.
scrollaway 4 hours ago
cedws 5 hours ago
pizlonator 7 hours ago
The pricing ($.68 in/$.07 cached/$2.09 out) makes it much cheaper than Kimi K3, GLM 5.3, and Meta Muse Spark 1.3. That's great!
But also much more expensive than GLM 5.3-flash and Spark 1.3 Contributor (the Meta-takes-your-data pricing of Spark 1.3).
So, I think it would have to be significantly better than GLM 5.3-flash to be worth it. GLM 5.3-flash is already very good.
drbscl 7 hours ago
Source https://artificialanalysis.ai/models/mistral-large-4?total-c...
james2doyle 7 hours ago
After that, it will be much closer to GLM 5.3, but you can also get 5.3 in their API! I dont see people really talking about that.
pizlonator 7 hours ago
barrell 7 hours ago
pacha3000 5 hours ago
GLM 5.3 is incredible because for the first time with an open-source model, it feels.. enough. I don't need much anymore, this model is great in everything. Except a thing : speaking french.
If the benchmarks are true, I'd be glad to switch entirely to Mistral.
rglover 8 hours ago
This was the era of the AI race I was waiting for.
netvarun 7 hours ago
rglover 6 hours ago
bartstp 7 hours ago
laserbeam 7 hours ago
volkk 7 hours ago