A misalignment of AI in mathematics (mathandai.org)

696 pointsby meredydd9 hours ago727 comments
https://terrytao.wordpress.com/2026/09/11/a-severe-misalignm...

https://www.economist.com/science-and-technology/2026/09/11/..., https://unwall.app/www.economist.com/science-and-technology/...

tmhn2 8 hours ago

As a mathematician maybe I am a little more optimistic than this declaration.

I am thinking of Mochizuki's abc conjecture: He worked in relative isolation, and dumped a huge incomprehensible proof on the community (to oversimplify a bit). That's not totally unlike what might happen if AI generates a huge, incomprehensible proof of let's say RH.

Well, what is the result? In the Mochizuki case, it was a lot of skepticism, but it also generated conferences, papers, talks in the hallway, discussions with students, and so on--a flurry of exactly that kind of community process that the declaration says is the main driver of mathematics.

Ultimately we think a fatal flaw was found in Mochizuki's proof, so it didn't lead anywhere in particular. But in our hypothetical "AI lean-verified proof of RH" situation, it would presumably generate substantially more of that community activity we saw in the Mochizuki situation. And if it's correct, that community activity would be productive (expository talks, students given problems to flesh out or generalize, etc).

Maybe mathematics just becomes a little more like other fields--relying on labs with lots of money for compute, digging through a corpus of AI-generated proofs, etc.

num42 7 hours ago

>Maybe mathematics just becomes a little more like other fields--relying on labs with lots of money for compute, digging through a corpus of AI-generated proofs, etc.

Dr. Tao said the same thing. Somehow, this letter came through. He wants to conduct Math competitions where participants who don’t have formal credentials can contribute to mathematical research through AI.

Title: Terence Tao - SAIR Competitions and the Future of Experimental Mathematics

https://www.youtube.com/watch?v=rB9YOi3lb7w

and this:

Daniel Litt - Working with LLMs to do high quality math

https://www.youtube.com/watch?v=0wL8NlhxXcU

cubefox 7 hours ago

> Dr. Tao said the same thing.

Apparently he has since changed his mind.

ksoped 7 hours ago

Did he say so somewhere? I don't think these ideas are contradictory. It's just an pro AI tooling but anti-slop stance.

cubefox 7 hours ago

Where is the difference?

SpicyLemonZest 6 hours ago

He sees value in mathematicians using AI to carefully study mathematics, develop an understanding of both old and new things, and help others understand the new things.

He doesn't see value in scrolling through unsolved problems asking an AI to please solve them. In his view, this is a fundamental confusion about what mathematical research is for. Knocking down unsolved problems without developing the community's understanding of them is like prompting Claude to go through a Jira board, write code for all the open tickets, and then close them without merging or deploying the code.

bluecheese452 5 hours ago

Isn’t it more like it merges the code without a dev reviewing or understanding it?

SpicyLemonZest 4 hours ago

No. Merged code can perform actions with effects on the world, even if a human being never saw it. Constructing a giant Lean formalization that nobody understands simply doesn't do anything.

dbmikus 4 hours ago

Pretty close, but IMO not quite. A math proof in and of itself is useless unless either:

   (A) it furthers human knowledge
   (B) it gets used in applied sciences, engineering, etc.
If you merge and deploy code, you have released a tool that can be used. If you ship a gibberish math proof, it's not useful unless someone else can understand and deploy it to some other means. Now, it's possible AI could understand and make use of the math proofs, even if we can't, which refutes some of my hair splitting :)

coderenegade 42 minutes ago

Not necessarily. That's the best case scenario, but proofs can be intrinsically useful in and of themselves. It's just that for problems of that nature, speculative work is often done ahead of time, e.g. the body of work that already exists assuming the Riemann hypothesis is true.

glitchc 2 hours ago

> He doesn't see value in scrolling through unsolved problems asking an AI to please solve them.

Yet that's exactly how the field works. A new grad student is tasked with finding a suitably difficult problem from a list of unsolved problems. The sweet spot is obscure, so that fewer people are working on it, but not too obscure that no one knows about it. It works the same way in theoretical physics and theoretical Comp Sci, and I speak from insider knowledge. The rosy view of mathematicians in the media is largely a product of marketing.

SpicyLemonZest 13 minutes ago

The authors of the declaration agree with you that this is how the field works today. They think that fact causes AI use to produce bad results, and they want to reformulate how the field works so that AI use will produce good results instead.

jrecursive 6 hours ago

taste

genxy 5 hours ago

"I believe I did, Bob" lives rent free, every time someone tells another person to fuck themselves using technology.

Thank You!

jrecursive 5 hours ago

lmao, thank you, i think

genxy 4 hours ago

Please consider making more comics. You have it.

darkstarsys 5 hours ago

Yes. I find it really interesting to consider what the machines do and will think of as intrinsically interesting to them. Will they develop their own theories of beauty, mathematical and otherwise?

318274 7 hours ago

So he got exuberant because he is funded by SAIR and the "AI for math" fund.

And embarrassingly they used him for a "coal miners should learn math" moment that just benefits the AI industry.

He has severely reversed course in the past week. Without concrete propositions it remains to be seen how much of the new resistance is for show.

SavageNoble 6 hours ago

I have zero formal math training beyond my Grade 12 Pre-Calculus class. Yet with an LLM I have recently devised an architecture with incredible math potential. Math is a language like any other, and without LLM's I never would have developed the techniques that I have.

AI is a tool. It speaks languages I don't (Math, Science, Code). I would love to participate in a Math competition without a hint of any formal advanced math training because my experience so far tells me I will do well.

breezybottom an hour ago

How could you possibly know if it has potential or not? Sounds like AI psychosis setting in.

SavageNoble an hour ago

Because every test I run with it is telling me so? The great thing about math is it can be externally validated.

throw567643u8 5 hours ago

>Dr. Tao

Professor Tao.

tzumaoli 7 hours ago

Is this the scenario described in Ted Chiang's short story https://en.wikipedia.org/wiki/The_Evolution_of_Human_Science where scientists are "catching crumbs from the table" trying to decipher the results generated by superhuman intelligence?

alkyon 6 hours ago

It's still an optimistic scenario. Artificial superintelligence may develop hypermathematics of a kind that never will be accesible to human mind, enhanced or not. One can't teach geometry to ants even if you put them on a Moebius strip.

It would be more like Lem's novel where it completely disappears from the human horizon: https://en.wikipedia.org/wiki/Golem_XIV

GPerson 6 hours ago

What’s optimistic or non-optimistic specifically about the machine having a system of mathematics beyond our comprehension within it? Why should we care about that in itself?

alkyon 5 hours ago

In the first case, we'd still have a chance to take a glance at the frontier of discovery (even if ordinary human mathematicians had to spent years translating what metahumans achieved).

In the second case all human-level maths would be solved and what lies beyond would be always out of our scope.

genxy 6 hours ago

Which is why if humanity had empathy, it would be working on how to make smarter ants, so that they can learn more advanced geometry.

CamperBob2 5 hours ago

What do you think we're doing?!

falcor84 5 hours ago

I love that as a goal!

Perhaps indeed a better understanding of what intelligence really is would allow this sort of Uplift (as in Brin's books).

tomrod 31 minutes ago

A fellow Brin aficionado!

alkyon 5 hours ago

Ant Intelligence will be next big thing in machine learning after this bubble bursts

_superposition_ 3 hours ago

It's quite possible ants understand a geometry more advanced than our own.

slopinthebag 3 hours ago

Why would this be optimistic? This sounds extremely negative to me…

well_ackshually 7 hours ago

So it wasted everyone's time, thousands of hours of research trying to disprove something said very loudly. What OpenAI is doing is a DoS of the scientific community: wasting your time trying to check if they're not wrong, and claiming glory in the mean time.

ksoped 7 hours ago

Agreed! Although, if done by a mathematician, it's not a ~complete waste. I think the community learns something along the way.

Is there an established term for the idea of "DoS"? I've taken to calling it slop fatigue.

SJMG 7 hours ago

Denial of service is the established term. Hammering their API (reviewer committees) would be an informal one

tmhn2 7 hours ago

That's true, but the story would have unfolded differently if Mochizuki had a lean-verified proof and was correct. I guess baked into my premise is that AI is producing reliable proofs (in the long term at least).

charcircuit 5 hours ago

OpenAI avoids this by formally verifying the proof.

https://github.com/openai/NavierStokesAndEuler

stabbles 7 hours ago

The maths community is now in the antithesis phase, synthesis will take a while ;)

Lee Sedol said in an interview that "losing to AI, in a sense, meant my entire world was collapsing. ... I could no longer enjoy the game. So I retired", and I think there will be folks in the mathematical community who would feel the same when the solutions pages to hard problems are suddenly available.

But on the other hand, people learned a lot from chess engines. After decades of chess computers beating humans, there was still a renewed interest in watching Leela beat Stockfish, with many people trying to understand the strategy Leela used.

If your happiness comes from grinding on a problem and making progress, the prospect of having to dig through a corpus of AI-generated proofs might be hard to swallow. But if you're willing to do that, you will still find beautiful things that only so many people can truly appreciate.

318274 7 hours ago

Chess is kept afloat by a couple of billionaires like Sinquefield, MBS and the guy who sponsors freestyle (Fisher random) chess.

Carlsen is bored by studying engine lines.

The popularity is boosted by YouTubers because chess is very suitable for somewhat higher class content.

I'm not sure we'd want that world for math. Positions will be cut just like archaeologist positions are cut now.

mna_ 7 hours ago

Chess is kept afloat by chess players, not by billionaires. If all the billionaire backers stopped sponsoring tournaments, people like me would still play, still pay for chess club memberships, still pay entry fees for tournaments, and still buy chess books, and so on.

magicalist 7 hours ago

I mean, this is the problem with analogies and trying to use them to prove things, right? People working through problems from an analysis book with their friend (or an LLM) is not the same as research mathematics. People playing in a chess club is not the same as what makes for a good chess tournament. Lumping everything together is just making this branch of the conversation less relevant.

debatem1 7 hours ago

Also, you learn to be a better chess player by... playing better players. The widespread availability of chess engines has made flawless opponents available to every player.

If your goals are understanding the game, self improvement, building thinking skills-- this is the best chess has ever been. It's only if your goal is to beat every opponent you can find that chess is in a bad place.

ArcHound 6 hours ago

It's not that easy. Playing stockfish is like playing tennis against the wall (for untitled players at least).

Even if you don't blunder anything, you'll still find yourself in a worse position without any clue as of what went wrong and why.

Whereas when playing humans, they can usually explain their approach and when they noticed errors in your play.

Quekid5 6 hours ago

> Playing stockfish is like playing tennis against the wall (for untitled players at least).

It's the same for Magnus Carlsen. Even with Queen odds, Stockfish is literally unbeatable for the best players in the world. It's just too strong at evaluating all kinds of random tangent moves (and ensuing positional advantage) which no human player can possibly pay attention due to the time required.

Stockfish vs any human is like Carlsen vs other players by about 3-5 orders of magnitude[0]. It's that stark.

[0] A wild pun appears.

EDIT: To avoid having to respond to each responder, fair comments about Queen odds. Maybe I was thinking Rook odds? Also, I kinda lumped Stockfish in with all the other engines, but I realize there are other engines with different properties ofc.

ArcHound 5 hours ago

No worries, I know stockfish is unbeatable by humans.

But sometimes, these GMs can flag it, which counts as a win (especially when it's proxied by a cheater). Sometimes they can also explain the idea that cost them the game, so they've learned something maybe.

Whereas us scrubs literally cannot do anything at all for reasons completely beyond our understanding.

jacobolus 4 hours ago

> Whereas us scrubs literally cannot do anything at all for reasons completely beyond our understanding

Computer moves are typically much more concrete than human moves: a human will play based on pattern matching ("intuition") and can only make explicit calculation of a small fraction of possibilities, after which decisions are guided by guesswork. The computers are unbeatable in practice because they can calculate concretely in seconds what might take an expert human long intensive study to notice, and they don't make the same kinds of oversights humans can make.

But if you stop and explore a particular position for an extended time, and if you have an intermediate level of chess skill, you too can probably often (usually?) figure out why it's doing something. Sometimes understanding the computer's reasons takes searching multiple branches of a tree several unlikely looking moves deep, but the collection of threats the computer was preemptively thwarting, traps it was setting, etc. are comprehensible to humans with enough effort, especially in games between the computer and a human.

The frustrating thing about playing against the computer is that it notices and thwarts every plan you might come up with, before you make up the plan yourself, and it doesn't make (human-apparent) mistakes, so the game ends up feeling hopeless. Nothing you try works on it, and if your idea is even slightly inaccurate it will be exploited.

VulgarExigency 5 hours ago

Well, that's actually not true at all. Stockfish is not a very good odds player and at queen odds is easily beatable even by bad players like me. It will just trade down into more trivial and easier to win positions that it perceives as "less bad", since everything is super-losing anyway when you start down a queen.

Leela odds networks, on the other hand, are an entirely different beast. I cannot beat Leela queen odds, much less rook or minor piece odds, and even GMs struggle against Leela knight odds.

Without odds though, yeah, Stockfish is just incomprehensibly strong by human standards. All top chess engines are, but Stockfish moreso.

debatem1 5 hours ago

I mean, I'm not a great player but I've learned a lot from working through games with stockfish using a git repo and a small script that lets me rewind to different moves and try different approaches. It may not explain its moves but if you're thinking through what's happened you can usually debug your game anyway.

psvv 6 hours ago

I think the parent comment meant professional, high-level chess. The kind people get played to play, not just do for a hobby. That's absolutely on life support.

I'm not sure what the equivalent would look like in the math field, but it probably involves a lot of mathematicians losing their jobs and the quality of human-produced math decreasing overall.

The quality of the math in general would be fine, since in this scenario cpus will keep producing it. The quality of cpu-cpu chess games is quite high, beyond human understanding in many cases.

Chess is a weird example because it doesn't really have any utility beyond itself. Even pure math sometimes ends up having use in the strangest places. Although if no one understands the frontier math (because no one is getting paid to), I'm not sure it even matters what the quality of the cpu math is?

It's a bit like a tree falling in a forest. If an LLM proves a theorem but no one understands it, did it make a sound?

fidotron 5 hours ago

> It's a bit like a tree falling in a forest. If an LLM proves a theorem but no one understands it, did it make a sound?

But in future most proofs will be for consumption by other AI models in the pursuit of yet other proofs.

It's kind of surprising so many mathematicians act surprised by this given this was clearly where automated proof assistants would lead. I guess they assumed they'd always be the ones guiding them.

mb7733 5 hours ago

> But in future most proofs will be for consumption by other AI models in the pursuit of yet other proofs.

What is the purpose of that?

Its like art being produced for AI to consume. What is gained from that?

fidotron 5 hours ago

Well if it proves useless presumably they'd stop doing it.

But if AI is to recursively self improve understanding and evolving its own foundations, which are clearly mathematical, is essential. There is no need for humans to grasp what is going on in that loop.

saghm 4 hours ago

I mean, was the point of math ever just because some humans enjoyed doing it? Even though a lot of it is theoretical, there's been all sorts of useful things that have come out of it as well due to an improved understanding of the universe through new ways of thinking about it. If it got to the point where no human could understand it and there were no ways to actually use it, I don't think anyone would bother having their computers doing it at all.

lacunary 4 hours ago

wouldn't AI solving problems in science, engineering, economics, etc be able to apply the new AI math?

tomrod 37 minutes ago

> was the point of math ever just because some humans enjoyed doing it?

Yes.

Friends and I often work on Putnam problems and this series:

The (Almost) Impossible Integrals, Sums, and Series by Cornel Ioan Vălean

TheOtherHobbes 4 hours ago

It's going to be hard to compete with something that has access to all of math at once and can find connections between elements that appear unrelated to humans.

And at some point AI will start suggesting - or doing - physical experiments.

charcircuit 5 hours ago

>but no one understands it, did it make a sound?

Does your "one" only contain humans or does it also contain other AI systems. AI math is not a single monolithic thing, but a distributed one. I see value in sharing proofs even among just AI.

stabbles 5 hours ago

Mathematics is more than establishing arbitrary facts (although some look like curiosities), it's also defining what interesting research directions are and establishing common language/notation. I think that will stay relevant?

vouaobrasil 4 hours ago

People become interested in things when they become invested in it personally, because they've contributed to it. So I don't think it will stay relevant...

Xirdus 4 hours ago

In theory you can automate finding interesting research directions by identifying conjectures with many dependencies. And notation has never been mathematicians' forte, with them trying to cram the entirety of universe into single letters.

onetimeusename 33 minutes ago

Why? How do you define interesting research directions? That used to be defined by testing the limits of human understanding i.e. some people can't figure something out. AI might have very different ideas about what is interesting and I am not sure what humans would get out of putting years into understanding AI proofs for what? What are we doing at that point? Like if you spend years understanding some AI proof of theorem 123456, why is that meaningful? I am actually asking why you think defining interesting research directions will stay relevant. In my opinion, people spend years acquiring knowledge so they can work on problems which is separate.

There's two points about this I am assuming 1) Mathematics actually has a significant subjectivity to it and is community oriented and not just climbing a never ending list of theorems that exists in the universe 2) A lot of mathematical research work is inside of a subfield and isn't directly motivated by applications. Sometimes it is but e.g. people don't work on obscure theorems about elliptic curves because of a dire need for that but more because the community found it interesting.

jasonfarnon 3 hours ago

"Although if no one understands the frontier math (because no one is getting paid to), I'm not sure it even matters what the quality of the cpu math is?"

Presumably AI will be connect the dots to the applications. As the declaration says, this isn't just about math. Human understanding is losing economic value. You can understand stuff on your own time, I guess.

The standard justification for pure math to holders of purse-strings is something like "it might lead to a useful application down the road, like crypto, who knows". That looks pretty inefficient now. We have to entertain the possibility that AI can develop the math needed for any application we put to it. Eg if number theory didn't exist, we could have asked AI for a way to transit messages securely and it would maybe come up with fermats little theorem as part of its solution or maybe come up with an approach we can't conceive of right now seeing as most of us are constrained to available number theory. Like how in the last year when I give an LLM a programming project I see it doesnt even bother with of the many software libraries I and others have written and just codes up the calls it needs on the fly or finds some other ad hoc solution.

cgio 2 hours ago

I would argue people getting paid to play chess was a short lived phenomenon anyway if you put it in context. The transition there is less related to the introduction of chess engines and more related to the shift in the media landscape.

Quekid5 6 hours ago

Chess is a fun game. That's why it's been around for 1000+ years.

There was a renaissance during Covid and due to 'The Queen's Gambit' where it gained much more mainstream popularity, but... Chess AI was already far far (like 1000+ Elo) ahead of human players at that point.

The thing is... chess is humans playing (communicating) with humans and that's what keeps it interesting. Check out the view counts of chess AI tourneys vs. human tourneys.

robotpepi 6 hours ago

I really hate this overly condescending takes. First of all, what do you know about the internals of math research that allows you to speak with so much confidence. Second, you're not even addressing the issues raised by the letter! This is not about "oh they made a bunch of problems easier". There are huge economical interest behind: who owns and has access to models? are these companies interested in developing research or they just grind PR stunts without worrying about externalities in how research is actually conducted? Etc etc.

Bourget 6 hours ago

If it's worth anything: I have a PhD in (theoretical) mathematics and I entirely stand by stabbles' comment.

There is a real, undeniable possibility of AI becoming better at mathematics in the same way that it became better at chess and Go, and in such a scenario, one may expect the community's response to be comparable.

3ash14 6 hours ago

Stockfish isn't owned by a club of three trillionaires. It does not cost $15 million to achieve a result in Stockfish.

Stockfish does not steal research or scoop researchers.

The concentration of computing resources and capital should be examined by the math community.

coderenegade 4 hours ago

This is the real problem. We're looking at a future where those who control AI have an insurmountable advantage in everything. They can control the amount of intelligence the masses have access to -- for their own safety, of course -- and they will never, ever be able to close the gap.

tomrod 32 minutes ago

Say they do. What then? Will people buy from them? What if we decide not to?

robotpepi 6 hours ago

It makes no sense to compare mathematics with chess. Chess is a sport. No one is interested in watching two machines compete. Chess doesn't have a practical impact. Etc. What you seem to suggest is that AI will be able to completely (or at least in a great part) replace mathematicians. It could be the case in the future, but no one knows right now, and more importantly: tech companies don't even think about it! they don't think on the externalities.

flatline 5 hours ago

Does mathematics still have a practical impact without humans in the loop? I don't think there's one single answer to that question, but I think it's worth considering exactly what that impact may be.

Tech companies are as much the topic of this post as AI, I think that's the immediacy.

jeremyjh 4 hours ago

To a significant extent, the pursuit of mathematics research is a pursuit of human understanding of mathematics, without knowing where it might lead, or whether it might lead anywhere at all. I don't see how the motivation for that goes away on its own, but the institution supporting it is certainly threatened by the potential loss of grant money and graduate student applications.

phoghed 4 hours ago

> No one is interested in watching two machines compete

I’ve watched quite a lot of YouTube videos where two machines compete, so you may not be completely right here

sebastiansm7 3 hours ago

https://tcec-chess.com/

Top chess engine championship is pretty fun to watch.

ksoped 7 hours ago

It's not just the isolated dumping, it's the fast, isolated, possibly untraceable dumping, without long term support.

It'll basically become slop fatigue if OpenAI starts dumping out proofs faster than the community can keep up, and some turn out to be wrong, never formalize it, don't stay to support it, etc.

tmhn2 7 hours ago

I wonder if they will continue to dump proofs, though? Their point has been made, the novelty will wear off, and it maybe won't be a priority use of their resources to spend however many millions on another big proof--they will move on to the next thing to show off I'm sure. At that point, the ones generating proofs will be, I hope, mathematicians (professional and otherwise) that are more interested in the results and community discussion.

(Well that's my hopeful, optimistic take, anyway.)

famouswaffles 7 hours ago

They aren't going to stop at one, that's for sure. They already claimed they have "made substantial progress" on another millenium problem. Let's say they bag another one (Hodge and/or BSD according to the rumors), if it looks like their internal model could solve P/NP or Riemann Hypothesis, you think they wouldn't take that chance ?

pizzly 6 hours ago

I think AI companies making a point is not the only thing at play. Discovering new maths ultimately leads to new technologies and applications. It may start theoretically but end up being of practical use in the future. Even if humans do not understand it (lose interest, too complex, or just way too many new proofs to go though) AI can use this AI derived math corpus which will help it in other fields.

demibabs 6 hours ago

They already told the NYT that they’ve made “substantial progress” on another one of the MP Problems (most likely the Hodge Conjecture).

That said, that’s probably just because of the drama miring their most recent one. After 2 I don’t see why they’d bother anymore.

famouswaffles 6 hours ago

>That said, that’s probably just because of the drama miring their most recent one. After 2 I don’t see why they’d bother anymore.

P/NP and the Riemann Hypothesis are part of the milllenium problems. They will 100% keep trying to crack those regardless.

deanCommie 6 hours ago

Even before AI we used to say if you write code that you only barely understand, then it will be to complicated to debug. (and/or maintain)

Mochizuki was still one human and it required legions of other humans to unpack and untangle to confirm that it didn't lead to anywhere in particular.

AI is now capable of constructions so complex that no human or human team can unpack. And its ability to increase that complexity is growing while our human ability is stagnant.

meta-AI analysis cannot help. We (software professionals who use AI regularly) already know that if you run into a situation where a Fable/Astra-generated analysis reaches the limits of our comprehension/complexity due to their subjectivity, throwing more AI at the problem doesn't always converge.

There are many reasons to feel optimistic about AI, and ultimately its general ability to help science and mathematics.

I see no reason to feel optimistic about the future of mathematics and AI based on the current path of frontier labs, unless the misalignment Tao is writing about can be reconciled.

zozbot234 6 hours ago

> AI is now capable of constructions so complex that no human or human team can unpack.

How can we possibly know this when we haven't even seriously started on the endeavor of actively reverse engineering these AI-generated proofs? That's a proper job for human mathematicians, because the AIs themselves are demonstrably clueless about what steps in a proof are genuinely interesting and load-bearing from a human POV. This is evidence of a limitation in AIs' capabilities, not of any kind of misaligned behavior. The fact that Tao actually uses that term in his complaint is deeply disappointing.

kfse 5 hours ago

Not to mention, there are already (pre AI) machine-generated proofs that we've pretty much agreed not to try to explain fully, like the four-color theorem which ends up with brute-force verification of 600+ cases (down from close to 2,000 when first demonstrated)

12_throw_away 3 hours ago

> there are already (pre AI) machine-generated proofs that we've pretty much agreed not to try to explain fully, like the four-color theorem

Algorithmic verification is a very unsatisfying answer to the problem (e.g., surely it's not just dumb luck that every single case happen to have this exact property), but that's an entirely different issue than saying that no one follows logic of the proof method itself.

mk_stjames 6 hours ago

For the interested; the saying I believe you are referencing in regards to writing code / debugging is from Brian Kernighan, specifically:

  Everyone knows that debugging is twice as hard as writing a program in the first place. So if you're as clever as you can be when you write it, how will you ever debug it?
(from, 'The Elements of Programming Style')

It's prescient.

antonvs 5 hours ago

It’s a cute statement, but it doesn’t really match reality. Programs written by humans can generally be debugged by humans.

Could AI write programs that humans can’t understand or debug? Probably, but that’s not what Kernighan was describing.

overfeed 6 hours ago

> AI is now capable of constructions so complex that no human or human team can unpack

Can you give an example of this?

antonvs 5 hours ago

The first major computer-assisted proof, of the four-color map theorem in 1976, was an example of this. It created a lot of controversy at the time. It used proof by exhaustion, i.e. essentially analyzing every possible relevant case, something that no human could do without the assistance of, at the time, a supercomputer.

matco11 6 hours ago

I get excited at the idea of a world in which advanced mathematical problems (and their solutions) become much more accessible to a much greater number of people. As a result, making mathematics much more loved at a societal level.

Imagine a world where these most complex mathematical problems are not accessible to a few hundred people, but a few hundred thousands people. ...Those original few hundred gifted mathematicians would have an even more prominent role, and their names and achievements would be known by orders of magnitude more people that they are now.

sghiassy 6 hours ago

As a laymen, I wish the same. But I also hope it doesn’t disincentivize those that dedicated themselves to the study

genxy 6 hours ago

Math department administrators fire Terence Tao.

Based on current reward models, the frontier AI labs will burn down mathematics as an impressive display of capabilities and in doing so, will make it impossible for people that get paid to do mathematics to stay employed.

If your job is literally to publish papers, and OpenAI and Anthropic decide that making an infinite-paper-printing machine is the best thing to show how effective their tech is, then as a demo, they destroy that industry.

DoctorOetker 5 hours ago

I wouldn't expect them to destroy that industry, imagine global squadrons of academics and mathematicians focusing their attention on LLM's, training algorithms, scaling laws, ... they're gonna try and beat the incumbent frontier AI labs, eye for an eye, tooth for a tooth

genxy 4 hours ago

The apocalypse being triggered by frontier labs picking a fight with mathematicians was unexpected.

coderenegade 3 hours ago

This is the hope, but I suspect the reality is that we see an ever widening gap between the fortunate and the unfortunate. We're looking at the automation and commodification of all knowledge, and the best models will be kept locked behind closed doors so that they can't be stolen. And, of course, "for our own protection".

onetimeusename 6 hours ago

But the work the Mochizuki case generated can also be done by AI. AI could generate a landmark proof and then people could use it to solve or simplify intermediate problems and you could use a different AI prompt to try to disprove it if you were really skeptical. From my memory I think they said it took 88 hours to solve a Millenium Problem versus the decades of time humans have put into it.

I don't like nuance here. I think progress is really measured by what humans are able to do and understand, not machines. It is significant if we find problems we struggle to solve. That tells us something. What does it take for humans to solve these problems is related.

The best analogy I can give is if you wanted to climb Mt. Everest you might ask someone for guidance. Would it be better to ask someone who has climbed Mt. Everest or someone who took a helicopter ride up near the top and then went to the peak? This is like the AI versus human gap to me. The helicopter is like using AI to generate a proof. The person who actually climbed Mt. Everest has firsthand knowledge of the experience. Same thing for a difficult proof. The struggle people have is actually valuable here. Likewise, we know people are actually capable of climbing Mt. Everest but if they had only ever rode a helicopter to the top, the knowledge of climbing it would not exist, and surely that is meaningful knowledge given the risks.

So if we rely on AI for proofs I think we lose a sense of what is difficult and why. We lose a sense of what human achievement is. Surely climbing Mt. Everest means more than taking a helicopter up? For students, why bother grinding through all the material of climbing Mt. Everest and then attempting it if the helicopter ride is how things are done now? This would have the affect of destroying knowledge.

(please do not nitpick the analogy because it's the best but perhaps a clumsy way to describe my thoughts)

visarga 6 hours ago

> Would it be better to ask someone who has climbed Mt. Everest or someone who took a helicopter ride up near the top and then went to the peak?

Depends on if I want to go by helicopter myself.

NateEag 5 hours ago

Tangent:

> From my memory I think they said it took 88 hours to solve a Millenium Problem versus the decades of time humans have put into it.

Keep in mind those ~88 hours were spread across ~10,000 simultaneous agent instances.

So, roughly 880,000 hours of compute.

Assuming a fifty-year career, and forty-hour workweeks, a human mathematician's career is about 100,000 hours of "compute".

I suspect that with six good mathematicians spending their whole careers primarily focused on it, and working together closely, Navier-Stokes might well have fallen already.

The perverse incentives of academia mean this has never occurred.

The perverse incentives of industry mean OpenAI intentionally scooped researchers who were getting close (granted, with AI help).

I'm not trying to dismiss the achievement - if the proof turns out to be solid, it's quite impressive (though much less so if the training data included the recent human breakthrough, which seems pretty plausible).

I'm just pointing out that "88 hours" is a very misleading way of framing this.

curt15 4 hours ago

> I suspect that with six good mathematicians spending their whole careers primarily focused on it, and working together closely, Navier-Stokes might well have fallen already.

> The perverse incentives of academia mean this has never occurred.

This. Mathematicians in their most energetic years are trying to get tenure or land a tenure-track job. They are disincentivized to go all-in on ultra high risk, high-reward problems. The potential downside is just too forbidding. It's much safer to develop a research program in a mainstream field that affords many opportunities for partial progress that can translate to a robust publication record.

onetimeusename 29 minutes ago

ok I realize this is a tangent but you're saying my post is very misleading and then also saying that a human mathematician's career is about 100,000 hours of compute and that Navier-Stokes could've had a solution by now if not for perverse incentives. You may be right but I don't think this is a great argument because in a year I would bet that those numbers change since computing power tends to increase or get cheaper over time. So I am taking the stance AI can outdo people if not now, perhaps soon.

CamperBob2 5 hours ago

I think progress is really measured by what humans are able to do and understand, not machines.

Building a machine that solves Millennium problems is pretty cool too. You wouldn't know it from reading these stories, though.

tmhn2 4 hours ago

I agree entirely with what you're saying, right up until your final question:

> why bother grinding through all the material of climbing Mt. Everest and then attempting it if the helicopter ride is how things are done now?

I think you answered this yourself earlier:

> I think progress is really measured by what humans are able to do and understand

People want to make this progress. Therefore people will "grind Everest" as a mathematical community, and that is maybe not so hugely different from a lot of previous mathematical work.

There's still ample room for creativity: simplifying, generalizing, asking new questions humans are interested in, ...

_superposition_ 3 hours ago

That grind is emotional. And it's something AI will never have. The desire to solve a problem.

unknownunknown7 an hour ago

Citation needed. Who are you to say large enough clusters of neurons can't develo emotions?

cevi 6 hours ago

Mochizuki's claimed proof of the abc conjecture was extremely unusual for the reason that nobody was able to extract a single useful idea from the argument. I was starting grad school when it came out, and my immediate visceral response was "if this is what number theory is going to look like in the future, then I will leave mathematics."

The current wave of AI slop mathematics might end up driving the next generation of mathematicians away from the subject for the same reason that Mochizuki would have convinced me to quit if his proof had been accepted by the community. Luckily, my professors had the taste to immediately recognize that it was garbage.

zahlman 5 hours ago

> Ultimately we think a fatal flaw was found in Mochizuki's proof, so it didn't lead anywhere in particular. But in our hypothetical "AI lean-verified proof of RH" situation, it would presumably generate substantially more of that community activity we saw in the Mochizuki situation. And if it's correct, that community activity would be productive (expository talks, students given problems to flesh out or generalize, etc).

This also sounds like a vector for trolling the community with complex putative proofs hiding a known flaw.

amelius 5 hours ago

Not if it's lean-verified.

bayindirh 5 hours ago

Didn't some of the recent proofs exploited a couple of blind spots of lean, and they were invalidated?

Edit: Yup. A bug report to Lean was disguised as a "Collatz" proof in a humorous way. Links below.

- https://x.com/gro_tsen/status/2082483878480977959

- https://infosec.exchange/@0xabad1dea/117002106099986943

amelius 4 hours ago

I don't know ... do you have a reference?

bayindirh 4 hours ago

Yup, found it:

https://news.ycombinator.com/item?id=49101465

amelius 4 hours ago

Cool, thanks!

Ethan_Barry 4 hours ago

There was a hash collision bug in the main Lean kernel that was patched, but AFAIK nothing relied on it. You'd have to know what you were doing to accidentally get there...

bayindirh 4 hours ago

The incident in the links I posted exploited several bugs AFAICS, so it's a different story than a single hash collision bug, it seems.

evenhash an hour ago

"Lean-verified" is not some magical incantation that makes a supposed proof irrefutable. Even disregarding potential bugs in the kernel as others have said.

Say that AI gives you a Lean proof and says it proves Theorem X. It could just as easily give you the same proof but claim that it proves (not X). How would you know the difference?

Nothing can really be considered proven unless a human expert can read the Lean proof and determine that (X as defined in the Lean proof) corresponds to X. The proof (at least the statement of the theorem) must be intelligible to humans to have value.

It's possible people will just start taking AI at its word. Maybe AI says "Here is a Lean proof of X" and we all just shrug and go "Okay, X is proven." But that's not how it works right now for human mathematicians. Why would we apply that standard for AI?

fph 5 hours ago

> Maybe mathematics just becomes a little more like other fields--relying on labs with lots of money for compute, digging through a corpus of AI-generated proofs, etc.

I think a better comparison is: mathematics just becomes like mining bitcoins.

sweezyjeezy 4 hours ago

I think you might have to explain that comparison a bit more to be honest. How are math proofs like bitcoins? A bitcoin has a pre-defined value, a math conjecture / proof is a bit more complicated.

roywiggins an hour ago

If there's a machine that automatically turns electricity into proofs then mathematics becomes a different thing altogether.

Timwi an hour ago

A Bitcoin does not have any pre-defined value. The value of Bitcoin keeps fluctuating, and historically has risen dramatically from its initial value of 0. If this weren't the case, there would be no investment/speculation in Bitcoin because there would be no potential for any ROI.

In reality, the value of Bitcoin is determined by humans (even if indirectly, not by planning), and I think the OP’s point may have been that maths proofs can be regarded similarly. No intrinsic value, just what humans find in it.

citizenpaul 4 hours ago

I've met a few Ph.D Mathematicians in Academia socially. My unfortunate experience was that they were insufferable,borderline hostile people. I tried to genuinely engage with them too. I've met one Ph.D Mathematician that left the industry whom was very enjoyable to talk to. I have a feeling that my experience was not unique and the Math world is mostly a bunch of too good for everyone on their high horse a-holes that are now being knocked down a peg. They don't like it obviously.

I'm not a fan of knocking down things that work, however I also find it hard to be against death of the gatekeeping old guard of any industry.

I think math is just gonna have to suck it up like every other industry now. Math productivity is longer out of reach of the average grad student. Like every other industry they are no longer untouchable and are gonna have to adjust to the new way of things or market forces will do what they always do which is refuse to fund ineffectiveness.

I've had to accept that tech/IT will never be the same. Just how it is. You can thrash against it all you want.

kzz102 4 hours ago

The difference is the scale. A few incomprehensible long papers per year, sure, we will study it. A flood of AI results closing research directions left and right, that will be a problem.

embedding-shape 4 hours ago

> closing research directions left and right

Why would research be closed in one direction? Even if AI or human says "Tried that, didn't work" or whatever, someone (or something I suppose) might very well retry it in the future, if nothing else to reproduce it didn't work, in theory at least.

kzz102 3 hours ago

AI tends to take nearly finished research directions and push it to the conclusion in one step. If deployed massively, it will pluck all the low hanging fruits causing a drought of near term promising research project. Because people who start promising research directions do not get to see it finish, over the long term fewer people will start new directions, causing the field to slowly whither.

_superposition_ 3 hours ago

I really like this take, and while I hate math I value it. Your position sounds extremely plausible and it fits with the pattern we see in the community here. Regardless if it's ai slop or not we still debate the value and attempt to understand. In the process generating new insights and ideas. Life will go on.

GPerson a minute ago

It’s frustrating that this comment is at the top because it absolutely misrepresents the actual declaration. The declaration is absolutely not making any statements about not using any AI in mathematics. The entire point is to push the use of the technology in a direction which is compatible with positive pre-existing features of the math community, and to make it better known what some of the current problems are.

pks016 3 hours ago

I never expected this many people (on this thread) arguing semantics and what not. I know that not everyone has morality and ethics, but I didn't realize it was this bad.

I'm afraid of the ripple effect of the agenda pushed by AI companies will have. In future and even now, they say AI has significantly progressed math and scientific research in general. There is truth to this, but the narrative has done more damage (so far) to the students, researchers, and the culture of knowledge transfer in academia. Many graduate students (I know) are having a crisis if any of their research worth it? If AI can (or will) do everything, what's the point of doing experiments and all? This will eventually deter a whole generation of curious minded students from research.

I guess, only time will whether this is for the good or bad. And how good AI models get without new data from research and experiments.

octoberfranklin 3 hours ago

This story's comments are heavily astroturfed.

Just compare with the comments on https://news.ycombinator.com/item?id=49639408

fc417fc802 2 hours ago

I actually feel like it's the other way around. The online discourse over the past ~day or so has seemed unusually irrational to me for a technical audience. Commenters making emotionally charged claims of wrongdoing that appear inconsistent with the published claims without justification of the discrepancies. Granted you might well doubt openai's version of events but there's a general expectation of clear evidence when advancing claims of malfeasance.

oliculipolicula an hour ago

I identify as neither mathematician nor "maker of things people want" (coder, hacker, engineer). But having friends who identify as those kinds of professionals, let me make some observations.

To use a metaanalogy from chess (once again), mathematicians play the opening game, and builders play the end game. AI is sort of a middleman connecting human understanding to applications.

I think there's a Technical argument to be made that openAI is a threat to the game itself. For example, could it have produced the navier-stokes counterexample without human inputs? since it seemed to have used the much gossiped research strategy "C" and "D", you can't absolutely be certain that Son of Astra (son of altman?) was magicking an unknown unknown from nothing (sorry to cue Rumsfeld). You have got to wait for the other five problems to be solved after general boycott

Subpar PR engine of the OpenAI leadership might kill the pipeline of inputs that they won't admit they still need in this dreamtime before "recursive self-improvement". You can call that emotional. Personally I would rather accuse mathematicians of "preferring local models that believe in the usefulness of unidentifiable individual contributors, and the uselessness of named generalist managers (ie the prompt writers at oAI)"

Big man tlb likes to say that science might be dead but engineering is just getting started. Navier-Stokes is the hammer of the nail in the science coffin. It kills science by killing the prestige of science. The engineers have to imagine that it's likely they will now get all their design ideas from the hypothetical future datacenters.

fc417fc802 an hour ago

> The engineers have to imagine that it's likely they will now get all their design ideas from the hypothetical future datacenters.

At the current rate of advancement that phase will last, what, all of 6 months if we're lucky?

gpt5 3 hours ago

I think that the main thing people in this thread are missing, is that it's not about math. AI progression is very likely to affect every single thing humans can do today. Mathematicians are feeling the blow this week, especially as there was wide spread denial in that math community over the capabilities of AI over the last few years, but it's the same problem everywhere.

pvab3 an hour ago

that is incredibly depressing

redox99 2 hours ago

> Many graduate students (I know) are having a crisis if any of their research worth it? If AI can (or will) do everything, what's the point of doing experiments and all? This will eventually deter a whole generation of curious minded students from research.

Those who think it's me or the machine will fail.

Those who realize how much you can accelerate your research with the help of AI will succeed.

timcobb 2 hours ago

What does success look like?

kibwen 2 hours ago

There will be no curiosity, no enjoyment of the process of life. All competing pleasures will be destroyed. But always—do not forget this, timcobb—always there will be the intoxication of power, constantly increasing and constantly growing subtler. Always, at every moment, there will be the thrill of victory, the sensation of trampling on an enemy who is helpless^W not also subscribed to ChatGPT. If you want a picture of the future [of math], imagine a boot stamping on a human face—forever.

timcobb 41 minutes ago

> Always, at every moment, there will be the thrill of victory

If only though, because:

> There will be no curiosity, no enjoyment of the process of life

FridgeSeal 2 hours ago

I _want_ to agree, but I fear this is too close to the old “do what you love for work and you’ll never work a day”.

It didn’t lead to a lot of people having a wildly successful career, it lead to a lot of people getting burnt out, exploited, underpaid and generally disillusioned.

There will be a lucky few, who have the benefit of being given the space to work alongside. The vast majority of people will (unless we change things) simply be made to take whatever the machine outputs and call it a day.

pks016 2 hours ago

That's the crux of it. In the current ecosystem of AI model usage, it's very hard to figure this out for research.

If you set a wrong foot and start trusting the model outputs, you can waste years searching for nothing.

How can someone realize this? By getting proper research training, failing, and learning from mistakes. For people beginning their research, it would be really hard to make decisions to move forward.

immmmmm 2 hours ago

Those who have token money will succeed.

It biases maths and theoretical physics towards the rich.

That one thing that was free.

coderenegade 30 minutes ago

I think the only thing that stops this from becoming true is what Chinese and European labs decide to do. If they can keep up and keep opening their weights, then we might see some kind of democratization. But right now it looks like the gap has increased, and those groups can't replicate research that isn't published, or distill models that are internal only, or for select (very wealthy) customers.

trimbo 2 hours ago

> Those who realize how much you can accelerate your research with the help of AI will succeed.

Yes, but, in the last week we saw an AI lab front-run[1] the research of mathematicians doing what you suggest. The lab threw something like $15M of compute at a problem and the researchers were able to spend nowhere near that. I think the authors are more concerned about that kind of asymmetry and race to publish the results.

[1] - I am not going to debate whether that was deliberate on the part of the lab or if it crept into training data, etc. I don't know and don't think it matters towards the point of the authors here.

redox99 38 minutes ago

This has always happened, way before AI. You'd spend months or years building and growing your business, and then Google would release a feature or product that would kill your business overnight because they can throw way more money at the problem, plus their branding. That's life.

AndrewKemendo 18 minutes ago

I got Sherlocked hard and it’s one of those things that when you realize it’s happening there’s nothing you can do you just gotta sit back and take it

ekidd an hour ago

> Those who think it's me or the machine will fail.

> Those who realize how much you can accelerate your research with the help of AI will succeed.

This is only true up until a point. If I treat a mid-sized model (say, Qwen3.8 Flash Next) like a pair programmer, then yes, it accelerates my work.

But I can already see the next stage with Fable: If I give it a couple of paragraphs of spec and $50, then I can just leave the room and go wash the dishes. I learn nothing, I participate in nothing, and I bring nothing to the process. I am no longer succeeding at all. Fable's succeeding without me.

Now, in this model generation, Fable starts getting sloppy after a few thousand lines. I can still build better at scale.

But I don't expect AI to accelerate humans or improve our productivity for long. I can already see the first signs of a future where the AI doesn't need us for anything at all.

redox99 24 minutes ago

Of course you're not going to get rich with the kind of software that LLMs can one shot these days. But that kind of software like to-do lists or basic CRUD have been saturated for over a decade, way before LLMs. People overestimate how much you can one shot, yeah a good prompt can get you 90% there but that 10% remaining often takes months of extra work.

Software has always progressed this way, lots of devs back then would work on business websites that have been 99% replaced by wordpress, squarespace and instagram.

I'm sure it's the same with research, you're going to tackle problems that would have not been worth the effort or outright impossible without AI. The old stuff that you'd work for months, yeah that's going to be a prompt away.

munificent an hour ago

Why would anyone pay you to do "your" research, when they can just cut out the middle man and ask the AI directly about whatever it is you're thinking about?

So many people who are excited about AI making them more productive are, I think, drastically overestimating how much value they are adding to that process.

LordDragonfang 2 hours ago

> I know that not everyone has morality and ethics, but I didn't realize it was this bad.

This is a very disrespectful way to make a point about acting with integrity.

You should consider that maybe your views on what makes something ethical or moral are not universal -- and that coming to a discussion with the assumption that your position is the only valid one is not conducive to convincing others who disagree with you.

pks016 2 hours ago

I agree. I was wrong and I had a naive world view of morality and ethics in research and academia.

I now realize many people have different tolerance level for this.

teravor 2 hours ago

    > Many graduate students (I know) are having a crisis if any of their research worth it? If AI can (or will) do everything, what's the point of doing experiments and all? This will eventually deter a whole generation of curious minded students from research.
should they not be deterred?

we stumbled into a way of brute forcing intelligence with gradient descent.

lbrito an hour ago

>I never expected this many people (on this thread) arguing semantics and what not.

You must be new here.

myng111 41 minutes ago

This really resonates with me. I'm early in my PhD and I'm researching a niche form of data compression and IC design. I don't use any AI at all in my research, I do it the super old fashioned way, I read papers cover to cover and sections of textbooks to familiarise myself with the field.

I genuinely enjoy doing this, it's really fun to think critically about what an author wrote or how a particular approach works.

But it does make you wonder, why bother? Probably a frontier model could one shot my algorithm in a day or less. It's incredibly depressing. At least I'm not forced to use it now, but I fear I will have no choice after I join academia or industry in the future.

afiodorov 27 minutes ago

I used to be a PhD student more than a decade ago, and I published a paper containing a solution to an open problem. Yet shortly after my first publication I became increasingly disillusioned, because I started to think that within my lifetime AI would reach and eventually surpass my ability to solve such problems—and that we only had a decade or two left.

So I started saying that it only made sense to focus on problems whose solutions would be useful immediately. I even emailed my supervisor about it, arguing that our efforts were “pointless” in the sense that AI-related problems were much more pertinent and had to be prioritised.

My supervisor thought I was bonkers. I still have the email, though. Quoting myself from April 2015:

> By 2030-2040 we will have enough computing power to simulate a human brain neuron by neuron. Once we manage to create a human intelligence we will be one little step away from super intelligence: just set the intelligence to modify itself and see the exponential growth in action. Our human intelligence is bounded by a number of biological factors (e.g. size of a skull) and even the smart human who has ever lived will appear to be a primitive ant to a supper intelligence (machine intelligence will also have perfect motivation). There is plenty of literature on this if you are interested in discussing this further. > > What does it have to do with research in pure maths? I can say that research in pure maths which won't come handy in the next 60 years is just wasted effort. The super intelligence will be able to do maths way better than humans. I believe a lot of current efforts should go into researching of artificial intelligence (or areas to do with AI) instead rather than the pure maths. I want to be proven wrong but most mathematicians I interact with are too narrow-minded to counter me and they just laugh about even contemplating the above. Frankly I am myself so perplexed that I take the above seriously, but I do and it's hurting my motivation.

I'm quite curious what my supervisor thinks of that email now.

hn_throwaway_99 22 minutes ago

I commented something similar on a bunch of other posts. The thing that scares me about a lot of the AI community in general is that their utopia is basically more frightening to me than their doom scenarios. They present this "incredible abundance" as the ultimate human endgame, but I agree, when we all end up like the humans in Wall-E, what's the purpose of it all?

And then to get retorts of "it's just just rich techies that want to find 'meaning' in their jobs while millions starve in the third world". Why would we expect the most wealth concentrating technology in history to lead to mass benefits for those in the 3rd world?

dmoniz 2 minutes ago

The rise of AI is going to lead to a lot of similar issues in many fields as it grows and develops further. This can be seen form 2 perspectives. The death of intelligence as we no longer need to think for ourselves or understand anything since AI can do it.

Alternatively, and this is what I choose to believe, it will lead to further intellectual enlightenment and advancement for use as a species as we start to discover new problems and areas of research that we had never conceived of before.

If we let AI take over all of our thinking then we are heading in the wrong direction. If we continue to ise it as the tool it is it will help us grow and advance as a species.

jeremysalwen 8 hours ago

To me it doesn't seem like what AI has destroyed is the ability for mathematicians to develop understanding and share it with each other, but rather it's destroyed the yardstick (solving open problems) that has traditionally been used to measure how much they have contributed to that understanding.

I do see how this is a problem in terms of assigning credit, but I think the cat is already out of the bag in terms of these models being capable. Even without AI labs spending millions of dollars to solve millennium prize problems, there are plenty of other people who will use them to pick low hanging fruit. I don't think any social solution is going to make things go back to the way they were, where you could share your progress towards a famous open problem without risking someone "scooping" you within a couple of days.

I think that the most likely outcomes are either mathematics becomes more secretive, or there is a more deliberative approach to assigning credit than who was "first" to solve some problem. In the former case, this may slow down progress, and in the latter case, this could mean that credit would become more subjective, and be a continual source of controversy.

HeavenFox 8 hours ago

Agreed. The problem is not with AI per se, but with the reward mechanism in academia in general.

cluckindan 8 hours ago

For those outside academia, the ”reward mechanism” is a choice between A) being a genius and working hard to become a leader in your field, B) becoming very good at writing grant applications, C) capitulating to corporations and living with the moral burden of their exploits

wbl 8 hours ago

Open problems are not that yardstick. Fermat's Last Theorem is the result of Wiles and Wiles-Taylor, but without key results from Serre, Ribet, Ihara, Langlands-Tunnels, and Frey's program none of what Wiles did would work. But Wiles did get the prize. Nowadays I think the inputs have shrunk a bit by doing more in the R=T theorem so less other cleverness needed.

cubefox 7 hours ago

Doing the last step of solving an open problem is the yard stick. But not for much longer: https://davidbessis.substack.com/p/the-fall-of-the-theorem-e...

jrflo 8 hours ago

Right. It seems like reading an AI proof (although it may not be well written) will provide the same insights as reading a proof from another mathematician, assuming it's been reviewed and edited, just like any human-authored publication. If the work is inherently valuable on it's own, I feel like that's mostly what matters.

tmhn2 8 hours ago

At issue is the fact that it doesn't typically work like: mathematician produces a proof in isolation, generates a PDF, and shares it with a bunch of people. There's a whole culture and community going on behind the scenes with conferences, seminars, lectures, chats in the hallway, advising students, etc. that AI-generated proofs bypass.

jrflo 8 hours ago

I don't think that AI would disrupt any of that. Even if AI solves a problem, you can still discuss the methods at conferences, seminars, lectures, chats in the hallway, advising students, etc.

reasonableklout 8 hours ago

Tao’s key point is that the way people are using AI today does disrupt that. Because people and companies are valuing the results over understanding. So we’re getting slop results rushed out that are automatically verified.

Moreover in the past, discussion and idea sharing would happen naturally to overcome the friction of the process. But now when OpenAI is stuck on a particular part of NS for example, they can just throw more capital & tokens at the problem.

Shaptse 6 hours ago

I won’t speak for Tao. But it is not “people” who have valued the result over understanding. “People” are a bit impressed that the machines are as good or better than the priests who have been praying at the inscrutable altar of pure mathematics. But it is mathematicians—I am speaking as a mathematician—who have adopted a culture of valuing results over exposition and transparency. The field has been rife with extremely opaque papers for many years and the character of research participation has been one of exclusion and competition over proof priority at the expense of understanding and transparency. It is unbelievably ironic to listen to mathematicians complain about “AI slop” when they have built careers upon human “slop” if slop means papers crammed with technical density that prevents all but specialists from reading the work.

ksoped 7 hours ago

Scale the problem down to an email. You're saying AI won't disrupt writing an email, when the text is vetted and authored by a person. True!

But now let chatGPT write lengthy emails unrestricted and now no one wants to read your slop anymore. That's what is being advocated against.

tmhn2 8 hours ago

The issue of credit is a relatively minor point in the declaration.

It's more about bypassing the culture and processes mathematicians have developed that lead to human understanding, generating new ideas, and bringing up new generations of mathematicians. (See also his article about "non-renewable mining" of good problems.)

Reducing mathematics to "let's just generate results through an isolated and automated system" is a misalignment since it bypasses those processes.

jimkleiber 8 hours ago

I wonder if this is a root of the complaints across fields, how AI is ruining the greater picture and process in writing, acting, drawing, filming, coding, and more.

That it makes life more ends and less means.

gritzko 8 hours ago

Let's take your argument one step further.

Suppose that tomorrow we learn that AI just exploited a bug in Lean and the proof is, in fact, bullshit. Or suppose it is the case, but we never learn that.

Where are "ends" and where are "means" here?

margalabargala 8 hours ago

Well, given a proof of something then a system can be built that relies upon the inviolability of that thing.

Should the proof turn out to be bullshit, then that system will be revealed to be unreliable. Maybe.

Xelynega 7 hours ago

It's also known as "commodification of labour", and AI is just the latest and greatest tool to do it.

Luddites complained that the trajectory of technology was to allow less skilled workers to mass produce goods via machines owned by factory owners, as opposed to helping skilled workers build up and use their skills while passing them on.

Now we have a lot of money and time focused on LLMs owned by a few companies, making it easier for them to monetize low skill labour(prompting versus art/research/artisanry)

jrflo 8 hours ago

I don't think it really attacks human understanding though. You can still read and understand an AI written proof. If another person comes up with a solution to a problem, you can read their methods and understand it. It doesn't matter if a human came up with that or not. It's really only attacking the "generating new ideas" part.

9question1 7 hours ago

That's precisely the problem though. You cannot still read and understand an AI written proof at the current skill level of the AI being applied, because they're orders of magnitude longer than human written proofs even when they don't need to be, and spend most of that length on the parts that aren't important. This has been really thoroughly documented by expert mathematicians who are engaging with AI in public like Terence Tao and showing in detail how much work it takes working alongside AI to figure out how to understand AI generated proofs. With human generated proofs that process is forced to happen before publishing the proof because the new style of AI generated proofs validated only by formal verification is supplanting the old human peer review process that forced the burden of understanding onto the publisher and not the reader.

jrflo 7 hours ago

That doesn't seem to be true. The OpenAI NS paper was 166 pages. Wiles-Taylor proof of Fermat's last theorem is 129 pages. The length is not unprecedented for a difficult unsolved problem.

To be honest, I feel like the difficulty of reading AI proofs is due to the fact that we are on the verge of being beyond human comprehension. This is a demonstrable fact as no human has figured this out despite the problem being open for almost 100 years.

jscd 7 hours ago

This is a very token-brained take. The length of a work has no bearing whatsoever on its comprehensibility.

pred_ 7 hours ago

> To be honest, I feel like the difficulty of reading AI proofs is due to the fact that we are on the verge of being beyond human comprehension.

I can see where that's coming from, but I really don't think it's the case. Even with Astra, the proofs you get are just off in a way that doesn't signal superhuman comprehension. As 9question1 says, a common theme is that they dwell on insignificant steps. Another one is that they'll often be full of terminology that either doesn't exist, or has this weird quality where it looks like it is trying to make some minor insight seem much greater than it is. At first glance, that'll often make it look like it knows more than you, but when it's really just doing the same thing but in a more complicated and worse fashion, that to me isn't a signal of comprehension at all. The bizarre thing is that despite all the "stochastic parrot" style nonsense you'll get in individual proof steps, they still often combine to something valid.

In either case, what all of this means is that the working mathematician still needs to go through, and generally completely rewrite, any proof output by an LLM. Otherwise you are passing the burden of unreadability onto the reader.

LegionMammal978 7 hours ago

Yeah, that mirrors what I've seen throwing some of the leading models at a set-theory problem that's stumped me (https://mathoverflow.net/q/511601): in this case, the problem does not easily yield to the standard tools, but the LLMs do not recognize it as a major open problem they should give up on. So they seriously try it, but typically end up in a loop of inventing certain classes of simple solution or counterexample attempts, defeating them, and trumpeting each one as a major result, each time inventing some new terminology.

It's definitely quite curious that the AI labs are able to push these results through seemingly with pure brute force. Perhaps it's largely a function of how many monkeys you have attempting various constructions on top of the known results and strategies the models have memorized.

paulhebert 4 hours ago

That matches my experience with AI writing in software engineering so I’m not surprised

curt15 4 hours ago

> This is a demonstrable fact as no human has figured this out despite the problem being open for almost 100 years.

That's not true. Alpoge and Buckmaster's related LLM-assisted blowup result (https://news.ycombinator.com/item?id=49605915) utilized a strategy developed recently by Cordoba and Martinez-Zoroa.

lelanthran 7 hours ago

> You cannot still read and understand an AI written proof at the current skill level of the AI being applied, because they're orders of magnitude longer than human written proofs even when they don't need to be, and spend most of that length on the parts that aren't important.

Just like how they write software, then :-)

ksoped 7 hours ago

Not "a" human's understanding; Humanity's understanding. Understanding the research problem, and the solution especially, is a lot more involved than simply "read their methods". That's the whole point being made.

It matters if a human came up with it because of everything mentioned in the article... A mathematician's solution is necessarily built on other's ideas that have been disseminated, internalized, pressure tested etc. Methodologies differ too. AI can abuse its compute resources and generate a true/false or counterexample statements, without laying the foundation that a decade of globalized research would have.

well_ackshually 7 hours ago

>You can still read and understand an AI written proof.

No you can't lol, they're multi million lines of Lean, which is already an obscure language to understand. It's an assault on your senses.

jrflo 7 hours ago

It's not only Lean code, there are English writeups too. To my understanding the pipeline for these problems is 1) solve in english 2) formalize systematically to check. No one is tackling problems purely in Lean, to my understanding.

https://cdn.openai.com/pdf/32d9f210-8b73-45e0-91bc-82a30aef8...

jeremysalwen 6 hours ago

I agree that the declaration doesn't focus on credit, but I think it's still at the root of the problem. Because ask yourself: if the AI generated proofs are not creating any new ideas or insight, just brute forcing a boolean true/false result, then why can't mathematicians simply ignore their results? Why does it matter if OpenAI or even amateurs with AI are "solving" these problems, without contributing to any deeper understanding?

I don't think "intellectual poisoning" is really the mechanism that harms the mathematics community.

The harm is if you have a community of mathematicians who are focused on expanding human understanding, then having instant access to a bunch of AI proved results muddies the water about who has contributed what. If someone could scoop any significant theorem at any time by pointing an AI at it, how do you really demonstrate that you have created new understanding? Or that your new understanding is about something important? How do you prove that the AI needed your new concepts to be able to solve it?

glitchc 2 hours ago

> The issue of credit is a relatively minor point in the declaration.

What a load of croc. This entire debate is fueled by a perceived lack of attribution. The AI learnt from researchers and did not give them a sporting chance of being first before scooping them. They were expecting some sort of fair play, instead they got a ruthless machine. Every other tangent to this debate is irrelevant, the culture, the community, the shared symbolic growth. Every mathematician I know is secretly trying to one-up their peers.

bz64 8 hours ago

Nothing will be lost if the credit system disappears. History of Science has many examples where wipe outs happen. The chimp brain cant survive without creating elaborate stories about how important it is, more as a cope to its own limitations and what it cant predict or control. Humility is good for health. 3 inch chimp brains didnt create the universe.

jameshart 8 hours ago

> destroyed the yardstick … that has traditionally been used to measure how much they have contributed

This goes much broader than mathematics or academia. This is the entire basis via which society distributes its wealth: based on a labour market derived valuation of ‘contribution’.

palmotea 8 hours ago

> This is the entire basis via which society distributes its wealth: based on a labour market derived valuation of ‘contribution’.

Correction: that's not how society distributes its wealth, it's how it throws some bones to the masses. I wouldn't be surprised if over half the wealth goes to people who don't sell their labor at all.

malwrar 8 hours ago

I feel a similar fate will befall engineers too. Your predictions anre quite interesting from that perspective.

Markets defined entirely by law have distorted our collective understanding of what can actually be built with the knowledge our species has accumulated thus far. How will traditional shields that have protected capital accumulation in tech to survive in a world where governments now realize control of technology is a national priority? Especially as we see its impact on modern warfare, and that such conflict looks like it’s only escalating over time.

Mathematicians appear to me (as an outsider) to exist in a field without such distortions, and I think offer engineers a preview of what’s to come. I certainly have completely ceased sharing original ideas online at this point.

glitchc 2 hours ago

Engineers are needed to build real things. AI isn't there yet.

sigmar 8 hours ago

>or there is a more deliberative approach to assigning credit than who was "first" to solve some problem

it feels like an unintended consequence of the millennium prize is that people view the [last contributor to the solution] as the only one to make progress on the problem. I've never viewed Poincaré as solved by one person and the objective of the prize was to encourage more people to make attempts and contribute towards progress.

this issue is independent, but in these circumstances perhaps interweaved, with the 'ai is taking over math' concerns

omnicognate 8 hours ago

The statement is not about AI but about the behaviour of AI companies. OpenAI have put vast resource into solving open maths problems: many millions of dollars of compute just on the Navier-Stokes result, plus whatever they spent on the broader Millenium Prize problems initiative and the other results they have published. Anthropic are doing the same. The statement is asking them to stop doing this.

AI companies are investing these resources primarily as a marketing exercise. There is no near term commercial value to a 100 page Lean proof of blow up in an extreme special case of Navier Stokes, besides the bragging rights. As the statement says any commercial value in this stuff comes a very long time later after new insights and techniques have been digested, integrated into the mathematical canon, expressed in ways that don't take a lifetime of study to understand, etc. (things that AI is not yet capable of doing itself). The bragging rights, on the other hand, are massively valuable. There is a mystique to maths that makes "our AI solved a Millenium Prize problem" an irresistable headline for a company like OpenAI.

What the mathematicians are saying is stop pouring resources that most mathematicians can only dream of accessing into projects that are actively damaging to their field. They face a massive challenge of figuring out how maths can evolve in the face of this new technology, and this is not helping.

dominotw 7 hours ago

yeah beacuse no one trust any of their benchmark results now they are scrambling to find a signal thats undeniable

pfdietz 7 hours ago

I expected this. They prove a millenium result, but it doesn't count because they are bad people.

It's this sort of thing that motivates people to burn down the institution you might be trying to defend.

magicalist 6 hours ago

> It's this sort of thing that motivates people to burn down the institution you might be trying to defend.

lol, yes, this sort of thing is what many people who voted for Trump were saying, and things are going great for them.

> They prove a millenium result, but it doesn't count because they are bad people.

OpenAI has only themselves to blame for this, and they know it. They could have handled this so much better. I'd bet there's more meeting time right now going into how to unveil future math results than on meeting about the actual math research.

mosura 7 hours ago

> What the mathematicians are saying is stop pouring resources that most mathematicians can only dream of accessing into projects that are actively damaging to their field. They face a massive challenge of figuring out how maths can evolve in the face of this new technology, and this is not helping.

Is it reasonable for any field to make such demands? If this were doctors objecting to AI becoming good at medical practice would you have the same concerns?

While any idea of OpenAI spying on people to pursue their goals is disgusting, the rest of this is par for the course, as Kasparov experienced with IBM in the 90s. Humans still play chess after all.

drivebyhooting 7 hours ago

Chess is even more mainstream and accessible now!

omnicognate 7 hours ago

It's not a simple matter of "becoming good at", and yes, I could very well have similar concerns, depending on how it impacts the field. The statement itself mentions that such concerns exist in many other fields.

I doubt OpenAI will take such a combative stance and accuse these mathematicians of "demanding" things, as you do. As I said, the purpose of this is marketing and the statement simultaneously undermines the value of that marketing (showing these projects as irresponsible) and gives these companies an even better piece of marketing in its place: "our AI got so good at maths the mathematicians begged us to stop". It's entirely possible they will stop pouring millions into these projects.

mosura 7 hours ago

So let’s say OpenAI cure cancer and put every cancer researcher out of work depriving them of intellectual satisfaction, this would also be a problem? It would certainly impact the field.

The fact is these fields are supported by society because of the benefits to everyone else. Once the same results can be achieved in a cheaper and faster way that is what will be done. We should mourn this in the same way we do buggy whip manufacturers. Again people still ride horses.

Xelynega 7 hours ago

Why don't we first assume OpenAI can create a cold fusion reactor and then give them a trillion dollars to get it done?

magicalist 7 hours ago

> The fact is these fields are supported by society because of the benefits to everyone else. Once the same results can be achieved in a cheaper and faster way that is what will be done.

Maybe it's worth double checking that you know how these fields benefit everyone else? Proving the blowup of the Navier Stokes equations in 3D isn't going to make your gas cheaper or make harvesting food easier or make drones easier to protect against. Maybe consider the deeper effects at work?

mosura 5 hours ago

Math people absolutely are important to things like the SpaceX landing control systems, hypersonic gliders, 5G networks and so on.

If you can make breakthroughs on such areas as fluid dynamics, control theory or information theory with AI then that absolutely is a big deal with real technological implications.

curt15 4 hours ago

> So let’s say OpenAI cure cancer and put every cancer researcher out of work depriving them of intellectual satisfaction, this would also be a problem? It would certainly impact the field.

Would you expect Fields medalists to cure cancer if you moved them from the math department to a medical research lab? This is precisely the fallacy that the frontier labs are counting on to inflate their valuation as their IPO approaches. They want to use headline-grabbing problems in pure maths to make their models look "smart" in the public eye. But what does "smartness" in mathematics really mean in terms of economic value? It is not at all obvious whether success in abstract mathematics should translate to successes and, more importantly, profitability, in more grounded endeavors.

Look at OpenAI's job postings (https://openai.com/careers/search/). Those roles involve far more pedestrian yet profitable duties than research mathematics. So why isn't OpenAI automating them with their vaunted models? Success in one field, no matter how "difficult", does not predict results in another field.

sobellian 7 hours ago

That solution (stop pouring resources in to proofs, stay in your lane) works today. How does it work 5, 10, 20 years from now? The software and hardware advances will continue.

causalmodels 7 hours ago

What do we do about the problems that don't require many millions of dollars in resources?

Last weekend I spun up a small agent swarm and pointed it at a field of math I have some affinity towards. Within four hours I had settled three conjectures, one of which is rather famous (for the field, not in general). It cost me about four hundred dollars.

I am at a loss about what to do with these results. On one hand I feel like the mathematicians working on these should know about them, but on the other I feel a bit like a barbarian who suddenly finds themselves sacking Rome.

lasersox 7 hours ago

How do you know the proofs are correct?

causalmodels 6 hours ago

The numerical results are trivial to check. I wrote the analytical results in Lean by hand before asking a former professor to confirm after asking him to keep this private.

They're valid.

ardme 6 hours ago

It's really tough to come to terms with it but making academic contributions in general from the outside (with or without AI tbh) is not often welcome and the whole process feels very gate-kept.

FartyMcFarter 5 hours ago

If that's the case, that makes me much less sympathetic, even though I can understand how it's very disturbing to see the field suddenly changing like this.

gammarator 3 minutes ago

(academic.) Unfortunately the overwhelming majority of outsiders are missing core knowledge (or are cranks), so the optimal prior from a time management perspective is to ignore them. AI just makes engaging more costly because there’s more volume and it’s harder to get signal on whether they know what they are talking about.

magicalist 6 hours ago

I mean, let's say you spun up a swarm of agents to rewrite a large component of a well used open source library to be memory safe. You could dump it in a big PR and walk away (we all know how that would go), or you could try engaging, see if they're interested, write something up and see where it goes.

The biggest problem is, IMO, drivebys uninterested in actual results, just getting a check mark, and the equivalent of dropping a 200k line PR on people and expecting them to be interested and do the work for you. These are things many on HN are familiar with and know how to do better :)

bwfan123 4 hours ago

> and the equivalent of dropping a 200k line PR on people

I can understand why the community is pissed. So now, lean proofs can be churned out at scale, and the community is left to decipher all of that slop into human understanding. There are bad actors with misaligned incentives coming in with drive-by proofs upending what the community holds dear which is to practice and propagate the art. I applaud them for this declaration.

To re-align incentives the following could happen. AI slop lean proofs are dumped unceremoniously into a lean dumpster, and what gets rewarded are results that could digested into human understanding - via the already followed human review process. Prizes are not given to lean proofs since anyone with sufficient compute can churn them out.

youoy 6 hours ago

It depends.

If you are trying to understand better the field, then do a good write up of the proofs so that people can learn from it.

If you want to earn the respect of people because you found interesting proofs. Then do a good write up of the proofs so thst people can learn from it.

If you want to plant flags and pollute peoples minds. Then please publish it anonimously, no one wants to correct LLM slop for you.

Probably we should build a repository of AI slop proofs that are only allowed to be publish anonimously. That way people may be more inclined to work on it because they would feel like they are cleaning your house for free.

causalmodels 6 hours ago

Maybe I'll end up doing the write ups pseudonymously. I have taken care to make sure the results can meaningfully contribute to field but I don't want to plant flags or really receive credit of any kind. I just think they are interesting.

I like my current life and don't want to get dragged into the current fracas surrounding the use of AI in math.

youoy 5 hours ago

If you have taken care, then share it with the community. People will be thankful and happy.

The problem is with people that may do it without contributing to the community.

causalmodels 5 hours ago

If this was limited to just three results, I would agree with you. But those three are just the ones I've managed to verify myself. The list of unverified results is quite a bit larger.

The field I've been investigating is not large. Even if I were to take the time and care to beat the interesting results into something meaningful, I'm afraid the pace at which I'm able to produce these results would not be well received.

tejstead 5 hours ago

> I am at a loss about what to do with these results.

I would recommend publishing them to Palomar (https://palomar-registry.org/) - I have no affiliation, this is an online registry of Lean-verified proofs created by Terrence Tao.

I have submitted a proof there that's also minorly important in an extremely niche field.

Anyway, I feel like it's a good place to dump AI slop lean proofs because the main point of the registry is that it verifies that: 1) your Lean challenge statement is the same as what you informally state you're trying to prove; 2) your Lean proof actually compiles.

This could be useful to future AI slop researchers who want to know if a given result has already been formalized, and they may be able to mine some lemmas from your work. Also, it's good to know for the field in general what has been proven.

I'm fairly certain you can set your publishing name to be whatever you want, so you could set it to be just the word "Anonymous", or the name of the model you used.

bubblegumcrisis 5 hours ago

It is interesting isn't it.

You asked for a painting. A robot made the painting. You looked at it and said, "well, I guess it's good. Should I put it online or something? Dunno. Hey Fred, what do you think of this?"

Meanwhile, your next door neighbor spends their entire life developing their understanding of life through art. They "understand" (maybe not in a way they can articulate) art. You go next door, you look at their painting and say, "well I guess it's good." But you also understand that your neighbor is just like you, and maybe you are a painter in another way.

I find it strange that, people can't see that, we don't need to solve hunger and poverty and work balance, and etc, by a round-about make-super-intelligent-AI. We could just solve it. It's pretty obvious how to, as well.

We can all be painters, if we put restrictions on the psychopaths.

pickleRick243 3 hours ago

what about cancer?

rurp 5 hours ago

I'm not a mathematician but it seems to me that if all it took to solve the problem was an enthusiast level understanding of the domain and a few hundred dollars of tokens then the result probably isn't that valuable. Even assuming you are the first person in the world to solve it, these kinds of LLM-friendly problems that are now easy to solve and easy to verify will almost certainly be picked off by one person or another in the near future.

It's also possible that the result is already known and you just weren't aware of it. It's easy for someone outside of a field, or even one steeped in it, to not be aware of certain solutions.

pickleRick243 3 hours ago

sounds a bit like cope. Crouzeix's conjecture was solved exactly under these circumstances and I wouldn't describe it as "not that valuable". In any case, AI capabilities will increase dramatically over the next few years while human math capability will not. That means that there's a fixed target regarding whatever is currently considered a "serious math problem" and its difficulty level. Soon the average problem solved with a few hundred dollars of compute will be at that bar.

smj-edison 3 hours ago

> while human math capability will not.

I actually partially disagree with this. What happened to all the excitement about Intelligence Augmentation (IA)? Now it's AI instead of IA. I think there's so much untapped potential for augmenting our intellect with the likes of https://dynamicland.org and https://folk.computer, as well as the work that's been going on in college math education, things like Lean, etc. I think the only reason human math capabilities haven't expanded that much is a failure of our imagination, not our potential.

adastra22 7 hours ago

Humanity is better off for knowing these proofs. This strikes me as academic NIMBYism.

Panoramix 7 hours ago

Your point is largely addressed in the article, did you try reading it? "In many fields and activities, years of training have traditionally served not only to produce a final answer or product, but also to develop understanding and the ability to formulate new questions and ideas. However, building on a vast body of previous human work, AI systems are becoming increasingly capable of producing the results of such work directly, and these goals cease to align."

The point is that these proofs are largely useless without the insights. The value of a proof is largely in the travel, not so much in the destination.

yieldcrv 6 hours ago

Different person here, I read the article and they are all wrong. Hope that helps.

Okay, to elaborate, substantively, their point is that the people using these AI models are not doing it for the love of the game, but for marketing. And instead of them - and nobody - spending millions of dollars to solve the problem, successfully, they want every problem of their academic industry to persist because even though they never solve the problem, they synthesize and solve lots of other problems nobody asked for. And get to boost their egos?

Yeah, stop that. Actual alignment is on the humans themselves, if they want to remain relevant as academics and mathematicians, they need to learn how to replicate the proofs and the steps that alluded humans for decades and don't worry about the narcissistic elements that slow their industry down.

SpicyLemonZest 6 hours ago

Their claim is only indirectly related to the motivations of the people using their models. What they're saying is that doing math in this way does not produce the same value as traditional mathematical research, and the people using these AI models aren't concerned about that because their marketing objectives don't depend on whether their results produce mathematical value. If people doing valuable work are made irrelevant by people doing a larger volume of non-valuable work, that's not a positive outcome.

ndriscoll 4 hours ago

But it does produce value. We now have an explicit solution and even a proof. Humans can then work on clarifying why it's true. Unsurprisingly, not all that different from software, where models can generate working code just fine. The details will all be there and all correct, but the architecture is currently not ideal, so a human guiding it can greatly improve the proofs.

I actually found this to be the case with some basic linear algebra notes I was recently doing in Lean (without using mathlib). The model could generate working proofs, but they obscure the basic ideas (actually I wonder somewhat if this is because the Lean code that's out there to train on doesn't make a huge effort to read like textbook proofs, which was my motivation in the first place). I give it a skeleton of a couple lines of `calc`, letting it fill in the reasoning for each line, and it does much better. Then ask it about making some macros to simplify "trivial" or "obvious" things, and it does even better. etc.

I suspect there's a good workflow where a big SOTA model makes an impenetrable proof (or code) and then a human works with a FIM model to simplify it (with the larger gnarly proof right there in context for FIM), but unfortunately everyone seems to only care about agents right now.

YeGoblynQueenne 3 hours ago

>> Humans can then work on clarifying why it's true.

Presumably you're a human. Are you going to do that?

ndriscoll 2 hours ago

There are vanishingly few research mathematician positions and it's one of the most competitive fields, so no. But I'm not sure how that's relevant. As the OP says, usually the value of a proof is not the knowledge that something is true per se, but the reasoning techniques to understand why. How can it be anything other than helpful then to have a truth oracle as you try to figure out why things are true?

SpicyLemonZest 3 hours ago

> But it does produce value. We now have an explicit solution and even a proof. Humans can then work on clarifying why it's true. Unsurprisingly, not all that different from software, where models can generate working code just fine. The details will all be there and all correct, but the architecture is currently not ideal, so a human guiding it can greatly improve the proofs.

To me this analogy points in the complete opposite direction. Imagine somebody takes a half-completed project design you're trying to figure out, vibecodes a rough prototype of it, emails your manager to announce that the project just launched in alpha, and then dumps it back on your lap for approvals and testing and productionization. Would you say that they've added value to this process? Or did they just strip away all the hard parts of the problem so they could claim credit for the easy part?

ndriscoll 2 hours ago

One of the awesome things about LLMs is they make it quick and easy to make PoCs, so yes. Proving that an approach will work before spending a bunch of deep design effort is absolutely valuable. Your exact scenario is something I've literally done: give a half-completed design to a team member and asked them to vibecode a PoC to prove the approach will work and figure out some of the details, explore scaling and failure characteristics, etc. Or I do the PoC vibecoding myself too. LLMs have been a gamechanger here.

SpicyLemonZest 2 hours ago

It's valuable for you, the person who's going to spend a bunch of deep design effort, to make POCs. Is it valuable for someone else to drive by, dump some POCs on your lap, and then leave you to do the deep design effort while they run away to study AI?

If that person then runs around telling people that they're the real author of your project, because they generated the original POC, would you consider that an accurate assessment?

ndriscoll an hour ago

Your analogy is far enough away from the way that the real world works that I'm not sure that I can really even strain my experiences to fit within it. Sure, I guess that would be annoying?

But mathematicians define their field. They're smart people. They're capable of recognizing when someone just did a vibecoded throwaway PoC and when someone has a well structured proof. Actually even before LLMs they'd publish new, clearer or more elegant proofs of old results. They can say that inscrutable proofs are exactly as valuable as they are, and that the first explanation people can actually understand carries its own prestige.

SpicyLemonZest 18 minutes ago

It's a real example that's happened to me twice in the past year, so I'm not sure what to make of the idea that it's far away from how the real world works.

I'm also not sure I understand what you're objecting to if we agree that mathematicians define their field. The source link is a declaration from 25 Fields Medallists with precisely that goal. They believe/define/declare that the type of AI-generated proofs we've seen are vibecoded throwaway PoCs; they feel that a well-structured proof must include factors such as "a proper writeup, the isolation of new methods and ideas, and citing relevant previous work of others", and the success criterion is not a true/false conclusion but rather "development and integration into the mathematical canon".

qlte 3 minutes ago

Yes, mathematicians will clearly need to rewrite the qualifying criteria for prizes to better align with the actual goals and value they were hoping to get from a solved problem. The field as a whole assumed good faith actors and collaboration, not expecting a few trillion dollar companies to walk in and start turning in piles of Lean no human understands to be able to claim "first".

This letter includes someone like Terrance Tao who publicly expressed a lot of optimism about AI for solving novel math like with the Erdos problems. It's not sour grapes but the first steps to define those new expectations for the future to reduce the perverse incentives.

And yet, predictably, people are accusing him of "gatekeeping" and ignoring the arguments he has made here and elsewhere about the benefits vs. harm in different ways of using AI.

valegrete 2 hours ago

Mathematical breakthroughs with commercial relevance are few and far between, and often depend on dusting off old results which were, at the time of discovery, "solutions nobody asked for."

The NS counterxample is actually, by any market measure, a "problem nobody asked for" in the sense that its existence doesn't have any commercial relevance (beyond juicing OpenAI's IPO). So the only long-term value solving it could have is by virtue of whatever reusable theory/insights were generated along the way to the counterexample itself. The letter is absolutely right on that point.

It's not actually clear that those insights will come faster from reverse engineering this LLM proof vs. humans building theory to solve the problem themselves. So what you're saying may or may not even be an efficient way of operating. Also, it implicitly depends on mathematicians to do the hard work of creating problems and then deciphering LLM hieroglyphics for essentially free while the only immediately profitable component gets outsourced to a frontier lab. In what world is that model going to work?

Reading between the lines, it seems like maybe you have a personal grudge for some reason and simply think the technology will advance enough to where we won't need academics at all. But you should say that in the first place.

yieldcrv 21 minutes ago

What irks me is the ego

My stance is that solving the problem is aligned with humankind

the rest is just hypothesizing a way that academics fit in this world at all

vector_spaces 6 hours ago

Do you "know", in any meaningful sense, any of OpenAI's recently publicized proofs? Do you suppose that there is any large community of non-academics that does?

One of the points the parent makes, along with the TFA, is that academia -- or more specifically, the "mathematical community"-- is a setting primarily for creating and ingesting mathematical knowledge, and disseminating it to the next generation and to other fields. Humans absorb this material slowly, through lots of discussion and collaboration -- it is necessarily a slow process. Facilitating this is one of the important functions of academia. Your usage of academic as a slur here is a bit silly for this exact reason.

I don't claim it is perfect, and we can argue about pedagogy in elementary courses till the cows come home. That's not really material. But this is one of the only settings in which such knowledge is broadly valued for its own sake, and in which there is a semblance of incentive to help others "know" this stuff as well, be they future generations of mathematicians, science and math educators and communicators, practitioners in other fields, or genuinely curious amateurs.

mypalmike 6 hours ago

I suspect that beyond just marketing, these pursuits yield plenty of useful information about model design that will likely lead to model improvements and optimizations for both mathematics and general reasoning going forward.

soerxpso 6 hours ago

Are they claiming that the only value in solving these problems was for their field's personal development process? I thought Navier-Stokes (and some of the other millenium prize problems) actually had implications for useful technology. It would be insane to demand that people avoid making progress on technology that can save lives or improve general quality of life, just to protect the sanctity of your karate belt system. Perhaps in lieu of open problems left to solve, mathematicians should be welcome to take up chess or sudoku to keep their minds spry.

Sharlin 6 hours ago

No, there's most likely zero practical benefit of having found a pathological edge case in which the Navier-Stokes equations do not work. We're most assuredly not talking about "saving lives" here. Unless advanced aliens show up and tell us they'll destroy Earth unless a counterexample to the N-S equations is provided within 24 hours.

SpicyLemonZest 5 hours ago

> I thought Navier-Stokes (and some of the other millenium prize problems) actually had implications for useful technology.

This is the core misunderstanding that the open letter is attempting to correct.

Developing a better understanding of the Navier-Stokes equations could have a number of implications for useful technology. They're fundamental to fluid dynamics, and turbulence in particular is something that many people feel we could work with more effectively if we better understood how and why it's generated. The Navier-Stokes smoothness problem is an interesting and long-standing benchmark for this understanding; we don't know why it should be so hard to answer, so we hoped that the process of developing a proof to the problem would produce more understanding. (We may still be able to extract this understanding after the fact, if OpenAI's proof is fully human-comprehensible.)

Simply knowing that there exists a finite-time blowup is not practically useful. We know that fluids in the real world don't produce random singularities, so the result can't really have much physical meaning. What it illustrates is that the Navier-Stokes equations fail to model physical fluids in some yet to be characterized way.

swid 5 hours ago

> There is no near term commercial value

How much do you think other AI companies would offer to get access to the transcripts of the generation that led to the proof? No doubt OpenAI will include it in their training data somehow and use it to build the next generation.

There is already economic value.

ralph84 5 hours ago

> primarily as a marketing exercise

Like the mathematicians working on famous problems in private until they could claim full credit for something interesting wasn't also a marketing exercise for their own careers. The commercial value (or lack thereof) of a proof doesn't depend on whether it was done by a human or a machine.

paulhebert 4 hours ago

These mathematicians dedicated their life to math and were working for a long time to achieve the pinnacle of their careers.

OpenAI just burned millions of dollars over a weekend after hearing that someone else was close to solving the problems. Their interest was in their AI system more than the actual math problems.

Don’t you see how that’s different?

ralph84 4 hours ago

I see how it can be devastating to their ego, but no, I don't see a particular difference in a company spending money for clout vs. a person spending time for clout. The underlying motivation is the same.

paulhebert 4 hours ago

Do you really think people dedicate their life to mathematics for clout?

If clout was the goal I don’t think becoming a lifelong mathematics academic would be the first step

pickleRick243 3 hours ago

for many fairly strong mathematicians, the career calculus is fame and status (mild though it may be) through mathematics or anonymity but financial reward in tech or finance. Yes, clout by becoming a lifelong academic is in fact rational for some and part of their motivation. Of course they really like what they do as well, but earning the respect of the peers they know are also respected by a large swathe of society is very important.

paulhebert 38 minutes ago

Yeah that’s totally fair!

I guess that type of “clout” feels different to me.

Wanting to be validated by peers for your talents in a niche field vs. using millions to try to solve a math problem that you don’t really care about with AI to market the gigantic company you work for.

phoghed 3 hours ago

Yeah bro, same for probably the majority of this site that write software for a living and/or hobby.

paulhebert 3 hours ago

Sorry I’m not entirely sure what you’re trying to say

phoghed 3 hours ago

How good were you at those “A is to B as C is to D” SAT questions?

paulhebert 3 hours ago

No need to be rude. I didn’t want to misinterpret when responding.

I can see how what is happening to mathematicians is similar to what is happening to coding.

I’m not sure what your broader point is? Mathematicians shouldn’t be upset? Coders should? Something else?

phoghed 3 hours ago

The point is that nearly every white collar worker is in the same boat right now, and the majority of them probably have already had much more profound impacts to their fields. So yeah, we can certainly imagine what it’s like for mathematicians.

paulhebert an hour ago

Okay. I agree with all that you said in your latest comment.

From your first comment it seemed like you were disagreeing with me but I’m not sure how.

That’s why I asked for clarification.

Edit: BTW I’m a web developer and designer, Not a mathematician

oytis 7 hours ago

Is it not an option to make academia more like a normal job, where people focus on collectively achieved outcomes rather than credit, priority etc.?

edgyquant 7 hours ago

How is that like a normal job?

Xirdus 7 hours ago

Almost every "normal" job has regular performance reviews where individual contributions, not collective outcomes, are reviewed and used as a sole input for raises, promotions, and firings. If you can't sufficiently document what you personally did, you might've as well not done anything at all.

oytis 7 hours ago

Yes, that's true, but it is still generally accepted that a completed project is a result of some collective effort where people of varying degree of seniority and ability contribute. There is also not a singular event of a project being completed with a list of heroes/geniuses making it happen, but rather a whole lifecycle of gradual development, maintenance and going out of relevance with contributors coming and going.

I can imagine mathematics of the future being more like that rather than history of discoveries with dates and names

dgellow 7 hours ago

> but I think the cat is already out of the bag in terms of these models being capable.

When there's a discussion about doing something against the damage of the AI industry: "whoopsy, sorry, another cat escape, nothing can be done".

When there's a concrete mention of an actual solution to avoid more cats escaping: "that won't happen, and even if it did, the damage is already done, and in fact it’s not that bad you all just have to go with the future we decided for you."

So the bag is wide open, more cats will escape, and nothing can be done about any of it. not about the ones that got out, and not about the ones still inside. Sounds more like a preemptive excuse for inaction, cosplayed as pragmatism

david-gpu 8 hours ago

Tao's critique of AI in the field of mathematics reminds me of what French art critic Charles Baudelaire said in the 19th century about photography [0].

Baudelaire argued that photography became a haven for failed painters, the sorts of hacks that could not finish proper training. Photography, as a mechanical rendering of the world, could only record what already existed; it couldn't transform reality the way a painting could.

He also criticized the public's craze for "rushing" into it, and complained that this technical "progress" was weakening the arts.

Do you see some parallels as well?

[0] https://fr.wikisource.org/wiki/Curiosit%C3%A9s_esth%C3%A9tiq...

matherial 8 hours ago

I'm really tired of these arguments (this and "it's just like calculators").

Photography decimated other forms of visual art, so the concern wasn't wrong. But AI threatens the entirety of human intellectual endeavors. I can make do without oil paintings in my home. I'm not sure I want to live in a future where we make do without brains.

piker 8 hours ago

Had me until the last word. Cameras did take away oil paintings, but not eyes.

astrobe_ 6 hours ago

It did. While you're busy recording an "instagramable" moment, you are not entirely enjoying that moment. Your eyes are on the screen, not on the subject.

In a way, if photography is an ersatz for painting that eventually made imaging available for the masses, then AI could become an ersatz for thinking. But it feels like I'm paraphrasing TFA.

What will happen when AI companies have spent their advertising budget on math problems and whatever else gives the maximum wow effect for the bucks? Probably customers hooked on the vain satisfaction of spending token$ to impress friends.

beambot 8 hours ago

Humans still have brains (and thus human intellect) even in the presence of AI... Even today, people make varying use of their brains.

jiggawatts an hour ago

But what use will be their brains to others if a machine can do the thinking instead for a fraction of cost?

CuriouslyC 7 hours ago

Lots of people still make bad music that other people still manage to enjoy (a lot of it has gone multi-platinum!) even though they're not Mozart or Bach.

Forgeties79 6 hours ago

This would be valid if we didn’t live in under capitalism, which reduces the way we spend our time down to dollars and cents.

anthonypasq 6 hours ago

you can still get oil paintings in your house. you just arent willing to pay for actual art

12qqtt 8 hours ago

No, there are no parallels. It is just cliche propaganda used by AI boosters.

azan_ 7 hours ago

Wait, what? Who’s the ai booster in this scenario?

warkdarrior 4 hours ago

Beaudelaire

keybored 8 hours ago

We can’t look back with perfect hindsight because both the past and present have deeply ingrained blindspots. They don’t know what it is like to live in a world with perfect edges. We don’t know what it is like to live in a world with no edges. We can read about someone who proclaims that “something will be lost”. We will just think “but I have no need for any of that.” But we don’t even know what it is.

pred_ 7 hours ago

Are you aware that Tao is among the largest proponents of using AI in mathematics? The usage itself is not the point here.

david-gpu 7 hours ago

Tao doesn't go as far as Baudelaire, but there are some similarities. In particular, Tao has criticized that AI is not being used to create new interesting conjectures, and that the rush to prove old conjectures is not giving human mathematicians enough time to carefully analyze and understand the proofs and the methods used in those proofs.

My answer to both is the same: nothing stops mathematicians from doing both of those things, with or without the help of AI. And we all understand that it will take time to do that. But complaining about the dawn of a new era of advancements seems counterproductive.

robotpepi 6 hours ago

> But complaining about the dawn of a new era of advancements seems counterproductive.

you completely misunderstood the critics. your analogy is awful. this is much closer to the industrial revolution in the uk: it brought a lot of progress, but also extreme inequality and concentration of power.

david-gpu 5 hours ago

I am very specifically addressing what I have read Terrence Tao say on AI and mathematics, as that is the subject of this submission.

Are we on the same page?

beepbooptheory 5 hours ago

It's not about whether the conjectures themselves are interesting or not, or whether people simply have enough time. It's that a bare proof made by a machine doesn't actually do much for us. There is not some set of problems that, once finished, will amount to some kind of final, correct system and we can call it a day and, like, utilize it. "Mathematics" is the people doing it (the "mathematical community" Tao references below). This other stuff is kinda just.. expensive exercises to render a result. They are only actually beneficial to us insofar as they exist in a context of research among peers.

https://terrytao.wordpress.com/2026/09/11/a-severe-misalignm...

david-gpu 4 hours ago

You have expressed a few different ideas, so I will address them separately.

> It's not about whether the conjectures themselves are interesting or not, or whether people simply have enough time.

Tao seems to think otherwise, if I am reading him correctly:

I wrote recently about how the collection of good, fruitful open problems is now being mined in a non-renewable fashion, leading to the potential scenario of these problems becoming scarce [0]

Often these solutions are announced in a rush, leaving no time for a proper writeup, the isolation of new methods and ideas, and citing relevant previous work of others. As in all creative professions, this raises severe attribution and plagiarism questions. [1]

> It's that a bare proof made by a machine doesn't actually do much for us.

I get it, and I think the same can be said about all sorts of human endeavors.

> There is not some set of problems that, once finished, will amount to some kind of final, correct system and we can call it a day and, like, utilize it.

Sure. Although there are certainly practical applications to be found along the way. E.g. proving P=NP would be potentially very significant in the real world. I think we agree.

> "Mathematics" is the people doing it (the "mathematical community" Tao references below)

Sure. And the same can be said again about all sort of human endeavors. But I don't see how that is a reason to stop using AI in those fields, either. It doesn't subtract anything, in the same way that chess engines didn't destroy the love of the game for chess.

And just like in chess, these AIs can be used to gain a deeper understanding. Including, but not limited to, explaining to humans the proof they just came up with.

[0] https://mathstodon.xyz/@tao/117237320796901560

[1] https://terrytao.wordpress.com/2026/09/11/a-severe-misalignm...

beepbooptheory 4 hours ago

Are you trying to argue toward some final verdict with regard to LLMs and mathematics? I thought you were just trying to make a comparison to Baudelaire? I think maybe being clearer on this point would help. Even if you argued sufficiently for the latter (which is going to be tough already), it wouldn't really speak to the former. Or at least: that would have to be a separate argument I think.

Also, how, in your words, do you feel like the first quote justifies your point (presumably with regard to the question of "interesting" or not)? And why do you think the second one is more about time itself rather than attribution? Do these things actually contradict the letter above (or the comment on it) in your mind or not?

In general, do you disagree with something here specifically? Or is it kind of a yes/and thing? Does any of this help, in your mind, with the Baudelaire comparison you were at least at one point trying to argue for? Its a bit hard for me to see the argument here, if there is one, just with what you have written. But I am sure I am just not knowledgeable enough to grasp the argument!

david-gpu 3 hours ago

My opinion is that the field will adapt. AI will not spell doom for mathematics. On the contrary, it is the beginning of a new era. Math is having its Deep Fritz moment, just as computer science is going through the same.

Some folks are struggling to adapt to this change. Tao actually sounds like he is doing alright compared to most, even if some of his arguments seem a bit weak, as I alluded to in other comments in this thread.

Hopefully it makes some sense. And if it doesn't, at least we had a nice chat.

beepbooptheory 2 hours ago

Haha yes I imagine we all know your opinion is something like that! But isn't it much more fruitful and interesting to form an argument for it? If he is still in the context here: is there something about Baudelaire's critique you were discussing originally that has helped you arrive at this opinion? Do you think he is wrong? Or is he right, but providing the right kind of nuance that we need in order to cope with the revolutionary changes? Genuinely curious!

david-gpu 2 hours ago

> But isn't it much more fruitful and interesting to form an argument for it? If he is still in the context here: is there something about Baudelaire's critique you were discussing originally that has helped you arrive at this opinion?

The historical record. That's why I am drawing some lose parallels with Badulaire. Incumbents being unhappy about a disruptive technology, lashing against the early adopters, and fearing that it signifies the end of their craft, when in reality it's just a period of change and adaptation. Without the advent of photography we would not have Impressionism nor all the movements through the 20th century. Photography forced painters to reinvent themselves, and LLMs will force mathematicians to do the same.

I thought the historical examples of photography and chess engines would be enough for people to connect the dots, but apparently not.

robotpepi 5 hours ago

I don't see any meaningful analogy with your cites of Baudelaire. First of all, Baudelaire discusses art, which is very different to science. Several new forms of art have emerged, and then slowly integrated, despite the strong opinions of some. Here, Tao's letter discusses mainly practical aspect of scientific research. It is not about what a "failed mathematician" would do, or if "LLM can only record what already existed"; in fact, it ackowledges that LLM could become a central tool. If you want me to be even more precise, what they're saying (without saying it out loud) is that tech companies have too much power and are being irrespondible with it because they don't even think about the externalities.

shimman 5 hours ago

Of course there's a meaningful analogy. The poster wants to make Tao sound unreasonable and against "progress" while ignoring the very real concerns Tao has.

It's what fanatics do when they want to enforce their view on the world, they have to attack anyone with a reasonable viewpoint because they can't imagine a world where someone tells them they don't like what they're doing.

david-gpu 4 hours ago

I do not want to make him sound unreasonable. He is a brilliant man. But I also think that he is a bit shell shocked, much like other luminaries in the past have been when confronted with rapid technological development, so I wanted to highlight some historical similarities, imperfect as they can only be.

As for the rest of your comment, I hope you have a great day.

GPerson 13 minutes ago

He’s not shell-shocked. He’s had an extremely consistent narrative for years now. You are experiencing a bias where you project your own views onto others, probably because you are not very informed on this.

david-gpu 4 hours ago

> It is not about what a "failed mathematician" would do, or if "LLM can only record what already existed"

This is what I was alluding to:

I wrote recently about how the collection of good, fruitful open problems is now being mined in a non-renewable fashion, leading to the potential scenario of these problems becoming scarce [0]

[0] https://mathstodon.xyz/@tao/117237320796901560

mkprc 7 hours ago

Baader–Meinhof phenomenon:

Baudelaire popped up in this article two days ago.

https://www.noemamag.com/a-new-kind-of-creative-poverty/

318274 7 hours ago

The AI people sing from the same sheet.

willmeyers 7 hours ago

Genuine Art versus Mechanism, from 1901, (https://www.jstor.org/stable/25505621) is another article that I read a few years ago that other people might find interesting.

Forgeties79 7 hours ago

It’s not a parallel. No one ever claimed photography to be the same as painting or some other medium - it was a new medium that wasn’t respected.

AI is being treated and pushed as a replacement for every medium.

onethought 5 hours ago

It literally was a replacement for portraits.

pllbnk 7 hours ago

I have listened quite a few interviews with Tao and I see him being very careful about criticizing AI. He very often emphasizes the usefulness of it. Where he is critical has a lot of merit. One of the points I clearly remember him saying that having AI be able to solve many of the open problems, regardless of how important there are (there are many open problems that are not that important) greatly reduces the problem space for mathematics students to give new problems to work on.

In a parallel thread omnicognate correctly pointed out that for AI companies it's a direct commercial loss to pour all this money into bruteforcing the solutions to these problems, and that a lot of times the solutions by themselves are not directly commercially valuable. They are doing it for stock price, trying to lure in private capital in preparation for IPOs.

Their models are good, but they are not the moat because Chinese models are good too, so what they are doing, in my opinion, is more harm than good. Mathematics is a science by humans for humans.

lbrito 6 hours ago

>I have listened quite a few interviews with Tao and I see him being very careful about criticizing AI. He very often emphasizes the usefulness of it.

This is tiresome. People should be able to flat-out criticize AI without the implied need to justify themselves all the time or "be careful". Its almost like AI has a trillion-dollar agenda backing it, to the point that you have to add a careful "its really great! But there's this little issue..." for any criticism.

Even those who are very pro AI should have the intellectual honesty of admitting that there are very valid reasons to criticize AI.

subroutine 4 hours ago

Based on the amount of low-effort criticisms out there, clearly nobody is afraid of "big AI". Tao seems careful about his critiques of AI because low-effort off-the-cuff criticism can lead to all sorts of future problems/hypocracy.

lbrito 2 hours ago

That's not really addressing the same issue though. I'm saying that there is a pressure to preamble any (good or weak) criticism of AI with an Apostle's Creed of AI positiveness, reassuring the reader that what follows isn't blasphemy against AI (to borrow Huang's terms).

glitchc 2 hours ago

> One of the points I clearly remember him saying that having AI be able to solve many of the open problems, regardless of how important there are (there are many open problems that are not that important) greatly reduces the problem space for mathematics students to give new problems to work on.

The crux of the argument perhaps. It suggests that too many people are currently studying mathematics without making much progress.

fg137 7 hours ago

We can start having a meaningful discussion when people use real reasoning instead of analogy.

archagon 7 hours ago

You don’t have to participate in the discussion, but it will take place regardless of your preferences.

ksoped 7 hours ago

Analogy is a meaningful way to discuss. Drawing parallels can illustrate a point and bring up nuances by pointing out where it falls apart.

david-gpu 5 hours ago

See https://news.ycombinator.com/item?id=49664505

BeetleB 7 hours ago

Keep in mind that his critique is very recent, and likely applying to a specific use of AI, as opposed to AI as a whole. If you've been following his Mastodon account, he's been happily using LLMs for math purposes for well over a year.

bustermellotron 5 hours ago

He posted about using ChatGPT to transcribe PDFs when it first became popular. So he’s been enthusiastic about LLMs for a while.

larodi 7 hours ago

Wonder what would've Baudelaire said about generative art as in diffusion-based imagery.

ksoped 7 hours ago

If he didn't understand photography, he wouldn't understand the nuance there either.

ksoped 7 hours ago

People aren't ready to discuss AI assisted imagery as art yet. Most discussions lack the nuance that Baudelaire lacks in that critique, which deals with the nature of art and the importance of human intention and input.

uonr 5 hours ago

Both of Baudelaire’s criticisms were reasonable, and the same thing happened again when AI image generation showed up. You think the opponents look ridiculous because you’re viewing the history from the winner’s side. As for “the public,” people had a real demand for photography as a way to record things, which is also why it won. There is no comparable public demand for proving Fermat’s Last Theorem.

echoangle 5 hours ago

Wasn’t the criticism of photography kind of correct though? I don’t think people think of photography as an art as much as painting is an art. There aren’t a lot of photographs that a layman couldn’t in principle also take, but only a few people could replicate a good painting.

vvilliamperez 4 hours ago

It's not exactly correct. Photography is just as much an art, but it's a very different medium and pieces are judged differently.

Realism isn't difficult, so critique shifts towards composition, narrative, context, process, and emotional impact.

Taking a clear photograph is easy with modern equipment, which means the bar for what's considered "good" is high.

atomicthumbs 4 hours ago

i think that's just you

scotty79 4 hours ago

Photography vs painting is the first thing that comes to my mind when I listen to AI critics.

GMoromisato 3 hours ago

I think there's a difference between "photography will change art--we need to be ready" and "photography will change art, therefore stop photography."

There is no doubt that AI has changed the practice of mathematics, just as it has changed the practice of software engineering (and will soon change almost every intellectual job).

Trying to deal with change by saying, "please stop the change" is foolish, IMHO. Mathematicians need to redesign the discipline with AI in mind. But I get that it's easy for me to say that and hard to actually do.

fasterik 3 hours ago

I don't think the analogy works because for the past few years Tao has been one of the most vocal advocates of AI in mathematics and has used it extensively in his own research. You can find several of his talks about this on YouTube. It's completely consistent to believe two things at once, that the tools are useful and that the companies are misbehaving.

Animats 40 minutes ago

This is a PR problem, not a mathematical problem. It's possibly the worst PR problem mathematics has faced since the execution of Hippasus for whistleblowing on the cover-up of the regular dodecahedron. It's still a PR problem.

So what went wrong? Mathematics education. Math below grad school is all about solving stated problems. Credit is given for solving puzzles successfully. Homework is problem sets. Everybody below a very advanced level is taught math that way. Even at the higher levels, puzzles remain important. Awards in mathematics are often tied to solving puzzle-like problems. That's still the criterion for becoming Senior Wrangler at Cambridge, "the greatest intellectual achievement attainable in Britain". This despite Polya's attempt at reform a century ago. Puzzle solving gets good grades and class rank. So it's the status indicator mathematics presents to the outside world.

Then reasonably good AI comes along. AI has become rather good at solving puzzles. So people aim powerful AIs at known hard puzzles, with some success. That blows up the status indicator system. Mathematics itself is fine. It's the status symbols that have a problem.

Maybe the Fields Medalists need to hire a crisis management team to reframe what success means in mathematics. That's what they're trying to do with that letter, but they're mathematicians, not PR people, and they don't know how.

gwd 7 hours ago

This sounds a lot to me like people in the 90's complaining that computers were destroying chess. Thirty years later, chess is more popular than it ever was, and chess players are better than they ever have been. I wouldn't be surprised if there are now more chess books now than there ever have been. Furthermore, it turns out that a lot of chess books written before computers were just wrong about a lot of things. It turns out having an oracle for the "right" answer in chess, even without an explanation, used properly, allows humans to develop broader, more accurate insights.

The argument here sounds similar. The fear, as I understand this statement to be saying, is that by being given the correct answer, in the form of a 100-page Lean proof, humans will be robbed of the chance to from insights about the structure of mathematics itself. I don't see any reason that humans can't continue to develop insights as they try to digest the 100-page Lean proof into something more manageable; but with more certainty and fewer false starts.

tossandthrow 7 hours ago

I like this approach.

A but like whenever the first sprinter hits a new world record other runners follow along.

Knowing that something is possible tends to strengthen our ability to work with it.

We will potentially see the same with math.

fantasizr 7 hours ago

=> If chess.com was worth billions and Daniel Rensch was threatening everyone to do what he says.

senfiaj 7 hours ago

Well, chess is a sport where humans are supposed to compete. But math, programming, science are mostly not, and AI might affect economy, careers, etc.

magoghm 5 hours ago

That's correct. As far as I know nodody builds any kind of science or technology on top of chess, but mathematics is at the base of most science and technology. It would be horrible if we prevented AI from solving mathematics problems, just because mathematicians want to solve them by themselves the "hard way".

NewsaHackO 4 hours ago

I guess part of the problem is that being against being against anything for economic interests doesn't really rally anyone to your cause; everyone has to make a living doing something productive for society, and professions have come and gone all the time due to technological advances. In fact, when one thinks about it, the people that are losing their professions now were major contributors to others losing their form of income. Often people talk about how they can use technological/programming/IT skills to make some secretary or administrative assistant's job obsolete. So most people just don't feel a lot of sympathy when people complain that AI are going to take those people's jobs.

sweezyjeezy 3 hours ago

As a chess fan, 100% this. We have known for the last ~15 years who the best human chess player is, and that he will lose against stockfish on his phone. But chess survives because of the human characters involved, the rivalries and dramas, watching two people trying to overcome each other under insane pressure, and sometimes coming up with something astonishing. In short - it's a sport.

There is no equivalent in math.

djierardi 6 hours ago

But computers have destroyed chess as a "sport". Nobody will sit to watch two chess programs compete, or analyze their tactics. Kinda like how now, anybody can construct a "game" over the weekend or a new song or a slop video. The value of each of these decreases to 0 as the slop overwhelms.

Bourget 6 hours ago

>Nobody will sit to watch two chess programs compete, or analyze their tactics.

I know nothing about chess yet I dare say that I'd doubt this. Surely chess enthusiasts would be interested in analyzing how a superior chess program came out victorious, no?

7373737373 5 hours ago

People do: https://tcec-chess.com/

latent-person 5 hours ago

Yes you are right, some people do watch chess engines play. TCEC (Top Chess Engine Championship) [1] streams them. Popular chess YouTubers goes over engine games from time to time too.

[1] https://tcec-chess.com/

throwaway27448 5 hours ago

That's not proof of interest so much as proof of commerce. It could easily be a money-laundering mechanism.

Mostly nobody cares about professional chess. The number of people who are actually interested in today's game and not the drama are a tiny sliver of that. This is actually great because it means we can train and evaluate both without interference from chess players, possibly even building a stable society.

derac 6 hours ago

erm actually, chess tournaments rose in popularity and more people watch now than ever

jaccola 6 hours ago

> I wouldn't be surprised if there are now more chess books now than there ever have been

Well yeah... how would there be fewer??

But the point itself is silly. Few people are putting effort into Maths for the fun of it (and of those that are many derive fun from being the only one who can produce a solution). Chess differs in that it never had any point but the game its self.

lll-o-lll 5 hours ago

> We are witnessing a general threat to intellectual work

This is the crux of it and goes far beyond Mathematics or Computer Science. To get a bunch of humans to do anything, you have to motivate them. Kleos and Timē; renown and stuff. These AI companies threaten to rip this away from everyone but themselves, and this recent millennium prize is the perfect example.

Solving this problem as a human would have led to tremendous Kleos; my name would be written in the annals of mathematics, lecture tours of praise were mine to be had for the rest of my days. This one victory would have earned my recognition throughout history. The greatest a mortal may hope for. Ripped away.

It would also have given me great Timē. The prize money, the professorships, the book deals. Gone.

If all hope of “renown and stuff” in the intellectual realm is now taken by the AI companies, they will remove all human motivation to pursue these endeavours.

Perhaps the glory will come from slaying these fell beasts.

dhbradshaw 5 hours ago

These are both big motivations but not the only ones.

But there's still play. There's still curiosity. And there's still the drive to understand something for yourself.

lll-o-lll 5 hours ago

> But there's still play. There's still curiosity. And there's still the drive to understand something for yourself.

Yes, but think about what that implies if those are the only motivations left. Gone are the professions. Gone are the ambitious.

There is plenty of space for people to work on intellectual pleasure pursuits (as there is with art and music), but the death of all intellectual based industries is still something to avoid. Or to mourn.

noslenwerdna 14 minutes ago

Good? Those people were in the field for the wrong reasons. No reason to mourn them leaving

lbrito 4 hours ago

Play and curiosity usually make poor breadwinners.

yzydserd 8 hours ago

> solving problems is only a tool and proxy for achieving the primary goal of conceptual understanding and insight.

This is the effect of AI on most intellectual disciplines, and it’s a real worry.

rbtprograms 8 hours ago

Tao basically reiterating "AI is making us dumber" a bit more eloquently and applied to math.

andriy_koval 8 hours ago

Its by choice. Mathematicians still can use AI to "achieve the primary goal of conceptual understanding and insight"

stymaar 8 hours ago

At the individual level any mathematicians can, but the social group of “mathematicians” will not unless drastic social changes are made, because that's not how the system works today. Drastic social changes usually take decades, and funerals, to occur, so the reasonable hypothesis is that it's not going to change fast enough.

andriy_koval 8 hours ago

Tao has enough moat (name, funding) to make shift happen and even capitalize on this.

nickysielicki an hour ago

Mathematics belongs to all of us. If mathematicians get depressed about it, I don't give a hoot. Five years of guided computer-aided math could improve billions of lives or solve global warming outright through cheap clean energy.

This is a litmus test for the effective altruists in AI labs: fashion or conviction? The social pressure is to slow down and let the guild keep its norms a few more years. But what's clearly best for humanity is to push forward violently and leave them in the dust. A Fields medalist's feelings are not more important than the wellbeing of the planet.

rbtprograms 8 hours ago

youre not wrong, but i think you are missing his base point. tao is saying that mathematics is a practice that the group collectively builds. students and new methods are built on top of what has previously been done, so its a community effort of research, experimentation, and knowledge sharing.

his core argument that is LLMs (and specifically LLMs owned and gated by corporations was my read), while able to solve problems, do not contribute to this practice of knowledge building. solving problems is just one piece, and the mathematics community ingests problems and new methods, iterates and thinks on them, and then produces new ideas, methods, etc. this is what he is defining as progress, and solving things like millenium problems are markers of this progress.

matherial 8 hours ago

Yeah. My other favorite example are books. Why do nonfiction books exist? There are some pathologies and corner cases, but fundamentally: to develop and share new ideas. Downstream from that, if it reads well and if you're lucky, you make some money.

But now, LLMs can generate hundreds of books per hour. They make up 80-90% of new arrivals in many nonfiction categories on Amazon. They short-circuit the system, allowing their "authors" to extract money from the system with zero effort by crowding out human work. And it's not even the question of whether these books are good or bad (although overwhelmingly, they're terrible). It's whether it's actually accomplishing anything worthwhile, or just destroying incentives for humans to write or go into any other sort of intellectual work.

In fact, I see many professions push back. Artists, writers, now mathematicians. And I'm amazed that our profession doesn't and that we have so many people who are hooked on vibecoding. I'm still waiting for that 10x payoff. All this velocity and somehow, the landscape of the software I want to use still looks the same as it did in 2021.

kolinko 8 hours ago

I would argue that creating an llm book that sells also requires human labour, just a different kind of one.

daveguy 7 hours ago

It also takes human labor to run an effective nigerian prince scam. That doesn't make it okay.

matherial 6 hours ago

It doesn't, that's the whole point. You can produce output that passes the smell test with naive buyers in a matter of hours. If you want to write good LLM books, then you gotta work closer to human speed - weeks, months - which allows others to produce 100 slop-books in the same timeframe. You still lose.

aswegs8 8 hours ago

Now you piqued my interest. Any good AI-written books out there? I thought it's all slop.

CuriouslyC 7 hours ago

Given the witch hunts around AI creative output, nobody who has put in work to avoid the stigma is going to out themselves.

suddenlybananas 7 hours ago

Why would they have had to put work in?

nomel 6 hours ago

If an author let AI do 90 to 95% of the work, cleaned up the remaining 5 to 10% themselves, would you say AI or the author wrote it? Because this is roughly the split between an author and an editor, with human writing, now.

nilkn 4 hours ago

The work is either: (1) the act of concealment: generating an AI book that cannot be detected as AI generated; or (2) actually just writing the book yourself.

sim04ful 7 hours ago

Playing a devil's advocate. Why do we need understanding ? To take an example i would say ~99% of the population do not understand how combustion engines or how semiconductors work, what say another 1% ?

Is the fear post-apocalyptic in nature ? We need some human priesthood to carry on tradition ? why ?

Let's assume in the next decade GPT-7 PRO Ultra is cheaply ubiquitous, inspectable, reproducible, transferable, reasonably un-constrained by any institutional interests.

What say the 1% ?

Just a thought.

yunwal 7 hours ago

It’s not going to be ubiquitous? There hasn’t been a single frontier model where generation n costs less than generation n-1 to run. So the reasonable thing is to assume that GPT-7 will cost even more than GPT-6, and more and more of the frontier of knowledge will be locked behind a giant paywall. Participating in any field will mean ponying up to the oligarchs that own the infrastructure that runs the model.

drdeca 7 hours ago

There ought to be more to life than sitting in a pod receiving sufficient nutrients from a tube, even if there was no doubt that humanity would in this way survive until the sun expands and makes the earth uninhabitable.

cevi 6 hours ago

I grew up in a cult. Based on my experience, I believe that the most dangerous thing a human can do is to allow someone else to do their thinking for them.

robotpepi 5 hours ago

> Let's assume in the next decade GPT-7 PRO Ultra is cheaply ubiquitous, inspectable, reproducible, transferable, reasonably un-constrained by any institutional interests.

I'd argue that this extremely extreme scenario is the only one in which it kind of makes sense to not have understanding. But let's be honest: no one knows if we'll be there (and it seems unlikely since everything reaches a plateau eventually). So, what happens if we allow ourselves to forget everything and then we don't reach the ideal scenario?

ceh123 7 hours ago

I published a substack about this just a few days ago [1], my core theory here is that we will absolutely have what I call a "highly productive dark age" in mathematics where knowledge vastly outpaces understanding driven by publish-or-perish incentives, but additionally this will lead to the loss of the skills necessary to understand.

The hopeful note is that I do think we are entering a golden age for the curious casual/semi-pro mathematician and for niche mathematics areas that won't get the attention of the top labs. Everyone is sprinting to solve the millennium problems, but this is a very exciting time to be in a sub-sub-field where you and 4 others are keeping things alive.

[1] https://substack.com/home/post/p-214740151

soerxpso 6 hours ago

The existence of solutions to problems does not prevent you from solving the problems yourself anyway, if your goal truly is personal development. You're free to go solve Navier-Stokes yourself right now. You're free to manually do all of the AI work in any field, actually. You won't get grant money or prestige (which are not a part of conceptual understanding and insight) but you will get all of the conceptual understanding and insight you're after. You have not been deprived of it.

pseudotensor 9 minutes ago

Tao et al. are effectively calling for diseases like childhood cancer to remain persistent for longer.

Physics and Biology will see major breakthroughs that WILL fundamentally alter our world. That is one key thing missing from alot of discussion here is the narrow focus on math (or parallels with software engineering). Doing well in math is key to doing well in physics and other sciences.

pseudotensor an hour ago

Job protectionism for elite mathematicians under guise of caring about student development. The glory of the super smart math person will need to shift to more creative modes, just like art had to handle photography. Attribution is legitimate issue but should not stall progress as it is easy to address via the same research mechanisms that agents already do.

tomrod 43 minutes ago

> Job protectionism for elite mathematicians under guise of caring about student development.

Huh. Weird. This hasn't been my take of mathematicians at all. The dozens I know are quite humble and dedicated to math and the beauty one finds in it.

pseudotensor 26 minutes ago

"guise" doesn't mean they don't care. It means they are shadowing their concerns when in reality they have concerns primarily about what AI will do to their success in math and the credit they will receive in the rest of their lifetime -- i.e. their legacy.

keeda 6 hours ago

I understand this stance and where they are coming from, but I can't help but think this sounds very analogous to engineers' arguments against AI-assisted and vibe-coding, especially with regard to cognitive debt. Yet the software industry is plowing ahead, reportedly pushing mountains of unreviewed code to Prod, and the world hasn't ended.

Of course, nobody's really comfortable with it, so this is also a forcing function for the industry to adapt and figure out new techniques to manage complexity and trust. I think the same will happen with Mathematics.

But it is also possible we will end up with three forms of Mathematics: the one we understand, the one we don't, and the one we don't understand but can prove to work. Kind of like magic -- with all the positive and negative connotations of the word.

It is pretty evident that these models will soon exceed our cognitive capabilities. Is it right to hold them back just because we can't keep up? Many of those discoveries will be so beyond us that we can't do anything with them, but that also means they can't hurt us. On the other hand, there could be many discoveries that we can parlay into practically useful applications, even if we don't understand them.

Just like LLMs.

Creamsicle47 4 hours ago

> Yet the software industry is plowing ahead, reportedly pushing mountains of unreviewed code to Prod, and the world hasn't ended.

Yes, this has gone so well

ls612 4 hours ago

It really has. All of the places facing an unusually high outage rate are places that have seen huge growth in their service usage (Anthropic, GitHub, etc) which is to be expected. The rest of the world has been happily chugging along with coding agents for almost a year now and things seem to still be working just fine.

Creamsicle47 2 hours ago

> All of the places facing an unusually high outage rate are places that have seen huge growth in their service usage (Anthropic, GitHub, etc) which is to be expected.

That's not true, many of these outages have been directly attributed to AI tooling.

> The rest of the world has been happily chugging along with coding agents for almost a year now and things seem to still be working just fine.

Many more services are now being attacked by AI agents that originate from all kinds of organizations including OpenAI, Anthropic, and many others. Unless you're purposefully being obtuse, I would not call that "working just fine".

ls612 40 minutes ago

So far for most organizations and services other than the ones I mentioned reliability has been the same this year as it was before. So I call that "working just fine".

boron1006 8 hours ago

To be honest I didn’t get that much outrage here: https://news.ycombinator.com/item?id=49662116

It seems like a lot of the issue here is that these problems aren’t interesting in and of themselves, but they lead down interesting roads. It defeats the purpose if you solve them without getting any real understanding.

It’s akin to saying you’ve solved “pancake flipping” problems with a waffle maker, or “travelling salesman” problems with a zoom meeting.

grufkork 8 hours ago

Going to the gym and using a forklift

cristoperb 8 hours ago

This is a good metaphor for treating the means as an end, thanks. And I agree with your parent comment that that is largely the misalignment that Terence is pointing out.

bpodgursky 8 hours ago

There is an implicit agreement that mathematics is funded, for the most part by the public, as a way to advance the state of knowledge and propagate (even if very indirectly) what has learned for the public good.

It was never about helping individual mathematicians demonstrate that they are individually good at math. It happened to work out that way, but it wasn't the goal.

criddell 6 hours ago

I've been wondering if mathematics is going to go the way of particle physics in that it will start to take a lot of money to make advancements.

grufkork 6 hours ago

Agree, my comment is rather about the fact that most work you do in school and (lower) university is not itself of interest, rather the purpose is to work out your brain. The more of your exercise you leave to the forklift, the steeper the cliff will be when you find a weight that doesn’t fit on forks. (That is, unless/until machines get good enough to completely replace people, at which point we have another few problems on hand, such as day-to-day purpose and justification of existence)

Art9681 8 hours ago

Nothing is stopping these folks from continuing to study the problems and arriving at their own solutions so they can continue having whatever insights along the way.

Well, one thing is stopping them. There will be no more adoration for their genius.

If you truly do it for understanding and not the attention, carry on. AI should change nothing about your motivations.

boron1006 7 hours ago

> Nothing is stopping these folks from continuing to study the problems

My understanding is they are? And literally everything in this world is based around incentives. If you say “well you can continue to work on understanding, but your kids are going to starve” that’s not nothing.

ThalesX 7 hours ago

It boggles my mind. Why cure cancer? The cancer researchers will be out of a job and their kids are going to starve!

daveguy 7 hours ago

The point isn't what happens after something is solved. Where is the motivation for humans to study/find a potential solution if your kids are going to starve while you do it? These things can brute force problems that have a clearly and feasibly searchable answer, but they can't make creative leaps. It will be a severe dark ages for mathematics if humans stop contributing.

ifethereal 7 hours ago

Maybe an analogy of "why not solve all medical problems at once" is closer.

We absolutely want to, of course. But you extinguish an industry and the systems of training that supply it. It's hard to know if letting it go that way is right.

anthonypasq 6 hours ago

methinks mathematicians may have a warped perception of their own importance.

boron1006 7 hours ago

You have misunderstood.

I have a the cure for cancer. Simply kill the host. Does it work? Yes. Have you learned anything from it? No.

gdsimoes an hour ago

Why not cure cancer and take care of the researchers’ kids?

robryan 5 hours ago

A lot of this would be public funding. If AI is reducing the need for mathematicians in an area it makes sense to fund it less and fund something else that isn't solved.

mastodon_acc 7 hours ago

Your comment perfectly illustrates why most are missing the point. Many people, including you, deeply believe that most mathematicians are chasing adoration of their genius, otherwise why would anyone care about such abstract work?

People like Grigori Perelman would baffle you, a mathematician who solved the Poincaré Conjecture, refused the monetary prize, field medal and continues to live a life of total recluse.

For most mathematicians their primary drive is chasing the unknown, not for anyone’s adoration, but to pursue their desire to see what lies in the beyond.

Strilanc 7 hours ago

This letter is complaining that human understanding has been crucial to advancing of mathematics, and AI companies are not bothering with it. But the promise (and horror) of AI mathematics is that, if it succeeds, human understanding becomes irrelevant. That's the goal. So this letter's message will fall on deaf ears.

Keep in mind employees at AI companies are publicly stating that they believe they're risking a >10% chance of human extinction. They're knowingly risking the lives of every man, woman, and child to continue the work. The lives of their own sons and daughters. A person already rationalizing that isn't going to shed a tear for the careers of mathematicians. Just a bug on the windshield.

philipwhiuk 7 hours ago

The point of mathematics is human understanding though.

Strilanc 5 hours ago

That's only one of mathematics' many purposes.

falcor84 5 hours ago

Exactly!

What got me interested in memes as a kid is precisely the fact that mathematics is true in a way that is wholy independent of our understanding of it.

robotpepi 5 hours ago

> Keep in mind employees at AI companies are publicly stating that they believe they're risking a >10% chance of human extinction.

Are we really going to take what they say in public seriously?

chrisjj 4 hours ago

> This letter is complaining that human understanding has been crucial to advancing of mathematics, and AI companies are not bothering with it.

Ironic or what. Mr Tao may be remembered as Mathematics' Canute.

bwfan123 2 hours ago

> So this letter's message will fall on deaf ears

AI companies are alienating the communities they serve. Instead of a win-win dynamic, they are keen on a win-lose proposition. You dont win trust by one-upping your customer. This is unfortunate and suggests a lack of adults in the room. It also reeks of hubris and is all good when making profits is not a concern. But watch the narrative shift when there is an AI slowdown which is inevitable.

Animats 7 hours ago

There's an unpopular branch of mathematics which does not have infinities - finiteism.[1] The constructive version of finitism takes the position that there is no such thing as infinity, just arbitrarily large upper bounds. You can have theorems about arbitrarily large numbers, but you never get

    1 + 1/2 + 1/4 + 1/8 ... = 2
The benefit of finitism is that it escapes undecidability.

The big objection to finiteism is that it's a lot more work. Infinity swallows many special cases. Proofs get longer without infinity, and most of the special cases are uninteresting. That's not a problem for AIs.

Someone may start up an AI and make it grind through Hilbert's program for putting mathematics on a fully consistent foundation, starting from a finiteism base. This is a huge, unrewarding job. Great for machine work.

[1] https://encyclopediaofmath.org/wiki/Finitism

feoren 5 hours ago

> 1 + 1/2 + 1/4 + 1/8 ... = 2

It's not necessarily clear that this statement requires infinity, if you're willing to treat "... =" as a shorthand. You might prefer something like "1 + 1/2 + 1/4 + 1/8 ... -> 2" if it's more clear, where "->" means something like "gets as close as you like without ever getting further away than that", but really the "=" sign is already overloaded in all sorts of subtly different ways anyway, so there's not really any trouble using it here. Almost any rigorous definition you can write down of exactly what that statement means would not rely on the use of infinity.

Animats 5 hours ago

See this introduction to limits.[1]

If you allow infinite recursion, you soon get to Godel and undecidable problems. Finite deterministic systems are decidable, because you can in principle enumerate all the states. The halting problem is decidable for deterministic systems with finite memory. It may be exponentially hard for some programs, but that's quite different from being undecidable.

(This is too long a subject to discuss here, and I haven't worked on constructive mathematics in many years. It's more practical than it was decades ago. You need power tools, which we now have.)

[1] https://www.mathsisfun.com/calculus/limits.html

cevi 5 hours ago

Finitism doesn't escape anything, it just gives you the illusion of safety. Any intellectually honest thinker should accept the possibility that 10 is a nonstandardly large number.

zeroonetwothree 3 hours ago

Limits can be defined within finitism as long as the end result is finite. Essentially it’s just a process which lets us get as close to 2 as we want.

A better example would be a limit that equals sqrt(2) which finitists would probably say cannot represent a real object because it is only defined as the end of an infinite process.

ianjbutler 8 hours ago

> We are witnessing a general threat to intellectual work, with misalignment between the outcome of the use of AI and its initial purpose. In many fields and activities, years of training have traditionally served not only to produce a final answer or product, but also to develop understanding and the ability to formulate new questions and ideas. However, building on a vast body of previous human work, AI systems are becoming increasingly capable of producing the results of such work directly, and these goals cease to align.

This is about how good taste in both research direction and in design are essential to steering AI, but we have no plan at all for instilling that taste in students or practitioners in a post-AI world.

> The issues the mathematical community faces now are similar to issues that other scientific and creative professions are facing, and indicate issues that all of humanity might face: how to make sure that, as AI changes the way work is done, we do not lose sight of what that work was meant to achieve in the first place.

Besides eroding taste and taste-building, this is about just how useful friction is as signal.

Everyone coding with AI knows it routes around difficulties like a river around a stone, which is not necessarily a good thing. It will do it tirelessly 1000 times instead of learning anything from it. AND if the AI does not fail in this, the human driver will get no signal, and never know it happened. This seems to be getting worse, not better.. my theory is that more models are cross-trained on cybersecurity stuff where the goal is success and the method doesn't matter. Fine for pen-testing, ultimately pretty bad for coherent code or math or physics.

Discrete tasks where we don't want to be bothered is a real use-case, but optimizing for it everywhere is terrible for the future of durable abstractions that we can build on and ratchet up our understanding with. Bad for the models too eventually! They can maintain a codebase with millions of special cases or juggle tons of free variables in equations, but that just encourages bad abstractions.. they have a ceiling for this too, even if it's higher than humans.

trillobyte 8 hours ago

OpenAI: "Our mission is to ensure that artificial general intelligence benefits all of humanity."

- Except the mathematicians who we'll scoop and cause existential dread among their entire field.

- Except the software developers. They'll need to become plumbers or live on UBI.

- Except the people in countries that can't afford the cost of AI tokens to keep up with the rest of the world.

Just keep picking off groups of humans for the "benefits of all humanity"... while building larger and larger disparities been the have a lots and the just have enoughs.

We're going to build humans a utopia but along the way we'll leave a trail of destruction because that's not our problem.

deaton 8 hours ago

They're building a utopia that nobody will live in.

fwlr 2 hours ago

Nick Bostrom:

   We could imagine, as an extreme case, a technologically highly advanced society, containing many complex structures, some of them far more intricate and intelligent than anything that exists on the planet today – a society which nevertheless lacks any type of being that is conscious or whose welfare has moral significance. In a sense, this would be an uninhabited society. It would be a society of economic miracles and technological awesomeness, with nobody there to benefit. A Disneyland with no children.

td2 7 hours ago

But its normal, and good, that technologal progress creates, and destroys some jobs. Imagin a cheap, 100% reliable, self driving car would be released. Death from Traffic incidence fall by orders of magnitude

Would you argue it didnt benefit humanity, because taxi/bus drivers are nolonger required

buellerbueller 7 hours ago

the thing with all this innovation is that it seems to be inextricably linked to rent extraction, so where is this "cheap" you speak of?

trillobyte 7 hours ago

I'm actually a AI optimist. I think it'd be great to have everybody getting around in self driving vehicles.

If all that AI brought resulted in just taxi/bus drivers being phased out of their jobs in a thoughtful way, then that would be more manageable at the society level. But we're talking about almost all sectors of the economy.

If the magnitude of changes that OpenAI and Athropic believe will be delivered with increasingly powerful AI (and robotics) comes in a time frame that significantly worsens a large proportion of people's lives, this is a different situation. Can super powerful AI not be developed in a way that minimizes such disruption?

yunwal 7 hours ago

What I’m hearing is “it’s ok when other people lose their jobs…”

pickleRick243 3 hours ago

lol right. The nonchalance in referring to taxi/bus driver jobs.

Remember when we all said it's a good thing when the coal mining jobs are going away and that they should all just learn to code? Maybe a little more of that energy right now.

anthonypasq 6 hours ago

dude, AI is the cheapest and most widely distributed technological revolution ever.

oytis 8 hours ago

Looks like mathematicians (like people in many other professions) have to redefine what their work means and how to define success. Hard to agree that a tool that can find a proof is detrimental by itself, rather it voids some assumptions people relied on previously

imdsm 8 hours ago

I am on this track too. If AI leads to advancements, objectively that's a positive (depending on the advancement I guess) but it's only when mathematicians realise they'll get beaten to every thing now that they're outraged.

sdeframond 8 hours ago

Say AI becomes the best at everything. Best at chess/go, best at maths, philosophy, economics, romantic advices ... and so on. Then what's the point of thinking by oneself? Of talking to one another?

What's the point of being human if we dont do human things but entirely rely on AI?

I believe this is more or less these mathematicians' argument.

oytis 8 hours ago

It has been best at chess for quite a while. Yet everyone knows who is Magnus Carlsen, even though at no point of his career he was stronger than the machine

globular-toast 7 hours ago

Magnus Carlsen plays fellow humans at the game because people still care about human competitions. What motivation would a mathematician have for solving already solved problems by hand? Can you imagine someone spending years working on a proof for an already-proved theorem just in case it leads to new insight?

oytis 7 hours ago

If insights and human understanding are more important than specific results, I don't see why people should stop producing insights and human understanding. "Working on a proof" in today's sense of trying to come up with a proof before others probably stops being useful, other ways of working will be needed

SJMG 6 hours ago

Of course not. I'm so tired of the chess analogy. It was always a game and it was always performative.

Research is fundamentally different. We are seeing the erosion of specific needs for thinking at depth. AI is the automobile for the mind. There will 100% be undesirable consequences and selective atrophy of cognitive abilities once prized. This is a loss. There's no getting the cat back in the bag at this point so long as the objective dimension of work, as in object opposed to subject, is held as the most important.

SWEs felt this same crisis late last year. Now it's the mathematicians. They won't be the last.

pickleRick243 2 hours ago

actually I think the whole point is we're finding that the chess analogy to math is in fact quite apt. Math is just an extremely complex game. Its rules can be written down and there are win conditions. It's really not far fetched to think that through information theory, a reasonable measure of depth and beauty to a definition or conjecture can be defined. Then the game is simply to maximize the number of such artifacts produced of high depth and beauty along with proofs of the relevant conjectures.

Philosophers I'm sure can debate this back and forth but it seems probable that some things we thought were ineffable are in fact quantifiable to a degree, and now we have the technology and the machines to bring that to the logical conclusion.

cute_boi 32 minutes ago

>> Once a robot can do everything an IQ 80 human can do, only better and cheaper, there will be no reason to employ IQ 80 humans. Once a robot can do everything an IQ 120 human can do, only better and cheaper, there will be no reason to employ IQ 120 humans. Once a robot can do everything an IQ 180 human can do, only better and cheaper, there will be no reason to employ humans at all, in the unlikely scenario that there are any left by that point. [1]

The answer may sound harsh, but we will not need human once AI reaches that threshold. Current IQ of AI is currently 130 as per Google.

[1] https://www.slatestarcodexabridged.com/Meditations-On-Moloch

andrepd 8 hours ago

Yes if you don't think about the matter for more than 6 seconds you would indeed conclude that, and retreat to the comfortable cliché of "it's just a tool". Meanwhile, I'm glad that there are still people who _think_ about issues and ponder the consequences and reflect on things before they become a reality.

forshaper 7 hours ago

afaict it all started with taxes, farming, administration, astronomy, & construction, perhaps they could return to seeking results there

fwlr 8 hours ago

In internet culture there’s this phrase “Hydrogen Bomb vs Coughing Baby”, meant to highlight the absurd power difference between two combatants.

In almost any scenario even tangentially involving mathematics, twenty-five Fields medallists uniting to denounce something would be a veritable Tsar Bomba.

It should give you pause that here they feel like the ailing infant.

kzrdude 7 hours ago

Well OpenAI took one step on the back foot at least, withdrawing from sponsoring this math hackathon event https://xcancel.com/danintheory/status/2098125701782372640

demibabs 8 hours ago

> I am proud to be among the list of 25 initial signatories — all Fields Medallists — to the declaration below

I wonder if there’s a Fields Medalist group chat.

akd 8 hours ago

There definitely is. Everything runs on WhatsApp

soperj 8 hours ago

most of my group chats have moved to signal.

stymaar 8 hours ago

Lucky you. I've only managed to move my family to Signal, but everything else is still on Whatsapp.

kmmbvnr_ 33 minutes ago

It sounds like we are not happy with getting answers like 42 to our questions about life, the universe, and everything

XTXinverseXTY 9 hours ago

Puts the onus on the AI companies to provide a specific replacement mechanism, no? Unless I'm unfamiliar with something else he's written that proposes something more specific and constructive

To Tao’s credit he obviously identified the problem very clearly and admits understandably "we did not have the time to have a more consultative process, as with Leiden; but we decided that the urgency of the situation was such that we needed to release a statement sooner rather than later".

pfdietz 8 hours ago

> Puts the onus on the AI companies to provide a specific replacement mechanism, no?

Why? If someone makes an innovation that undercuts the underpinnings of some existing institution, why are they are responsible for cleaning up its failure?

vatsachak 8 hours ago

Exactly. Mathematicians are trying to cope that math wasn't slop this whole time.

Grothendieck was anti-slop but most papers are slop.

I don't think AI is going to rewrite bourbaki anytime soon

_flag 8 hours ago

With your understanding of mathematical "slop", do you believe that Deligne and Scholze, two signatories, produced or produce mathematical slop?

vatsachak 8 hours ago

Grothendieck accused Deligne of "slop" with his proof of the weil conjecture.

Scholze's math is definitely not slop.

But you're taking two of the best mathematicians of the last century against my claim about averages

tossandthrow 8 hours ago

Yes it seems odd.

Out sourcing construction jobs was great for the economy while leaving entire cities in rubbles.

But as soon as it hits the privileged class there is a call to "provide a specific replacement mechanism".

vatsachak 8 hours ago

You mean manufacturing

tossandthrow 8 hours ago

Yes

keybored 7 hours ago

Whose privilege are you sticking up for here in effect?

azan_ 7 hours ago

Math researchers are privileged class? I can guarantee - construction workers are extremely rich compared with math phds and most of postdocs. Talking about privileged class is absolutely laughable here.

yuuu 3 hours ago

I would bet that the parents of math PhDs are generally more privileged than the parents of construction workers.

anthonypasq 8 hours ago

The entire western world is anti-progress and pro-incumbency, and its very tightly linked to gerontocracy.

Older people are desperately trying to keep a grasp on their current power and lifestyles at the expense of younger people and technology.

We need to ban Waymos because taxi drivers need to be protected.

We need to block housing because it would lower my property values, and eliminate property taxes while we're at it! I don't use the local schools so why should I be taxed to pay for it.

We need to spend recklessly to pay my pension and have the next generation foot the bill.

Its just a repulsive ideology.

dullcrisp 8 hours ago

They’re not asking them to stop working on AI, but to stop publishing mathematics.

pfdietz 8 hours ago

In other news, evangelical christians ask scientists to stop publishing about evolution.

We apparently have a moral obligation to protect existing power structures?

Tao should maybe consider there are people who are indifferent to, or actively want to tear down, his institutions; why should they cooperate in preserving them? Whatever happens has to be resilient in the face of defection; any scheme where everyone is expected to agree to not use AI in a way he doesn't like will not qualify.

I think he's in the "bargaining" stage of dealing with loss right now.

tonfa 7 hours ago

> In other news, evangelical christians ask scientists to stop publishing about evolution.

not sure it's comparable, but the issue is that for a lot of those mathematical results, they don't really have utility by themselves. The utility is the new branches/understanding that's being developped.

dullcrisp 6 hours ago

You have a practical obligation to obey existing power structures. That’s what makes it a power structure. What do you all think is happening here? There have been people who decided what math gets published for centuries.

pred_ 8 hours ago

How specific and constructive it is might be debatable, but he has tried to make concrete recommendations earlier; see e.g. slides 46-51 from the ICM talk: https://teorth.github.io/tao-web/slides/age-of-ai-icm-2026.p... – obviously there's some way to go still.

cevi 7 hours ago

Nothing makes me respect Terry Tao more than the line "I hate Jean Bourgain" handwritten into the margin of one of Bourgain's papers. IYKYK

datakan 8 hours ago

Everyones outraged all the time. It doesn't mean anything anymore. It's that meme from years ago about the red ants and the black ants living in a box peacefully until someone shakes the box and they start trying to kill each other. They go after each other and not the one shaking the box.

OpenAI/Anthropic are shaking the box.

Yajirobe 8 hours ago

The ants ARE going after the one shaking the box

datakan 8 hours ago

This time, until they dont anymore. The next stage is people attacking the mathematicians and accusing them. The OpenAI fans come out to attack, then the masses will pick sides and it all just becomes a mess

scottyah 8 hours ago

I don't understand how that metaphor applies here, you seem to contradict yourself. The mathematicians are mad at openai. If the mathematicians are the black ants and OpenAI is shaking the box, who are the red ants?

datakan 8 hours ago

Everyone that will now pick sides. Just wait, I’m already seeing it on Reddit.

scottyah 8 hours ago

Gotcha. In that case, the "ants" part of the metaphor is truly fitting since these people just like to comment online and are mostly insignificant. Unlike the ants, they can evolve and become the mathematicians or get a job at a frontier AI lab but then they wouldn't be spewing so many comments.

Don't worry about what they write, they just want to feel emotions from any news article.

throwaway27448 8 hours ago

You're looking too closely. Whether or not mathematicians are actually outraged makes little difference to how this article plays. Hell, lying would likely INCREASE revenue.

jrowen 8 hours ago

Everyones outraged all the time. It doesn't mean anything anymore.

Everything is sensationalized, every super niche happenstance is sold as earth-shattering drama, the outrage arms race is so tiresome.

feoren 5 hours ago

Yeah, I'm so tired of people moaning about how huge swaths of Earth's ecosystem are being destroyed and rendered uninhabitable to humans, how secret police are murdering Americans in the streets, how the president is a child rapist who openly accepts bribes, how unfettered capitalism is destroying tens of millions of lives, how the United States is rapidly falling into facism, how civil rights are being systematically destroyed, consumer and environmental protections are being gutted, and entire generations are being intentionally cut out of the possibility of economic prosperity. YAWN! Give it a rest!

Man, the people who want to just get away with open corruption sure love you.

2snakes 5 hours ago

Hyperpolitics is one term for this. There was an article in NYT about the concept. Attention economy for political info causing burnout possibly same phenomenon as Gen Alpha surrealism and viral adoption, but no institutionalization to preserve the lessons.

jrowen 4 hours ago

The thing is, very little of the outrage is directed at those things.

izacus 8 hours ago

What a bizarre thing to say - is that your attempt to discredit anyone that complains about wrongdoing?

"Everybody's enraged, why don't you like this unethical thing being done to you by a company?"

What's going on here?

adastra22 7 hours ago

Mathematics is about discovering and understanding the logical implications of assumed axioms under various inference rules.

Alternatively, some claim that mathematics is about understanding these implications.

Under the first definition, AI is already, and forevermore will be faster and better at proving theorems. Just like it is better at checkers, chess, and now go.

The author asserts that AI proofs are incomprehensible to humans, and so under the second definition AI is merely a tool to overcome one hurdle on the way to understanding.

So which is it? The author seems to claim the second definition, but bemoan the end of mathematics under the first.

loglog 7 hours ago

> Mathematics is about discovering and understanding the logical implications of assumed axioms under various inference rules.

That's like saying that programming is about producing valid programs in various programming languages.

youoy 5 hours ago

Dont confuse mathematics with the formal system. If you beleive mathematics = formal system then AI is obviously better at it, and we dont need humans.

But then who decides why a statement is mor important than another? In the eyes of a formal systems all statements are born equal.

doubtfuluser 7 hours ago

Im surprised about the sentiment in this discussion.

I totally see the problem Terence is describing. We are loosing a lot in understanding and focus if it continues like that. The solution found for Navier Stokes doesn’t have much „real value“ - but what almost always happened in the past when people worked on the difficult problems, these sparked new ideas / new theorems that broadened our knowledge. Think back at your grad studies, figuring out a proof as homework was hard, sometimes incredibly hard, but while doing it we gained a lot of understanding how things work. Now asking AI for the solution and „just“ getting it, risks our understanding, our creativity and our ability to connect the dots with other territories. I see it in students nowadays, there is much less understanding, much less creativity in finding solutions. I truly think this „short-path“ solution with the „death of struggle is one of the biggest risks with AI already for human development

kansface 5 hours ago

Some people are just out there doing automated math proofs for reasons. If you want understanding, value, and edification, that is on you, not them.