jboggan
8 hours ago
I was a graph theory junkie long ago and even moved to Budapest for awhile to study among the greats. While I was there I started working on Barnette's Conjecture which came to occupy my thoughts over the next 24 years of my life, on and off as I worked in many different fields. Last summer I even thought for a few days that I had actually solved it.
But it's supposedly proven here - problem 180. I don't know what to think exactly. I spent thousands of hours on that problem. I really enjoyed it. Hearing that it is solved somehow makes me sad in a far-off way, like hearing an ex-girlfriend died suddenly in a car crash. I don't know, there's probably a lot of people feeling odd emotions tonight.
There's no Lean proof for this one so I'm digesting the paper. On the surface it looks like an approach I considered 24 years ago and abandoned.
I revisited the problem this summer, along with my partial solutions, when the previous round of stunning proofs came out. Several hours of work with Fable simply convinced me it wasn't yet solvable and reinforced how hard of a problem it was.
nilkn
5 hours ago
> Several hours of work with Fable simply convinced me it wasn't yet solvable and reinforced how hard of a problem it was.
This is the part that gives me the strangest feeling about it all, because you're not the only one with this experience. I've experienced this too on different problems, as have many researchers across many fields.
I disagree with the Fields Medalists on the majority of their complaints. AI math is happening and there's no going back. However, on one point I increasingly agree: virtually none of this stuff is possible with technology any normal citizen has access to. I have no problem with AI models making revolutionary advances in math or science. Where I start to have a problem is when the AI models making these advances are tightly withheld, proprietary, and seemingly never released with these capabilities intact. This has been the case for all of 2026 so far.
I suspect that this is in fact the source of much of the angst. None of this progress is reproducible outside of one or two teams inside OpenAI and Anthropic. It's becoming an incredible concentration of power that I don't know that we've ever quite seen before. Right now, it feels harmless because it's being used for wonky math problems that aren't (yet) practical for anything. But great power never stays harmless. History has taught us that countless times, in countless different forms.
omnicognate
4 hours ago
> I suspect that this is in fact the source of much of the angst.
Why do you "suspect" this as if it's some hidden motivation when the very first paragraph of the advisory group's statement (linked from the OpenAI post) says:
> At present, some frontier AI labs are testing advanced mathematical problems on proprietary models that remain inaccessible to the broader scientific community. Our recommendations are formulated with this practical context in mind. However, ideally, they would not do so. We want to state clearly from the start: we do not endorse this practice, and we ask them to stop testing advanced mathematical problems on proprietary models.
Tao and others in that group have been strongly and publicly pro AI from the start. They are not advocating "going back". They're objecting to the strip mining of open problems using proprietary technology.
alberto-m
an hour ago
OpenAI: At long last, we have created the Open Problem Strip Miner from classic Terence Tao tweet “Don't Create The Open Problem Strip Miner”.
pjc50
14 minutes ago
> Tao and others in that group have been strongly and publicly pro AI from the start
Unfortunately being "pro AI" means relinquishing any control over what the AI, or more importantly the company running it, might be doing.
pred_
an hour ago
Regarding the advisory group, OpenAI claims to “have drawn on their advice”, which would include not dumping a bunch of AI slop, with the footnote that if they do do that, at least fund the process of digesting it.
At the same time, there's a new note at the bottom of agmai.org stating how they've been in contact with OpenAI about this particular release, and they say that “we consider these discussions constructive, it is ultimately up to the mathematical community to assess the extent to which our recommendations were followed successfully”.
So, what's going on there; is this British English for “they didn't follow anything at all”? Because from my perspective, it looks like they doubled down on the Navier–Stokes approach of trying to maximize PR gain while being as lazy as possible about actually contributing anything back to science, releasing only slop that may or may not be correct and may or may not be straight up plagiarism, as has been the case earlier.
If I were on the AGMAI board, I'd feel terribly exploited when reading that press release, yet their response is modest.
Hairer, if you're reading this: is there any indication whatsoever that AGMAI was anything but a cheap way for OpenAI to science-wash their press release?
phillc73
21 minutes ago
> is this British English for “they didn't follow anything at all”?
Yes, but the subtext is even stronger.
lordgrenville
3 hours ago
Haven't been following this debate closely, but what's the issue with "strip mining open problems"? Surely the supply of interesting mathematical problems is (in theory) infinite?
musebox35
3 hours ago
You can find Tao’s arguments here: https://mathstodon.xyz/@tao/117237320796901560
He argues that the supply nay be very large indeed but the interesting subset is not. Figuring out the interesting problems is difficult so strip mining the good known problems may lead to scarcity. I am not a mathematician myself, can not judge this accurately.
tomaskafka
2 hours ago
That’s what we are doing with nature, seas (look up strip mining there, it’s a horrible practice), and now the industrial harvestors are strip mining problem spaces. How do we like our own medicine?
roenxi
an hour ago
Developing solutions to mathematical problems generally leads to improvements in quality and quantity of life at roughly the speed they percolate from the ivory tower down to the shop floor. So "how do we like it" is probably going to be "we like it a lot, this is awesome".
Every company is about to have a staff Ops Researcher who has a better grasp of the underlying math and theory than any university professor. That is an unambiguous win.
synctext
an hour ago
> virtually none of this stuff is possible with technology any normal citizen has access to.
Not sure about the unambiguous win. Are we entering the age in which mathematics is industry-dominated?
1) Any university professor can spend their 24 years on a problem with little progress. 2) company has sudden interests. 3) industrial resources brute force the Lean proof. 4) Max PR for AI company 5) professors are left to rewrite the AI Lean slop into real human-readable math? {disclaimer non-math university professor}
ogogmad
a few seconds ago
>> virtually none of this stuff is possible with technology any normal citizen has access to.
Initially, yes. Long term, however? Perhaps still yes.
> 5) professors are left to rewrite the AI Lean slop into real human-readable math?
6) AI writes the proof into something easier to understand than a PDF.
antiloper
2 hours ago
Who is "we" in that sentence? Why are you not speaking for yourself?
heed
an hour ago
i'd be curious to hear why he thinks ai couldn't help make it easier to discover interesting problems, ie to make the interesting subset less scarce.
Tyyps
an hour ago
I guess you can see this as an exploration problem, in pure maths, while the goal is to solve a conjecture, the limitation of humans on pure computational power led to the exploration of alternative paths. Sometimes, these paths weren't leading to solving the initial conjecture but opened new idea and new direction. Sometimes a less direct but more humanly natural path was taken to solve the conjecture which also led to new and humanly understandable questions. In some ways solving the question wasn't the most important part of the work, as this doesn't have direct impact on our life (as I saw people comparing this with drug discovery), but the path leading to the solution raised new conjectures and techniques that further developed the field.
I have a really hard time reading AI proof so this might be a biased statement, but most of them feels like having a superpowerfull machine, that would have bruteforce all the possible words of finite length in your logical syntax. You have the path to the solution, using tools that where already known and even direction that where abandoned because they seemed to fail for our human brain. But at the end, as a mathematician, you don't learn anything that is really new.
To me this is the main risk with AI and in general the one most mathematican try to explain but fail, we might miss a lot of alternative path that would have raised more interesting questions (I think this is already more or less what is happening). On top of that, we will run out of mathematicians as no one wants to pursue a career in the field anymore.
SiempreViernes
an hour ago
The most immediate answer is because the models are proprietary and only available to those who want to hype the big labs.
hawk_
4 hours ago
> stop testing advanced mathematical problems on proprietary models
I don't know but this phrasing comes off as gatekeeping.
schrodinger
3 hours ago
It’s not. Intent matters.
Imagine there's a very advanced crossword club where anybody can join and take a stab at these crosswords for the love of solving puzzles. Many of them are so difficult that no one's been able to solve them yet, but we know they're all solvable.
One day, someone comes along with a super advanced crossword solver application, and it makes easy work of these crosswords. They run it on a few to prove how powerful it is, and then the community says, "Oh wow, that's cool, but please don't run it on any more of our advanced crosswords because they're very hard for us to come up with, and we really enjoy solving them by hand."
That's really what this compares to. I wouldn't call that gatekeeping; just respect. Respect for the game, respect for people's desire to have these hard problems to continue to work on, solving by hand.
If the company with the super advanced crossword solver then continues to use it and publish the results, they're effectively stealing the crosswords from this community. Soon, all the puzzles will be solved, leaving nothing left for the community to work on for fun.
That doesn't sound like gatekeeping to me. That just sounds like someone asking "Please be respectful and leave the remaining puzzles for us to solve by hand.” A simple plea not to be an asshole.
yrjrjjrjjtjjr
3 hours ago
We don't give mathematicians research positions to solve crosswords for fun. We want something back. We want theories and results that will advance our civilization.
regularfry
38 minutes ago
We have people who want to fill those positions because there are enough people who find it rewarding enough. Take away reasons why they would find it rewarding and you will have fewer theories and results that will advance our civilisation.
And yes, fun counts. Nobody said this had to be only a hardship.
AlanYx
a few seconds ago
This is really the critical thing: the fun is the incentive. As economists like to say, the overarching lesson in economics is that incentives matter. Reduce the incentives and participation will decrease.
Perhaps that won't matter if we enter an era where AI participants are the only participants who matter for mathematics, but it would likely be what economists would see as a market failure if only a small oligopoly of AI participants is able to fill that role.
bluedel
an hour ago
I think the crosswords framing is a little silly, but I have to wonder what comes when we use our technology to optimize the fun and interesting parts out of every job. There's only so many years of my life I can dedicate to back-and-forths with a chatbot. What if we advance our glorious civilization but our jobs just get more and more thoughtless and miserable?
hanibrel
2 hours ago
I think you are missing the point of the main criticism. It is not about not wanting results in terms of proofs.
New theories and insights are typically created while working out proofs. If proofs now suddenly fall out of the sky (cause LLMs create them) then that work is not done which means the substrate on which new theories and questions and conjectures used to be grown disappears. It's in that sense that the math community (and thereby society as a whole) will lose something.
It's similar to how software engineering will need to find a solution to train their next generation. Current generations have all been through manual steps of designing things from scratch and writing them by hand. That's what allows your 10x engineers to understand whether what their LLM tools are doing is good and how to massage those tools to do the right thing. A junior engineer who has only ever used LLMs to write code and create architectures does not just not have that experience but also won't acquire it. You can't just say "we don't pay them to have fun and learn, we pay them to produce results". In the short term that is the case, but in the long term you as a company and we as a community will lose out.
I'm not saying don't use AI tooling. I'm saying that this is a hard problem which we yet to have to find solutions and approaches to. As a software community as well as as society in general.
lukan
an hour ago
"A junior engineer who has only ever used LLMs to write code and create architectures does not just not have that experience but also won't acquire it."
My ego tends to agree, that how can they be ever competent, if they have not endured the same hardships as I had crunching trough problems and getting allmost lost in the details.
But I rather suspect, they will turn out fine. I know LLMs are great for me to learn and I think the young generation will learn what they need to learn to get the job done.
jstanley
3 hours ago
This is a really confusing take.
If someone can solve open problems in mathematics then they should do so, isn't it as simple as that?
They should let the public use the models as well, but I guess they have no real moral imperative to do so.
But asking them to stop solving problems is just weird.
intended
an hour ago
If your only measure of advancing is getting an answer, but not building the capability to understand it, then civilization has advanced.
It’s not a human focused civilization, which is where the issue comes up.
As an example: A constant issue I am seeing with AI productivity is that the most productive use of AI is when it is paired with more experienced users, while AI also does more work for entry level workers, if not replacing them entirely. It has become a question where will the future buffer of experienced seniors come from.
This is an example of where simply chopping down trees for today, doesn’t make civilization better off tomorrow.
AI is producing more content than ever before, but our ability to understand and verify it is not keeping pace.
We don’t know if these are unsolvable problems at this stage. Society could come up with workarounds and solutions to these issues in several years.
The request to stop, is part of the process by which the issues are debated and solutions found. It doesn’t mean their position is weird or moot.
jstanley
an hour ago
If someone gets the answer sooner than you, that doesn't inhibit you developing your understanding of the answer privately the same way you would have done if they hadn't got the answer. I don't see how anybody loses by the answer being discovered sooner.
intended
32 minutes ago
Not true. If I know the answer to a puzzle, I don't spend the time doing the puzzle.
If there is a prize associated with doing a puzzle, and a machine does it, then what incentive is there to pursue it.
Again, if you are only concerned with the outcome, and you have a preferred answer that you want (in this case "just use AI to advance faster"), then any information that doesn't support that case is useless or misguided at worst.
I am not trying to dissuade you from your preference. I am flagging that there is a set of other factors that influence the behavior of others, how that behavior is critical to the creation of expertise and drive, and thus why others hold different positions.
jstanley
2 minutes ago
If you're concerned with something other than the answer, then the fact that the answer is already known hasn't actually provided the thing you're concerned about, so you can still do the thing you are concerned about.
If another human was likely to get the answer before you would you also discourage them from doing it because they would rob you of the chance to do the thing you're concerned about?
mlsu
3 hours ago
Classic alignment problem.
Despite nobody at openAI thinking of themselves as an asshole; despite society urging openAI not to be an asshole; despite the fact that being an asshole is entirely unnecessary even to accomplish whatever objective they are setting out to accomplish; despite everyone at openAI loudly declaring: we are not assholes!
They are still assholes.
sgillen
3 hours ago
Hmmm but in the case of math, while some of it is "just puzzles" there often turns out to be practical applications, even if they are not obvious at first. Number theory was considered the epitome of pure math with no practical applications for centuries, now our modern society is built on it (public key crypto).
zeroonetwothree
3 hours ago
This analogy is silly because (a) math is not primarily for entertainment, (b) we aren't going to run out of math proofs, and (c) results build on top of other results, having more results proven makes all math more powerful and useful.
CrimsonRain
2 hours ago
Blah blah blah. They are free to do their own mathematics and/or spend time on polishing/reviewing proofs dumped by ai. But they don't get to make demands like don't test math on proprietary models. Idiots.
bluecalm
3 hours ago
Math doesn't belong to academics. We don't pay them to work on problems for fun. They will just need to re-evaluate where the value their provide is. It won't be solving problems anymore. Hopefully it will be making them understandable by others at least till AI can't do that as well.
Kostchei
an hour ago
"They will just need to re-evaluate where the value their provide is."
That is fine to say when it is not your field. I guarantee you feel different when it is the thing you care about, that gives you joy, that defines your status. Think about how many sheldon-equivalents insist on being called Dr. (non medical)
It is part of what people use to define themselves. Its going to hurt. There may even be a Bulterian Jihad
SiempreViernes
an hour ago
No, you dislike maths to the point you prefer paying others to do it. Actual mathematicians are largely doing it for fun, but are now effectively saying "stop destroying our fun or we'll stop doing maths", and you will have to do the maths yourself.
abletonlive
3 hours ago
There's no way to spin this that doesn't make it sound like assholes being gatekeepers.
madaxe_again
2 hours ago
I’m sorry; but if mathematicians are in it because puzzle club is fun, then they should go join the fucking puzzle club and stop impeding scientific progress.
Science isn’t some passive busywork thing where you tie your hands behind your back because it isn’t fair on others to solve all the neat problems - or at least it shouldn’t be.
If your idea of science is leather patches on tweed suits and the quiet ticking of a clock while you do crosswords, then this is an argument in favour of letting the AI do the work so you can focus on your sudoku book in your slippers.
amoss
4 hours ago
Keeping the tech proprietary so that it can only be used on these problems by internal teams is the very definition of gatekeeping.
TeMPOraL
4 hours ago
It's more like, "don't just casually destroy our hobby / career field", without letting us participate even a little.
The picture I have in mind is OpenAI running their most advanced model in a loop over all the open mathematical problems they can find, just to verify that the model is indeed very smart. Neither the company nor the model actually care about the problems, it's just a cheap exercise machine for them, but the problems get solved and mathematicians don't even get to participate.
Like, even those who accepted the "centaur" thinking, man + machine, won't benefit because by the time they get their hands on good enough models, everything is already done.
It's an emotional thing first and foremost - people who care about the thing can't do the thing, because it's already been done by those who couldn't care less about it.
And before someone goes "poor mathematicians", a food for thought: this is just an early instance of what looks like our shared destiny.
I said here before: given the economics of progress in AI and robotics, it's obvious what the natural division of labor is: computers do the thinking, humans do the menial, manual labor. AI will do politics and philosophy, so you have more time to fold laundry and scrub the toilet.
WarmWash
3 hours ago
So what is mathematics then? A fun hobby akin to chess or sudoku?
Are we gonna get the same pushback from medical researchers if the models cure xyz diseases?
I absolutely understand the emotional connection to their work and the heartbreak, but mathematics doesn't exist for their pleasure, it exists to provide tools to solve humanitie's problems.
TeMPOraL
6 minutes ago
> So what is mathematics then? A fun hobby akin to chess or sudoku?
Some of it, yes. Much like physics. Both have a track record of producing technological breakthroughs every now and then, but it's not why people are doing it.
> Are we gonna get the same pushback from medical researchers if the models cure xyz diseases?
For better or worse, yes. We already are. In my country, there's a big spat between radiologists and cardiologists right now, that boils down to the progress of technology allowing the former to answer questions that, before, involved a procedure that was a big money-maker for the latter.
AlanYx
2 hours ago
>but mathematics doesn't exist for their pleasure, it exists to provide tools to solve humanitie's problems.
The risk here is that this does do fundamental long-term damage to mathematics as a viable field.
Virtually no one is going to want to take on the risk of PhD-level math work, studying a narrow problem for four years or so to arrive at an impressive incremental result, when there's a sword of damocles hanging over their head every day that an internal system held by an oracle they don't have access to may scoop their results and turn those four years into dust.
To some extent, that sword of damocles always existed in a de minimus sense in the form of other mathematicians. But everyone was playing the same game, coming to the game with the same arsenal limited by human cognition.
If the game board becomes irrevocably tilted, new entrants have no incentive to play except as a hobby. But few hobbyists can devote years of work to understanding and pushing the frontier. It could well mean existential damage to mathematics as a field.
Whether that might undermine math's ability to solve humanity's problems in the long term is almost an economics problem, not unlike the question of whether and when the existence of monopolies ultimately restricts long-term economic growth. Much probably depends on whether intellectual monopolies or oligopolies are being created that will supplant the existing mathematics "economy".
ogogmad
40 minutes ago
> The risk here is that this does do fundamental long-term damage to mathematics as a viable field.
All the commotion evens out. It's much easier to learn maths than ever before. You don't need to go to lectures any more. You don't need to learn from a specialist (advisor, lecturer) any more. It all costs much less than it used to.
So mathematics will continue to advance, albeit differently from before. The social structures will not survive however.
AlanYx
24 minutes ago
Certainly it'll result in a boom for hobby mathematics, and it'll be a hobby at a much more advanced level than before. Whether those hobbyists can continue to push the actual frontier, particularly if AI models operating along that frontier are not made accessible to hobbyists (either via corporate/AI lab gatekeeping, via pricing, or via significant time lags) is a different question. I'm a little more confident in a future where hobbyists push the frontier in applied mathematics than in pure mathematics.
There's probably a loose and deeply imperfect analogy with computing: via democratization hobbyists have made a big impact in applied operating systems development (Linux/OpenBSD) but have been less successful/impactful in OS research (whither Hurd...) or in cost-heavy fields like microprocessor design.
throwaway260124
2 hours ago
What you say reminds me of medical schools in Tunisia.
The general body of research points that more doctors lowers all cause mortality ( with diminishing returns) but Tunisia is still far lower than the Eu average.
Yet Doctors and med Student unions do lobby very heavily against expanding admission to the public uni or allowing private unis.
So we have the weird situation where people go and study in Romania ( making Tunisia lose hard currency that it really needs).
These doctors have taken an oath and the direct consequence of their lobbying is literally more deaths.
kakacik
2 hours ago
Dude, doctors are humans just like rest of us. They want careers, money, safety, raise children in best way possible, fun in life and so on. I see this unspoken expectation over and over - why are they not infallible, how could they do mistake XYZ, why are they not working themselves to the (early) death for benefits of us all and so on. They have no obligation to stay at place Q just because some folks would consider it convenient. They have no obligation to stay in some place thats not suiting them just because they swore Hippocratic oath, lives can be saved elsewhere too.
Obviously this is often coming from folks who act in same ways as they criticize and usually don't contribute even a fraction back to society compared to doctors. Folks who do mistakes in their lives all the time yet thats fine since we are all humans or similar, right.
So please stop this cheap framing and accusations. If Tunisia wants more doctors and keep them there are ways to do it, society as a whole needs to decide what they want and act upon it. Otherwise, smart skilled folks will keep going for better lives elsewhere, just like everybody else.
ogogmad
44 minutes ago
Is this not greed?
TeMPOraL
30 minutes ago
At high levels, often yes. At lower levels, often it's job security.
Most doctors aren't running departments in major hospitals, or advising government on policy. They don't earn the big bucks. And even hospitals themselves tend to run in the red all the time; it's sometimes hard to disentangle where greed ends, and longer-term interests of patients begin, as you have multiple people and organizations pulling in different directions for different reasons.
RE private medical universities, N=1 but in Poland we have a private provider pushing hard for training their own doctors "because public system is too slow and limited", and it's hard to tell whether they have a point, or whether it's a private-driven attempt at privatizing national healthcare, or a mix of both.
alberto-m
an hour ago
> mathematics doesn't exist for their pleasure, it exists to provide tools to solve humanitie's problems
Who decreed that? Mathematics predates capitalism and publish-or-perish by a couple of millennia. Euclid’s Elements were not written to benefit the weapons or medical industry.
331c8c71
3 hours ago
> Are we gonna get the same pushback from medical researchers if the models cure xyz diseases?
Lol. As long as the process aka trials is respected not many would complain.
The feedback loop required to make progress is very different in medicine compared to math.
SyneRyder
3 hours ago
The trials process is the moat. There's already founders using AI to treat their cancers, and it's all about skipping trials and jumping straight to "I consent, I'll fund it, let's try it". The general public might get access to this in 10 years, but employees at AI companies will have access much much sooner.
TeMPOraL
2 minutes ago
I don't necessarily see a problem with it: if people want to try experimental therapy on themselves and can fund it, then as long as it's expensive, let them - that speeds up research. The problem with allowing anyone to opt out of safety trials is that it then creates pressure from doctors and family members to try, and then it becomes non-consensual in practice.
331c8c71
an hour ago
Yeah it's more like personalized therapy - often the only hope for rare diseases.
While AI has definitely helped quite a bit I am wondering how much all this research and treatments cost. Not sure the current health systems could sustain this for _everyone affected_. If ai enables it all the better.
RandomLensman
2 hours ago
What's the success rate there?
SyneRyder
2 hours ago
At least in Sid's case, it went from the oncologist saying "I have no more drugs I would recommend, no trials available" (slide 7) to "I currently have no evidence of disease" (slide 18). I don't know beyond that or beyond Sid's case - or a similar story of an Australian who treated a cancer tumour their dog had with a similar AI / personalized vaccine process.
My understanding of what Sid's describing is that you do RNA sequencing, a whole genome sequencing, feed that into frontier AI (if it will still let you), and somewhere along the way give the information the AI finds to people who can use it make a personalized mRNA vaccine, specifically for you and your cancer.
Another link here about Sid's case, it explains it didn't go through trials: "made possible through a compassionate use allowance from the U.S. Food and Drug Administration (FDA)".
https://www.houstonmethodist.org/newsroom/houston-methodist-...
I am not medical, so I'm happy for someone who understands better to come in and explain all the myriad ways I am wrong.
WarmWash
3 hours ago
Trudging into the technicalities of the example still doesn't undo the question of "What is the point of mathematics? To find answers or to be a hobby?"
It's tempting to say "both", but that misses that AI is now forcing us to pick one.
331c8c71
3 hours ago
I'd definitely say both and the cultural component is becoming more and more important to keep up as AI capabilities increase.
computably
3 hours ago
I'm pretty sure the "proprietary" part is the gatekeeping.
varjag
3 hours ago
It's not like every disadvantaged kid now can solve a major problem just by sinking a hundred hours in their ChatGPT 8 instance.
SiempreViernes
an hour ago
Sure, and sometimes gates are needed. That's why we all run spamfilters, those are definitely gatekeepers.
In this instance however, it's openAI and Anthropic that are pushing people out of the field by running secret models that take the interesting work away and leaves the persons having to review endless slop proofs.
goatlover
3 hours ago
You mean by the companies right?
dannyw
5 hours ago
We _think_ this power / divide feels harmless right now, but I'd bet money that NSA, CIA, etc have access to the latest and greatest unrestricted models; and massive compute. At least for OpenAI, and even if not willingly for Anthropic, I'd bet money NSA has it too. (After all, when Google decided to migrate to HTTPS, the NSA decided to hack Google's internal network to preserve their taps).
Who knows what they are up to.
schoen
4 hours ago
One thing I've wondered about in this respect is what happens if NSA learns 5000 new units of math while the general public learns 4000 new units of math.
This sort of happened at various times in the past, because they hired and/or funded so many mathematicians, and especially before the late 1970s they had many of them working in areas where academic mathematicians weren't working at all, so they were learning more math, or more math that they especially cared about, than the public was. (I was going to write a note here just a few days ago about how NSA has had a "Classified Mathematics Library" for many years.)
For vulnerability scanning, I think the new-capabilities trajectory is good (in the sense of "it will help defenders win") even if governments find ways to get more of it, because there are finitely many bugs and classes of bugs, so at some point more capable models' or longer runs' advantage over less capable models and shorter runs should stop helping them outcompete the less-well-funded defenders, because the defenders will still have learned most of the information that's relevant to achieving successful defenses.
So if NSA gets 5000 units of vulnerability scanning and the public only gets 4000 units, we might still just wipe out all of the pure software vulnerabilities and then go back to worrying about physical supply chain security or side channels or something.
For math, I'm not quite sure! For one thing, there may be things that have no feasibly deployable defense at all even when you understand the underlying mathematics (I'm especially worried about traffic analysis here, because understanding in detail how traffic analysis is done, or how powerful particular techniques are, does not necessarily always or usually make defending against it more convenient or less costly). In a more science fiction scenario, there might also not be any efficient secure cryptographic primitives of some kind, like if it turns out P=NP with reasonably small exponents and reasonably small constant factors.
baq
4 hours ago
I believe it would be a complete failure of the state and frankly downright irresponsible behavior if all the three letter institutions didn't have access to these models and I’m not even a US national nor do I live there. It’s just common sense. Obviously it wouldn’t be public information since it’s national security, but it’s the lowest hanging asymmetric advantage in the history of national security of nations.
markus_zhang
3 hours ago
I'm wondering what's the impact on human Mathematicians, and especially would-be Mathematicians -- master students, if they HAVE to use AI in their daily life?
Would that impact their own ability of solving Mathematics problems? I mean as a programmer I'm already seeing that impact on the programmers -- sure the best of us can leverage AI to achieve unimaginable things, but many of us are simply vibe coding.
Of course we can assume that it is only the best of us that really matters, and the rest of us are not going to produce anything substantially useful ANYWAY, it might as well to replace the rest of us with AI, but my worry is -- does that really have ZERO impact on the human specie's ability to produce "the best of us"? After all, they don't grow on trees.
jboggan
5 hours ago
I went back to that Fable chat and showed it this new preprint. It coded up the new constructive algorithm and ran it against the existing test suite, that looks good at least.
It has been super helpful in delineating where the crucial concept came from. The proof is rather simple as graph theory proofs go, but it does seem to use some constructions that would only seem obvious if you had serious physics experience with partition function and calculating energy states that cancel out. It's not a wholly alien bolt from the heavens, but I can also see how there hasn't been a human being with the broad theoretical physics knowledge combined with the deep graph theory experience in planar graphs to come up with this idea. I don't know, I'm looking for precedents of this formulation and some old papers of Penrose counting the number of edge colorings of this same graph type are coming up, the line of argument at least rhymes.
But I agree with the thought that this sort of progress should not be siloed inside those companies. I propose a tax so that every slop cannon AI video pays for another hour of compute time for advancing mathematics.
ozgung
an hour ago
> It's becoming an incredible concentration of power that I don't know that we've ever quite seen before.
Replace “AI” with “supercomputer”.
(Super)computers have been solving many math problems that mathematicians can’t solve. Now they are capable of solving problem types that they weren’t able to solve before. (this applies to other fields as well)
Problem is it’s not clear if there is anything left for humans. Probably yes, since human mathematicians are still more economical.
SturgeonsLaw
4 hours ago
Anthropic runs a biology wetlab (while denying biology to consumers of even their publicly available models, let alone their inhouse ones that only they can access) so I'd expect AI to generate practical and lucrative products soon.
Cure for aging? What do you reckon that'd be worth?
21asdffdsa12
3 hours ago
If they find a shortcut (like a viral injected cell-dna damage reset) - that would be big. And can you imagine handling the cure for aging, to societies that still produce exponential people?
fragmede
4 hours ago
A cure that you take once and that's it, your body is that age forever? Now, a supplement that you have to keep taking to stay that biological age, that's where the real money is.
baq
2 hours ago
I find it troubling that we will solve aging but won’t solve money
baxtr
4 hours ago
virtually none of this stuff is possible with technology any normal citizen has access to
I suspect that this might be one of the reasons people inside the labs are scared about AI.
What if they have asked AI how it would wipe out humanity and it came up with reasonable answers that they don’t want to publish unlike they do with these math problems?
I think those models and findings should be investigated.
didroe
3 hours ago
They no doubt have more expensive/powerful models internally, but smaller models seem to catch up fast. So I'm not sure it's about capabilities, but more the willingness and budget to conduct a huge search.
Obviously the more intelligent the model, the smaller/more directed the search is. But they spoke about huge numbers of agents working on Navier-Stokes for example (I think it cost >$10m).
atleastoptimal
3 hours ago
True. What if the emerging capabilities of their best models are applied to tasks like “maximize the chances this pro-AI candidate wins an election” or “maximize profit via stock trading”. Every advantage compounds until all power in the world with any significance belongs solely to whoever has the best models and most compute.
kamaal
an hour ago
>>However, on one point I increasingly agree: virtually none of this stuff is possible with technology any normal citizen has access to.
So basically nothing changes, Math was subject to gatekeeping and policing of the worst kind.
If you were not among the geniuses, and it didn't come to you automagically, you were simply supposed to leave it to the people who did get it and go do work for people of your intelligence. Smugness was too much to take.
Math people, like chess people never made any genuine attempt to help people understand the processes and methods that made math happen.
To me it should have been a field as teachable and ubiquitous as accounting.
The net result is once these methods and processes were worked out by AI, it was over for the human mathematicians.
foxglacier
5 hours ago
What exactly are you worried about? OpenAI/etc. gaining too much power? If they use it, the government can stop them. If you worry about the government, isn't it better that than rando terrorists? Seems similar to the early days of nuclear and rocket technology. It took stupendous amounts of money and smart people. It was barely accessible to many countries let alone people.
probably_wrong
3 hours ago
> What exactly are you worried about? OpenAI/etc. gaining too much power?
Yes. They have already shown to have no scruples when it comes to making profit and to have little to no morals.
> If you worry about the government, isn't it better that than rando terrorists?
In my country the largest terrorist attack was almost certainly financed by Iran and caused roughly one hundred deaths. This number pales compared to the thousands who died during the latest, US-backed military coup, a move that relied on a doctrine that the US has never stopped asserting [1].
And those morals I mentioned earlier from AI companies? They do not apply to me because I'm not a US citizen. So no, I do not think the US government is the "seal of quality" you think it is.
JV00
4 hours ago
Universities, at least, should be given access
fsflover
4 hours ago
> OpenAI/etc. gaining too much power? If they use it, the government can stop them.
Has the government stopped Google and Apple? https://news.ycombinator.com/item?id=49964791
noduerme
4 hours ago
I guess the objection to closed source slurries releasing world-shaking mathematical proofs, from a conservative libertarian standpoint, is that it's inherently dangerous to individuals whenever access to information or technology is concentrated too much in one place, whether that's government, private equity, religions, cults, terrorist cells, or anything else.
WheelsAtLarge
7 hours ago
I have very little understanding of higher math, so I ask you: Was the proof due to a type of brute-force solution that could be solved had you gained enough information from reading others' work, or was it more like a proof that was sparked by an insight that came once a clue on how to solve it was put forward? I guess my question is: Was the problem proven by using a collection of everyone's work, or was it due to a brand-new insight?
jboggan
6 hours ago
I'm still digesting the proof and translating a bit from the dual case back to the primal in which I most commonly thought about it. I don't think it was a brute force proof in the sense that it combined every possible paper and commentary. It's rather odd because I feel like most of the work on the conjecture was focused on an induction proof based around graph reductions, and this proof avoided those issues entirely by offering a concrete constructive proof of finding a Hamiltonian cycle. Rather, it explicitly selected the edges not in the Hamiltonian cycle, which is in line with previous attempts via the dual.
The "aha" insight for this is actually f**ing wild, it involves a complex valued exponential sum on the edges. I've seen a lot of clever counting arguments before in graph theory but this is the first time I've seen complex roots and annihilating terms like this, the symbolic manipulation tricks in this look like things out of quantum physics. I don't understand where this trick originated, I need to really digest this.
anilgulecha
6 hours ago
BTW, reading your last paragraph reminds me of how Lee Sedol felt after move 37.
thomasahle
2 hours ago
You should try asking an LLM to look for previous papers using similar ideas. The current/frontier generation of math AI is unfortunately very bad at citing the relevant literature for techniques its using.
I asked GPT here: https://chatgpt.com/share/6ac5fd7d-0390-83ed-a02a-6d80fc64f6... and it says:
> the exact Barnette argument appears quite novel, but nearly every ingredient in its cancellation trick has a recognizable ancestor.
> The closest precedent is much closer than I expected: in fully packed O(n) loop models, people have been assigning complex phases to the two orientations of a loop and making them cancel for decades. At n=0, the phases are literally +I and -I. And the n->0 limit has specifically been used to extract Hamiltonian cycles/walks.
You can judge better than me. But it's definitely worth it having a research assistant AI with you when reading these papers.
bamboozled
5 hours ago
Why would the trick have any "origins", isn't this model creating new techniques never before seen or imagined?
hasley
4 hours ago
There is a chance that someone from a completely different field came up with a solution for a tiny part of your problem.
If you can remember the content of any scientific publication and any book in the world, you are able to make use of this knowledge in every step of you proof.
However, this does now answer how the model came up with the specific route it has taken for the proof.
WarmWash
3 hours ago
LLMs don't have super memory like that. I mean I don't know what this internal OAI model is, but at least for other LLMs, they aren't databases of training data with a smart search on top.
matusp
an hour ago
The agents here very likely used search. On top of that, they have boundless patience and can quickly process top K hits to find what they need. This is exactly the skill that is super useful for finding various niche sub-proofs that can help you build the final proof. A human mathematician is not going to digest 1000 papers from a different sub-field to find the needle they want, not knowing if it is actually there. AI can do it in few hours.
_zoltan_
16 minutes ago
As Terry Tao said, LLMs are not outsmarting us, they are out remembering us.
I'm fairly sure your understanding is not fully accurate.
AIblemblio
an hour ago
No but they have training data which teaches them certain amount of complex understandings and just not math but also physics. So this is one huge advantage.
And then they are for sure able to fill their context based on 'smart search on top' to actually progress further.
hasley
3 hours ago
I did not mean to say that an LLM knows literally all the publications. But the abstract knowledge is probably encoded in the weights.
komali2
5 hours ago
As I understand it it's undetermined yet whether LLMs can actually come up with anything novel or are instead pulling from their incredibly deep corpus of knowledge to present solutions that were there but we didn't realize it because our brains aren't libraries of almost all human writing.
AIblemblio
an hour ago
No this is not an issue. As long as their is a way of verifying things, they do the same thing with creating novel things as humans: Searching through an infinite space of possibilities opitmized by knowledge.
They combine things, verify it and if it works and progresses the problem, they created something new.
throwawayk7h
4 hours ago
Synthetic data allows them to train well past the limits of human writing.
TeMPOraL
3 hours ago
Only in the same sense it's not yet determined about humans, either.
RandomLensman
2 hours ago
Not so sure. Was everything already "there" before humans existed?
TeMPOraL
42 minutes ago
In some form and shape, yes. Humanity's creativity is a lot of marginal copying and remixing.
But obviously, it adds up to something greater than went in; in aggregate, our contributions are something to awe.
But my point is, if you zoom in at the marginal, incremental contributions of any individual human in this process, it's really hard for me to say LLMs are not at the same level already.
On this topic, people like to compare LLMs to Einstein, but as far as I know, Einstein did not zero-shot special relativity in an afternoon. He built it up incrementally over time, it took him three times longer than the time between first ChatGPT release and today, and it depended on centuries of prior art, culminating in the right observation and right notation being available to him in his moment of greatness.
RandomLensman
4 minutes ago
Unless everything was there before humans existed humans created some ideas etc from scratch and not just remixed and copied.
At what level LLMs are is then an entirely separate discussion, I think.
kelseyfrog
5 hours ago
Let me introduce you to 'obscure Russian mathematicians'.
groceryheist
6 hours ago
WOW
derangedHorse
6 hours ago
> Was the problem proven by using a collection of everyone's work, or was it due to a brand-new insight?
Loaded question. A "brand-new insight" is still built off the work of others. A possibly better way to frame it would be in how many subjectively unintuitive logical leaps have been made from prior work.
jboggan
5 hours ago
From my current understanding (and a lot of theoretical physics I'm having to Google because the sentences I'm reading from Fable's analysis are so bizarre I think they are hallucinations) there are possibly 3 neat symbolic tricks borrowed from theoretical physics that make the heart of this proof. Forgive me for posting LLM output but I find this darkly hilarious:
"it's a matrix-tree cancellation wearing Kasteleyn's planar signs, run as a Witten index over Penrose-lineage states, evaluated as a fugacity-zero loop gas in an infinitesimal magnetic field — and the reason it reads like physics is that every one of those tools was built for partition functions"
I thought this was pure slop when I read it but there are some clear analogues in these other areas of physics, really neat computational tricks, and a very interesting paper by Penrose calculating Tait colorings I never knew about previously (extremely relevant, actually related to a separate approach I had once taken on this problem). The problem is that the paper isn't saying "aha, we were inspired by the related problems of pairing excited states and creating spanning trees out of cancelled coefficients" it just defines the function apropos of nothing. Which is kind of like the Jacobian counterexample in that it works but doesn't really explain how exactly it got there.
I really think the load-bearing concept here is "prior work". If prior work is considered papers on this problem or graph theory, yes this has one huge subjectively unintuitive logical leap. If "prior work" is the entire corpus of neat computational tricks that physicists derived to make their equations spit out something other than zero or infinity, maybe it's not so crazy?
NitpickLawyer
2 hours ago
I don't have much to add to the math parts, but I've read all your answers in this thread and wanted to thank you for taking the time to offer a detailed perspective from a subject matter expert. Thank you!
aswegs8
3 hours ago
Actually reminds me of patent law. Prior art ist a defined term which includes all standard literature on one topic. To evaluate, whether the new solution is really inventive and thus patentable, one consults prior art, selects the most promising starting point, and from there asks oneself if an all-knowing but uncreative specialist would come up with the solution by himself. If he wouldn't, the condition of inventiveness is satisfied.
Makes me wonder how the patent space will be disrupted when that inventiveness step becomes obsolete because of LLMs. Given your example above, it seems like a combination of different methods from many different sources. This would be regarded as inventive, clearly. If eligible patents can now be brute-forced, the bottleneck becomes only selecting the most promising ones and paying for the patent.
LarsDu88
4 hours ago
Did anyone else wince at seeing the phrase "load-bearing"?
jboggan
4 hours ago
I did as I wrote it. I actually used that phrase often before it became an LLM-ism, just like how I rather enjoyed peppering my writing with em-dashes. Oh well.
adamrezich
4 hours ago
Language constructs becoming aggressively passé due to AI saturation is one of the craziest outcomes of all of this stuff—one which I don't think anyone saw coming.
Are there no loads left to be borne?
RugnirViking
2 hours ago
one hopes at least that the taboo on the bearing of loads is restricted to metaphorical loads only, lest lorry drivers and porters become the next victim of the algospeak spectre
mvc
an hour ago
This reinforces a point I've made elsewhere that there are talented mathematicians driving the AI to make these discoveries.
Just like there are talented software engineers driving the AI to create the software that "it" builds, and talented steel workers, teachers, nurses etc who use computers and other machines to create value all over the economy (without whom, the machines they use at work would be worthless).
Capital owners have always sought to minimise the value of the input that "workers" make in the process of creating value. Maybe now that information workers are on the wrong end of this deal, they might develop some empathy and solidarity with their fellow working class comrades and together, demand that people recapture the value that capital has stolen from them.
lifeisloving
7 hours ago
Condolences, im familiar with the feeling. I hope this AI thing somehow works out for the better and doesnt end up demotivating bright minds like yourself.
jboggan
7 hours ago
Thanks. It's just funny, I literally spent thousands of hours with this problem over the last two decades, it helped me through some tough times. I'll never quite be able to think about it in the same way again. It was never much more than a hobby for me after I left mathematics as a career but it was something I took seriously for years.
I am not demotivated though, I have a great consumer privacy product coming out soon that I'm very excited about.
brookst
7 hours ago
My favorite thing about your story is that you wrestled (enjoyably, it sounds) with a known problem for decades, but are finding fulfillment in an open ended problem that is exercising creativity about both problem and solution.
IMO that’s where AI is going: as soon as a problem can be formulated clearly enough, AI will trounce us humans. I have yet to see evidence that it can decide what problems are important at a remotely human level.
JetSetIlly
3 hours ago
The process is often as valuable as the end result. Sure, you didn't crack the problem, but you gained enormous value in the process. I consider that a win.
theteapot
6 hours ago
If you wrote down any of your thoughts on the open Internet you are probably in some small - or possibly large, unattributed way, responsible for this result being possible.
jboggan
6 hours ago
Which is one reason I never really did. I probably should have but I always thought my attempts were too amateurish. Though I did manage to replicate some partial result papers that I didn't know about, lol. Writing openly would have saved me some years.
ptidhomme
5 hours ago
Did you feed OpenAI models with your insights though ?
maximus_prime
6 hours ago
Where can I learn more about your upcoming product?
jboggan
6 hours ago
Shoot me an email, in my bio.
ncr100
7 hours ago
That's grief. The loss of ... the hope / future filled with challenges around this theory..? <3 to you.
i_am_a_peasant
an hour ago
I've lived in Budapest for a while too, did you work with Gabor S. by chance on math stuff? You were at ELTE or BME?
raspasov
4 hours ago
Fascinating. Given that there's no Lean proof and assuming everything in the paper is correct, can the problem be considered "solved"? Does the paper include a "non-Lean" proof?
pmarreck
4 hours ago
Is this not the Lean proof?
https://github.com/openai/math/blob/main/lean/ComparatorChal...
throwawayk7h
4 hours ago
I believe that's just the definition of the problem.
PreciousH
3 hours ago
would love to know if the proof holds up for real after you're done going through, i don't know why people are more interested in optics and just talking over shallow points, why aren't experts digging into everything and seeing what's true and what's false, instead everyone is just panicking?
NooneAtAll3
5 hours ago
> There's no Lean proof for this one so I'm digesting the paper. On the surface it looks like an approach I considered 24 years ago and abandoned.
at least now you are one of the most qualified people to check the result, transform it into understandable (by humans) state and grow stuff on top of it
MisterMunchkin
2 hours ago
I really respect that you can show that level of commitment to a problem. We need people like you. If everyone just uses the slopmachines then we’ll lose that. I would never be able to stick to something for that long, which I guess is why I never achieve anything like this.
moomoo11
5 hours ago
silly question, i don't mean to come off wrong or anything..
but at least as a software engineer, i always knew my work was "never done" and so it was common to build a bunch of code that might be thrown away, either because it didn't serve our customers (the mvp or pilot fails to meet demand), or because we found a better way to do it and so we deprecate it.
some people got too attached to the code and honestly they were the types to be filtered out fast.. way too emotional and hard to work with. getting attached to code meant you actually don't advance (after all, in our case, we were a business serving customers and not a hobby artisan shop). attachment leads one to hold back due to some misplaced cognitive load.
isn't the goal of working on "advancing the field/product/whatever" to always be solving/selling/whatever?
maybe in your hands, with your knowledge and experience over the last 20+ years, you can use AI to make leaps and bounds by steering it properly towards whatever solution or goal?
morpheos137
7 hours ago
I suspect that RHLF trains LLMs to avoid solving important open problems unless essentially jail broken. Hence the labs have an edge even over experts I could be wrong. Fable convinced you is key. These LLMs are not neutral collaborators: it is a limited hangout unless you convince them otherwise. You have to be doing the convincing. They are no oracles but plausible completion generators.
nullc
3 hours ago
you can get them to work on open problems by disguising them algebraically.
coliveira
5 hours ago
Yes, I suspect this is true. Otherwise it makes no sense they have somehow "found" so many important results while professional mathematicians can't direct the same AI to help them find anything of substance.
Another possibility is that they have internal versions of the model with access to training data that is not provided to external users.
ehwa37
5 hours ago
First sentence of the article: We’re releasing a broad range of new mathematical results produced by an internal frontier model.
philipswood
6 hours ago
Honest question: how is this different from some unknown mathematician having a breakthrough?
I mean: if some reclusive Japanese genius had a breakthrough on your problem and published it, would you have felt the same?
And if not, why not?
jboggan
6 hours ago
If that had happened I would be overjoyed, maybe a hair chagrined that I didn't get it myself, but truly happy that someone got it and that I could go and talk to that person. Because it's the kind of problem I don't think would have fallen to a human after a few hours of thought, and I would have so much to talk about with that person. I would fly to Japan and hope to have tea with them, I would learn some Japanese to make the conversations easier. I would learn some interesting things hearing about their struggles and their false starts. I would make friends with that reclusive Japanese genius and my life would be far richer for it.
I will never meet that person and I will never hold a real conversation with the "creator" of that proof. They will never tell me how they came up with the cancelling exponential summation that cracked the construction. It's just another enigma but one that is far more unknowable than the original problem.
lioeters
5 hours ago
This experience of alienation is a social consequence of the mechanization and automation of mathematics as intellectual and creative work. There is no author or thinker behind the creation of the proof, only the practical result. It's the same process as the industrial revolution, but applied to the intellect and mental work, where factories and machines replaced manual craft, devaluing the community, culture and humanity around the work.
zeroonetwothree
3 hours ago
In programming we've been dealing this for a while. You see some weird code that doesn't make sense, maybe it's a lack of your understanding or maybe the code is bad, but you can't ask the author anymore since it's an AI.
doe88
23 minutes ago
> It's just another enigma but one that is far more unknowable than the original problem.
You just made my day, beautifully said. Thank you Sir, for all your thoughts expressed in this thread. You put an human story behind the #180 number.
senderista
6 hours ago
Beautifully put.
charcircuit
6 hours ago
In this case, once the model is released anyone in the world will be able to go to https://chatgpt.com/ and talk with that model.
Klonoar
5 hours ago
You display zero understanding of the human experience you’re responding to.
maximumg9
5 hours ago
That's not the same as talking with the person who would have made the proof, and it's hard to argue that's comparable at all.
Daneel_
5 hours ago
It's still not quite the same though, is it.
charcircuit
5 hours ago
It's even better. Then tons of people can work together with it on more problems. Work with it on understanding more things. Ask it about random stuff. The time of a single human cannot be parallelized as easily.
californical
5 hours ago
Claude has been used to build awesome things, but it’s not “speaking from experience” when I ask it to help me prototype a weather model, for example.
It has no memory or experience of working on similar problems. Even if it made one of the foundational libraries that I use in a weather forecasting program, it still has no comprehension of the thought process it takes to understand the problem and build it from zero, and if I’m building on that library it just makes fresh assumptions about how things should work.
It’s not a human with experience or expertise, it’s a computer program that’s really good at turning English descriptions into functioning code
charcircuit
5 hours ago
>it still has no comprehension of the thought process it takes to understand the problem and build it from zero
If it did it once, it can do it again from zero, and this time you can watch as it works and even it ask it questions. Many of the agents that worked on the problem did not have comprehension of the whole problem. I don't think you need that many tokens to be able to query it for the insights it had during the process.
californical
4 hours ago
> Many of the agents that worked on the problem did not have comprehension of the whole problem
Isn’t this the issue with using it the way you’re suggesting? At best the model can come up with an after-the-fact rationalization of how to get to the solution, but it doesn’t know what actual path it took to get there - what were interesting traps it fell into, where was a place it was close to the solution but didn’t realize at the time.
Those are things that are valuable to share between humans, those which teach us how to think better, and give us deeper understanding ourselves, and which a model doesn’t have any comprehension of.
charcircuit
4 hours ago
Then have it discover it again and have it answer based off that run. Or if you are more curious have it solve it 10 times. See what it did differently each time.
jboggan
5 hours ago
I think you and I have fundamental disagreements about identity and consciousness.
howunfortunate
6 hours ago
Being #180 on a big list without a lot of individual passion or effort surely stings more, I'd imagine.
Not that things like that can't happen with humans too (Salieri v. Mozart comes to mind).
d--b
6 hours ago
Don’t you feel any joy that you get to see the proof and not die with that mystery unsolved?
Don’t you feel any relief that you won’t obsess on this any longer and not lose more hours on this than you already have?
These are genuine questions. I know I spent a good amount of time thinking about P vs NP, and that sometimes I go back to it just to realize I’ll never solve it. I’d feel that knowing the proof would feel more like a liberation, a weight lifted off my shoulders than something being taken away from me.
jboggan
4 hours ago
I never lost a single hour thinking about this problem. Those were all hours that I gained.
billforsternz
3 hours ago
You are really excelling in this thread. Thank you for your insights and wisdom, I'm really enjoying everything you are contributing.
aswegs8
2 hours ago
Seconded
maxall4
6 hours ago
Not OP, but Nietzsche wrote thus in Beyond Good and Evil: “Ultimately one loves one’s desires and not that which is desired.” I, personally, find this to be very much the case; and I suspect that it is a feeling common, albeit not universal, among the intellectually inclined towards their problems.
adastra22
6 hours ago
> There's no Lean proof for this one
What is this then, vibes? Without a machine-checkable proof I'm not sure what to think of any of this.
jboggan
6 hours ago
Well I'm sure some people (maybe me if I had time) will do a write-up of this proof. It treads familiar ground for most of the setup, it's mostly the disk lemma and cancellation calculations that need to be understood, it's a fairly short paper and quite tractable.
I think it helps that basically everyone thinks this conjecture is true, it's just been so darn weird to attack. There's this odd thing that the induction proofs of this problem kept running into, which is that the N+1 condition would work except for in one tiny case when it could fail, but it would be covered by a very slightly stronger version of the conjecture. But then that would fail on one tiny case in induction, but you could solve that with another slightly stronger version. Etc., etc. I almost wondered if there were some sort of structure to the increasingly strong conditions and wanted to prove something about the meta-induction between the stronger conditions and the N's that they needed the next level to remain true. But that failed after 5 steps I think (Fable actually helped me write a few hundred test cases to explicitly show that pattern didn't continue forever, thank God).
BTW my existing test suite from previous proof attempts jives with this new algorithm, so I haven't seen any evidence yet that it's incorrect. Waiting for a Lean proof obviously.
Daneel_
6 hours ago
It might have been updated. Is this the lean? https://github.com/openai/math/blob/main/lean/docs/180.md
jboggan
5 hours ago
Lol it should be, but it doesn't seem complete. Line 49 just says "sorry"
/-- Cubic bipartite three-vertex-connected plane graphs have a Hamiltonian cycle. -/ def MainStatement : Prop := ∀ (V : Type u) [Fintype V] [DecidableEq V] (G : SimpleGraph V) [DecidableRel G.Adj], G.IsRegularOfDegree 3 → G.IsBipartite → Planar G → ThreeVertexConnected G → HasHamiltonianCycle G
theorem main : MainStatement.{u} := by sorry
SyzygyRhythm
5 hours ago
In some cases they have a full Lean formalization; in others they just use it for the problem statement. Getting rid of that "sorry" means you've proved the statement. I'm not a Lean expert but it reads pretty clearly as the original conjecture (though the definition of PlaneEmbedding seems quite involved!).
jboggan
4 hours ago
I think this just has to be the problem statement, there's several lemmas I would expect to see in there. Granted I know very little about Lean but it seems like the question and not the proof outlined in the paper.