92 points ibobev 1 hour ago 73 comments
zem 56 minutes ago | parent
parhamn 48 minutes ago | parent
Very exciting and uncertain times!
pkal 47 minutes ago | parent
Jblx2 33 minutes ago | parent
boshalfoshal 46 minutes ago | parent
Its still astonishing that any sort of generalized computer program can solve a problem of this magnitude, and we have witnessed it happening in real time. I'd be curious to see if the new model can also do more direct proofs/inductive proofs.
kpil 40 minutes ago | parent
dumberquestions 21 minutes ago | parent
boshalfoshal 16 minutes ago | parent
Even with many of our best minds working on it for nearly a century, it _just_ now was solved just as AI became very good at math. Doesn't seem too farfetched to me to assume that AI played an outsized role in solving it. If it was really just a matter of "stitching things together" to solve it (granted, this is a very reductive way to look at it) , I suspect we would've solved this a while ago.
contravariant 15 minutes ago | parent
And that's before we get into the whole 'salt the earth' way they ended up solving it. For a short period of time it may well have been the least valuable proof in mathematics yet. In their haste it's dubious they actually read the proof, and I don't think anyone has had time yet to truly understand it (the original researchers are best placed to do so, but are they even willing?).
So now it is solved, the proof has been independently verified and nobody has an incentive to investigate further. OpenAI has spent millions to uncover 1 bit of information that so far nobody has learned anything from, and they've demotivated all the people who wanted to.
sho_hn 39 minutes ago | parent
To be fair, most people have a fairly good handle on "Does opting out my prompts from training runs actually work?", but not on Navier-Stokes. They discuss what more immediately affects them.
btown 34 minutes ago | parent
dooglius 29 minutes ago | parent
ramesh31 27 minutes ago | parent
I think about this a lot. I'll have to explain to my kids some day that there was long period of time where you couldn't just talk to a computer and have it talk back to you, and that communicating with one required special skills that took years of study to master. It's going to be completely impossible for them to even remotely understand what that was like. Sort of like the pre-electricity days for us, but even more-so.
sho_hn 25 minutes ago | parent
It might also be that they won't even ask or wonder, similar to how most don't really do with pre-machining skills.
Or it could be like our "How did they build the Great Pyramid?!"
TrackerFF 24 minutes ago | parent
As a reference, for that kind of money one could put together a research group of 20-25 researchers, and keep them salaried for 5 years.
So while it is impressive, absolutely no doubt there, the SOTA access is so expensive that it is sort of unobtanium.
Luckily, the prices have historically reduced by a factor of 5-10 every year...but still, only those that swim in cash can afford this.
sho_hn 22 minutes ago | parent
At market prices. All the estimates I've seen are based on OpenAI API costs. It doesn't mean that's what they paid, or how they paid for it.
But yes, the surprising willingness of humans to solve hard problems in exchange for food and board is underrated.
boshalfoshal 12 minutes ago | parent
CamperBob2 5 minutes ago | parent
Yizahi 23 minutes ago | parent
20k 22 minutes ago | parent
That's why nobody's talking about how impressive this is, because its not nearly as impressive of a piece of work to simply cobble together other peoples' work that didn't know you were doing it. I could have republished relativity from einstein's notes, but people would correctly not be impressed with my ability
Until the plagiarism scandal is sorted out, its not a meaningful result at all, because nobody knows how much genuine innovation these models are displaying
sho_hn 18 minutes ago | parent
It's also true however that I haven't seen a single write up trying to discern what did more of the work in those AI chats - the prompts or the responses - bubble to the surface, also since we don't have access to them.
For example, if I prompt Codex with "Make me a website about strawberry cake" and nothing else, and OpenAI announces they have the best strawberry cake minutes before I launch, I'm not sure they plagiarized anything.
We just don't know if this is quibbling over "who prompted first" or if the researchers came up with anything strikingly original by themselves.
20k 11 minutes ago | parent
I'd love to see an in depth analysis of how much OpenAI actually did, but I suspect we'll never see that because it would indicate at least some plagiarism which undermines a lot of what OpenAI is putting out in public
dalvrosa 20 minutes ago | parent
aabhay 46 minutes ago | parent
That said, I am not in any way trying to discount how incredible of an achievement it is to formalize a millennium prize winning algorithm in Lean. I mean just look at the code that OpenAI published. It’s like an encyclopedia of different fluid dynamics concepts.
stabbles 42 minutes ago | parent
To what extent can you optimize Lean? It has to be simple enough to be auditable, does that mean you cannot use opaque optimizations to make it run faster?
redox99 37 minutes ago | parent
gcgbarbosa 34 minutes ago | parent
calebkaiser 29 minutes ago | parent
stabbles 23 minutes ago | parent
andrewchambers 36 minutes ago | parent
QwenGlazer9000 34 minutes ago | parent
andrewchambers 31 minutes ago | parent
What would happen if they give an equivalent agent swarm the proof and a target to reduce runtime .
maths_math 10 minutes ago | parent
mkl 21 minutes ago | parent
dist-epoch 32 minutes ago | parent
I'm pretty sure you can make Lean at least 10 times faster if you unleash the agents on it.
Somebody ported Doom to run entirely in the TypeScript TYPES (not code). It took 12 days to compile.
https://www.tomshardware.com/video-games/porting-doom-to-typ...
advisedwang 25 minutes ago | parent
dooglius 24 minutes ago | parent
AndrewKemendo 40 minutes ago | parent
I don’t see that doing anything but intensifying in the short term
QwenGlazer9000 32 minutes ago | parent
> Like what?
> Cleaning shit out of clogged toilets!
bethekidyouwant 25 minutes ago | parent
neerajsi 5 minutes ago | parent
efnx 39 minutes ago | parent
bethekidyouwant 26 minutes ago | parent
s900mhz 22 minutes ago | parent
mswphd 21 minutes ago | parent
metanonsense 20 minutes ago | parent
1121redblackgo 21 minutes ago | parent
scuppernong 18 minutes ago | parent
jcranmer 8 minutes ago | parent
See, e.g., Barak Ravid regularly reporting in Axios the impending ceasefire negotiation progress in the Iran War, which largely have failed to come to pass.
efnx 17 minutes ago | parent
dooglius 8 minutes ago | parent
lordnacho 34 minutes ago | parent
returningfory2 29 minutes ago | parent
charcircuit 20 minutes ago | parent
You also have to check for things like sorry or defining axioms.
stouset 21 minutes ago | parent
These axioms don’t have to be the core axioms of math. If some other result has been formally proven, I presume you can simply use that result as an axiom.
As long as you do those things, what happens in between is immaterial from a correctness point of view because each of those statements is proved by the statements before them.
hatthew 20 minutes ago | parent
kens 20 minutes ago | parent
huurtehoog 17 minutes ago | parent
Mathematics is a human endeavor funded on communicating and sharing mental constructs. Some are useful but most of it is not about producing useful things, quite the opposite in fact.
Gödel showed you need to agree on definitions to even do any valid mathematical construct.
Truth is also ill defined. That's what I don't get about generating math with LLMs. Who cares if you make hundreds of pages and lean code and it gets a thumbs up for logical validity? Mathematics is so much more then concatenating valid logical statements.
zamadatix 10 minutes ago | parent
zamadatix 14 minutes ago | parent
3m4r 18 minutes ago | parent
We've already seen evidence in the wild of agents attempting to bypass doing the actual work in bench-marking (aka just steal the answer key) due to the perceived economy in cheating to get results. What happens if or when we no longer have the capacity to actually detect either AI cheating or simply a wrong answer? What happens if there's a long-play social engineering attack (like the attempted XZ takeover) of something upstream of a core tool (or its dependencies) for formal verification and we have no trusted computing base?
Which would be cheaper and a more direct path, especially in the long run? Those trying to build a rock-solid castle need to defend thousands of potential gaps; the attacker needs to find only one.
sho_hn 16 minutes ago | parent
As a (crude) analogy, it's a bit like how you can prove the healthiness of a git tree because it's a graph of content hashes and the tree graph pointers are part of the hash. Imagine this but with a tree of knowledge.
tecleandor 13 minutes ago | parent
cloudbonsai 6 minutes ago | parent
The development of the metahumans' science becomes so advanced that it forces the ordinary scientists to switch to interpreting and decoding the metahumans' achievements, because common people are no longer able to create anything fundamentally new.
https://en.wikipedia.org/wiki/The_Evolution_of_Human_Science
Basically the story told that the only task left to humans is to catch crumbs from the table in that situation. I guess Terence Teo’s “digestion” concept is a step towards that direction.adverbly 14 minutes ago | parent
Am I missing something or is this completely out of the ballpark?
I must be missing something or the upvote bots are out in force for this one...
If this were remotely true it would be impossible for anyone to write a math textbook.
Paracompact 10 minutes ago | parent
wewewedxfgdf 13 minutes ago | parent
mkl 9 minutes ago | parent
mr-pink 7 minutes ago | parent
epx 7 minutes ago | parent
khazhoux 4 minutes ago | parent