113 points gmays 1 hour ago 48 comments
andsoitis 49 minutes ago | parent
Analysis paralysis stifles not just human intelligence, but other intelligences too.
monkey_monkey 49 minutes ago | parent
Also, did I miss a memo? Suddenly every article on AI seems to be talking about the Pareto frontier - or have I just not been paying attention?
AnodicElegy 36 minutes ago | parent
DonsDiscountGas 20 minutes ago | parent
user43928 9 minutes ago | parent
Kimi K3 with less reasoning tokens isn't exactly exciting either, and particularly so if the license is less open than original Kimi K3.
tomrod 47 minutes ago | parent
The pareto frontier needs clearer distinction. Benchmarks miss half the story. What, if any, capability is lost by the token reduction (for example, was it like super awesome at Golang before and now kind of sucks? that kind of distinction).
drob518 21 minutes ago | parent
esafak 47 minutes ago | parent
jamienk 47 minutes ago | parent
segmondy 36 minutes ago | parent
andsoitis 36 minutes ago | parent
I suspect the advantage that catapulted Linux ahead of the establishment was less technical potential and talent and more organizational advantage. That's not to diminish the technical talent of the Linux crew, but them being unencumbered gave them more degrees of freedom. The rest is history.
So as long as the AI companies don't succumb to "big company" dynamics, they can outlead. To wit: Open AI and Anthropic are kicking Google's ass.
jamienk 29 minutes ago | parent
andsoitis 27 minutes ago | parent
Indeed. And when you have freedom to play, you are able to find new stepping stones that you didn't anticipate. And you can combine stepping stones in new ways to make new discoveries.
Greatness cannot be planned.
zeroq 12 minutes ago | parent
swagatkonchada 11 minutes ago | parent
ls612 46 minutes ago | parent
spijdar 39 minutes ago | parent
That's just vibes, though.
tdhz77 44 minutes ago | parent
intothemild 43 minutes ago | parent
DonsDiscountGas 23 minutes ago | parent
swagatkonchada 13 minutes ago | parent
reactordev 11 minutes ago | parent
makeramen 11 minutes ago | parent
Not suggesting this is right or wrong, but is sort of the nature of the technology.
kingstnap 7 minutes ago | parent
> task and environment feedback
> on-policy planning and learning
> feedback connects decisions to their consequences
These are deliberately the least informative phrases you could possibly use to describe what you have done, while still being in the realm of words that go over a generic investor who has no idea whats going on and may be dazzled by sciencey sounding language.
Cursor compose 2.5 article where they used and described on policy self distilation was actual alpha.
netvarun 37 minutes ago | parent
drob518 22 minutes ago | parent
nostrebored 14 minutes ago | parent
logicallee 28 minutes ago | parent
nostrebored 13 minutes ago | parent
dbuxton 26 minutes ago | parent
nico 24 minutes ago | parent
This is partly the appeal of Jev et al; having a quick model for simple tasks, that doesn’t require that much thinking
It’s amazing all the workflows that models like that can unlock. And yes, classifiers and other ML models have been around for a while for these types of tasks, but Jev has made it easy and cheap to play and experiment. This in turn, is incentivizing people to try them for a bunch of stuff, unlocking creativity and producing a lot of new cool (and eventually potentially very useful) applications
demibabs 15 minutes ago | parent
neosat 9 minutes ago | parent
1. Evals (once you have your rubric defined and tuned using a reasoning model, jev can be great for running periodic evals especially those that run daily.
2. e-commerce catalog classification 3. quick search using anything as context and query mapping to a pre-defined set.
themgt 24 minutes ago | parent
"Pareto": 8 hits
"Opus 5.5": zero hits
wmf 17 minutes ago | parent
GodelNumbering 8 minutes ago | parent
shriphani 2 minutes ago | parent