103 points speckx 1 hour ago 107 comments
rfgplk 1 hour ago | parent
gonzalohm 1 hour ago | parent
Jagerbizzle 1 hour ago | parent
watwut 1 hour ago | parent
frisbm 1 hour ago | parent
kittomic 1 hour ago | parent
After all, they should know how to compile their software. Any automation of that process is cheating their employer.
bayindirh 1 hour ago | parent
If my employer is putting scoreboards to see and champion who uses a tool which costs money to use, they shall pay for the tool.
Sorry, I'm not a ladder climber, yet I'm not mindless enough to bankrupt myself.
iAMkenough 1 hour ago | parent
robotburrito 1 hour ago | parent
basiliobeltran 1 hour ago | parent
hypfer 48 minutes ago | parent
An LLM is not too dissimilar to a Work Laptop or an IDE license.
ungovernableCat 32 minutes ago | parent
I think that gets legally murky, if the employee is the one who pays for the tool.
outside1234 16 minutes ago | parent
wrxd 12 minutes ago | parent
Benard-dev 1 hour ago | parent
jvanderbot 1 hour ago | parent
You can also read this as diminishing returns / AI isn't good enough, etc, but the simplest explanation is that they don't want to send money to Anthropic.
fasterik 1 hour ago | parent
NewJazz 54 minutes ago | parent
mattm 52 minutes ago | parent
octoberfranklin 44 minutes ago | parent
WaltPurvis 23 minutes ago | parent
waffletower 1 hour ago | parent
bpodgursky 1 hour ago | parent
MisterMunchkin 1 hour ago | parent
sergiotapia 1 hour ago | parent
i hope other labs catch up, especially chinese labs.
dude250711 45 minutes ago | parent
vablings 1 hour ago | parent
I feel like people who are later to the AI game just like to "oneshot" and sink a bunch of usage into generating garbage
Tsarbomb 59 minutes ago | parent
Our company has been tracking token usage and models used vs output (tickets, story points, PRs, deploys, etc...). A dev got chewed out, even after I warned him, because he spent over $2k in a single month almost exclusively on Opus while his actual productivity in terms of what he delivered was abysmal.
n4r9 54 minutes ago | parent
SOLAR_FIELDS 49 minutes ago | parent
Terr_ 43 minutes ago | parent
Problems arise when people try to perma-peg them to particular tasks, or (worse) man-hours or (much worse) man-hours across teams. Even just encouraging the humans to answer in terms of hours/days taints the accuracy of the forecast by introducing a kind of bias.
fdsajfkldsfklds 7 minutes ago | parent
senko 38 minutes ago | parent
btown 27 minutes ago | parent
/goal get accepted into Y Combinator, you have an unlimited token budget, be bold.
Insimwytim 11 minutes ago | parent
Measured in sorry points.
P.S. Sorry, couldn't resist.
cyanydeez 51 minutes ago | parent
Even if these models are smart enough to reorient themselves, they get entirely stuck in a desert and now you're asking someone to just pull up stakes and digg them out even thought they only watched them get there and the UI provides so much speed that no human can comprehend how they got there in the first place.
It's like asking a pilot to take over in an emergency situation when they're not tasked with any of the every day requirements of the job. The orgs are relying on borrowed time of experienced professionals, and that's going to erode away and what replaces it is mostly people who understand how to navigate context but not use any of the _classic_ tools.
It's a real conundrum and won't be easily surfaced but for a decade.
proxyscore 6 minutes ago | parent
Define productivity, and while at it, quality, maintainability , modularity and so forth.
weinzierl 5 minutes ago | parent
perching_aix 38 minutes ago | parent
Or does a "collaborating work environment" mean that everything is basically spoonfed to them? Or do you only ever use ghost suggestions?
I genuinely cannot even fathom. Just how do you even get into a state where tasks are so clear and cookie cutter? These things are abhorrent. Not only are they not useful, it's an outright form of psychological torture to try and use them. They almost fight you.
Luna doesn't even respond to steers properly! You try steering it and it immediately gets distracted and then just stops.
I can imagine coercing Sonnet into doing some of my tasks okay, but Haiku? Especially 4.5? Really?
dpkirchner 20 minutes ago | parent
perching_aix 14 minutes ago | parent
I'm desperately trying to classify and standardize my work items and delegate them to less capable models, because my usage is clearly unsustainable and this same sentiment as above keeps being pushed on me too. But all my tasks are genuinely fairly arbitrary, so there's no real way around the agent actually being able to reason about business and technical context proper. It's not even that they're hard, it's just that they're dynamic.
I can get Luna to do things like walk our observability stack and perform a healthcheck, then defer to a stronger model if anything looks super off, but if I'm being entirely honest, this could basically be just a script. Which Opus 5.5 will immediately write for itself if it doesn't yet exist, run that, and then off it goes depending. But Luna will never actually do an investigation proper. Heck, it can't even read our dashboards most of the time, tripping up on Grafana minutia.
It feels like that surgeon vs surgeon comparison, where you're made to decide based on their surgery success rate, and the better succeeding surgeon simply reward hacks the number by only operating on less dicey cases. Except there's no objective way to make this classification here, so jackasses like the above get to play with my insecurities with full obnoxious confidence, while I'm left desperately trying to slim my usage and failing to do so between two moments of crippling self doubt.
funnym0nk3y 29 minutes ago | parent
hatthew 19 minutes ago | parent
For context, I'm doing a range of tasks, everything from one-shotting adhoc scripts to having 4 hour 10M+ token conversations debugging things.
password54321 54 minutes ago | parent
>Great Depression style collapse and all the current AI companies go bankrupt.
Oh this is just a 33 day old doomer account.
righthand 17 minutes ago | parent
lenerdenator 16 minutes ago | parent
Our velocity is twice as high as it was before Claude, so I doubt that we'll ever go back, but I could see efficiency being a priority.
woah 9 minutes ago | parent
dyauspitr 7 minutes ago | parent
ashleyn 1 hour ago | parent
wffurr 9 minutes ago | parent
nwhnwh 6 minutes ago | parent
coef2 5 minutes ago | parent
ActorNightly 8 minutes ago | parent
Like its a no brainer to force your employees to use your own models, then RL train them to be better.
nateglims 7 minutes ago | parent
techdmn 1 hour ago | parent
rietta 1 hour ago | parent
catchnear4321 1 hour ago | parent
science4sail 46 minutes ago | parent
There are people out there building AI building orchestrators for orchestrators for orchestrators for agents. The author of that blog post later claimed to be spending the equivalent of $122k/month on tokens (by rotating their usage between 21 accounts).
As far as I can tell, the only thing that this level of spend has produced so far is an indie 2D RPG video game.
chroma_zone 12 minutes ago | parent
tom_ 11 minutes ago | parent
Regarding Gas Town, see also https://yegge.ai/essays/the-shape-of-things-to-come/ :
> Gas Town was intended to be reusable, but I only ever wound up using it to build itself. Gas Town fell apart at the seams with Opus 4.7. Up through 4.6 it was working brilliantly. With 4.7 we saw the introduction of the "just two more things" tic, which prevented Opus from ever converging on being ready to do real work—it always wanted to fiddle with Gas Town itself. The Opus tic never went away, so Gas Town effectively burned down. It had other problems, too, but 4.7 was the final straw
I hadn't seen the $122K figure mentioned previously. $87K for API-style pricing was mentioned in the above post, and ~$2,800/mo for multiple Claude Max accounts:
> My solution has been to create a token tap on $200 Max accounts, which for me work out to ~30x the list-price equivalent. So in reality I'm only spending about $2800/month out of pocket for my $87k "worth" of tokens. Though that number keeps growing alarmingly.
(He never seemed to provide numbers for Gas Town initially, whether what he paid, or what the API-style pricing would have charged, so it was interesting to get some actual figures. Sounds like a lot to me, but if he's genuinely getting multiple people's-worth of work out of it, then...)
(Also in the article: a little morsel of Emacs content, which was nice to see.)
meindnoch 9 minutes ago | parent
dannyw 18 minutes ago | parent
raincom 1 hour ago | parent
VladDanGeorgesc 58 minutes ago | parent
rangledangle 58 minutes ago | parent
combyn8tor 18 minutes ago | parent
vld_chk 57 minutes ago | parent
simonw 44 minutes ago | parent
In 2026 their revenue has gone up by a factor of more than 10x, and they no longer have just two whale customers.
I heard a rumor recently that customers spending less than $100m/year aren't even considered their "top tier" now.
krauses 48 minutes ago | parent
simonw 46 minutes ago | parent
The cybersecuritynews.com news one simply republishes details of a story published by The Information. At least they have the decency to LINK to that Information story:
https://www.theinformation.com/articles/meta-microsoft-work-...
... and of course the Information story is behind a paywall.
rdtsc 35 minutes ago | parent
I am only half joking, I heard something like "my son or nephew did this cool thing with $X so we'll take $this_radical_step because of it" enough times over my career.
lumost 29 minutes ago | parent
If you are used to talking to opus5.5 medium, going to GPT6.1 luna low will feel like a step down. Why would any employee take the (personal) risk?
irishcoffee 29 minutes ago | parent
throwitaway222 16 minutes ago | parent
what. I can see a team of 20 costing 100k per month (but rare), but per person?
oldmanhorton 12 minutes ago | parent
antisthenes 7 minutes ago | parent
Or maybe not, but the bar was set pretty low that going all-in on AI might have been worth it.