Hacker Newsnew | past | comments | ask | show | jobs | submit | x-complexity's commentslogin

That's just a CDN, but now it's not HTTP accessible (thus strictly worse).

> What I expected instead was a lot more discussion about use cases, benchmarking, possibilities, limitations (that aren't about git history) and the scope of future development.

HN had an Eternal September. Such discussions have been drowned out by the rest of the mob.


All of the following are true:

- They have internal models that still show there's gas in the tank, in terms of improvements

- Open models are eroding their customer base on the low end of the curve, & also you need not use Astra Max to classify images.

- The true demand for high-intelligence tasks is lower than what they need to justify their spending

- It would be wonderful for them to kick up a scare & give reasons as to why high-int open models should be regulated out of existence, opening up the lower end for them to take again

- Setting up an oligopoly for themselves gives them a captive audience, just like it did for (insurance, banking, telecoms, etc.) via regulatory capture


The demand for human-level intelligence is obviously at least the same order of magnitude as all office jobs combined.

I'm personally skeptical of LLMs getting there without at least one new invention. But it's also empirically clear that the pace continues to increase, even if we ignore that exponential growth is the default in economics in general.


> The demand for human-level intelligence is obviously at least the same order of magnitude as all office jobs combined.

1) That level of demand is likely true, but that doesn't mean that the total spending demand is to match to that level.

2) Most of said office work can be automated via normal software, and need not use LLMs. If they were to be automated, the amount of work left available for Anthropic to routinely work on would be (lower than before the automation took place).


> Is there any data you can provide to support your claim, or any result you can contribute here?

By the time we can show you data that convinces you that it does work, the next generation would already be out & incrementally dismantling the old conjectures that were true in the previous generations.

You're fundamentally asking for a violation of how information passively disseminates amongst humans: To go any faster requires more effort on the receiver's part to move up on the adoption curve.


Wouldn't this also mean that all previous generations that were proclaimed as intelligent and working were in fact... not?

It doesn't matter what comes tomorrow, with the next generation, if the claims now can't be proven.

To preempt the response: The math proof, regardless of them using non-disclosed user data or not, they spent $30M do do something closer to a 1000 monkeys approach, rather than a singular inference being very intelligent.


To preface -- I try not to be dogmatic/politicized on AI, so I will genuinely consider your arguments! Please try to convince me. (indeed, I am the grandparent commenter)

I agree with the meat of your statement, but am very interested in the pre-emption, "they spent $30M do do something closer to a 1000 monkeys approach, rather than a singular inference being very intelligent". First, I think the $30M number is inflated -- that's what the general public would have paid, but presumably the internal cost is lower, perhaps it's more like $10M. But it is still expensive. Second, I'm curious if it's really the case that they did a 1000-monkeys approach? I haven't read much in-depth reporting about the proof, so it's totally possible I just don't know. What is it that they did which is more like 1000-monkeys? Also, I wonder if that distinction matters -- if 1000 monkeys can reliably make ground breaking proofs, and the approach generalizes to other tasks, I will happily become a circus owner. Maybe you're claiming that it won't yield other proofs? Or the proofs are too opaque to be useful to humans? Or it can handle proofs but not other tasks?


I'll respond/comment on the parts I hope are relevant to you, in no particular order:

Yes, a proof is a proof regardless how you get there. We however don't hear about when they fail, and I doubt their 10000 agents (from their own statement) would necessarily reach another solution/proof (this by leaning towards using user data after finding out others were close). They could as well have attacked another Millenium problem, but they didn't. In whichever case, we will have to wait and see if they (either company) can reach novel solutions/proofs without significant amount of human provided data for the LLM to bridge the gaps.

Further, and this is more of a policy opinion/prediction: If the numerable obtainable (albeit very hard) problems are solved, assuming training data is needed, will it push out future researchers from entering the field due to lack of reachable goals, thus cutting off future training data? LLMs have been great at replacing gateway jobs. But those jobs are what leads to frontier training data (be it maths, physics, chemistry, economics, graphics, prose, etc).


Ah it's interesting they legitimately used 10k agents, I didn't realize that. I do agree that it's significant they solved the Millenium problem only once humans had made significant headway. I'm not sure I believe the relevant training data will be produced at a significantly lower rate due to AI -- Millenium problem solutions weren't generated at a very high rate before anyways. But I think all your points hold nonetheless. Thanks for elaborating on them!

As you said yourself:

> > If you believe in civil liberties as rights and ideals, your belief shouldn't stop at borders.

If you believe in civil liberties as rights and ideals, your belief shouldn't stop at money.

Either the right extends everywhere, including money, or it doesn't.


That doesn't follow.

It's OK to say you wish a politician was dead. It's not OK to pay someone to kill them. It's OK to say that you support some criminal activity (let's say Luigi), but it's not OK to offer direct financial support for their crimes.


> It's OK to say you wish a politician was dead. It's not OK to pay someone to kill them.

Both actions are equivalent, as such words are (social and/or political) capital bounties offered to the masses in return for fulfilling the death of said politician.

Capital always exists, even in the form of words. To deny a variant of capital whilst allowing another variant is inconsistent/doublespeak.

Just because it's not printed onto a piece of denominated paper, doesn't mean that such capital doesn't exist.


> Both actions are equivalent, as such words are (social and/or political) capital bounties offered to the masses in return for fulfilling the death of said politician.

Neither the law nor common interpretation of words agrees with this. There is a significant and crucial difference between words and actions.


> There is a significant and crucial difference between words and actions.

The difference is artificial, and delusionally blinds itself to the base reality that social/political capital exists.

Again, just because there isn't a cash figure attached to that bounty, doesn't mean that a reward hasn't been attached to it in the form of social/political prestige.


Setting up an LSP relies on 2 requirements to properly work:

(a) The LSP's tools being competently built & consistent, and

(b) the LLM using it having been properly trained to use LSPs in general.

Using a native tool like grep has the same 2 assumptions, but

(a) is satisfied due to ossification of grep's core features (a good thing), and

(b) is extra-satisfied because the LLM can be trained to properly use grep specifically, and not "20th variation of grep wrapped behind an LSP, but just different enough to throw curveballs".

On top of that, grep is almost always present in default Linux environments, so its presence is assumed & can be relied upon when needed.


> I assume that by this you mean something more nuanced than "we should abolish the FAA and let the airplanes crash until private enterprise solves it."

Strawman fallacy: It is in the interests of private enterprise to not have their $100-million+ airplanes crash, because it means $100-million+ going up in smoke. The owners of those planes would therefore have an interest in not having them crash.

> Or maybe "Let drug dealers (pharmaceutical companies) sell whatever the free market will bear without testing, safety, or the ability to hold them responsible."

Under this scenario, if the public finds out that a drug is harmful/lethal, the public jerks away from that drug towards solutions that have a reputation of working.

Is the resulting rate of fatality/worse-QOL lower than (a central authority filtering out drugs & holding them to standards)? No, definitely not. But the gap between the two is not infinite, and in most cases falls into an order of magnitude of difference.

-----

> I only ask, because there are some people who unironically think that we'd be more safe going back to the time before OSHA, or medical licensing, and some of them hang out here.

An axiom of (feeble-mindedness in the average person) is proposed here. People are as good/competent as the standards that they are held to.

> medical licensing

There's a difference between a certification of competency, and a deliberate constriction of the supply of medical residencies.

https://thehill.com/opinion/healthcare/5550556-residency-cap...


Following this hypothesis, it would be rationally beneficial to release a wild bear every 6 months or so, just to make sure that the 'fear bank' is occupied by things that warrant actual fear.


We do have this whole global climate crisis going on. Maybe people could try being afraid of that


> We do have this whole global climate crisis going on. Maybe people could try being afraid of that

It's too conceptually big for some people to put into the 'fear bank' to be relied upon.

For the threat to be near-universally relied upon, it needs to:

(a) have a physical presence, and

(b) be within 2 orders of magnitude of size to trigger physical alertness (tick -> herd of elephants), but also

(c) not be (immediately dangerous/lethal that the only solution is to never meet it)


> Nevertheless you don't have to lie to kids in any field, science, art or otherwise.

...if and only if the kid in question has the mental capacity to take on the whole truth without confusion.

Some kids have the potential to go hog wild on multivariable calculus, and they should be given the chance to handle it all. But they shan't be the benchmark that others with differing capabilities must hit to understand a concept.


> On a spectrum between tulip wilting and fiber build out, we know deprecation cycle of DC hardware leans towards tulips i.e. <10 years (very generous) vs 20+ years for fiber layout, a lot of which is actually infra/earth works etc.

Counterargument: As advancements in transistor densities slow down, the rationale for increasing depreciation cycles makes more sense. As the performance gap between new & 5-year-old hardware continues to shrink, then the need to replace older hardware similarly shrinks, justifying longer depreciation cycles.


It's not just about node advancement, which is coupe de grace condition. Even if hardware advancement freezes, its about IC premium that fed current tranch of AI buildout. Current players paid $10 for a $2 hammer due to premium, a better future hammer might cost $3 but does twice the work of $2 hammer. That is like ball park the premiums we are talking about - from gpu to memory to other components getting inflated due to exuberate AI demand.

The economic logic is if current spend vs revenue gap is not sustainable... hardware prices / margins will revert towards mean. That $10 hammer will be compared against a $2 identical hammer (margin reversion/compression)... or worse, a $3 future hammer that does $4 / past $20 of work. The future player who only paid $2 can charge much less... i.e. simply paying $10 limits ability to price competitively. The future player who pays $3 has 50% more compute than incumbent who paid $10. The important DC TOC consideration, is in world where DC cost regress towards mean, opex > capex... so merely continuing to use that old $10 hammer is losing MORE than buying a $3 better hammer, i.e. the asset is economically stranded, it is COSTING MORE to run old hardware than simply buying new hardware. It's MORE than economically useless and $10 past purchase price not just sunk cost but dragging down balance sheet as amortized liability aka it is full write down / loss.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: