Hacker Newsnew | past | comments | ask | show | jobs | submit | epolanski's commentslogin

Judging the quality of this blogpost/presentation, AI is centuries away from being able to replace humans.

You're naive if you're thinking the scumbags running the US companies aren't using your data.

In any case old rules apply: if privacy is a concern don't share the data. I share all my work-related code because it's worthless, but I don't and would never share company business and process details, access to production/user data, etc.

Meanwhile I know of people connecting all the kind of MCPs for datadog/sentry/jira/concluce/production databases to their harnessess..lol.


I've occasionally got chinese characters in anthropic/openai's responses too, locally on codex/claude.

Hasn't happened in a while, last time was when I was testing fable 5 in june.


I don’t know what model codex uses for session summarization (I use Pro subscription, no third party models), but I get Chinese summaries from time to time, when the only Chinese that could have appeared in the session would be an i18n strings file that it may or may not have loaded. Very puzzling. Last happened yesterday.

What do you mean?

In any case mathematics is humanity's oldest open source project going on for millenia, it never belonged to a single country, institution or class.


Only the idle and curious rich

Ramanujan was dirt poor.

How many Ramanujan didn’t get the same opportunity to show their work to what was considered the intelligentsia of that time?

Irrelevant in 2026, you can share your results with a click.

Ramanujan was a phenomenon

If model X fits your need, you don't need to upgrade.

I have released applications on Gemini 3.5 flash that make real money and I don't see any particular reason to upgrade.


It feels cheap plastic, true, but it's also light and resists a lot of damage.

Mine v10 has 6 years, has seen some important falls, but bears no damage of any of those.

I wouldn't recommend it because price/quality Dyson vacuums feel like you're insanely over paying.

I've bought my grandma a 150$ Levoit LVAC-200-WEU, and I have used it a lot. I swear it lasts more than the v10 and it cleans better. It also feels much higher quality to touch.

It's cons are that it didn't come with a narrow tip to suck odd corners and that it's slightly heavier than the v10. But it costs also less than half and I see no victory at all for the Dyson product.


I think this whole distillation argument is between fully overblown and bogus.

In any case, highly misunderstood.


Interesting, I liked to experiment with a second model "simplifying" and summarizing the previous messages and continue.

Needless to say, it improved output on following messages by whatever metric I cared for.

Not sure why would they prevent it.

I give you a chain of messages, what do you care for what the origin is?


While I also agree that Opus 4.6, in some ways, was the last model that truly felt an assistant, all the following ones seem to have inverted the role, even a blind person can see that throwing difficult problems, and complex bugs at this model achieves more than predecessors.

I don't think there's nothing ground breaking, but sure it achieves and finds more, sooner.


Just the other way I was thinking that if I asked "What does Lamborghini do?" the only correct way to answer is a single sentence "Which Lamborghini are you referring to?".

But LLMs will fail at this question: they will tell you about Lamborghini's latest car and mix some history in it. Just try.

Which is the wrong answer anyway, because there's at least two major companies called Lamborghini, one making cars, one making agricultural equipment and at least one famous person (Elettra) with that family name.

This very simple test/question makes me realize how much do I hate LLMs in a sense: while I agree that the answer it gives is the most plausible for 90% of the users, it's ultimately both wrong and long. And that 90% compounds.

But there's no "correct" answer in my eyes than "who are you referring to?". Possibly without listing all the possible Lamborghinis.


This is ... unnecessarily pedantic. Anybody in my social universe who asked me that question would undoubtedly expect "they make cars".

If you're picking nits, why not focus on the word "do" and (wrongly) expect an answer like "Lamborghini (either of the two main companies of that name) does not 'do' anything - the companies employ humans who 'do' things. Lamborghini is a legal entity established to allow humans to 'do' things, such as make cars, or agricultural equipment."

Shared context is a thing. Reducing every conversation to first principles is not always required. Get a grip.


This is not a nit. This is a real problem in the technology.

It assumes the average and plausible answer token by token.

And this tendency shows in every single field it's applied to.

At the end of the day I want *correct* answers, to the point.

Instead LLMs, no matter if it's version 3500, are bound to producing average results: slop.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: