Hacker Newsnew | past | comments | ask | show | jobs | submit | a_wild_dandan's commentslogin

Guessing SH meant Steven Hawking, who kicked the bucket. Metaphorically.


Maybe everyone will migrate to Hugging Bay for downloading models via torrent?


Wow, you weren't kidding. I looked at their chart, and the cost-per-task for Fable is more than double Sol's. And DeepSeek absolutely stomps. Four cents per-task vs Sol's $1 and Fable's $3.

I might need to check out DeepSeek more. I had no idea the difference was this obscene. Makes me wonder if something's off with the benchmark. A 70x cost reduction vs. Fable seems too good to be true.


In my benchmarks of security auditing abilities of models, DeepSeek was roughly an order of magnitude cheaper than either GPT 5.5 or Opus 4.8, less than ten cents per task vs. roughly a buck each for the best American models at the time.

GPT 5.5 Pro was ~230x at almost $23 per task.

DeepSeek is my go-to when I need an API, and local Gemma 4 won't do because it's either too slow or not capable enough. DeepSeek isn't at the frontier but it's good enough for a lot of things, very cheap, and quite fast. Flash is even faster and cheaper, and still better than anything I can host locally.


solve p=np make no mistakes


n=1 or p=0


Have been having an awful day today, this really cheered me up. Thank you so much for sharing your dumb joke!


It's hard to solve P?=NP due to P!=NP.


Backward compatibility with current meatspace tooling.


You’re not stupid. That’s terrible UX. The button is completely disconnected from its modal, and is placed in a bizarre/nonstandard location.


It's placed like one of those chat services on sites. Which we've been trained to ignore.


Sounds like a bug in their css layout related to the smaller screen size


Speaking of tricks, does anyone here know how many angels can dance on the head of a pin?


Taiwan’s geopolitical position is vastly more complex than the fantasy where invasion would follow merely from fab parity.


It won’t. But again you’re missing the point. It’s one less incentive not to, a big one too.


> Unlike the previous GPT-5.1 model, GPT-5.2 has new features for managing what the model "knows" and "remembers to improve accuracy.

Dumb nit, but why not put your own press release through your model to prevent basic things like missing quote marks? Reminds me of that time an OAI released wildly inaccurate copy/pasted bar charts.


It does seem to raise fair questions about either the utility of these tools, or adoption inertia. If not even OpenAI feels compelled to integrate this kind of model-check into their pipeline, what's that say about the business world at-large? Is it that it's too onerous to set up, is it that it's too hard to get only true-positive corrections, is it that it's too low value for the effort?


> what's that say about the business world at-large?

Nothing. OpenAI is a terrible baseline to extrapolate anything from.


I always remember this old image https://i.imgur.com/MCsOM8e.jpeg


Their model doesn't handle punctuation, quote marks, and similar things very well at all.


It may have been used, how could we know?

Mainly, I don't get why there are quote marks at all.


Humans are now expected to parse sloppy typing without complaining about it, just like LLMs do. Slop is the new normal.


Maybe they did


Businesses do whatever’s cheap. AI labs will continue making their models smarter, more persuasive. Maybe the SWE profession will thrive/transform/get massacred. We don’t know.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: