Hacker Newsnew | past | comments | ask | show | jobs | submit | hobofan's commentslogin

Why? This just follows the scheme they have established in 5.6, just that GPT 6 introduced Astra as another tier above Sol.

How will GPT-5.6 Sol and GPT-6 Sol be differentiated?

>How will GPT-5.6 Sol and GPT-6 Sol be differentiated?

You literally just differentiated them in this sentence?


...by the version number? How are Claude Opus 4.8 and Claude Opus 5 differentiated?

I'd wager it's more due to decision making culpability.

Now, for the next "teen committed suicide after talking to AI therapist" headline, Claude can deflect the blame to the parents for not supervising their child enough.


If you have an ineffective system for preventing a thing from happening, but one where the "victim" has to override the system, the victim is at fault.

Your honor, the user manual in the box did say "don't put children in the microwave."


So in your example a microwave company should be held liable if someone puts a child in a microwave?

> Why use a limited version?

Because in most of those API, even many implementations of the Responses API, you lose a lot of control of where your data is going. e.g. an Agent or the Responses API may automatically invoke a tool call that leaks your data to an external service on the internet, without having an option to intervene.

If you want to have control over your data, you have to have control over your harness.


Photographing bikes or electric scooters does nothing without enforcement through the city. It allows tracking down the person who placed it wrong, but in the absence of fines through the city, the rental companies surely won't start fining their customers and drive them away.

The Lime app (bikes and scooters) refuses to close down a trip if you enter an area where parking is forbidden or if the picture shows that the vehicle is blocking the sidewalk or pedestrian paths.

It will also 100% forward you a fine if they get one and their records show you were the last user who misplaced it.


In a world were there is a significant amount of people that do things just to take photos of themselves doing things rather than experiencing the thing, this may be the lesser evil.

At least they actually showed up

> Improving models in a holistic way sounds a lot like training to me.

I think that's quite a leap. Using de-indetified data to improve the products is what everyone has been doing since the dawn of web analytics.


The distiction they are trying to make is: "One of our employees or the model was able to verbatim read the chats when they were actively tackling the problem" vs "The chat of someone working on the problem may have ended up in the training set of the model".

https://simonwillison.net/2026/Sep/1/codex-libreoffice/ - https://news.ycombinator.com/item?id=49527396

(and I can also just add that it's my go-to tool for that exact use-case, and used it in a dozen different client projects)


Regulation != censorship. (Apart from that, Mistral has Shieldstral[0] for adding policy enforcement).

With it's general approach of Open Weights, and being able to be deployed on-premise (/private/public cloud), they are a viable solution for your typical European enterprise, that has to comply with traditional (non-AI) regulation. With the NIS2 directive the amount of those companies is also significantly expanding.

We[1] are in the same market as Mistral, and among our customers, the go-to-solution of MS Copilot is typically performing badly, and the typical SaaS solutions are not even given a consideration, which is why they are reaching for on-premise-first solutions.

[0]: https://mistral.ai/news/shieldstral/

[1]: https://github.com/EratoLab/erato


Are they, though? I would say, let's find out when the IPO settles, but then again the market also sustains a incredibly crazy valuation for Tesla.

If not, I’m buying stocks

Better to buy GPUs and RAM with current rate

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: