> In 2006, SoftwareOnline was sued by The Washington State Attorney General's Office for alleged violations of the Consumer Protection Act after complaints were made about two products called "Registry Cleaner" and "InternetShield".
Asking an LLM to remove an Idea from a Document often times results in an edit which explicitly states that this Idea is not relevant, instead of just removing all references to that Idea.
That might be useful for some form of "evolving" Documentation (so future readers know that this part of the search space was covered and deemed irrelevant) but is just overly verbose and confusing to read in most situations.
You can get better results from less aligned models like Kimi K3. Still not actually funny, but at least it’s able to produce some unhinged stuff and I guess shock and twists are kinda related to humor?
Still missing the human connection of cause, so im not sure if this is a technical / skill issue in the first place.
That's a Claude thing btw. Try using a harness which does not inject half a novel of instructions in combination with a different model. I would recommend Pi + GPT 5.6 Luna for a very capable and cheap test.
After using Claude (paid by work) for a couple of months, I was amazed how well instruction following works in other setups.
Yes. The orgs that have gone all in on Claude specifically (Claude Enterprise lets say) have an extremely distinct smell to them. It is basically one of things have gone quite off the rails and no one seems to know how to clean it up.
I don't see a lot of value in that, but one could supply a lot of semi structured extra information which could aid the readers LLM to answer questions the reader might have with more accuracy / factual grounding.
If you're unsure about spending the time to learn Zig, I really recommend watching the following interview with the creator of Zig https://www.youtube.com/watch?v=iqddnwKF8HQ convinced me more than any design doc or blogpost could
One explanation would be that more load could mean higher (absolute) variance in queue length, and therefore higher latency especially at higher percentiles. It doesn't work out that way (for reasons that Erlang actually writes about in one of his original works), but it's not an entirely unreasonable intuition.
I'm mostly just surprised the graph starts at 5 seconds for a mean value for all datapoints. I would have assumed it starts much closer to 1s. Which just makes the poll responses even crazier. Who is picking B when you have 25% more capacity than you need?
But I suppose the question is underspecified. How does the load balancer know which systems are busy? What happens to a request if the load balancer routes a request to a busy server?
I thought the same thing. But, should we be surprised about what people believe in these days?
I think that the issue is in part due to the variables. Plotting the mean request time is less intuitive than plotting throughput.
If you plot throughput vs number of servers, it'll be a straight line. And asking people that, I think most would agree on a straight line. But who knows!
> In 2006, SoftwareOnline was sued by The Washington State Attorney General's Office for alleged violations of the Consumer Protection Act after complaints were made about two products called "Registry Cleaner" and "InternetShield".
https://en.wikipedia.org/wiki/Dave_Plummer#:~:text=16%5D-,In...