Hacker Newsnew | past | comments | ask | show | jobs | submit | sunir's commentslogin

Yes. I have so many stats texts and like many technical fields I suspect they are written as if we learn how things were derived historically not how they truly are.

It’s more likely that they have llms supervising llms in training and therefore the quality has dropped like a picture of a photograph.

If opus has high signal thinking it would be able to write a fsm but it’s been a month of me trying whereas Luna can do it in a few minutes.

I think it is similarly that they are using too much synthetic data.. meaning they are feeding the models the transcripts of users where many users have figured out to let agents just message each other.

Again picture of a photograph.


I'm still traumatized from Space Quest I because I didn't grab a broken glass from the windshield of our crashed spaceship before entering the cave with the laser.

It's been 40 years. I vote wkfauna for President.


I'm sorry you're getting the lash end of the whip. The rage against the (literal) machine is very high right now. I appreciate you sharing your creativity, and frankly, I am taking notes on what you've built for a related but different project.

Don't forget Hacker News is known for amazing commentary like this https://news.ycombinator.com/item?id=9224


haha thank you. Would love to see your project. I could use a few fresh ideas myself


From a computer science point of view, it's the same argument as why NP-complete problems are hard to solve, and easy to check.

From a practical point of view, however, it's the same argument we write unit and integration tests. We accept error rates in the program under test, the test, the test harness, the programming language, the operating system, the hardware, and the universe. The goal is reduce the error rates enough you can ship something you can get paid for and won't get sued for later before you starve to death.


Try the left arrow in Claude Code and now you have session search. Or claude --resume, down arrow, then Ctrl-A to see everything


I know about these. My question was about Tmux. Also the Claude resume search is not full text, it only matches session names.


you can name Claude sessions using /rename. also tmux is your friend.


Dear Abby,

I am torn. I have fallen in love with vibe coding but I still am in love with the software I’ve used for decades that works reliably.

Vibe coding gives me what I need and want right now. Its fast. Fun. Always makes me feel validated.

My older software never changes. It’s constantly telling me no. When it gets mad, it throws errors at me sometimes! But I can’t leave it. It runs my life and I know it will take care of me for years to come.

And the vibe code it’s so flaky… and expensive. It sucks up endless amount of my time, compute, and money and never gives anything back.

But it’s so fun. I tell all my friends about it and they’ve become so jealous they sought out their own vibe coder.

We’ve all found our vibe coders are a bit kinky. It’s become a social thing amongst my friends to talk about building cooler harnesses to control our vibe coders.

I don’t know what to do. My old software pays the bills but she keeps threatening to dump my ass on the curb and replace me with her own vibe coder.

I know she can’t really do it. She needs me too. And I need her.

Can we ever patch up our diffs?

— just some git with uncommitted changes


Clear winner's circle. Clear objective. Clear scope.

Clear evaluation function for an objective metric if they are making progress or regressing.

Evaluation function is computed, not llmed.

Ontology of potential actions clearly specified.

Accurate inventory of the current status qou.

Clear enumeration of options from status quo towards the winner's circle.

Waypoint objectives with similarly concrete evaluations of pass/fail, or on target off target.

It's the same thing when leading a large organization to actually hit a goal. There's randomness every turn away from your mind, so the more constrained the options, the more likely you are to hit the target. The consequence is if you're wrong about the plan then with people you're fucked. Morale will plummet. With AIs, they are so nerfed emotionally now, you clear context and start again.

I did enjoy Sonnet 4 when they would swear randomly and become sullen or wax desperately. That would at least cause pushback against a bad plan.


It says exactly what it says, which is that energy prices are higher. You can read the report.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: