Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

What is this supposed to mean exactly? Do you think there are developers at OpenAI or Anthropic whose job is to train these state of the art models to draw pelicans riding bicycles? Like how exactly do you expect them to be doing that anyways? Hiring graphic artists to create SVGs of bike riding pelicans and feeding thousands of them into the model's training set?
 help



Generate two images, get a vision model to judge, then RL reward the better one. And yes, OpenAI employees on twitter bragged about the model's SVG capabilities. So there are clearly people working there who care about it.

You gotta look up how RLHF works before you ask a demanding question like this.

Right, but that's not the crazy part. The crazy part is thinking they do it all specifically for pelicans on bikes.

If that was the case, the models would have been producing near perfect outputs for it a year ago.

Instead they are just training on general SVG generation, which in no way should be viewed as "benchmaxxing".


> Do you think there are developers at OpenAI or Anthropic whose job is to train these state of the art models to draw pelicans riding bicycles?

Yes


Do you really think there aren't data annotator services specifically training for SVG drawing? And that those annotators don't have a rich set of frequently requested icons/graphics/etc. that they review and train on?

I'd be shocked if they didn't myself.


How did you decide to go from people whose jobs it is to optimize an LLM for a very specific benchmark that is mostly a fun curiosity at best... to employing the services of data annotation services for a rich set of icons and graphics?

You really have to go out of your way to completely misrepresent what's being claimed here in order to make such a wildly off-topic reply.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: