mvanveen an hour ago
There’s this sort of local optima all these models pick which is instantly recognizable and hardly digestible for human consumption. Perhaps too much internal RL against benchmarks during chain of thought? The older non reasoning models had they’re own problems but at this point I can’t get an LLM to summarize data for humans, which is troubling.
As much as I’d like to read this piece it follows suit and I don’t have the patience to try and extract anything valuable from it.
VariousPrograms 21 minutes ago
wewewedxfgdf an hour ago
It is possible to get the AI to write stuff that does not sound like AI you just have to ask it to properly.
addaon an hour ago
jondwillis 43 minutes ago
>Six steps, no jargon. The note beside each one is the precise technical version, for anyone who wants to check the work.
I think this is called "concept leak" or "prompt leak" (please someone correct me, I can't recall at the moment.) It is one of my least-favorite failure modes. I notice it a lot when drafting landing pages, and it eats up a lot of time to edit away.
aappleby an hour ago
HeartStrings an hour ago
gmerc an hour ago