For sure, something happened. Grok 3 was awesome to work with. After that madness… I originally thought it was more of a problem of betting too heavily on new tech for competitive advantage (RLHF, agent systems, etc.) and accepting worse results in the process. But in the meantime, the usefulness of the LLM has gone downhill. Way slower, way more steps, and you're getting something worse than Grok 3—at least in my day-to-day experience :(
Yep also a grok 3 supporter. I actually liked GPT-4 Turbo and Claude 3, and have found each successive update substantially more useless. Grok 3 came out and it was a bit of that original magic... but seems to have went the way of the other models.
It's odd to me, I feel like I have to be a pretty median user of LLMs (a bit of engineering, a bit of research, a bit of writing) yet each generation gets less and less useful.
I think they all focus way too much on finding a 'right' answer. I like LLMs for their ability to replicate divergent thinking. If I want a 'right' answer, I'm not going to even have an LLM in my toolbox :/
https://www.businessinsider.com/elon-musk-xai-layoffs-data-a...
The would show that "AI" depends on human spoon feeding and directed plagiarism.