3 ms·
That probably plays into it as well. I have yet to see any convincing evidence that contradicts LLMs being mere pattern parrots. My personal benchmark for thes
by colonial 1y ago
That probably plays into it as well. I have yet to see any convincing evidence that contradicts LLMs being mere pattern parrots.
My personal benchmark for these models is writing a simple socket BPF in a Rust program. Even the latest and greatest hosted frontier models (with web search and reasoning enabled!) can only ape the structure. The substance is inevitably wanting, with invalid BPF instructions and hallucinated/missing imports.
- tough 1y agoimho these tools are great i fyou know what you're doing, becasue you know how to smell test the output, but a footgun otherwise. It works great for me, but it is necessarily an aid learning tool more than a full on replacement, someone's still gotta do the thinking part, even if the llm's can cosplay -reasoning- now