4 ms·
This article is about things which aren't limitations anymore! You are applauding it as pushback for pushback's sake, but it's an article about limitations in
by comeonbro 2y ago
This article is about things which aren't limitations anymore!
You are applauding it as pushback for pushback's sake, but it's an article about limitations in biplane construction, published after we'd already landed on the moon.
- suddenlybananas 2y agoIs there any evidence that these fundamental issues with compositionality have been resolved or are you just asserting it? Has the paper been replicated with a CoT model and had a positive result?
- dr_dshiv 2y agoWell, yes — because modern models can solve all the examples in the article. The theory of compositionality is still an issue, but the evidence for it recedes. I think most of the issue comes from the challenge of informational coherence. Once incoherence enters the context, the intelligence drops massively. You can have a lot of context and LLMs can maintain coherence— but not if the context itself is incoherent. And, informationally, it is just a matter of time before a little incoherence gets into a thread. This is why agents have so much potential—being able to separate out separate threads of thought in different context windows reduces the likelihood of incoherence emerging (vs one long thread). Actually, maybe “cybernetic ecologies” are closer to what I mean than “agents.” See Anthropic’s “Building Effective Agents.” https://www.anthropic.com/research/building-effective-agents https://www.anthropic.com/research/building-effective-agents
- anon84873628 2y ago>I think most of the issue comes from the challenge of informational coherence. Once incoherence enters the context, the intelligence drops massively. You can have a lot of context and LLMs can maintain coherence— but not if the context itself is incoherent. As a non-expert, part of my definition of intelligence is that the system can detect incoherence, a.k.a reject bullshit. LLMs today can't do that and will happily emit bullshit in response. Maybe the "gates" in the "workflows" discussed in the Anthropic article are a practical solution to that. But that still just seems like inserting human intelligence into the system for a specific engineering domain; not a general solution.