4 ms·
Is a model so huge that’s only at the level of GPT 3.5 actually good? That seems incredibly inefficient to me.
by OkGoDoIt 3y ago
Is a model so huge that’s only at the level of GPT 3.5 actually good? That seems incredibly inefficient to me.
- fwlr 3y agoOpenAI is valued at 90 billion and all they do is make GPT; Twitter is valued at 40 billion and this was essentially a vanity side-project by a cowboy CEO. Presuming that benchmarks and general “it’s about the level of 3.5” is accurate, it’s inefficient, but not incredibly inefficient imho
- pelorat 3y ago> Twitter is valued at 40 billion WAS vaulued at 44B. Now? Maybe 5 billion.
- wongarsu 3y agoLast I heard they lost 15% of their users, so let's call it 36 billion.
- mceachen 3y agoMore like $13b. https://arstechnica.com/tech-policy/2024/01/since-elon-musks-twitter-purchase-firm-reportedly-lost-72-of-its-value/ https://arstechnica.com/tech-policy/2024/01/since-elon-musks...
- wraptile 3y agoTwitter didn't have direct competitors other than Mastodon when it was taken at 44B. Now there's Threads, Bluesky and bigger Mastodon.
- _ea1k 3y agoHonestly, none of those look like meaningful competitors at the moment.
- squigglydonut 3y agoNone of these matter
- dilyevsky 3y agoThey weren't even 44B when elon took the keys - he specifically tried to back out of the deal because 44B was insane peak '21 asset bubble price. In truth they were probably like 10-15B at that moment. And now that bunch of advertisers left due to we know who it's probably about 10B
- Lewton 3y agotwitter was valued around 30 billion when musk tried getting out of buying it (then the market cap went up when it became clear that he would be forced to pay full price)
- alvah 3y agoLOL @ $5 billion, but if it that was the valuation, you'd be making parent's point stronger.
- thekhatribharat 3y agoxAI is a separate entity, and not a X/Twitter subsidiary.
- xcv123 3y agoAccording to their benchmarks it is superior to GPT-3.5
- cma 3y agoSince it is MoE, quantized it could be able to run on cheaper hardware with just consumer networking inbetween instead of needing epyc/xeon levels of PCI-e lanes, nvlink, or infiniband type networking. Or it could even run with people pooling smaller systems over slow internet links.
- drak0n1c 3y agoIt’s designed to be actively searching real-time posts on X. Apples and oranges.
- hn_20591249 3y agoThe data pipeline isn't included in this release, and we already know it is a pretty simple RAG pipeline using qdrant, https://twitter.com/qdrant_engine/status/1721097971830260030 https://twitter.com/qdrant_engine/status/1721097971830260030. Nothing about using data in "real time" predicates that the model parameters need to be this large, and is likely quite inefficient for their "non-woke" instructional use-case.
- lmeyerov 3y agoAgreed. We have been building our real-time GPT flows for news & social as part of Louie.AI, think monitoring & and investigations... long-term, continuous training will become amazing, but for the next couple of years, most of our users would prefer GPT4 or Groq vs what's here and much smarter RAG. More strongly, the interesting part is how the RAG is done. Qdrant is cool but just a DB w a simple vector index, so nothing in Grok's release is tech we find relevant to our engine. Eg, there is a lot of noise in social data, and worse, misinfo/spam/etc, so we spend a lot of energy on adverserial data integration. Likewise, queries are often neurosymbolic, like on a data range or with inclusion/exclusion criteria. Pulling the top 20 most similar tweets to a query and running through a slow, dumb, & manipulated LLM would be a bad experience. We have been pulling in ideas from agents, knowledge graphs, digital forensics & SNA, code synthesis, GNNS, etc for our roadmap, which feels quite different from what is being shown here. We do have pure LLM work, but more about fine-tuning smaller or smarter models, and we find that to be a tiny % of the part people care about. Ex: Spam classifications flowing into our RAG/KG pipelines or small model training is more important to us than it flowing into a big model training. Long-term, I do expect growing emphasis on the big models we use, but that is a more nuanced discussion. (We have been piloting w gov types and are preparing for next cohorts, in case useful on real problems for anyone.)
- pests 3y agoIsn't that... the same thing as search?