4 ms·
People are focused on the drama but the problem showing is the data. This is the elephant in the room and I am surprised that openai can be that stupid with it.
by _bobm 25d ago
People are focused on the drama but the problem showing is the data. This is the elephant in the room and I am surprised that openai can be that stupid with it.
How can openai do this, what is being claimed, at the scale of their entire userbase? If they do this only for particular sessions then how do they sieve through sessions for the good stuff?
How are sessions stored, how are they processed, how much storage and how much compute is used in these pipelines, how economical is it, how fast are the requirements on the storage on the compute growing as userbase grows and generated data grows.
All these questions are far more pertinent than the navier-stokes, but i can only imagine all at openai doubling down on this "very important" mathematical milestone.
- jonathanstrange 25d agoIt seems completely trivial to feed sessions to their own LLM and ask it to look for various things in them, from detecting problematic use cases to finding interesting mathematical work.
- _bobm 25d agolet's say that they ask a single question for each session they get. they are immediately doubling the compute they need in processing and then post-processing the same session twice. nothing trivial about it. not saying they cannot feed "their own LLM" saying it isn't trivial especially at scale. if you do not trust me try it without the "at scale" part.
- jonathanstrange 25d agoIt's trivial and a solved issue for the companies developing frontier AI models. Obviously, you don't even need AI for searching every prompt every user has ever written to find interesting topics, but you can create automated summaries and use AI on them if you want. There is no "scaling issue" here for companies who are used to processing almost everything that has ever been written anyway. I didn't want to insinuate that it's trivial for small companies or individuals to do big data mining at that scale, sorry if I made that impression.
- _bobm 25d agoI not only think it is not trivial, I know it is not solved.
- jonathanstrange 25d agoI don't trust your judgment, it's in my opinion even hilarious given that we're talking about companies worth almost a trillion dollar (4 trillion in the case of Google). Be that as it may, it was nice chatting with you!
- _bobm 25d agoI also find it hilarious. I wonder what will happen if they don't solve this.