6 ms·
This is almost certainly what they're doing and rebranding the original o3 model as "o3-pro"
by CSMastermind 1y ago
This is almost certainly what they're doing and rebranding the original o3 model as "o3-pro"
- behnamoh 1y ago> rebranding the original o3 model as "o3-pro" interesting take, I wouldn't be surprised if they did that.
- anticensor 1y ago-pro models appear to be a best-of-10 sampling of the original full size model
- Szpadel 1y agohow do you sample it behind the scenes? usually best of X means you generate X outputs and you choose best result. if you could do this automatically, it would be game changer as you could run top 5 best models in parallel and select best answer every time but it's not practical because you are the bottleneck as you have to read all 5 solutions and compare them
- joshstrange 1y agoI think the idea is they use another/same model to judge all the results and only return the best one to the user.
- anticensor 1y agoI think the idea is they just feed each to the RLHF reward model used to train the model and return the most rewarded answer.
- anticensor 1y ago> if you could do this automatically, it would be game changer as you could run top 5 best models in parallel and select best answer every time remember they have access to the RLHF reward model, against which they can evaluate all N outputs and have the most "rewarded" answer picked and sent
- spott 1y agoI believe it is a majority vote kinda thing, rather than a best single result.
- tedsanders 1y agoNope, not what we’re doing. o3 is still o3 (no nerfing) and o3-pro is new and better than o3. If we were lying about this, it would be really easy to catch us - just run evals. (I work at OpenAI.)
- arresin 1y agoNot quantized?
- tedsanders 1y agoNot quantized. Weights are the same. If we did change the model, we'd release it as a new model with a new name in the API (e.g., o3-turbo-2025-06-10). It would be very annoying to API customers if we ever silently changed models, so we never do this [1]. [1] `chatgpt-4o-latest` being an explicit exception
- ant6n 1y agoIt was definitely annoying when o1 disappeared over night, my impression is that was better at some tasks than o3.
- thegeomaster 1y agoGoogle could at least learn something from this attitude, given their recent 03-25 -> 05-06 model alias switcharoo with 0 notice :)
- johnb231 1y agoThat is a preview / beta model with no expectation of stability. Google did nothing wrong there. No one should be using a preview model in production.
- thegeomaster 1y agoHard disagree. Of course technically they didn't do anything explicitly against the public guidance (the checks and balances would never let them), but naming a model with a date very strongly implies immutability. It's the same logic of why UB in C/C++ isn't a license to do whatever the compiler wants. We're humans and we operate on implications, common-sense assumptions and trust.
- mliker 1y agoWhere are you getting this information? What basis do you have for making this claim? OpenAI, despite its public drama, is still a massive brand and if this were exposed, would tank the company's reputation. I think making baseless claims like this is dangerous for HN
- beering 1y agoI think Gell-Mann amnesia happens here too, where you can see how wrong HN comments are on a topic you know deeply, but then forget about that when reading the comments on another topic.