8 ms·
This is incredible. In April I used the standard GPT-4 model via ChatGPT to help me reverse engineer the binary bluetooth protocol used by my kitchen fan to int
by bartman 2y ago
This is incredible. In April I used the standard GPT-4 model via ChatGPT to help me reverse engineer the binary bluetooth protocol used by my kitchen fan to integrate it into Home Assistant.
It was helpful in a rubber duck way, but could not determine the pattern used to transmit the remaining runtime of the fan in a certain mode. Initial prompt here [0]
I pasted the same prompt into o1-preview and o1-mini and both correctly understood and decoded the pattern using a slightly different method than I devised in April. Asking the models to determine if my code is equivalent to what they reverse engineered resulted in a nuanced and thorough examination, and eventual conclusion that it is equivalent. [1]
Testing the same prompt with gpt4o leads to the same result as April's GPT-4 (via ChatGPT) model.
Amazing progress.
[0]: https://pastebin.com/XZixQEM6 https://pastebin.com/XZixQEM6
[1]: https://i.postimg.cc/VN1d2vRb/SCR-20240912-sdko.png https://i.postimg.cc/VN1d2vRb/SCR-20240912-sdko.png (sorry about the screenshot – sharing ChatGPT chats is not easy)
- losvedir 2y agoWow, that is impressive! How were you able to use o1-preview? I pay for ChatGPT, but on chatgpt.com in the model selector I only see 4o, 4o-mini, and 4. Is o1 in that list for you, or is it somewhere else?
- hidelooktropic 2y agoI see it in the mac and iOS app.
- authorfly 2y agoYes, o1-preview is on the list, as is o1-mini for me (Tier 5, early 2021 API user), under "reasoning".
- MattHeard 2y agoIt appeared for me about thirty minutes after I first checked.
- m3kw9 2y agoLikely phased rollout throughout the day today to prevent spikes
- natch 2y ago“Throughout the day” lol. Advanced voice mode still hasn’t shown up. They seem to care more about influencers than paying supporters.
- vidarh 2y agoIt's available for me. Regular paying customer in the UK.
- rovr138 2y ago> lol. It's there for a lot of people already. I can see it on 3 different accounts. Including org and just regular paid accounts.
- taberiand 2y agoIt's my understanding paying supporters aren't actually paying enough to cover costs, that $20 isn't nearly enough - in that context, a gradual roll-out seems fair. Though maybe they could introduce a couple more higher-paid tiers to give people the option to pay for early access
- guiambros 2y agoNot true; it's already available for me, both O1 and O1-mini. It seems they are indeed rolling out gradually (as any company does).
- bartman 2y agoLike others here, it was just available on the website and app when I checked. FWIW I still don’t have advanced voice mode.
- sroussey 2y agoI don’t have either the new model nor the advanced voice mode as a paying user.
- michelsedgh 2y agou do just use this link: https://chatgpt.com/?model=o1-preview https://chatgpt.com/?model=o1-preview
- sroussey 2y agoThat worked. Now can you do that for advanced voice mode??? Pretty please!
- michelsedgh 2y agoHaha I wish, although I saw the other one i forgot its name which makes music for you, now you can ask it for a soundtrack and it gives it back to you in your voice or something like that idk interesting times are ahead for sure!
- fivestones 2y agoWait what is this? Tell me more please
- michelsedgh 2y agoI heard on X suno.com has this feature but couldn’t find it maybe its coming soon? Idk but there are ways u can do it, maybe it was a different service suno is pretty cool tho
- rahimnathwani 2y agoI think they're rolling it out gradually today. I don't see it listed (in the browser, Mac app or Android app).
- accidbuddy 2y agoAvailable on ChatGPT Plus signature or only using the API?
- cft 2y agoit's in my MacOS app, but not in the browser fir the same account
- obmelvin 2y agoThe linked release mentions trusted users and links to the usage tier limits. Looking at the pricing, o1-preview only appears for tier 5 - requiring 1k+ spend and initial spend 30+ days ago edit: sorry - this is for API :)
- romeros 2y agois it better than Claude?
- bartman 2y agoNeither Sonnet nor Opus could solve it or get close in a minimal test I did just now, using the same prompt as above. Sonnet: https://pastebin.com/24QG3JkN https://pastebin.com/24QG3JkN Opus: https://pastebin.com/PJM99pdy https://pastebin.com/PJM99pdy
- hmottestad 2y agoI think this new model is a generational leap above Claude for tasks that require complex reasoning.
- natch 2y agoWay worse than Claude for solving a cipher. Not even 1/10th as good. Just one data point, ymmv.
- antman 2y agosecond is very blurry
- bartman 2y agoWhen you click on the image it loads a higher res version.
- avodonosov 2y agoSolved here: https://news.ycombinator.com/item?id=41525164 https://news.ycombinator.com/item?id=41525164
- jazzyjackson 2y agoIsn't there a big "Share" button at the top right of the chatgpt interface? Or are you using another front end?
- bartman 2y agoIn ChatGPT for Business it limits sharing among users in my org, without an option for public sharing.
- fshbbdssbbgdd 2y agoI often click on those links and get an error that they are unavailable. I’m not sure if it’s openAI trying to prevent people from sharing evidence of the model behaving badly, or an innocuous explanation like the links are temporary.
- arunv 2y agoThey were probably generated using a business account, and the business does not allow public links.
- fshbbdssbbgdd 2y agoIn context, a lot of times it’s clear that the link worked at first (other people who could see it responded) but when I click later, it’s broken.
- coder543 2y agoThe link also breaks if the original user deletes the chat that was being linked to, whether on purpose or without realizing it would also break the link.
- OutOfHere 2y agoEven for regular users, the Share button is not always available or functional. It works sometimes, and other times it disappears. For example, since today, I have no Share button at all for chats.
- baal80spam 2y agoThanks for sharing this, incredible stuff.
- GaggiX 2y agoDid you edit the message? I cannot see anything now in the screenshot, too low resolution
- fwip 2y agoWhat's the incredible part here? Being able to write code to turn hex into decimal?
- fwip 2y agoAlso, if you actually read the "chain of thought" contains several embarrassing contradictions and incoherent sentences. If a junior developer wrote this analysis, I'd send them back to reread the fundamentals.
- CooCooCaCha 2y agoWhat about thoughts themselves? There are plenty of times I start a thought and realize it doesn't make sense. It's part of the thinking process.
- fwip 2y agoWell, it doesn't "correct" itself later. It just says wrong things and gets the right answer anyways, because this encoding is so simple that many college freshmen could figure it out in their heads. Read the transcript with a critical eye instead of just skimming it, you'll see what I mean.
- soheil 2y agoGreat progress, I asked GPT-4o and o1-preview to create a python script to make $100 quickly, o1 came up with a very interesting result: https://x.com/soheil/status/1834320893331587353 https://x.com/soheil/status/1834320893331587353
- fsndz 2y ago> Asking the models to determine if my code is equivalent to what they reverse engineered resulted in a nuanced and thorough examination, and eventual conclusion that it is equivalent. Did you actually implement to see if it works out of the box ? Also if you are a free users or accepted that your chats should be used for training then maybe o1 is was just trained on your previous chat and so now knows how to reason about that particular type of problems
- bongodongobob 2y agoThat's not how LLM training works.
- fsndz 2y agoso it is impossible to use the free user chats to train models ??????
- bartman 2y agoThat is an interesting thought. This was all done in an account that is opted out of training though. I have tested the Python code o1 created to decode the timestamps and it works as expected.
- jeffpeterson 2y agoVery cool. It gets the conclusion right, but it did confuse itself briefly after interpreting `256 * last_byte + second_to_last_byte` as big-endian. It's neat that it corrected the confusion, but a little unsatisfying that it doesn't explicitly identify the mistake the way a human would.
- avodonosov 2y agoThe screenshot [1] is not readable for me. Chrome, Android. It's so blurry that I cant recognize a single character. How do other people read it? The resolution is 84x800.
- rovr138 2y agoWhen I click on the image, it expands to full res, 1713x16392.3
- deathanatos 2y ago> it expands to full res, 1713x16392.3 Three tenths of a pixel is an interesting resolution… (The actual res is 1045 × 10000 ; you've multiplied by 1.63923 somehow…?)
- rovr138 2y agoI agree, But it’s what I got when I went to Inspect element > hover over the image Size it expanded to vs real image size I guess
- Jerrrrrrry 2y agoPixels have been "non-real" for a long time.
- deathanatos 2y agoIn some contexts. In this context (a PNG), they're very real.
- Jerrrrrrry 2y agoThis context is the moreso the browser, complete with it's own sub-pixels, aliasing, simulated/real blurring, zooming, etc. But file-format context, yes, PNG, BMP, and TFF are the real lossless image kingpins.
- 2y ago
- guiambros 2y agoFYI, there's a "Save ChatGPT as PDF" Chrome extension [1]. I wouldn't use on a ChatGPT for Business subscription (it may be against your company's policies to export anything), but very convenient for personal use. https://chromewebstore.google.com/detail/save-chatgpt-as-pdf/ccjfggejcoobknjolglgmfhoeneafhhm https://chromewebstore.google.com/detail/save-chatgpt-as-pdf...
- andraz 2y agoWhat is the brand of the fan? Same problem here with proprietary hood fan...
- bartman 2y agoInVENTer Pulsar
- smusamashah 2y agoWhat if you copy the whole reasoning process example provided by OpenAI, use it as a system prompt (to teach how to reason), use that system prompt in Claude, got4o etc?
- azeirah 2y agoIt might work a little bit. It's like doing few shot prompting instead of training it to reason.
- fragmede 2y agoI'm impressed. I had two modified logic puzzles where ChatGPT-4 fails but o1 succeeds. The training data had too many instances of the unmodified puzzle, so 4 wouldn't get it right. o1 manages to not get tripped up by them. https://chatgpt.com/share/66e35c37-60c4-8009-8cf9-8fe61f57d30c https://chatgpt.com/share/66e35c37-60c4-8009-8cf9-8fe61f57d3... https://chatgpt.com/share/66e35f0e-6c98-8009-a128-e9ac677480fd https://chatgpt.com/share/66e35f0e-6c98-8009-a128-e9ac677480...
- 8thcross 2y agoThis is a brilliant hypothesis deconstruction. I am sure others will now be able to test as well and this should confirm their engineering.
- deleted 2y ago[deleted]