Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
akadeb
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
31.
▲
by
akadeb
1y ago
I agree, it's still pricy. The cost works out better with `gpt-4o-mini-realtime-preview-2024-12-17`. Yep its constrained to the system prompt but I pass in conversation history with each new session to keep it relevant. It also support
32.
▲
by
akadeb
1y ago
The willow team has iterated fast. I think ESP-IDF is more advanced and using Arduino makes it easier for people to jump on and tinker with Speech-to-Speech AI which is why i created this repo
33.
▲
by
akadeb
1y ago
Currently our device is a toy accessory. And for children we are strictly focusing on `Story mode`. Where adventure stories / fairy tales feel more engaging. I think there's value in getting the AI to create epic stories consisten
34.
▲
by
akadeb
1y ago
Thanks for the feedback. I have attached the raw unedited video here: https://drive.google.com/file/d/1kEmbVInvUrYFwjddyGL8Rz03c0N... (sorry the video is a bit long ~5min with some intro about my company :-)
35.
▲
by
akadeb
1y ago
Yeah the way I am handling this is turn detection which feels unnatural. I like how Livekit handles turn detection with a small model[0][1] [0] https://www.youtube.com/watch?v=EYDrSSEP0h0 [1] https://docs.livekit.i
36.
▲
by
akadeb
1y ago
The long connections are ultimately handled by Deno Edge so the site isn't used there. The NextJS frontend (which also could be an iOS/Android app) helps provide an interface to select character, create AI characters, set ESP32 vo
37.
▲
by
akadeb
1y ago
For parents we added a `Story mode` option (similar to Yoto toy / Toniebox). The idea is: the AI crafts a story and invites the child to craft the story together in a more engaging way. The story prompt keeps the story focused and in s
38.
▲
by
akadeb
1y ago
Currently we connect to a Wifi network to reach the Deno edge server. Some popular toys doing it: Yoto, Toniebox
39.
▲
by
akadeb
1y ago
thank you! The nextjs frontend is to set things like device volume, selecting which character you are interacting with, viewing conversation history etc. I just tried it and for a 15 minute chat, it's roughly 20c. Roughly 570 input
40.
▲
Show HN: I open-sourced my AI toy company that runs on ESP32 and OpenAI realtime
(github.com)
177 points
by
akadeb
1y ago
|
94 comments
41.
▲
Show HN: I Open-Sourced My AI Toy–ESP32, OpenAI Realtime API on Deno, Supabase
(github.com)
2 points
by
akadeb
1y ago
|
0 comments
42.
▲
by
akadeb
2y ago
Saving this for my next trip
43.
▲
by
akadeb
2y ago
We are all LLM-generated
44.
▲
by
akadeb
2y ago
Thanks for this feedback. I hear what you are saying. If I understand correctly its like LLMs are a black box doctors dont understand and while it talks back in a friendly voice it can cause harm if it says something awful. While this is no
45.
▲
by
akadeb
2y ago
Thanks for the useful feedback. I just updated our product page with the subscription price. Since you can use docker for our backend, you can self-host your own service with our hardware. You can use our subscription only if you want us to
46.
▲
by
akadeb
2y ago
Those books made good props for the demo but I should add that we aren't exactly experts in it either. I have a background in computer engineering and math at UIUC and my cofounder has a background in data science and machine learning
47.
▲
by
akadeb
2y ago
thanks for your words of encouragement! > A lot of the subscription based pull ins could be replaced by networking into a machine running whisper/ollama etc anyway. could you clarify this point? I think local LLMs are great for us t
48.
▲
by
akadeb
2y ago
Neat! How can I help you? I think our devkit's components will help you get started. Here is the main.cpp code for this ``` /* * @file streams-i2s-webserver_wav.ino * * This sketch reads sound data from I2S. The result is prov
49.
▲
by
akadeb
2y ago
I have yet to watch AI (2001). I was very inspired by asimov's book the positronic man where the main family robot Andrew NDR has a positronic brain and finds himself "feeling" human emotions.
50.
▲
by
akadeb
2y ago
will I think the devkit on our website will speed up your project. And if you would like extra parts (like a servo) for a moving teddy head, I am happy to send it to you free of charge. email me at akash at starmoon dot app
51.
▲
by
akadeb
2y ago
Dean I think that's important feedback. It's important for us to be platform agnostic. Like Junru stated, we are supporting ESP32 models: WROOM/WROVER and S3. But down the line would like to support rasp pi 4B, 5, zero etc. W
52.
▲
by
akadeb
2y ago
I think I heard this on a joe rogan podcast. Will add this to our roadmap along with DMT as an alternative.
53.
▲
by
akadeb
2y ago
For our device currently, conversational audio is only used by the LLM when a button is pressed like a "push-to-talk" feature. In the future we want to either make it a capacitive touch sensor or listen for a wake word, like "
54.
▲
by
akadeb
2y ago
Oh I see that. what inspection scenario you are thinking of? Have you faced something similar as the inspector and needed structured data immediately?
55.
▲
by
akadeb
2y ago
hey there I am one of the founders. this is our project which we are trying to grow through open-source. I agree our wording can be better so its backed by data and not just a marketing stint. > Do you have a single psychologist on your
56.
▲
Show HN: Build a conversational AI companion device for $10
(github.com)
11 points
by
akadeb
2y ago
|
1 comments
57.
▲
by
akadeb
2y ago
Could you clarify what running models multi-tenant means?
58.
▲
Tool Use (function calling)
(docs.anthropic.com)
222 points
by
akadeb
3y ago
|
99 comments
59.
▲
by
akadeb
3y ago
Hi HN! I recently tried to build a habit to learn more about Banking and Finance with ChatGPT. However I felt frustrated after having to go to ChatGPT and typing “New lesson” in the chat daily. To make it easy to build a daily habit of lear
60.
▲
Show HN: I made an app to email daily newsletters using your own AI prompt
(onbloom.app)
5 points
by
akadeb
3y ago
|
1 comments
More ›