Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
kraddypatties
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
kraddypatties
5mo ago
I think the only thing that gives me pause is the fact that they SFT on Opus 4.5 explanations as a pertaining step. But, generally I agree, especially given the auto encoder is only seeing a single token activation!
2.
▲
by
kraddypatties
5mo ago
I believe that’s _part_ of the point (or at least a side-effect) of the KL divergence loss term they have on the AV. That and training stability.
3.
▲
by
kraddypatties
7mo ago
I can believe that in the long run. Does the agent have access to arxiv (a brief skim of the README didn't have an answer)? If not, it could be that the current approach of relying on the model's weights only is resulting in the p
4.
▲
by
kraddypatties
7mo ago
Hm, that's fair. It does feel like there's low hanging fruit in combining "old school" methods for conducting a hyperparameter sweep efficiently _with_ the higher level architecture edit ability of Autoresearch. Probably
5.
▲
by
kraddypatties
7mo ago
I feel like most of this recent Autoresearch trend boils down to reinventing hyper-parameter tuning. Is the SOTA still Bayesian optimization when given a small cluster? It was ~3 years ago when I was doing this kind of work, haven't ke
6.
▲
by
kraddypatties
8mo ago
Glad you liked it! Currently the avatar does it based on the text, which maps the incoming audio to one of our emotion codes, biasing the generation to that emotion. It's not foolproof, but we've found it works pretty well in prac
7.
▲
Show HN: Emotional photoreal AI humans at $0.06 / min
2 points
by
kraddypatties
8mo ago
|
4 comments
8.
▲
by
kraddypatties
9mo ago
I'm interpreting this as "uv was built off of years of PEPs", which is true; that being said the UX of `uv` is their own, and to me has significantly reduced the amount of time I spend thinking about requirements, modules, et
9.
▲
by
kraddypatties
10mo ago
Running into "no healthy upstream" when navigating to the link -- hug of death maybe?
10.
▲
by
kraddypatties
10mo ago
been lurking for most of my adult life (and it shows :-)) Thanks HN! You make me smarter every (other) day.
11.
▲
by
kraddypatties
10mo ago
Thanks for trying it out! Yea that latency makes sense; "listening" includes turn detection and STT, "thinking" LLM + TTS _and then_ our model, so the pipeline latency stacks up pretty quick. The actual video model start
12.
▲
Show HN: Realtime, expressive AI personas that you can video call
(playground.keyframelabs.com)
4 points
by
kraddypatties
10mo ago
|
2 comments
13.
▲
by
kraddypatties
11mo ago
We've been tinkering with building realtime talking head models (avatar models, etc.) for a while now, and finally have something that works (well enough)! Operates at ~2x realtime on a 4090, significantly faster than that on enterpris