Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
heresalexandria
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
Pawvis: Control your Mac via camera and train gestures (local FOSS)
(github.com)
2 points
by
heresalexandria
1mo ago
|
1 comments
2.
▲
by
heresalexandria
1mo ago
I made Pawvis because I wanted to use my computer the way we've been promised for decades that we would in the future. Pawvis reads your hand through the webcam and lets your control your mouse, manage windows, and bind custom actions
3.
▲
by
heresalexandria
2mo ago
I'm working on Pawvis, visual hand-tracking mouse & voice control that's open-source and fully local on your Mac (with optional handoff to Codex or Claude Code for complex computer use tasks). We live in the future, and I want
4.
▲
Vision based touch-less macOS control via hands (FOSS)
(pawvis.app)
2 points
by
heresalexandria
2mo ago
|
1 comments
5.
▲
by
heresalexandria
2mo ago
Built a tool to let you control your Mac cursor including click, right click, and scroll with your hands a la Minority Report. Beta voice commands also involve control via voice, including invoking computer use via Claude Code or Codex CLI.
6.
▲
by
heresalexandria
2mo ago
Been using this and it’s really a phenomenal model, big step up in quality and capability. This feels like the serious inflection point for high quality full length feature film productions using this tech.
7.
▲
by
heresalexandria
2mo ago
That's fair and I agree with this framing.
8.
▲
by
heresalexandria
2mo ago
I'm not suggesting that fine-tuned models don't have their place, all I'm saying is that the constant drumbeat of "cheap model X beats more expensive model Z" completely misses that the more expensive model is capab
9.
▲
by
heresalexandria
2mo ago
Ahh that makes sense, I was under the impression that OS dictation used exclusively on device models for dictation now but you're right - they do still send this data for transcription. That definitely changes my take and I'm inte
10.
▲
by
heresalexandria
2mo ago
This continuous cycle of fine-tuned open models beating frontier on (often vaguely labeled/defined) benchmarks doesn't provide an accurate comparison to the expanding generalized capabilities of the SoTA, which makes them effectiv
11.
▲
by
heresalexandria
2mo ago
I don't think get it unless I'm missing something - how is this different from just using the built in Mac OS dictation feature? At first I thought it was going to be a CLI/package to interface with that API which sounded int
12.
▲
NYC Roam: 3D world with real transit and building data
(nycroam.com)
5 points
by
heresalexandria
3mo ago
|
2 comments
13.
▲
by
heresalexandria
3mo ago
An explorable Manhattan built from real map & transit data. Ride any subway, bus, or bike along its true routes & stops, or take helicopter mode and fly above the city to explore. Walk up to any building's address plaque for it
14.
▲
by
heresalexandria
3mo ago
This is super fun! Could be neat to add keyboard controls for auto routing (i.e. select plane number n and auto route it to strip x or have it fly a go-around to wait). Would also be sick if you added models of real world airports to play.
15.
▲
by
heresalexandria
3mo ago
My impression was that they made this editor with Fable, and its JSON project structure would only serve well for manipulation by lesser models.
16.
▲
by
heresalexandria
3mo ago
Cool concept, will try it out! I've had decent results with computer use operating conventional editing tools, but being able to directly edit JSON project files is a solid optimization and opens up a lot of opportunities with things l
17.
▲
by
heresalexandria
3mo ago
For sure, I appreciate your comment - this is a tough crowd, but it's their loss.
18.
▲
by
heresalexandria
3mo ago
That's fair, I do agree that you don't need a harness or ultra-high thinking mode for many problems. Many folks evaluate without those things on a task that would benefit from them leading to the sort of attitudes in this article
19.
▲
by
heresalexandria
3mo ago
Never said I was bad at math, but I am aware of the fact that computers can do math better and faster than me - and with our powers combined...
20.
▲
by
heresalexandria
3mo ago
This was pretty cool, knocking out a problem that the best minds in maths couldn't for 80 years: https://openai.com/index/model-disproves-discrete-geometry-c... Also this is a remarkable (and realistic) evaluation
21.
▲
by
heresalexandria
3mo ago
Offline models are becoming increasingly more capable - merely a few years ago it would've been unthinkable to run the LLM I have on my phone even on my MacBook Pro. Are you suggesting that losing electricity in the modern age (entirel
22.
▲
by
heresalexandria
3mo ago
It shouldn't be a surprise that the baseline for "best" shifts as better tech comes out, but that doesn't make dated models any less capable than they were when they came out. Skeptics continue to move the goalposts on w
23.
▲
by
heresalexandria
3mo ago
That's exactly what the people in my orbit and whom I'm watching are doing, and some of their outputs are fueling the excitement. If you aren't seeing remarkable things being done with this tech, I'd argue you aren'
24.
▲
by
heresalexandria
3mo ago
Qwen is a lightweight locally hosted model that's many months behind the SoTA available from the big three - while the crowd here (myself included) is excited for locally hosted models to catch up to the usable baseline, regardless of
25.
▲
by
heresalexandria
3mo ago
This sounds like you may be using subpar models and/or tools - have you had this experience using Codex with GPT-5.5 on at least "high" reasoning or on Claude Code using Opus 4.8 (both with ability to browse web and sufficien
26.
▲
by
heresalexandria
3mo ago
Totally agree that the lack of a common base of evaluation is terrible for the discussion, and benchmaxing only contributes to this. The only way to get a sense for these systems is to use them on things you know well, and everyone knows di
27.
▲
by
heresalexandria
3mo ago
Did you try providing it documentation for the respective formats (via browsing/tool use or input to the prompt)? And were you using a modern thinking model from Anthropic or OpenAI? The crucial breakdown here sounds like either lack o
28.
▲
by
heresalexandria
3mo ago
Something a lot of folks struggling with these systems don't get is that the instruction and management of them is often quite important - just because they're capable doesn't mean they're mind readers. Most of the skept
29.
▲
by
heresalexandria
3mo ago
They're literally doing novel research. The smartest mathematicians in the world couldn't solve Erdős' planar unit distance problem for 80 years, and OpenAI's models knocked that out a couple months ago. This stuff is mo
30.
▲
by
heresalexandria
3mo ago
The same attitude has been directed at points through history for people "who depend on the internet," "who depend on computers," and "who depend on machines." I was told growing up "you won't always
More ›