Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
dr_blueberry
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
dr_blueberry
14d ago
Working on real-world computer vision demos with a focus on fitness. So far, they are based on ViTPose+ Large pose estimation in the context of different sports. My demos include: - Comparing dancers' sync performing the same choreogra
2.
▲
Show HN: Run open-weight OCR, VLM and vision models behind one API
(vlmrun.com)
5 points
by
dr_blueberry
24d ago
|
0 comments
3.
▲
I made a rock climbing tool using computer vision [video]
(youtube.com)
5 points
by
dr_blueberry
1mo ago
|
2 comments
4.
▲
by
dr_blueberry
1mo ago
I prompted VLM Run’s visual agent Orion to segment all of the blue bouldering holds, and it did a good job! It is interesting that now we can prompt VLMs to segment all of the holds, rather than creating a new dataset from scratch to train
5.
▲
by
dr_blueberry
2mo ago
Doing some initial testing of Gemini ER 2 within the Orion 2 visual agent harness. I'll share a few initial results and chat threads here. I'm impressed with how fast Gemini ER2 is. 1. Crop the segment when adding rice to the rice
6.
▲
Mm-ctx: multimodal context for agents
(colab.research.google.com)
3 points
by
dr_blueberry
5mo ago
|
0 comments