Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
xenova
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
Bonsai 27B: A 27B-Class model that runs on a phone
(prismml.com)
706 points
by
xenova
3mo ago
|
250 comments
2.
▲
1-Bit and Ternary Bonsai Image 4B: Image Generation for Local Devices
(prismml.com)
3 points
by
xenova
4mo ago
|
0 comments
3.
▲
by
xenova
5mo ago
yep! :) https://huggingface.co/spaces/webml-community/bonsai-ternary...
4.
▲
ML-intern: open-source ML engineer that reads papers, trains and ships models
(github.com)
3 points
by
xenova
5mo ago
|
0 comments
5.
▲
by
xenova
1y ago
This demo runs Voxtral-Mini-3B, a new audio language model from Mistral, enabling state-of-the-art audio transcription directly in your browser. Everything runs locally, meaning none of your data is sent to a server (and your transcripts ar
6.
▲
by
xenova
2y ago
We have released a bunch of speech recognition demos (using whisper, moonshine, and others). For example: - https://huggingface.co/spaces/Xenova/whisper-web - https://huggingface.co/spaces/Xen
7.
▲
Kokoro WebGPU: Real-time text-to-speech 100% locally in the browser
(huggingface.co)
227 points
by
xenova
2y ago
|
53 comments
8.
▲
by
xenova
2y ago
It took some time, but we finally got Kokoro TTS (v1.0) running in-browser w/ WebGPU acceleration! This enables real-time text-to-speech without the need for a server. Looking forward to your feedback!
9.
▲
by
xenova
2y ago
NPM package: https://www.npmjs.com/package/kokoro-js GitHub: https://github.com/hexgrad/kokoro
10.
▲
by
xenova
2y ago
For those interested in learning more, the source code is available on GitHub: https://github.com/huggingface/transformers.js-examples/tree...
11.
▲
Transformers.js v3: WebGPU Support, New Models and Tasks, and More
(huggingface.co)
1 points
by
xenova
2y ago
|
0 comments
12.
▲
SAM 2: Segment Anything in Images and Videos
(github.com)
824 points
by
xenova
2y ago
|
147 comments
13.
▲
by
xenova
2y ago
It uses OpenAI's set of whisper models, which support multilingual transcription and translation across 100 languages. Since the models run entirely locally in your browser (thanks to Transformers.js), no data leaves your device! Huge
14.
▲
by
xenova
2y ago
We have some other WebGPU demos, including: - WebGPU embedding benchmark: https://huggingface.co/spaces/Xenova/webgpu-embedding-benchm... - Real-time object detection: https://huggingface.co/spaces
15.
▲
by
xenova
2y ago
Odd, the links seem to work for me. What error do you see? Can you try on a different network (e.g., mobile)?
16.
▲
by
xenova
2y ago
We’ve put out a ton of demos that use much smaller models (10-60 MB), including: - (44MB) In-browser background removal: https://huggingface.co/spaces/Xenova/remove-background-web . (We also put out a WebGPU versio
17.
▲
by
xenova
3y ago
The 8-bit quantized version of the RMBG-v1.4 model is ~45MB, which makes it perfect for in-browser usage (it even works on mobile)! Link to model: https://huggingface.co/briaai/RMBG-1.4
18.
▲
PhotoMaker: Customizing Realistic Human Photos via Stacked ID Embedding
(github.com)
1 points
by
xenova
3y ago
|
0 comments
19.
▲
Mamba: New SSM arch with linear-time scaling that outperforms Transformers
(github.com)
6 points
by
xenova
3y ago
|
2 comments
20.
▲
by
xenova
3y ago
Paper: https://arxiv.org/abs/2312.00752 Models: https://huggingface.co/state-spaces
21.
▲
by
xenova
3y ago
Hi everyone, Joshua from Hugging Face (and the creator of Transformers.js) here. Starting with embeddings, we hope to simplify and improve the developer experience when working with embeddings. Supabase already has great support for storage
22.
▲
Count tokens used by GPT-4 and Llama for large texts (> 50k characters)
(huggingface.co)
2 points
by
xenova
3y ago
|
1 comments
23.
▲
by
xenova
3y ago
This web-app fixes the two main problems of OpenAI's tokenizer playground: (1) being capped at 50k characters, and (2) not supporting GPT-4/GPT-3.5 tokenizers. Everything runs in-browser thanks to Transformers.js.
24.
▲
Making real-time ML-powered web games with Transformers.js
(huggingface.co)
2 points
by
xenova
3y ago
|
1 comments
25.
▲
by
xenova
3y ago
Demo: https://huggingface.co/spaces/Xenova/doodle-dash Source code: https://github.com/xenova/doodle-dash
26.
▲
MMS: Scaling Speech Technology to 1000 languages demo
(huggingface.co)
1 points
by
xenova
3y ago
|
0 comments
27.
▲
by
xenova
3y ago
Whisper Web is a web-app which allows you to run OpenAI's whisper models directly in your browser, with no need for a server. This comes with the release of Transformers.js v2.2.0, which now supports multilingual transcription and tran
28.
▲
Whisper Web: ML-powered speech recognition in the browser
(twitter.com)
3 points
by
xenova
3y ago
|
1 comments
29.
▲
by
xenova
4y ago
Haha very interesting! I assume it's because that type of image is only found on computer screens, so, the model thinks the grass "contributes to it's idea of what a computer screen is". ... and of course, the library on
30.
▲
by
xenova
4y ago
Once ONNX runtime releases their WebGPU backend, we will add support for it! :) It should also be noted that browser support for it isn’t very high at the moment… so, unfortunately, we are stuck with WASM (CPU) for now.
More ›