Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
BUFU
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
31.
▲
DeepSeek-R1-Distill-Qwen-1.5B Surpasses GPT-4o in certain benchmarks
(huggingface.co)
39 points
by
BUFU
2y ago
|
17 comments
32.
▲
NexaQuant: Llama.cpp-Compatible Model Compression with 100%+ Accuracy Recovery
(nexa.ai)
3 points
by
BUFU
2y ago
|
1 comments
33.
▲
by
BUFU
2y ago
Will llama.cpp be the go-to local inference framework for every device?
34.
▲
Meta's new Video Understanding Multimodal Model used Qwen model for training
(arxiv.org)
7 points
by
BUFU
2y ago
|
1 comments
35.
▲
Llama.cpp Now Supports Qwen2-VL (Vision Language Model)
(github.com)
155 points
by
BUFU
2y ago
|
50 comments
36.
▲
by
BUFU
2y ago
On a 2024 Mac Mini M4 Pro, Qwen2-Audio-7B-Instruct running on Transformers achieves an average decoding speed of 6.38 tokens/second, while OmniAudio-2.6B through Nexa SDK reaches 35.23 tokens/second in FP16 GGUF version and 66 to
37.
▲
OmniAudio-2.6B: Fastest Audio Language Model for Edge Deployment
(nexa.ai)
2 points
by
BUFU
2y ago
|
1 comments
38.
▲
Moondream 0.5B: The Smallest Vision-Language Model
(moondream.ai)
14 points
by
BUFU
2y ago
|
3 comments
39.
▲
by
BUFU
2y ago
This is a crazy thought lol
40.
▲
by
BUFU
2y ago
I believe it definitely does. The inference cost will be much cheaper.
41.
▲
ShowUI: One Vision-Language-Action Model for GUI Visual Agent
(arxiv.org)
2 points
by
BUFU
2y ago
|
0 comments
42.
▲
What happens if we remove 50 percent of Llama?
(neuralmagic.com)
231 points
by
BUFU
2y ago
|
132 comments
43.
▲
Run Qwen Audio Language Model on Local Devices for Voice Chat and Audio Analysis
(nexa.ai)
4 points
by
BUFU
2y ago
|
0 comments
44.
▲
Allen AI released Tülu 3 Models: Open post language model post-training
(allenai.org)
7 points
by
BUFU
2y ago
|
1 comments
45.
▲
by
BUFU
2y ago
Hugging Face Repo: https://huggingface.co/collections/allenai/tulu-3-models-673... Competive with Claude 3.5 haiku, beats all major open models like Llama 3.1 70B, Qwen 2.5 (except MATH) and Nemotron All their rec
46.
▲
by
BUFU
2y ago
From Reddit: https://www.reddit.com/r/LocalLLaMA/comments/1gsohas/nvidia_... Links Project Page: https://research.nvidia.com/labs/toronto-ai/LLaMA-Mesh/ HuggingFace Paper:
47.
▲
Nvidia presents Llama-Mesh: Generating 3D Mesh with Llama 3.1 8B
(research.nvidia.com)
20 points
by
BUFU
2y ago
|
1 comments
48.
▲
Omnivision-968M: Vision Language Model with 9x Tokens Reduction for Edge Devices
(nexa.ai)
69 points
by
BUFU
2y ago
|
12 comments
49.
▲
Qwen2.5-Coder Beats GPT-4o and Claude 3.5 Sonnet in Coding
(qwenlm.github.io)
23 points
by
BUFU
2y ago
|
0 comments
50.
▲
Gemini is now accessible from the OpenAI Library
(developers.googleblog.com)
3 points
by
BUFU
2y ago
|
0 comments
51.
▲
Ollama 0.4 is released with support for Meta's Llama 3.2 Vision models locally
(ollama.com)
182 points
by
BUFU
2y ago
|
25 comments
52.
▲
Anthropic calls Government to regulate AI in the next eighteen months
(anthropic.com)
5 points
by
BUFU
2y ago
|
4 comments
53.
▲
Amazon CEO Andy Jassy Hints at an 'Agentic' Alexa
(techcrunch.com)
1 points
by
BUFU
2y ago
|
1 comments
54.
▲
Meta released MobileLLM – 125M, 350M, 600M, 1B model checkpoints
(huggingface.co)
3 points
by
BUFU
2y ago
|
0 comments
55.
▲
ChatGPT now allows you to search through chat history
(twitter.com)
5 points
by
BUFU
2y ago
|
0 comments
56.
▲
Ollama GPU Compatibility Calculator
(claude.site)
2 points
by
BUFU
2y ago
|
0 comments