Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
varunvummadi
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
Giga Launches Realtime Hallucination Correction
(giga.ai)
2 points
by
varunvummadi
5mo ago
|
0 comments
2.
▲
Phi-3 Technical Report
(arxiv.org)
411 points
by
varunvummadi
2y ago
|
130 comments
3.
▲
by
varunvummadi
2y ago
The easiest is to use vllm ( https://github.com/vllm-project/vllm ) to run it on a Couple of A100's, and you can benchmark this using this library ( https://github.com/EleutherAI/lm-evaluation-ha
4.
▲
by
varunvummadi
2y ago
It beats the old GPT4 version in lmsys benchmark you can check it out here https://huggingface.co/spaces/lmsys/chatbot-arena-leaderboar... but Command R is commercially licensed We can assume that mistral will do
5.
▲
by
varunvummadi
2y ago
Not sure trying to download the torrent and checking it out
6.
▲
by
varunvummadi
2y ago
They Just announced their new model on Twitter, which you can download using torrent
7.
▲
Mistral AI Launches New 8x22B MOE Model
(twitter.com)
379 points
by
varunvummadi
2y ago
|
153 comments
8.
▲
by
varunvummadi
3y ago
Should appreciate the dedication of the person who made this website, haha
9.
▲
Is Mistral Large Trained on Open AI Synthetic Data?
(twitter.com)
2 points
by
varunvummadi
3y ago
|
0 comments
10.
▲
by
varunvummadi
3y ago
So please let me know if I am wrong are you guys running a batch size of 1 in 500 GPU's? then why are the responses almost instant if you guys are using batch size 1 and also when can we expect bring your own fine tuned models kind of
11.
▲
Mistral AI launches Mixtral-Next
(chat.lmsys.org)
204 points
by
varunvummadi
3y ago
|
49 comments
12.
▲
Giga ML (YC S23) Is Hiring
(ycombinator.com)
1 points
by
varunvummadi
3y ago
13.
▲
by
varunvummadi
3y ago
That is true cursor extension is awesome