Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
ozenhati
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
by
ozenhati
2y ago
100%. we've found that llama-3.3-70b-versatile and qwen-qwq-32b perform exceptionally well with reliable function calling. we had recognized the need for this and our engineers partnered with glaive ai to create fine tunes of llama 3.0
2.
▲
by
ozenhati
2y ago
can you reach out to us via live chat on console.groq.com with your organization id?
3.
▲
by
ozenhati
2y ago
do you happen to be trying this out on free tier right now? because our rate limits are at 6k tokens per minute on free tier for this model, which might be what you're running into.
4.
▲
by
ozenhati
2y ago
amazing! and yes, we'll have maverick available today. the reason we limit ctx window is because demand > capacity. we're pretty busy with building out more capacity so we can get to a state where we give everyone access to lar
5.
▲
by
ozenhati
2y ago
hi! i work @ groq and just made an account here to answer any questions for anyone who might be confused. groq has been around since 2016 and although we do offer hardware for enterprises in the form of dedicated instances, our goal is to m