Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
ndr_
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
9 ms
·
31.
▲
by
ndr_
2y ago
With gpt-4-32k viewed as deprecated and generally only available through Microsoft Azure until mid next year, this development may be reassuring to some users. Tentative pricing from the website: In: $6.00 / 1M tokens, Out: $18.00 
32.
▲
by
ndr_
2y ago
Different approach, which also works with Llama 3.0 8B: let the LLM write the tool invocation of choice in Python syntax, parse the LLM response with the Python AST package. (PoC'y hackathon demo) code here: https://github.c
33.
▲
by
ndr_
2y ago
The only thing close are the "logprobs": https://cookbook.openai.com/examples/using_logprobs However, commenters around here noted that these have likely not been fine-tuned to correlate with accuracy - for p
34.
▲
by
ndr_
2y ago
Prompts in the background: const systemPrompt = ` Convert the following PDF page to markdown. Return only the markdown with no explanation text. Do not exclude any content from the page. `; For each subsequent page:
35.
▲
by
ndr_
2y ago
There used to be Winpooch Watchguard, based on ClamAV. Stopped using it when it caused Bluescreens. A "Killer" indeed.
36.
▲
by
ndr_
2y ago
Yes, Crowdstrike specifically disabled Linux systems in the past: https://news.ycombinator.com/item?id=41005936
37.
▲
by
ndr_
2y ago
If you're on macOS, check out the MLX community: https://huggingface.co/mlx-community - Ollama not needed, potentially more efficient(?). The Gemma 2 27B model was broken as of a few days ago, but a fix is already comm
38.
▲
by
ndr_
2y ago
I have first noticed logprob fluctuations in GPT-4o. Perhaps the same phenomenon is also going on with Turbo. I din‘t recall specifics but it was naming inconsistencies with variable names, meaning: same variable name got a typo somewhere,
39.
▲
by
ndr_
2y ago
Python here. And like they said, only noticable in the last few weeks.
40.
▲
by
ndr_
2y ago
Low-value, high-effort work. Currently: create a mockup server in C# from a HAR recording. For this specifically, I use my Amazon Bedrock workbench [0] with Claude 3 Opus. But I also use GPT-4 Turbo and GPT-4 Classic a lot [1] (in that orde
41.
▲
by
ndr_
2y ago
If you prefer an existing webapp over an elisp function, https://huggingface.co/spaces/ndurner/oai_chat . Choose Whisper as the model, upload your 25 MB chunks, hit Send, choose GPT-4 Turbo, ask it to clean up, h
42.
▲
by
ndr_
2y ago
ref: https://x.com/stevenheidel/status/1777794809270575135?s=61&t...
43.
▲
by
ndr_
3y ago
This news piece is, first and foremost, about the Model, not the ChatGPT System. (More about the difference between “Model” and “System”: https://ndurner.github.io/antropic-claude-amazon-bedrock ). Not sure what their upgrad
44.
▲
by
ndr_
3y ago
Apparently not. But here: https://huggingface.co/spaces/ndurner/oai_chat
45.
▲
by
ndr_
3y ago
https://huggingface.co/spaces/ndurner/oai_chat (bring your own API key)
46.
▲
by
ndr_
3y ago
Well, there is two Claudes: first, there is claude.ai. Second, there is the Anthropic API. (And there is also AWS Bedrock, but it doesn’t offer Opus). Signing up for the Anthropic API from within the EU works without hassle. The reason for
47.
▲
by
ndr_
3y ago
With OpenAI, you can first build Question & Answer pairs derived from your documents and use the OpenAI fine-tuning feature to build yourself a custom model. This method is more than just learning behavior in that facts do get recalled.
48.
▲
by
ndr_
3y ago
It‘s generally available in the EU to AWS Bedrock customers. Just in the Frankfurt region, and with a limited context window AFAIK, but it does exist.
49.
▲
by
ndr_
3y ago
Miles Brundage of OpenAI offered a categorization of „AI things“ into Models, Systems, Platforms and Use-Cases: https://www.youtube.com/watch?v=5j4U2UzJWfI&t=5728s Bard is a System, PaLM 2 would be the model (presumably
50.
▲
by
ndr_
3y ago
LumaFusion/LumaTouch is one counter-example, and I‘m certain there are others. At the surface, you may describe it as a „video editor“, but I totally agree that it‘s really about „Story Telling“. I use it on iPad, and I am mesmerised a
51.
▲
by
ndr_
3y ago
This article is rooted in the junk science by James Zou et al, published in their paper "How is ChatGPT's behavior changing over time?" ( https://arxiv.org/abs/2307.09009 ). Initially, one of their benchma
52.
▲
by
ndr_
3y ago
So RAG (retrieval augmented generation) is all the rage, but there are problems with it that make it appear almost conceptually flawed - to the point where results are flat-out poor? There‘s the paper about non-uniform attention („Lost in t
53.
▲
by
ndr_
3y ago
Is this update made visible somewhere? The language models offered on my Playground are still the ones from March, same with ChatGPT.
54.
▲
by
ndr_
3y ago
I get this: „ERROR. Quota exceeded for aiplatform.googleapis.com/online_prediction_requests_per_base_model with base model: chat-bison. Please submit a quota increase request.“ Has anyone gotten this fixed?
55.
▲
by
ndr_
3y ago
Check out llama_index at https://github.com/jerryjliu/llama_index . What it does: it creates an index over your data using OpenAI embeddings vectors, using the OpenAI Ada model. When querying, it compiles as much contex
56.
▲
by
ndr_
3y ago
In multiple entities I have worked for over the past 15-20 years, there was a sentiment that we wanted to support the fine folks behind Qt, thus licensed each and every developer seat, to the tune of 4000€. At some point, they started bitin