Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
jug
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
16 ms
·
211.
▲
by
jug
2y ago
Hopefully, you’ll be able to avoid the whole X Premium Plus thing in the near future with OpenRouter. It’ll still use xAI backend but via your OpenRouter API key. Then you can use it with any web or mobile app that supports OpenRouter. Pers
212.
▲
by
jug
2y ago
Ugh, this is such a trigger for me and my high PSA levels. Currently rocking 3.0 and 46 years old and on a yearly checkup routine since several years.
213.
▲
by
jug
2y ago
I spontaneously feel like this is bad news for open AI, while playing in the hands of corporate behemoths able to strike expensive deals with major publishers and top it off with the public domain. I’m not sure this signals the end of AI an
214.
▲
by
jug
2y ago
With the horror stories I've heard, my max range from a national dealership would be around 80 km.
215.
▲
by
jug
2y ago
I've seen lots of reports in Sweden where the climate control stops working when it's too cold. It's funny to read posts in a thread where they post every winter like clockwork. Of course, if you're subject to this, Tesl
216.
▲
by
jug
2y ago
Cool! IIRC, his game Alpha Centauri also had procedurally generated music.
217.
▲
by
jug
2y ago
I only skimmed the text, but wonder if this study had controls for their social lives and mental health. These behaviors are also on their own indicative of depression, because it often interferes with sleep patterns. It becomes a causation
218.
▲
by
jug
2y ago
I'm happy to see this because it's my experience with Gemini too. Google did terribly with Bard (clearly an emergency launch to say "Hey, we're here too!"), Gemini 1.0 and even Gemini 1.5 was only decent in the top
219.
▲
by
jug
2y ago
Yeah, they need something in their system prompt to tell their name or else they have absolutely no idea what they are and will hallucinate to 100% based on training data. If you're lucky, the AI just might guess right based on these c
220.
▲
by
jug
2y ago
Google is the least confusing to me. Old school version number and Pro is better than Flash which is fast and for "simple" stuff (which can be effortless intermediate level coding at this point). OpenAI is crazy. There may be a da
221.
▲
by
jug
2y ago
I think a counterpoint to this is that SQL has a specific and well-defined meaning and it takes effort to get what you actually want right. However, communication with an AI can sometimes request a specific context or requirements but also
222.
▲
by
jug
2y ago
App without iPad version. This is so weird, Apple.
223.
▲
by
jug
2y ago
Gemini 2.0 Experimental is now a leading LLM. They started out poorly, but after a more reasonable 1.5 Pro, 2.0 is in another class entirely and a direct competitor to o1 (or o1-mini as for Gemini 2.0 Flash). They've made quick strides
224.
▲
by
jug
2y ago
Even the latest AI up scalers will have a 384x384 look pretty terrible when put against e.g SDXL @ 1024x1024 native. It's just too little to work on.
225.
▲
by
jug
2y ago
Yes, that's what OpenAI o1 does, and DeepSeek R1. Also Google Gemini 2.0 Thinking models. It's a way to significantly improve benchmark scores, especially in math. It's funny to watch too. I played with Gemini 2.0 on Google A
226.
▲
by
jug
2y ago
The man behind Apple's modern design philosophy is Jony Ive.
227.
▲
by
jug
2y ago
Wow, this was even enabled on our otherwise fairly locked down corporate phones... Normally these features are both disabled and greyed out via "group policies" or whatever Apple calls them.
228.
▲
by
jug
2y ago
I liked the SimpleQA benchmark that measures hallucinations. OpenAI models did surprisingly poorly, even o1. In fact, it looks like OpenAI often does well on benchmarks by taking the shortcut to be more risk prone than both Anthropic and Go
229.
▲
by
jug
2y ago
Not sure this is an AI limitation. I think you'd be better off here with the Gemini Code Assist plugin in VS Code rather than that. Sounds like the AI is provided with unstructured information compared to an actual code base.
230.
▲
by
jug
2y ago
Speaking of which, I wonder how they'd do on SimpleQA. OpenAI is an outlier there in the negative sense vs Anthropic. This benchmark also deals with hallucination and "inappropriate certainty".
231.
▲
by
jug
2y ago
True that - and I think Gemini-Exp-1206 is Gemini 2.0 Pro in testing. I noticed how they only replaced the "experimental" moniker for one of their experimental models, and it turned into 2.0 Flash. And that still experimental mode
232.
▲
by
jug
2y ago
I feel like these are test versions of Gemini Pro 2.0. The changes are too foundational to be mere iterations/break date updates for 1.5 Pro.
233.
▲
by
jug
2y ago
This year seems to finish on the same note as it began -- that most AI evolution happens in the smaller models. There's been a true shift as corporations have started to realize the value of training data and massively outsizing the re
234.
▲
by
jug
2y ago
My favorite feature about properties is that I can set a breakpoint on the setter. It'll now break on anything that sets it with a single breakpoint. Or use "Go to calls" on the setter and I instantly get everything that sets
235.
▲
by
jug
2y ago
Yeah, I went into the article thinking this because I expected someone had created waypoints right on top of each other and in the process also somehow generating the same code for them.
236.
▲
by
jug
2y ago
Even o1-preview hallunicates more than e.g. Claude Sonnet 3.5 before it and only slightly better than GPT-4o, according to the paper on OpenAI's own SimpleQA benchmark. This is precisely the problem o1 tried to tackle at great effort i
237.
▲
by
jug
2y ago
So, OpenAI has created a whole new team very late in the process to explore methods of post-training and enhancing the output. Screams of desperation to me. It was supposed to launch on Azure perhaps already this month, and more widely in d
238.
▲
by
jug
2y ago
One thing that makes FLUX so special is the prompt understanding. I now gave FLUX 1.1 a prompt "Closeup of a doll house built to resemble a famous room in the TV show Friends" and it gave me one with the sign "Central Perk&qu
239.
▲
by
jug
2y ago
Ugh. Code reviews and helping out with tedious code comments. That's great stuff for software developers. And will be a headache to control for our company. This is taking increasingly more restraint from developers to not send code as
240.
▲
by
jug
2y ago
He's specifically mentioning less social friction than iPhone, so I assume not a smartphone but a wearable. Glasses or some Star Trek/Rabbit-like button? I wouldn't mind something that works. I still want a sleek and function
More ›