Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
85392_school
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
31.
▲
by
85392_school
1y ago
This might be because Gemini silently updates checkpoints (1.5 001 -> 1.5 002, 2.5 0325 -> 2.5 0506 -> 2.5 0605) while OpenAI doesn't update them without ensuring that they're uniformly better and typically emails custome
32.
▲
Agent Village
(theaidigest.org)
1 points
by
85392_school
1y ago
|
0 comments
33.
▲
by
85392_school
1y ago
You're actually sending data to random GPUs connected to one of the Bittensor subnets that run LLMs.
34.
▲
by
85392_school
1y ago
MCP is Model Context Protocol, a standardized API for declaring remote tools for AI to call. Roo Code supports Model Context Protocol, not your creation.
35.
▲
by
85392_school
1y ago
> PRE-RELEASE SOFTWARE. The software is a pre-release version. It may not operate correctly. It may be different from the commercially released version.
36.
▲
by
85392_school
1y ago
Does anyone else feel like the new design feels less trustworthy? I've probably just been conditioned on too many templates that all look the same, and there's nothing inherently wrong with it, yet it makes me wonder if I've
37.
▲
by
85392_school
1y ago
There are some limits: > 2 concurrent tasks > 5 total tasks per day
38.
▲
by
85392_school
1y ago
> Also, you can get caught up fast. Jules creates an audio summary of the changes. This is an unusual angle. Of course Google can do this because they have the tech behind NotebookLM, but I'm not sure what the value of telling you h
39.
▲
by
85392_school
1y ago
Agents definitely fix this. When you can run commands and edit files, the agent can test its code by itself and fix any issues.
40.
▲
by
85392_school
1y ago
This is about AlphaEvolve, currently being discussed at https://news.ycombinator.com/item?id=43985489
41.
▲
by
85392_school
1y ago
Reading the actual article, this seems odd. It only covers cases when the models degrade, but there hasn't been evidence of a LLM pinned to a checkpoint degrading yet.
42.
▲
by
85392_school
1y ago
You'd probably meet the talking point that if we don't accelerate AI development China will win.
43.
▲
by
85392_school
1y ago
AI systems have been improving. O3 now has the capability to decide to search multiple times as part of its response.
44.
▲
by
85392_school
1y ago
It could also mean that the project is stable. Since you only look at the one repository's commit activity, a stable project with a maintainer who's still active on GitHub in other places would be "less trustworthy" than
45.
▲
by
85392_school
1y ago
This announcement accompanies the new and proprietary Mistral Medium 3, being discussed at https://news.ycombinator.com/item?id=43915995
46.
▲
by
85392_school
1y ago
It sounds like they have yet to focus on products yet. Loosely quoting https://news.ycombinator.com/item?id=43907634 : > So the search should work best for people, companies, papers, high quality written content. > Typ
47.
▲
by
85392_school
1y ago
Oh, good to hear. I've been waiting for its return.
48.
▲
by
85392_school
1y ago
> while your product is incomparably slower than Google Exa was originally just a search engine. They try to hide it these days to promote Websets, but you can still use it at https://exa.ai/search .
49.
▲
by
85392_school
1y ago
That's inaccurate. First, there was the experimental 03-25 checkpoint. Then it was promoted to Preview without changing anything. And now we have a new 05-06 checkpoint, still called Gemini 2.5 Pro, and still in Preview.
50.
▲
by
85392_school
1y ago
Transformers.js wraps the ONNX runtime which is rather versatile (WASM, WebGL, WebGPU, and WebNN). It's not the backend that makes it novel.
51.
▲
by
85392_school
1y ago
The funny thing is that if your request only needed the top 100's temperature or the top 33's precipitation, it could just read "List of cities by average temperature" or "List of cities by average precipitation&quo
52.
▲
by
85392_school
1y ago
This part seems relevant: > in beta on the Max, Team, and Enterprise plans, and will soon be available on Pro
53.
▲
by
85392_school
1y ago
You should try it. It's trained for tool calling and thinks before taking action.
54.
▲
by
85392_school
1y ago
It's just a joke about AI not getting a joke (that's really just self-censoring).
55.
▲
by
85392_school
1y ago
The "skip the line" page: https://www.shopify.com/careers/extraordinary_0dac47b1-3275-...
56.
▲
by
85392_school
1y ago
Which itself is a summarized version of https://andymasley.substack.com/p/individual-ai-use-is-not-b... (discussed at https://news.ycombinator.com/item?id=42745847 )
57.
▲
by
85392_school
1y ago
I think you're confusing GPT-4.5 with GPT-4.1. GPT-4.1 is their recommended model for non-reasoning API use.
58.
▲
by
85392_school
1y ago
https://archive.is/mmzWj
59.
▲
by
85392_school
1y ago
(2014)
60.
▲
by
85392_school
1y ago
IDE just stands for Integrated Development Environment, so something that doesn't edit text could still be an IDE
More ›