Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
paradite
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
9 ms
·
91.
▲
by
paradite
2y ago
My understanding is that they have a base model checkpoint for Behemoth from pre-training. This base model is not instruction-tuned so you can't use it like a normal instruction-tuned model for chatbots. However, the base model can be
92.
▲
by
paradite
2y ago
I might be missing the point of the author, but what's the difference between Steam Deck and a normal Windows gaming laptop in terms of software freedom? Wouldn't you have more software freedom on Windows? Because you can run both
93.
▲
by
paradite
2y ago
This is not a good comparison for real world coding tasks. Based on my own experience and anectodes, it's worse than Claude 3.5 and 3.7 Sonnet for actual coding tasks on existing projects. It is very difficult to control the model beha
94.
▲
by
paradite
2y ago
Maybe it is. Maybe it's not. I prefer to state facts and leave interpretations to others.
95.
▲
by
paradite
2y ago
You can turn it off. This is just a poorly worded FAQ. At the beginning, there is an FAQ that says you can't turn off ads in scenic mode. Q: Can I turn Scenic Mode ads off? A: No, not at this time... But in fact you can turn
96.
▲
by
paradite
2y ago
For anyone not in LLM field, OpenAI API is the de facto standard for almost all AI companies and labs nowadays. DeepSeek docs just point users to use OpenAI SDK: https://api-docs.deepseek.com/ Anthropic recently did some so
97.
▲
by
paradite
2y ago
Ok, so the certificate used to sign the package is generated by Apple, why can't I just use that to prove my identity for notarization? Or maybe simpler, why can't Apple just do code sign and notarization with one single cli call,
98.
▲
by
paradite
2y ago
It's true. One example I can give is how Gmail used to automatically recognise flights and hotel bookings and add them to calendar. It was suddenly completely broken and stopped working a few years ago. I tried every setting to try to
99.
▲
by
paradite
2y ago
As someone who actually signs, notorizes and distributes desktop apps for macOS, I can safely say their documentation is less than ideal. Maybe because I'm using Electron framework which makes things more complicated, but I don't
100.
▲
by
paradite
2y ago
I think this is an interesting pivot, but Cursor's project level rules and custom modes will probably quickly evolve to cover all the aspects listed on your hub. (Maybe it already does) This also allows developers to switch between pro
101.
▲
by
paradite
2y ago
MCP is basically commoditizing SaaS and software by abstracting them away behind the AI agent interface. It benefits MCP clients (ChatGPT, Claude, Cursor, Goose) more than the MCP servers and the service behind the MCP servers (GitHub, Figm
102.
▲
by
paradite
2y ago
I see what you mean. It is a paradigm shift indeed if you look from the user's perspective.
103.
▲
by
paradite
2y ago
I think this is a good explanation on the client side of MCP. But most developers are not building MCP clients (I think?). Only a few companies like OpenAI, Anthropic, Cursor and Goose are building MCP client. Most developers are currently
104.
▲
by
paradite
2y ago
I actually tried it a few months ago and I still using it (mainly for testing its capabilities), despite it not performing up to my standards. I wrote my first impressions of Devin here: https://thegroundtruth.substack.com/p
105.
▲
by
paradite
2y ago
Missing a few. Check out mine visualized in 2D quardants: https://paradite.github.io/ai-coding/ Also there are a lot of cli tools in this space: https://prompt.16x.engineer/cli-tools
106.
▲
by
paradite
2y ago
Yes. This is my point I'm trying to make. Thank you for explaining it.
107.
▲
by
paradite
2y ago
Ok looks like people are not getting my comment. Being a judge in a hackathon is one of the criterion for O-1 visa. https://www.linkedin.com/pulse/getting-o-1-visa-easier-than-...
108.
▲
by
paradite
2y ago
O-1
109.
▲
by
paradite
2y ago
You are asking the right question, but to the wrong person.
110.
▲
by
paradite
2y ago
Normalizing mediocrity is not something I want in STEM fields. We need exceptional people who can push boundaries and do exceptional work in order to progress.
111.
▲
by
paradite
2y ago
I don't know. As someone not from US, it looks like Sam Altman is also in good relationship with Trump due to Stargate.
112.
▲
by
paradite
2y ago
My burning question: Why not also make a slightly larger model (100B) that could perform even better? Is there some bottleneck there that prevents RL from scaling up performance to larger non-MoE model?
113.
▲
by
paradite
2y ago
Interesting. This is the first time I am hearing about intrinsic positional bias for LLM. I had some intuition on this but nothing concrete.
114.
▲
by
paradite
2y ago
Thinking from the retrieval perspective, would it make sense to have two layers? First layer just describes on high level, the tools available and what they do, and make the model pick or route the request (via system prompt, or small model
115.
▲
by
paradite
2y ago
Hi. This is very helpful. Thanks for sharing!
116.
▲
by
paradite
2y ago
Hi. I'm an electron app developer. I use electron builder paired with AWS S3 for auto update. I have always put Windows signing on hold due to the cost of commercial certificate. Is the Azure Trusted Signing significantly cheaper than
117.
▲
by
paradite
2y ago
It's actually fairly easy to setup a 3rd party app to use Claude via API, to get extremely generous limits. I wrote a step-by-step guide for the app I built: https://prompt.16x.engineer/guide/claude
118.
▲
by
paradite
2y ago
It's there in the docs (Model comparison table) https://docs.anthropic.com/en/docs/about-claude/models/all-m...
119.
▲
by
paradite
2y ago
DeepSeek also has a FIM (Fill In the Middle) completion model via API, if anyone is interested to try out: https://api-docs.deepseek.com/guides/fim_completion
120.
▲
DeepSeek R1 speed benchmark across different providers
(github.com)
3 points
by
paradite
2y ago
|
0 comments
More ›