Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
ofou
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
9 ms
·
61.
▲
by
ofou
2y ago
This is one of the reasons I've been advocating to use UTF-8 as a tokenizer for a long time. The actual problem IMHO are tokenizers themselves, which obscure the encoding/decoding process in order to gain some compression during t
62.
▲
by
ofou
2y ago
Who would have known that BitTorrent, shadow libraries, and seeders will help to train the best AI models out there, that adds a whole new meaning to a "seed".
63.
▲
by
ofou
2y ago
It’s called DeepSeek. The founder just confirmed a few days ago that he got the data from Anna's to train on, I think for their latest vision model.
64.
▲
by
ofou
2y ago
This is a wonderful submission to Anna's archive [1]. I really love people pushing the boundaries of shadow source initiatives that benefit all of us, especially providing great code and design. Can't emphasize enough the net plus
65.
▲
by
ofou
2y ago
I find quite interesting they're releasing three compute levels (low, medium, high), I guess now there's some way to cap the thinking tokens when using their API. Pricing for o3-mini [1] is $1.10 / $4.40 per 1M tokens. [1]:
66.
▲
by
ofou
2y ago
competition is beneficial for all of us, this is great
67.
▲
by
ofou
2y ago
Where can I get the actual tokenizer data? Nevermind, it's here https://api-docs.deepseek.com/quick_start/token_usage
68.
▲
Fill your GitHub calendar with fake commits to confuse recruiters
(gist.github.com)
1 points
by
ofou
2y ago
|
0 comments
69.
▲
Portrait of the Hilbert Curve (2010)
(corte.si)
54 points
by
ofou
2y ago
|
14 comments
70.
▲
by
ofou
2y ago
By the way, I think this is a fantastic reading list for creating AI products, and especially for staying updated on the latest in the AI space. However, it feels a bit scattered and might be hard for beginners to follow, IMO. I read your b
71.
▲
Intelligence in the Age of Mechanical Reproduction
(charleseisenstein.substack.com)
2 points
by
ofou
2y ago
|
0 comments
72.
▲
by
ofou
2y ago
Dive into Deep Learning is implemented using various libraries such as PyTorch, NumPy/MXNet, JAX, and TensorFlow. Here’s an example: https://d2l.ai/chapter_natural-language-processing-pretraini...
73.
▲
by
ofou
2y ago
I believe that most of the papers presented here focus on acquiring knowledge rather than deep understanding. If you’re completely unfamiliar with the subject, I recommend starting with textbooks rather than papers. The latest Bishop’s &quo
74.
▲
How Many Cells Are in Your Body?
(nationalgeographic.com)
3 points
by
ofou
2y ago
|
0 comments
75.
▲
Information Is Surprise
(plus.maths.org)
3 points
by
ofou
2y ago
|
0 comments
76.
▲
Fast, Robust Google Flights Scraper (API) for Python
(github.com)
3 points
by
ofou
2y ago
|
0 comments
77.
▲
by
ofou
2y ago
Learn to pronounce German sounds accurately, and then read the entire Harry Potter series out loud. By the time you're halfway through, you'll be well on your way to fluency. Many focus too much on understanding meaning, but what&
78.
▲
by
ofou
2y ago
From the City of London wiki: In December 2012, following criticism that it was insufficiently transparent about its finances, the City of London Corporation revealed that its "City's Cash" account – an endowment fund built u
79.
▲
EmergentMind: AI research assistant to Discover and Learn about papers
(twitter.com)
2 points
by
ofou
2y ago
|
0 comments
80.
▲
by
ofou
2y ago
this is a list of books mentioned https://www.goodreads.com/list/show/121496.SWEBOK_Consolidat...
81.
▲
Why Abercrombie and Fitch's Stock Exploded Faster Than Nvidia's [video]
(youtube.com)
1 points
by
ofou
2y ago
|
0 comments
82.
▲
Face Search Engine Reverse Image Search
(pimeyes.com)
2 points
by
ofou
2y ago
|
0 comments
83.
▲
by
ofou
2y ago
Whether you're doing research or just keeping up with AI news, EM is doing great work. We've gathered millions of data points for each paper. Going forward, we plan to improve the capabilities and expand the sources behind every p
84.
▲
The fight to save Chile's white strawberry
(atlasobscura.com)
109 points
by
ofou
2y ago
|
52 comments
85.
▲
What I've Learned Building Interactive Embedding Visualizations
(cprimozic.net)
2 points
by
ofou
2y ago
|
0 comments
86.
▲
The Old-Fashioned Library at the Heart of the A.I. Boom
(nytimes.com)
2 points
by
ofou
2y ago
|
1 comments
87.
▲
by
ofou
2y ago
Awesome! I'm trying out this again, I'd be awesome that you share your methods. Semantic chunking it's a pretty cool thing too. Nice work
88.
▲
by
ofou
2y ago
Compare with using ChatGPT (GPT4o) for this https://chatgpt.com/share/66f3a5c6-4d60-8009-af96-a3aea066f3...
89.
▲
by
ofou
2y ago
This is an output example from the raw transcription of "10 Programmer Stereotypes" ( https://www.youtube.com/watch?v=_k-F-MMvQV4 ) [ "the programmer an offshoot of the great ape family closely related to chi
90.
▲
by
ofou
2y ago
From the company, they explained that “the large database of Cramer, which covers thousands of formulas for its fragrances, allows the AI of Notco, Giuseppe, acquire this information and generate top -quality aromas at once, doing the pro
More ›