Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
afiodorov
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
8 ms
·
61.
▲
by
afiodorov
2y ago
More software is written than kept. It's harder to write useful software than to configure CI/CD. The latter is a problem that has been solved before, whilst chances of any software codebase being useful enough that it is even wor
62.
▲
by
afiodorov
2y ago
A 100 mb is not much - about 2 mins on youtube
63.
▲
by
afiodorov
2y ago
Bollywood is the second largest film industry!
64.
▲
by
afiodorov
2y ago
First load takes some time, it is downloading a parquet that's a 100mb and has more than 1.5 million titles of series and movies.
65.
▲
by
afiodorov
2y ago
Ah 7 might have been a typo. May be it was meant to be *7? Then that's exactly the Laplace estimator. We start with a prior of 1_000_000 votes of 7.0 and then adjust as we accumulate real votes.
66.
▲
by
afiodorov
2y ago
The data was produced using this script: https://github.com/afiodorov/imdb-sql/blob/main/notebooks/im... It is on hosted on http://www.imdb-sql.com/imdb01-11-2024.parquet In fact th
67.
▲
by
afiodorov
2y ago
That's awesome. Care to explain? Looks like Laplace estimator of some kind? P.S. Query sharing is supported: https://www.imdb-sql.com/?query=SELECT+*+EXCLUDE+%28titleTyp...
68.
▲
by
afiodorov
2y ago
Whole dataset http://www.imdb-sql.com/imdb01-11-2024.parquet loaded (100mb, 1.6 million titles). It has these columns: https://www.imdb-sql.com/?query=SELECT+*+FROM+%27imdb01-11-2... It is produced using h
69.
▲
by
afiodorov
2y ago
doesn't appear possible from https://developer.imdb.com/non-commercial-datasets/
70.
▲
by
afiodorov
2y ago
I initially loved looking for obscure stuff, e.g. setting region to soviet union. It surely is the case that 99% of the users want 10% of the data at most. I'll have to work ability to select the file and download & cache it only i
71.
▲
by
afiodorov
2y ago
Shouldn't be hard to find since the repo name on github is same as the domain :). I just run npm run dev locally. You'll need the parquet file of the data in public/ though, can get it from http://www.imdb-sql.com&
72.
▲
by
afiodorov
2y ago
It'd be trivial to ask LLM to generate the query; but since it's a client-side app there's nowhere to store the api key - so each user would have to supply one, which is a bit of an awkward experience.
73.
▲
by
afiodorov
2y ago
Browsers are becoming OS'es :).
74.
▲
by
afiodorov
2y ago
Will think about improving the ui in this aspect...
75.
▲
by
afiodorov
2y ago
Thanks for the suggestion, will try to improve soon. If anyone with actual css experience wants to open a PR in the meantime I'll be very happy!
76.
▲
by
afiodorov
2y ago
Good spot, will deduplicate in the next iteration. However titles are repeated often due to the region/language variations.
77.
▲
by
afiodorov
2y ago
Sorry about that. I prefer for my UI to be discoverable. P.S. We can share links too https://www.imdb-sql.com/?query=select+1 The SQL flavor is https://duckdb.org/docs/sql/query_syntax/select
78.
▲
by
afiodorov
2y ago
https://github.com/afiodorov/imdb-sql
79.
▲
by
afiodorov
2y ago
It's slow on the first load as you download a 100mb parquet file with all the IMDb data. It'll be cached on the second load though (using browser's indexeddb to cache). Not sure what you mean by scraping in parallel.
80.
▲
Show HN: IMDb SQL Best Movie Finder
(imdb-sql.com)
129 points
by
afiodorov
2y ago
|
81 comments
81.
▲
by
afiodorov
2y ago
I've observed a phenomenon in corporate accountability resembling quantum behavior: 1. Macro level: Departments claim broad accountability. 2. Micro level: Pinpointing task ownership causes accountability to vanish. 3. Indefinite state
82.
▲
by
afiodorov
2y ago
> What this tells us for AI is that we need something else besides LLMs I am not convinced it follows. Sure LLMs don’t seem complete however there’s a lot of unspoken inference going on in LLMs that don’t map into a language directly alr
83.
▲
by
afiodorov
2y ago
Idea is to add enough constraints to the function so that inverting it is non trivial (where inverts finds ANY input). Like the video said inverting multiplication with the output of 15 could yield 15 = 15 * 1, so just make the function f(a
84.
▲
by
afiodorov
2y ago
Reminds me of GTA3 radio - can we retrofit this somehow? I miss driving around mindlessly and now we can get actual quality podcasts too. I wonder which successful game will make use of AI generated content next.
85.
▲
by
afiodorov
2y ago
No, all of them do
86.
▲
by
afiodorov
2y ago
probably not the answer you’re looking for but iPads match the description (usb-c headphones for the newer models)
87.
▲
by
afiodorov
2y ago
No, in this case it is much more pragmatic. Elon has commented on this aspect himself. When he started he thought of just buying rockets and related equipment from Russia on the cheap but quickly found the exuberant prices states want to ch
88.
▲
by
afiodorov
2y ago
The browser runs the code, which is open-sourced. You can look at the network tab of the debugging console or isolate the browser completely by sand boxing it. I agree that UX can be improved but that’s more on the browser side.
89.
▲
by
afiodorov
2y ago
I made two one-page react apps https://www.livetranslate.net/ https://www.subsgpt.com/ Second one is more refined but both are functional. I was pleasantly surprised somebody made a YouTube tutorial in Japa
90.
▲
by
afiodorov
2y ago
I absolutely think that OpenAI or Anthropic should provide such integration. It’s very similar to how Apple Pay centralises your subscriptions and makes payments secure and simple. Would be nice if AI labs had an equivalent portal where eac
More ›