Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
joelschw
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
31.
▲
by
joelschw
5y ago
dbdiagram.io for ERDs these days: https://dbdiagram.io/home
32.
▲
by
joelschw
6y ago
Kedro, Quantumblack Labs | Python Software Engineer | London | REMOTE currently, ONSITE ? | Full-time Kedro is an open-source Python framework for creating reproducible, maintainable and modular data science code. It borrows concepts from s
33.
▲
by
joelschw
6y ago
Darn
34.
▲
by
joelschw
6y ago
Is it Big Sur only?
35.
▲
by
joelschw
6y ago
I think the big difference here is that when comparing SQLite you are at least using the same query language. In the graph space you have Gremlin, Cryper, GQL and many other proprietary query engines (which also looks to be the the case her
36.
▲
by
joelschw
6y ago
They sort of already have https://en.wikipedia.org/wiki/Mailbox_(application)
37.
▲
by
joelschw
7y ago
This is a wider point for anyone looking to take advantage of machine learning, but reproducibility is also a problem which needs to be catered for.
38.
▲
by
joelschw
7y ago
I think one of the big differences is that during development the pipeline DAG is inferred from the data catalog and not explicitly coded in the same way you need to do in something like Airflow. The logic being that once you've finish
39.
▲
by
joelschw
7y ago
The ruling government will outlast any presidential administration as well
40.
▲
Show HN: Kedro – A Python library for building production-ready ML pipelines
(github.com)
3 points
by
joelschw
7y ago
|
0 comments
41.
▲
by
joelschw
7y ago
We're passed the early peak of the hype cycle, but now the marketers have calmed down the real world applications are maturing. If you think of Data Science as AI sure, but if you frame it as applied statistics + good software engineer
42.
▲
by
joelschw
8y ago
Say you have transactions which follow a complex supply chain... Sure you can reconstruct the path taken using recursive SQL, but you're also joining lots and lots of things together at runtime. In a graph database, you've effecti
43.
▲
by
joelschw
8y ago
Mermaid!
44.
▲
by
joelschw
8y ago
There are some things which Pandas is just better at, such as: extracting content via RegEx and pivoting... However, there are also some situations where you should use SQL such as UPSERT or date-range joins.
45.
▲
by
joelschw
8y ago
Whilst it has its uses, I think we should encourage people to do things in a reproducible way
46.
▲
by
joelschw
8y ago
The direct link: https://www.imperial.ac.uk/media/imperial-college/administra...
47.
▲
by
joelschw
10y ago
The problem with Neo4j is that the end results are great, but the ingestion pipeline (especially for unstructured data) is very hard to make general purpose. The ICIJ used a combination of Apache Tika, Nuix, Tesseract and a bunch of other c
48.
▲
by
joelschw
10y ago
It already is - SwiftKey has been doing it right on my phone for a while!
49.
▲
by
joelschw
11y ago
Why should I use this over Spyder?
50.
▲
by
joelschw
13y ago
This is brilliant! I've been using "refiddle.com" for the last few months and I am blown away by how much better this is. The only thing I could think of improving is including VB code generation as I tend to use that a lot f