Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
RobinL
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
35 ms
·
421.
▲
by
RobinL
5y ago
Love it! It'd be nice to have a couple more example datasets, and I'd promote them to the top level (rather than in 'more'). I think the first thing many users will want to do is just try to tool with some pre-provided
422.
▲
by
RobinL
5y ago
I don't think it's correct to say Python is far behind in the viz space at all. It's just different. I primarily use Altair within Python. ggplot is ahead of Altair in some respects, but behind in others. For example, here
423.
▲
by
RobinL
5y ago
Thanks, I was curious about that. Key quote for others: > [...] Tesla would need a 600 kWh to 1,000 kWh battery pack to produce the Semi’s 300-mile (483-km) and 500-mile (805-km) variants. Since a 600 kWh pack weighs about 8,000 pounds
424.
▲
by
RobinL
5y ago
I agree that industries will not do this voluntarily. They will just chase a profit. Climate change is a market failure so we need governments to change incentives to make it in industries' interest to change. A carbon tax is one op
425.
▲
by
RobinL
5y ago
The video argues that, with current battery technology, to achieve the same range as a current semi, most of the weight of the truck will be batteries, leaving little weight for cargo (3 tonnes relative to 19 tonnes). I've noticed ther
426.
▲
by
RobinL
5y ago
Checkout the R bindings for DuckDB[0]. You should find that it does the same thing (i.e. run SQL against a dataframe/file on disk) much faster for many SQL operations. [0] https://duckdb.org/docs/api/r
427.
▲
by
RobinL
5y ago
DuckDB is a newer but much faster alternative that lets you run SQL on csv files (and also directories of parquet files etc.) It's so fast/good that I've actually been using it for a lot of data cleaning and transformation th
428.
▲
by
RobinL
5y ago
For me, the biggest reason to use a programming language instead of Excel is its capacity for abstraction, and hence ability to organise code and separate concerns. Attempting to solve a complex analytical problems in Excel is similar to a
429.
▲
by
RobinL
5y ago
Pilates. Quote: "here’s what Pilates really is: a systematic sequence of exercises designed to bring functional strength and flexibility to the whole body."
430.
▲
by
RobinL
5y ago
Birdnet is great! Once you've identified the bird, you can then listen to a variety of additional recordings on https://www.xeno-canto.org/ which I believe is one of the sources used to train the machine learning model
431.
▲
by
RobinL
5y ago
'Overestimates' because at the start of the time series people were using cheaper alternatives to the very-expensive li-ion batteries.
432.
▲
by
RobinL
5y ago
This can help build interactivity: https://vimeo.com/453645616 http://vega.github.io/lyra/
433.
▲
by
RobinL
5y ago
Really interesting, thanks! Great that it accepts colunar data, that was a question that had immediately sprung to mind. There's something very satisfying about processing data using Arquero's grammar of data transformation and th
434.
▲
by
RobinL
5y ago
As soon as I saw this, I thought: 'how is this different from vega lite'. An answer is here: https://observablehq.com/@observablehq/plot-vega-lite It'll be very interesting to see this develop. My initi
435.
▲
by
RobinL
5y ago
You can get a long way nowadays with Arquero[0] and Observable[1]. Arquero allows columnar based data storage and processing, with a grammar of data processing verbs similar to e.g. dplyr. Not as fast as vectorized computations in e.g. Pyt
436.
▲
by
RobinL
5y ago
I did this in the UK through https://schools-energy-coop.co.uk/ . The main thing I was looking for was something where I felt my money would genuinely provide additional capacity (i.e. they weren't just going to build i
437.
▲
by
RobinL
5y ago
If you're not convinced by this article, or simply want to know more about effective altruism, a good place to start is Peter Singer's free book The Life You Can Save, available at https://www.thelifeyoucansave.org/
438.
▲
by
RobinL
6y ago
yeah, agreed - a good understanding of the model's statistical assumptions can often help you make the model more robust and also give you ideas for what types of feature engineering are likely to work.
439.
▲
by
RobinL
6y ago
Yes - this is pretty much exactly how I explain the difference between machine learning and statistics. Despite using similar models, the expertise required for 'doing statistics' (statistical inference) is actually very different
440.
▲
by
RobinL
6y ago
Looking at the costs of e.g. Hinkley Point c, this feels right. I do wonder, though, what the costs would look like if a big country like the US built, say, 400 Hinkley Point Cs. What economies of scale would you see? You could probably ge
441.
▲
by
RobinL
6y ago
I'm not sure that is true. As I understand it, it's a cohort of people of that age who live in a geographical area of Israel where vaccination was started early. So it includes people of that age who declined the vaccine/di
442.
▲
by
RobinL
6y ago
Yes - nowadays I think pretty much everything that uses a lot of energy comes under scrutiny, and, in particular people look to see if the same outcome can't be achieved with less energy
443.
▲
by
RobinL
6y ago
Yes. A second important point is the recognition that data tooling often re-implements the same algorithms again and again, often in ways which are not particularly optimised, because the in-memory representation of data is different betwee
444.
▲
by
RobinL
6y ago
I wrote a bit about this from the perspective of a data scientist here: https://www.robinlinacre.com/demystifying_arrow/ I cover some of the use cases, but more importantly try and explain how it all fits together, jus
445.
▲
by
RobinL
6y ago
Now available: https://arrow.apache.org/blog/2021/01/25/3.0.0-release/
446.
▲
by
RobinL
6y ago
As an occasional R user, I think R markdown one of the things that R does really well. For data scientists who want to output reports (mix of text and calculations), I haven't come across anything as mature or easy to use in the Pytho
447.
▲
by
RobinL
6y ago
Here's some material from the official website: https://arrow.apache.org/powered_by/ https://arrow.apache.org/use_cases/ I've also written a bit about this here, from the point of view o
448.
▲
by
RobinL
6y ago
Is there a summary of the most important new features anywhere? The releases page has a list of changes, but it's very lengthy! https://arrow.apache.org/release/3.0.0.html
449.
▲
by
RobinL
6y ago
Indeed. Presumably if it is effective, you could study people who are using Fluvoxamine for its anti-depressive properties, and you'd find that (other things equal, using something like propensity score matching), they have less serio
450.
▲
by
RobinL
6y ago
I was more thinking about migrating away from Excel fully rather than interfacing with Excel from Python. I agree that to interact with Excel programmatically VBA is a better choice (and no doubt C#/VB.NET as well, but I have no direct
More ›