Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
glaugh
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
14 ms
·
121.
▲
by
glaugh
12y ago
Thanks. For now, the best way to deal with confounds and interactions is to run an analysis a couple times, filtering for a third variable in various ways to see how that affects the analysis. Obviously it'd also be nice to also have m
122.
▲
by
glaugh
12y ago
Yeah, sorry, just looking at this now. These are all represented in the raw data as "10,000" and "5,000", etc., as though they were actual numbers; they should actually be ranges, though, as indicated by that post. Thank
123.
▲
by
glaugh
12y ago
Point taken about the scrollbar only showing on hover, definitely an issue. For context, there's some tricky issues that come up in building a way to view tables that can scale to 10,000s by 10,000s (e.g., you can't just plop down
124.
▲
by
glaugh
12y ago
Pre-analyzed results are available here, also downloadable: http://blog.stackoverflow.com/2014/02/2013-stack-overflow-us... http://blog.stackoverflow.com/2013/01/2012-stack-overflow-us...
125.
▲
by
glaugh
12y ago
Hey, can you shoot me an email at contact @ statwing? I'd love to hop on a screenshare for 10 minutes. Agreed that the gigantic help screen thing upon close is dumb. But a couple of the things you mention sound more like browser-relate
126.
▲
by
glaugh
12y ago
. Yup, you can make filters universal with the "Enable Filters" button in the upper left (only applies to future analyses, not retrospectively). . The data is actually freely available for download via Stack Overflow, who ran the
127.
▲
by
glaugh
12y ago
Ha. We apologize. Oversight on our part, we're normally more loyal to our good friend cynicism. For what it's worth, this demo is a bit less fun than typical because all the data is categorical (i.e., almost all of it stems from m
128.
▲
Results of Stack Overflow survey of 20,000 developers
(statwing.com)
179 points
by
glaugh
12y ago
|
97 comments
129.
▲
by
glaugh
12y ago
I think he's referencing "Lifestyle diseases" or "Diseases of civilization", like diabetes and heart disease, which appear to be quite rare in tribal settings (and not just because people die first of other things),
130.
▲
by
glaugh
12y ago
So, this dataset isn't broken down by state (just region), but it looks at what religion Americans say they have (including "None"), how that has changed over time since 1972, and how that differs by region: https:/
131.
▲
by
glaugh
12y ago
Here's a smattering of example apps built off this kind of data (e.g., data in Socrata-powered open data portals like this one): http://www.socrata.com/civic-apps/ (Disclosure?: I'm friendly with the Socrata
132.
▲
by
glaugh
12y ago
Fwiw, here's a fun little blog post on Socrata's blog, pulling from the crime dataset[1], looking at which crimes happen at which time of day/week/year: http://www.socrata.com/blog/crime-time-visuali
133.
▲
Crime Over Time: Visualizing Crime Data in Chicago
(socrata.com)
25 points
by
glaugh
12y ago
|
9 comments
134.
▲
The Early Origins and Development of the Scatterplot (2005) [pdf]
(datavis.ca)
3 points
by
glaugh
12y ago
|
0 comments
135.
▲
Results of HN poll: Half think bootcamp grads as good as fresh CS Majors
(blog.statwing.com)
1 points
by
glaugh
12y ago
|
1 comments
136.
▲
by
glaugh
12y ago
My brother works in Hollywood, and would make the distinction that while musicians can make money on tour and with merch, movies basically only make money from selling the right to watch the thing. I'd like to justify pirating movies,
137.
▲
by
glaugh
12y ago
Yup! Here's details: https://www.statwing.com/overview/integrations API docs: https://www.statwing.com/docs/api
138.
▲
by
glaugh
12y ago
Nope. Folks who sign up for the API generally have some awareness of what size files work and what don't. And since they're uploading for their users, we're aligned around wanting those users to have a good experience. We hav
139.
▲
by
glaugh
12y ago
Yeah, agreed. We do actually handle xls files in our main product (this is sort of a demo of an API connection). Probably should have enabled that for this little implementation. Can't yet take gz/zip files, probably should though
140.
▲
by
glaugh
12y ago
Things will definitely start slowing down pretty linearly after 100k lines, but we often see millions, and most files shouldn't break us as long as they're not over ~500MB. Edit: Fleshed out explanation
141.
▲
Visualize any public CSV on GitHub in a few clicks
(blog.statwing.com)
69 points
by
glaugh
12y ago
|
11 comments
142.
▲
Heat and Violence in Chicago (Statistical Analysis)
(nbviewer.ipython.org)
1 points
by
glaugh
12y ago
|
0 comments
143.
▲
Churnalism: US Tool for Jounalistic Accountability
(churnalism.sunlightfoundation.com)
4 points
by
glaugh
12y ago
|
2 comments
144.
▲
by
glaugh
12y ago
Weird. Not sure why that would be the case for you but not me. Well, in any case, we're planning on doing a lot of checks on the back end around this (b/c after all, anyone can submit a survey multiple times from multiple device
145.
▲
by
glaugh
12y ago
It was an intentional choice, based on my (incorrect) impression that it was exclusively more advanced folks. Though now I'll act like vitno's logic was mine, too :) I think if I were to do it over I might include it, but at thi
146.
▲
by
glaugh
12y ago
Middle of next week. (I'm leaving town for a wedding tomorrow morning, sorry).
147.
▲
by
glaugh
12y ago
Pretty sure that's because I didn't have that setting on before your first submission, as I'm not able to do multiple submissions. Either way, we'll watch out in the results for suspicious behavior. Thanks for bringing t
148.
▲
by
glaugh
12y ago
Appreciate that feedback. Actually working on improving that output view shortly.
149.
▲
by
glaugh
12y ago
Thank you. Fixed that issue.
150.
▲
by
glaugh
12y ago
Yup, they'll be posted for download by anyone.
More ›