4 ms·
GitHub Repository Collaboration Network
- gingerlime 13y agoWow this looks awesome. I wish I could search for other repositories not on the graph already to discover other related open source projects.
- icco 13y agoPretty neat paper linked from the bottom (https://github.com/coyotebush/github-network-analysis/raw/master/github.pdf https://github.com/coyotebush/github-network-analysis/raw/ma...), it would be neat to see how the correlations changed if you tweaked some of the constants.
- coyotebush 13y agoAnalyzing a larger dataset would be neat indeed, though somewhat more challenging. Especially for the layout algorithm to produce something nice in a reasonable amount of time. There are ~11k repos with >= 100 stars, compared to the 825 I had here (fewer after filtering for the giant component).
- icco 13y agoFair, I'd say you might just want to have a searchable db of relationships instead of the visualization (or maybe do something google-maps-esque, where as you zoom you load relationships that you can see...).
- rane 13y agoCan't help it, reminds me of EVE Online.
- Decent 13y agoReally wish I could zoom in to take a closer look at specific links.
- coyotebush 13y agoScroll to zoom. Sorry that's not noted anywhere.
- cachvico 13y agoDoesn't seem usable with a Magic Mouse - scrolling is way too sensitive.
- zekenie 13y agowhat is x and y? what is the color?
- Laremere 13y agoI /think/ color is the language it's programed in. A key would be really handy.
- coyotebush 13y agoYeah, per the footer, color is the primary programming language as identified by GitHub. A key felt like a bit too much clutter; you can always click on a repository to open its page and note the language listed.
- skylan_q 13y agoBeen waiting to see this for a while...
- drorweiss 13y agoSo Twitter Bootstrap is the king of github...
- coyotebush 13y agohttps://github.com/popular/starred https://github.com/popular/starred
- ivanist 13y agoI wonder how much BigQuery quota did you use while testing this stuff?
- coyotebush 13y agoHeh, nearly 80GB... and that's with mostly testing out the queries on the smaller publicdata:samples.github_timeline dataset.