Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
Tpt
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
31.
▲
by
Tpt
9y ago
You should have a look at [1] that outputs an HTML rendering of pages with a lot of metadata. [1] https://en.wikipedia.org/api/rest_v1/#!/Page_content/get_pag...
32.
▲
by
Tpt
10y ago
There is https://github.com/livingbio/syntaxnet_wrapper that does the job fairly well (I also spent days trying to be able to pass to SyntaxNet different textes without having to reload the model). Warning: installatio
33.
▲
by
Tpt
10y ago
> Can you throw machine learning at these changes to identify future abusive ones? There are already filters mostly based regular expression that can forbid page saving and there is an ML tool that flags edits as "probably bad"
34.
▲
by
Tpt
10y ago
Wikidata (the wiki used) is not a regular MediaWiki but host special pages with structured data. See https://www.wikidata.org
35.
▲
by
Tpt
10y ago
I use MongoDB for a Wikidata replica and index performances are quite good. I use some hacks in order to keep size of indexed values low (see https://github.com/ProjetPP/WikibaseEntityStore/blob/master/..
36.
▲
by
Tpt
11y ago
There is also Platypus that is a small query answering engine based on Wikidata: http://askplatyp.us