Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
phowon
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
9 ms
·
1.
▲
by
phowon
7y ago
My roommate has a cat who mostly hangs out in the living room. I never had pets growing up, so I never interact with it. Every so often, when I'm walking past it to get around the apartment, it swipes at/scratches me, even drawing
2.
▲
This Erotica Does Not Exist (NSFW)
(erogenerator.gitlab.io)
6 points
by
phowon
7y ago
|
0 comments
3.
▲
by
phowon
7y ago
India is generally not considered to be in South-East Asia.
4.
▲
by
phowon
7y ago
As someone who aggressively uses conda, I do think one of its downsides is how heavy it is. I agree that it's a good one-stop-shop if you want to get an environment up and running with no issues. But if you're not doing any scient
5.
▲
by
phowon
7y ago
And here is a tweet thread on why using NLP models to "fill in" the redacted portions is a horrendously terrible idea. https://twitter.com/emilymbender/status/1119081131234611201
6.
▲
by
phowon
7y ago
If anything, I've found Airpods to be extremely visible given their stark white color. As opposed to darker bluetooth headpieces that I often missed.
7.
▲
by
phowon
7y ago
>I've definitely noticed a much stronger emphasis on wowing visual effects in today's RPG's and a much weaker focus on the story and plot lines as was the case in past games. This basically started in the PS3 era, where it
8.
▲
by
phowon
7y ago
>As an example, there are many people who like grinding in JRPGs and some even consider it as a defining element of JRPGs (in that a JRPG is not real JRPG if it doesn't have grinding) I am one such person who is happy to argue for t
9.
▲
by
phowon
7y ago
I know this talking point has been brought up by different people, but it's worth pointing out that Transformers were already covered in the class in 2018. https://web.stanford.edu/class/archive/cs/cs224n
10.
▲
by
phowon
8y ago
The success of Transformers aside, I'm not sure you should be relying on model titles for anything, lest we forget papers like "One Model To Learn Them All" [1]. [1] https://arxiv.org/abs/1706.05137
11.
▲
by
phowon
8y ago
Aren't causal convolutions basically CNNs with masking?
12.
▲
by
phowon
8y ago
Putting aside the whole Schmidhuber debate - where are people getting this idea that causal convolutions are anywhere near the prominence of RNNS/LSTMs? As far as I'm aware, causal convolutions were used in WaveNet (and subsequent
13.
▲
by
phowon
8y ago
The Hug of Death seems to have killed the page so I can't really tell what it's about, but here's an example of inserting new trainable layers in BERT: https://arxiv.org/abs/1902.00751
14.
▲
by
phowon
8y ago
There is no constraint that companies must maximize shareholder value. Companies can be set up for any number of reasons (including non-profit reasons).
15.
▲
by
phowon
8y ago
What is it generally used for though? To me, a "k.v" store sounds like a very generic but nice thing to have, but I still don't have a good sense of what it is and what people think of it. The one place I've run into it
16.
▲
by
phowon
8y ago
As someone who is fairly tech literate but not familiar with this tech stack - in practical terms, what is Redis, and what is it used for?
17.
▲
by
phowon
8y ago
The BERT paper also introduced BERT Base, with is 12 layers with approximately the same number of parameters as GPT, but still outperforms GPT on GLUE.
18.
▲
by
phowon
8y ago
With ELMo, the pretrained weights are frozen. Only the scalars for ELMo layers are tuned (as well as the additional top-level model, of course).
19.
▲
by
phowon
8y ago
It's relatively small modification of BERT with multi-task fine-tuning and slightly different output heads. It should be easy for any NLP researcher to replicate.
20.
▲
by
phowon
8y ago
Put another way, even if you take the strict definition of a rational (optimal) agent, in probably any non-trivial or non-contrived corner case, for any arbitrary set of actions, there exists some utility function and information constraint
21.
▲
by
phowon
8y ago
I am aware of the origin of the word. I am distinguishing between the generic notion of a meme (which as you pointed out simply means "transmittable idea", which is an incredibly general concept) and the specific notion of an Inte
22.
▲
by
phowon
8y ago
That's not what's being discussed here. We're talking about when the word 'meme' was repurposed to refer to "Internet trend".
23.
▲
by
phowon
8y ago
Outsiders to 4chan would probably also be surprised at the site-wide outpouring of genuine sadness, nostalgia and gratitude when moot announced that he was stepping down.
24.
▲
by
phowon
8y ago
Yes! I needed someone to confirm that people used to call it "fads" before memes.
25.
▲
by
phowon
8y ago
>For example, in jointly learning to align and translate, the attention is certainly not invariant to number of context vectors. You train the attention to take in a fixed number of context vectors and produce a distribution over the fix
26.
▲
by
phowon
8y ago
So I'm curious - is this considered blackmail if the object isn't money?
27.
▲
by
phowon
8y ago
It really is underappreciated how much 4chan has shaped Internet, and now mainstream, culture.
28.
▲
by
phowon
8y ago
I'm not sure if you're understanding me correctly. Attention is generally length invariant. You take some transformation on the hidden representations (/+ inputs) at that each time step, and then you normalize over all the tr
29.
▲
by
phowon
8y ago
ULMFiT is LSTM-based. https://arxiv.org/pdf/1801.06146.pdf
30.
▲
by
phowon
8y ago
And convolution-based models still find use in all sorts of cool applications in language, such as: https://arxiv.org/abs/1805.04833 With regards to adversarial discussions, it's one thing to argue about whether m
More ›