Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
lgessler
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
8 ms
·
31.
▲
by
lgessler
2y ago
I liked Yoav Goldberg snarky's quote tweet: > next paper: transformers can solve any problem but on some of them they may compute indefinitely and never provide an answer > (and you cannot tell in advance which is which!!) https
32.
▲
by
lgessler
2y ago
Most LLMs determine their token inventories by using byte-pair encoding, which algorithmically induces sub-word tokens from a body of text. So even in English you might see a word like "proselytization" tokenized apart into "
33.
▲
by
lgessler
2y ago
A division of the NIH (NCBI) maintains a repository of open access publications in life sciences called PubMed Central ( https://en.wikipedia.org/wiki/PubMed_Central ) and this article was in a qualifying journal.
34.
▲
by
lgessler
2y ago
Yep, came here to say this. The big thing about the results here that might not be obvious to someone not in AI is that the models being trained in this paper are very many orders of magnitude smaller than the LLMs we've all heard so m
35.
▲
by
lgessler
3y ago
So you're thinking of expressions like "Winston said he had noticed a continued decrease in his suicidal ideation" (real example). I don't know if I'd agree that it's clear a rule-based system would be able to
36.
▲
by
lgessler
3y ago
I don't personally have a glottal stop in "look cool". I think it's the same as "lukecool" except for one vowel, at least in my English: * lukecool /luk kul/ * look cool /lʊk kul/
37.
▲
by
lgessler
3y ago
This might have something to it, but there are plenty of common expressions that have the same phonetic properties you're describing (IPA transcriptions provided): * fast track /fæst træk/ * swim meet /swɪm mit/ * F
38.
▲
by
lgessler
3y ago
I tried to get into emacs a few years ago and just got fed up with the incredible regularity with which packages broke during routine updates. It was at least an hour a week on average, I feel, conducting bug hunts for either minor (this or
39.
▲
by
lgessler
3y ago
The gold standard is immersion, everyone agrees on that. Most people can't have that, so a mix of language instruction, conversation with native speakers, and exposure to appropriately challenging texts with memorization tools like Duo
40.
▲
by
lgessler
3y ago
IMO this is a misunderstanding of the product. This might be a cynical take, but I think that for most users of Duolingo, it functions more as a game than as a tool which produces substantial gains in language knowledge. I'm a computat
41.
▲
by
lgessler
3y ago
> The sad truth is - all of classical nlp is Dead, with a capital D. If your language is well served by LLMs. Which, generously, is true for maybe 20-50 of the world's 7000 or so languages. For all the rest, "classical" NL
42.
▲
by
lgessler
3y ago
Sure, but I don't think even Bender or Marcus would deny their potential for practical utility, right? I think they mean to say that LLMs are not exactly as miraculous as they might seem, and very capable of misbehavior.
43.
▲
by
lgessler
3y ago
Idk if this work really bears on that argument. I can imagine Bender reacting to this by observing that this is a clever way of finding the needle (a good heuristic algorithm for an NP hard problem) in a haystack (millions of other heuristi
44.
▲
by
lgessler
3y ago
Tl;dr in my words after a skim: this is a method for using LLMs to find new, SOTA (or at least very good?) heuristic algorithms for NP hard combinatorial optimization problems. They achieve this not by making the LLM itself smarter but by f
45.
▲
by
lgessler
3y ago
In a sense you're not wrong, but would you say the same thing about refined sugar? Both of these substances, in terms of immediate physiological effects, are known to be unambiguously bad for you past a certain low threshold. But immed
46.
▲
by
lgessler
3y ago
It's been a while since I worked with any of them so it'd take some review form me to even write a blurb confidently. Sorry! On your latter point though, NLG from AMR is a topic that's been studied since its inception. See th
47.
▲
by
lgessler
3y ago
AMR: https://github.com/nschneid/amr-tutorial UCCA: https://aclanthology.org/P13-1023.pdf And yeah, you could consider UMR an improved version of AMR. AMR is often lampooned with the false name "A
48.
▲
by
lgessler
3y ago
CC BY SA 3.0: https://github.com/amir-zeldes/gum/blob/master/LICENSE.txt I didn't know about that project, that's really cool! I'd be curious to know whether the person who devised this sc
49.
▲
by
lgessler
3y ago
Nobody asked, but here's a syntactic parse of a portion of this article in Universal Dependencies done by yours truly :) https://github.com/amir-zeldes/gum/blob/master/dep/GUM_bio_g...
50.
▲
by
lgessler
3y ago
I think you're right. Here's my math: https://www.wolframalpha.com/input?i=%28%281+square+mile+*+1...
51.
▲
by
lgessler
3y ago
Not surprising, imo—voice assistants got big at a time when there was no way to implement them except by what I would call a massive catalog of hacks. Now that LLMs with really good NLU for major languages exist, Amazon must have made a det
52.
▲
by
lgessler
3y ago
How are conflicts resolved?
53.
▲
by
lgessler
3y ago
YMMV but this doesn't seem true to what I've witnessed in my own life. I can think of many friends (I'm 29) who found either an LTR or their spouse using dating apps. In many cases they were single for long stretches and comp
54.
▲
by
lgessler
3y ago
I'm a linguist, let me try to explain. It's true that in some cultures, especially very economically developed ones, it's common for adults like parents or schoolteachers to teach kids "language". But this is more a
55.
▲
by
lgessler
3y ago
As someone who's casually been eyeing up CRDTs for two years I've wondered continually what the story for authorization is, and this article seems to suggest (in line with my own understanding) that a CRDT library on its own like
56.
▲
by
lgessler
3y ago
IMO, as an NLP researcher, it's a product of the "fast science" culture of AI where there are more papers than anyone could possibly read even in a given subfield at a particular conference, and an "old" paper is an
57.
▲
by
lgessler
3y ago
SEEKING FREELANCER | USA | REMOTE We are a couple of academics who run a popular Olympiad-style competition for high schoolers nation-wide. Our current software system for administering this competition is quite old and idiosyncratic, and w
58.
▲
by
lgessler
3y ago
Idk, that isn't the sense I got from "It is absolutely trivial to show Hyp2 is false", but sure, I agree with you that this evidence certainly ought to tip the scales one way and not the other.
59.
▲
by
lgessler
3y ago
This is a false dichotomy. It's not the case that models are truly capable of reasoning if and only if they are insensitive to irrelevant perturbations to input. In other words, the mere fact that sensitivity to names sometimes causes
60.
▲
by
lgessler
3y ago
Used to write MUMPS at a certain company, and there were lots of "wtf" moments but this is one I'm particularly fond of. In the system, all dates are represented as an integer starting from an arbitrarily defined earliest dat
More ›