Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
goldemerald
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
8 ms
·
31.
▲
by
goldemerald
3y ago
Hire a grad student for a year to write a research paper on it, haha. I've just gotten into the LLM space for work, so I wouldn't know
32.
▲
by
goldemerald
3y ago
ACL had a recent tutorial about state of the art for this topic. https://acl2023-retrieval-lm.github.io/ My favorite takeaway was that purely fine tuning your model on your documents (without extra document context during i
33.
▲
by
goldemerald
3y ago
For those interested, a modern version (vision transformers) was just published this year at CVPR https://openaccess.thecvf.com/content/CVPR2023/html/Park_RGB...
34.
▲
by
goldemerald
3y ago
I love your content, the only channel I've turned the bell on for. I am curious about your approach to explaining coding and machine learning, do you think it's possible to make "coded for 5 days" engaging in the same wa
35.
▲
by
goldemerald
3y ago
I've been reading some Buddhist texts lately, and this I feel this article would fit right in towards the appreciation of all lives. I know LLMs are fundamentally matrix multiplication text calculators, but attributing a Buddha nature
36.
▲
by
goldemerald
3y ago
I did an oral final exam for my theory of computation class. From the student perspective, it was nicer than an in person final as I knew I didn't have to have perfect prose in my answers (and that it only took 30 mins rather than 2 ho
37.
▲
by
goldemerald
4y ago
This article is not about ChatGPT, but GPT-4. I've found the latter to be significantly better at being a coding assistant in all areas: the code actually compiles and often works out of the box. Even better than chat, I can paste it&#
38.
▲
by
goldemerald
4y ago
Ahh I fell for it. I was fairly convinced by the 100+ 40GB files as weights to download, but when I saw the paper was written in a word doc I remembered to check the date. One interesting thing to note is that I could tell they used ChatGPT
39.
▲
by
goldemerald
4y ago
"Why put your happiness in other people's hands?" I ask myself that every time I've submitted a paper. It probably doesn't help that GPA is an easy surrogate for self-worth, which smoothly transformed into publishin
40.
▲
by
goldemerald
4y ago
Thank you!
41.
▲
by
goldemerald
4y ago
I've never cold connected to anyone. I wouldn't know how to go about doing it. My first idea is to find members of teams that cited one of my papers, but that makes me sound neurotic about papers, doesn't it. I've had th
42.
▲
by
goldemerald
4y ago
It's strange, choosing to stick out a Ph.D. (and not just get a masters and be happy) really cemented the competitiveness in my mind for years. I knew from the start I'd never make it to Deep Mind, but I had figured my experience
43.
▲
by
goldemerald
4y ago
Thank you for your perspective! I could definitely talk about all my work in extreme detail, though I have quite a hard time reaching the interview stage. I have not put much effort into my GitHub (I only have the code for 1 of my 5 project
44.
▲
by
goldemerald
4y ago
For some extra context, I have been applying every ML research engineer/scientist job that would take a new grad. I don't have any sense for what a hiring manager wants, it's especially hard with all the layoffs, I'm sur
45.
▲
by
goldemerald
4y ago
The example here is a bit worrying for the peer review process. I am not looking forward to my "peers" reviewing my paper by putting it through LLMs and blindly copy pasting the output. I can already imagine emailing the Area Chai
46.
▲
by
goldemerald
4y ago
I am going to write a lament about being a machine learning PhD (near) graduate. Whenever I see these very successful new PhD's, I just have an overwhelming sense of unfairness and sadness. Many of my papers were borderline rejects, de
47.
▲
by
goldemerald
4y ago
I am at loss towards the tech-landscape right now. A year ago, everyone was talking about hot the market was and how you could easily switch jobs for a 40% raise. Now all the megacorps are doing lay-offs and here comes Nintendo with the opp
48.
▲
by
goldemerald
4y ago
Using pixel-space or latent-space distances to measure if a generative model has simply memorized the training data is a common evaluation metric in ML literature. If the website is using the full training set of LAION 5-B used to train St
49.
▲
by
goldemerald
4y ago
This is a great website, but not in the way the authors intended. Based on some of the examples they explicitly provided, it is clear to me Stable Diffusion creates novel art. Here's a random example https://www.stableattri
50.
▲
by
goldemerald
4y ago
Wow, these songs were impressively similar to my queries. I would love for a interpolation playlist feature, where I put in 2 different songs (from say, classic rock and EDM), and I could get 10-20 songs that slowly change from the start so
51.
▲
by
goldemerald
4y ago
It seems to me that research scientist is a post-postdoc in industry, so all these textboxes for cover letter (or "introduce yourself") are the way for a researcher to say "Here's how my extremely specific research area
52.
▲
by
goldemerald
4y ago
I am currently looking for ML scientist jobs (just about to get my PhD) and this tool is a life saver for me. I have been procrastinating cover letters for weeks, and I just wrote 10 letter tonight with this. Obviously the output is a littl
53.
▲
by
goldemerald
4y ago
I also got GPT vibes too, so I checked the first two paragraphs with a couple AI detectors. And yep, it's likely not written by a human.
54.
▲
Victorian Hacker News
(victorianhackernews.com)
374 points
by
goldemerald
4y ago
|
106 comments
55.
▲
by
goldemerald
4y ago
I don't suppose you have a way of converting these models into a pytorch usable version, do you?
56.
▲
by
goldemerald
4y ago
The shape analogy doesn't really apply with modern language models. Each word gets its own context dependent high dimensional point. With everything being context dependent, simple transformations like rotations are impossible. A more
57.
▲
Do a cost-benefit analysis of your technology usage
(lesswrong.com)
2 points
by
goldemerald
5y ago
|
0 comments
58.
▲
Optimal AI agents tend to seek power
(twitter.com)
15 points
by
goldemerald
5y ago
|
18 comments
59.
▲
by
goldemerald
5y ago
This work is quite interesting, and I'm always happy to see such large improvements for memory usage and computation over existing high performant models. I imagine transformer-based models will be with the ML community for a very long
60.
▲
by
goldemerald
5y ago
You'd be interested in learning about the origin and application of Chi-Squared tests (or Fisher's exact test). Detecting changes in population means, and making sure those changes aren't just random noise, is quite straightf
More ›