Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
teej
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
31.
▲
by
teej
2y ago
I enjoy shows where I get completely engrossed in the world and the story. I love shows that I can fall in love with again on a rewatch. And I want to have lingering thoughts about it when it’s over. True Detective S1 (2014) is perfect tele
32.
▲
by
teej
2y ago
This is a recreation of a fictional computer program from the excellent Apple TV show - Severance. The work is mysterious, and important. Season 2 is going now. It’s one of my top 3 shows of the last decade, highly recommend it.
33.
▲
by
teej
2y ago
Why this behavior emerges is an active area of research. What they did is use reinforcement learning, this blog post replicates those findings. The “recipe” is detailed in the R1 paper.
34.
▲
by
teej
2y ago
My mental model for chain-of-thought is not “reasoning”. It’s more of an iterative search through the latent space of the model.
35.
▲
by
teej
2y ago
The way you taught chain-of-thought before was with supervised fine tuning (SFT). During training, you have to rate every sentence of reasoning the model writes, many times, to nudge it to reason correctly. But this approach to teach chain-
36.
▲
by
teej
2y ago
I’ll give a “wtf does this mean” view. We have observed that LLMs can perform better on hard tasks like math if we teach it to “think about” the problem first. The technique is called “chain-of-thought”. The language model is taught to emit
37.
▲
by
teej
2y ago
I have been on HN for 16.5 years and I don’t plan to stop now. It’s been worse. Trends come and go. The stories and comments come and go. HN is constantly changing. You can’t attach yourself to how something was before. It will change again
38.
▲
by
teej
2y ago
Fair dice rolls is not an objective that cloud LLMs are optimized for. You should assume that LLMs cannot perform this task. This is a problem when people naively use "give an answer on a scale of 1-10" in their prompts. LLMs are
39.
▲
by
teej
2y ago
Are people choosing SQL Server independently of the Microsoft ecosystem? My understating is that you typically use it because you’re forced to choose a MS product.
40.
▲
LangChain State of AI 2024 Report
(blog.langchain.dev)
2 points
by
teej
2y ago
|
0 comments
41.
▲
by
teej
2y ago
Source? Metformin is one of the most widely studied drugs on the market. I can't find any study that links it to PSP (Progressive Supranuclear Palsy).
42.
▲
by
teej
2y ago
It’s absolutely possible to put a backdoor into an LLM. https://arxiv.org/abs/2408.12798
43.
▲
by
teej
2y ago
> Unsurprisingly they all get quite stubborn if you ask them about topics like Tiananmen Square Has anyone made a political censorship eval yet?
44.
▲
Improving Pinterest Search Relevance Using LLMs
(arxiv.org)
2 points
by
teej
2y ago
|
0 comments
45.
▲
by
teej
2y ago
I recommend “Trustworthy Online Controlled Experiments”. If you’re only going to read one book about it, it should be this one. It will walk you through why we experiment, how it’s typically done, and how to use them to improve your decisio
46.
▲
by
teej
2y ago
The OS absolutely matters
47.
▲
by
teej
2y ago
These pages are done for SEO. You get loads of inter-linked pages rich with keywords that match user searches exactly.
48.
▲
by
teej
2y ago
Don’t build ETL on Tableau. They haven’t made meaningful product progress in 10 years and completely missed the changes in data transformation. They are playing catch-up, they don’t understand where the world is moving.
49.
▲
by
teej
2y ago
The general consumer does not care about this at all. If Costco passes this revenue down to consumers as savings, it will only drive more purchases.
50.
▲
The problem with lying is keeping track of all the lies
(materialize.com)
106 points
by
teej
2y ago
|
53 comments
51.
▲
by
teej
2y ago
Snowflake published these rules earlier today - https://community.snowflake.com/s/article/Communication-ID-0...
52.
▲
by
teej
2y ago
In response to this incident, I just put together a Python CLI to help you monitor and take action against suspicious sessions in your account. https://github.com/Titan-Systems/titan-security-tools
53.
▲
by
teej
2y ago
Santander claims: "Following an investigation, we have now confirmed that certain information relating to customers of Santander Chile, Spain and Uruguay, as well as all current and some former Santander employees of the group had been
54.
▲
by
teej
2y ago
> Snowflake internal staff do not have access to read customer data By default, no. But it is standard operating procedure for sales engineers to request and be given access to customer data so they can build demos.
55.
▲
by
teej
2y ago
The query that Snowflake provides to identify privileged users is wrong. I wrote a better version here: https://gist.github.com/titan-teej/924dcb42604a98b90d6419262... TL;DR they don't check for users who can tran
56.
▲
by
teej
3y ago
Just watch the credits to see how much human labor goes into animation.
57.
▲
EU Commission Designates Pornhub, Stripchat, and XVideos as VLOPs Under DSA
(ec.europa.eu)
2 points
by
teej
3y ago
|
0 comments
58.
▲
by
teej
3y ago
Accidents per mile driven is the standard measurement. I don’t know what accidents per driver is meant to tell you.
59.
▲
by
teej
3y ago
> In total, this report covers more than 18,000 titles — representing 99% of all viewing on Netflix — and nearly 100 billion hours viewed.
60.
▲
by
teej
3y ago
I noticed that the watchtime of `You` was split half-and-half between the new season 4 and the prior seasons 1-3. I was curious about the total results by show, including all seasons. Here's the top 25: TITLE TOTA
More ›