Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
antiochIst
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
antiochIst
8mo ago
The story spread from Bloomberg at 12:01 PM as first to report, then amplified by x.com and 9to5mac, followed by feeds.macrumors, machash, and appleworld.today. Yandori tracked 45 sources with 10 citations showing how the news propagated:
2.
▲
by
antiochIst
8mo ago
You can track how this story spread from CNN at 3:32 PM as first to report, then whistleblower.org, peters.senate.gov, and eventually scienceblog.com and democracydocket.com. Yandori shows 54 sources picked it up with 48 citations: https:&
3.
▲
by
antiochIst
8mo ago
Interesting to see how this news spread across 45+ media outlets. Yandori tracked the news flow from AP News at 9:31 AM as first to report, then blog.google as the amplifier around 8:07 AM, followed by mashable and other tech sites. You can
4.
▲
by
antiochIst
8mo ago
Interesting to see how this story developed across 60 news sources: https://yandori.io/news-flow/story/2026-01-21-10020950-takea...
5.
▲
by
antiochIst
8mo ago
You can see how this story spread across 52 sources over 21 hours here: https://yandori.io/news-flow/story/2026-01-22-10318781-amazo...
6.
▲
by
antiochIst
9mo ago
Full news story coverage/flow here: https://yandori.io/news-flow/story/2026-01-07-6507910-federa...
7.
▲
by
antiochIst
9mo ago
Full news coverage/flow here: https://yandori.io/news-flow/story/2026-01-07-6507910-federa...
8.
▲
by
antiochIst
10mo ago
Here is a clear example (sad news) but you can see how it breaks (via twitter here) and then flows: https://yandori.io/news-flow/story/2025-12-18-3003874-nascar...
9.
▲
Show HN: Improved real-time news propagation tracking via source citation graphs
(yandori.io)
1 points
by
antiochIst
10mo ago
|
1 comments
10.
▲
by
antiochIst
10mo ago
I have a bunch of heuristics - but broadly speaking a new domain + a new good looking site. ie: some legit pages, legit socials, and some real owner/purpose behind it.
11.
▲
by
antiochIst
10mo ago
I don't know if drop shipping is still viable, but there are lots of those types of sites so...
12.
▲
I tracked 677k newly launched websites in November. Here's the breakdown
(websitelaunches.com)
2 points
by
antiochIst
10mo ago
|
6 comments
13.
▲
by
antiochIst
10mo ago
I run a system that tracks newly launched websites worldwide. In November it detected 677,544 new sites across 392 countries. Some findings from the dataset: Geography (top countries): United States: 253,589 India: 34,127 Canada: 20,263 Uni
14.
▲
by
antiochIst
10mo ago
Thanks - I got fix for this.
15.
▲
by
antiochIst
10mo ago
FYI - I'll integrate this if anyone want to pay for twitter api fees.
16.
▲
by
antiochIst
10mo ago
I think this could be done, but would require paying more than I want to for the highest level of api access...
17.
▲
by
antiochIst
10mo ago
Yea there is some spam stuff for sure... working on improving filtering it out... I get most of it, but I think especially around the holiday some stuff is getting through... Some black friday deals were actually hitting like news does...
18.
▲
by
antiochIst
10mo ago
You can kinda tell based on the distribution. Organic spread has less similarity between articles, less syndication, more spread out in timeline of releases... Some stories are very clearly manufactured
19.
▲
by
antiochIst
10mo ago
I'm polling rss feeds from a bunch top 200k sites in the world. Thanks for that bug feedback - ill get fix.
20.
▲
by
antiochIst
10mo ago
ehh, timezones handles just with some basic parsing logic... I'm not pulling from social media yet.
21.
▲
by
antiochIst
10mo ago
Yea I feel you... Honestly I kinda just whipped this thing up in context of a larger project I'm working on.. so i have not given much thought to who it will serve. "rooting out bias" is interesting idea... But a bit negativ
22.
▲
by
antiochIst
10mo ago
It's a whole thing... I run a project called websitelaunches, so I have index of basically the whole internet (500M+) sites. I took the top ~200k news related sites from there that had rss feed.
23.
▲
by
antiochIst
10mo ago
I'm basically throwing away non english articles for now... I'll pry get them in later, but I want to get english right first before trying to move to other languages... The embeddings themselves will (pry) cluster ok in different
24.
▲
by
antiochIst
10mo ago
yea, what im currently doing is pretty simple check on published at date from the rss feed (with some small validation checks)... but its causing issues bc it can be wrong and mess up everything... I think checking source in story is next s
25.
▲
by
antiochIst
10mo ago
Yea not all major have rss feeds, but it seems like the majority still do. No translation yet. I think the biggest problem is im relying on published date from the news source itself too much and its wrong sometimes... not super often, but
26.
▲
by
antiochIst
10mo ago
Currently I'm using Snowflake’s Arctic embedding model on the whole story not just the title, to cluster stories. There are still some issues, but its not as simple as looking at title publish date. Yea, I need to do some work on impro
27.
▲
Show HN: Real-time system that tracks how news spreads across 200k websites
(yandori.io)
256 points
by
antiochIst
10mo ago
|
74 comments
28.
▲
We Tracked Every Website That Launched in September 2025. The Data Is Wild
(websitelaunches.com)
1 points
by
antiochIst
1y ago
|
1 comments
29.
▲
Tool that tracks all new website launches
(websitelaunches.com)
4 points
by
antiochIst
1y ago
|
1 comments
30.
▲
by
antiochIst
1y ago
I built a system that detects (almost all) new website launches. Right now it only covers .com and English sites, but support for more TLDs and languages is coming. You can browse launches, filter by category or location, and upvote/bo
More ›