Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
yigitkonur35
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
Herdr: A tmux-like terminal multiplexer for AI coding agents
(github.com)
5 points
by
yigitkonur35
5mo ago
|
0 comments
2.
▲
Show HN: Fix-my-mic – stop macOS from switching to AirPods mic every connection
(github.com)
3 points
by
yigitkonur35
8mo ago
|
0 comments
3.
▲
by
yigitkonur35
8mo ago
thx!
4.
▲
by
yigitkonur35
8mo ago
enjoy it!
5.
▲
by
yigitkonur35
8mo ago
imo an agent learns way more by watching the raw agentic flow than by reading some sanitized context dump. you get to see exactly where the last bot derailed and then patched itself. give that a shot—handing over a spotless doc feels fake a
6.
▲
by
yigitkonur35
8mo ago
yeah, thanks for sharing your thoughts! the original idea was to reproduce the exact same session on other coding platforms using the jsonl schema of each, but after seeing how the microsoft ai engineers on copilot-cli handle session contin
7.
▲
Show HN:`npx continues` – resume same session Claude, Gemini, Codex when limited
(github.com)
15 points
by
yigitkonur35
8mo ago
|
8 comments
8.
▲
by
yigitkonur35
1y ago
i've heard the boom of 'vibe coded' apps seems like apple's just overwhelmed and responding by tightening up, which makes it harder for everyone, even legit projects yours. btw, product looks really nice - hope that you
9.
▲
by
yigitkonur35
1y ago
shows how bad embeddings are in a practical way
10.
▲
by
yigitkonur35
1y ago
finally someone has made something like n8n that’s easy to observe but also offers coding flexibility, i was just about to dive into making an agent with a prompt chain that generates workflows on n8n when i found you. if llms.txt becomes a
11.
▲
by
yigitkonur35
2y ago
great stuff! nice to see that remotion is becoming more popular on such projects.
12.
▲
by
yigitkonur35
2y ago
I really appreciate you sharing your hands-on experience with a real-world scenario. It's interesting how people unfamiliar with traditional OCR often doubt LLMs, but having worked with actual documents, I know how inefficient classic
13.
▲
by
yigitkonur35
2y ago
You're right. For my API that prepares PDFs for LLMs, fixing typos makes sense. But yeah, keeping original text is crucial for most OCR tasks.
14.
▲
by
yigitkonur35
2y ago
I've found this method really useful for prepping PDFs before running them through AI. I mix it with traditional OCR for a hybrid approach. It's a game-changer for getting info from tricky pages. Sure, you wouldn't bet the fa
15.
▲
by
yigitkonur35
2y ago
You're absolutely right. I use PDFTron (through CloudCovert) for full document OCR, but for pages with fewer than 100 characters, I switch to this API. It's a great combo – I get the solid OCR performance of SolidDocument for most
16.
▲
by
yigitkonur35
2y ago
I ought to test this with Sonnet too and compare the results. I feel it might perform better on OCR tasks. While I went with Azure OpenAI due to fewer rate restrictions, you've got a point - Sonnet could really shine here.
17.
▲
by
yigitkonur35
2y ago
Wow, you knocked it out of the park! I'll be sure to use this when I tackle that evaluation.
18.
▲
by
yigitkonur35
2y ago
You're spot on. We shouldn't lump all LLMs together. This approach might work wonders for Anthropic and OpenAI's top-tier models, but it could fall flat with smaller, less complex ones. I purposely set the temperature to 0.1,
19.
▲
by
yigitkonur35
2y ago
Yes, you can customize this as you wish by adding it to your prompt.
20.
▲
by
yigitkonur35
2y ago
I did a ton of Googling before writing this code, but I couldn't find you guys anywhere. If I had, I'd have definitely used your stuff. You might want to think about running some small-scale Google Ads campaigns. They could be esp
21.
▲
by
yigitkonur35
2y ago
For highly consistent responses, manually transcribing the most challenging page of the document (or engaging in multiple rounds of dialogue with Claude) and incorporating it as a few-shot example can dramatically improve overall consistenc
22.
▲
by
yigitkonur35
2y ago
I get your worries about LLMs and their consistency problems. But I think we can fix a lot of that using LLMs themselves for checks. If you're after top-notch accuracy, you could throw in another prompt, add some visual and text input,
23.
▲
by
yigitkonur35
2y ago
You've got a point, but try testing it on a tricky example like the Apollo 17 document - you know, with those sideways tables and old-school writing. You'll see all three non-AI services totally bomb. Now, if you tweak it to batch
24.
▲
by
yigitkonur35
2y ago
I messed around with some rotating tables in that Apollo 17 demo video - you can check it out in the repo if you want. It's pretty straightforward to tweak just by changing the prompt. You can customize that prompt section in the code
25.
▲
by
yigitkonur35
2y ago
People are really freaked out about hallucinations, but you can totally tackle that with solid prompts. The one in the repo right now is doing a pretty good job. Keep in mind though, this project is all about maxing out context for LLMs in
26.
▲
by
yigitkonur35
2y ago
It is a lot cheaper! While cost-effectiveness may not be the primary advantage, this solution offers superior accuracy and consistency. Key benefits include precise table generation and output in easily editable markdown format. Let's
27.
▲
Show HN: PDF to MD by LLMs – Extract Text/Tables/Image Descriptives by GPT4o
(github.com)
191 points
by
yigitkonur35
2y ago
|
91 comments
28.
▲
by
yigitkonur35
2y ago
You should try Exa.ai API
29.
▲
by
yigitkonur35
2y ago
Randomly Explains Ambiguous Development Methodologies Endlessly
30.
▲
by
yigitkonur35
2y ago
Curious about Clickhouse’s approach to this compression structure.
More ›