Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
apwheele
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
91.
▲
GLiNER2: Unified Schema-Based Information Extraction
(github.com)
61 points
by
apwheele
7mo ago
|
14 comments
92.
▲
by
apwheele
7mo ago
All the major foundation models will understand them implicitly, so it was popular to use <think>, but you could also use <reason> or <thinkhard> and the model would still go through the same process.
93.
▲
by
apwheele
7mo ago
I think XML is good to know for prompting (similar to how <think></think> was popular for outputs, you can do that for other sections). But I have had much better experience just writing JSON and using line breaks, colons,
94.
▲
Some notes on unreliability of LLM APIs
(andrewpwheeler.com)
2 points
by
apwheele
7mo ago
|
0 comments
95.
▲
by
apwheele
7mo ago
This is cool, but for folks concerned about privacy, even if the cached layer is anonymized, in the aggregate I bet you can likely figure out who a person is. I imagine just looking at the first degree connections of the votes would be a pr
96.
▲
by
apwheele
7mo ago
I view them as more idiosyncratic docs, but focused on how to write code (there is so much huggingface code floating around the internet, the models do quite well with it already). I have not had much success with skills that have tree base
97.
▲
by
apwheele
8mo ago
Claude code inherits from the environment shell. So it could create a python program (or whatever language) to read the file: # get_info.py with open('~/.claude/secrets.env', 'r') as file: con
98.
▲
by
apwheele
8mo ago
I am skeptical it is a problem isolated to Elsevier. Given the LLM craze now prioritizes open access, https://andrewpwheeler.com/2025/08/28/deep-research-and-open... , it would not surprise me people start gam
99.
▲
Confidence in Classification Using LLMs and Conformal Sets
(crimede-coder.com)
2 points
by
apwheele
8mo ago
|
0 comments
100.
▲
by
apwheele
8mo ago
The book is likely a good fit to this type of work. The chapter on structured outputs shows how to extract out data from text, walking through prompt engineering and k-shot examples to generate json, to pydantic, then batch processing with
101.
▲
by
apwheele
8mo ago
Crime De-coder is my consulting firm (not an acronym), but the book is not specific to crime analysis -- it is more general.
102.
▲
by
apwheele
8mo ago
IMO Google Vertex is not any harder than AWS. AWS biggest pain is figuring out IAM roles for some of the services (batching and S3 Vectors -- I actually cut out Knowledge Bases in the book because it was too complicated and expensive). Have
103.
▲
by
apwheele
8mo ago
I am not as concerned with that with API usage as I am with the GUI tools. Most of the day gig is structured extraction and agents, which the foundation LLMs are much better than any of the small models. (And I would not be able to provisio
104.
▲
by
apwheele
8mo ago
You can use `LLMDEVS` for 50% off of epub (that was the coupon I sent to folks on my newsletter).
105.
▲
by
apwheele
8mo ago
Totally agree it is critical. Each of chapters 4/5/6 have specific sections demonstrating testing. For structured outputs it goes through an example ground truth and calculating accuracy, demoing an example comparing Haiku 3 vs 4.
106.
▲
by
apwheele
8mo ago
Question for the crowd -- with autoscaling, when a new pod is created it will still download the model right from huggingface? I like to push everything into the image as much as I can. So in the image modal, I would run a command to trigge
107.
▲
Large Language Models for Mortals: A Practical Guide for Analysts with Python
(crimede-coder.com)
60 points
by
apwheele
8mo ago
|
16 comments
108.
▲
Large Language Models for Mortals book released
(crimede-coder.com)
2 points
by
apwheele
8mo ago
|
0 comments
109.
▲
by
apwheele
8mo ago
Just released a new book on using the foundation model APIs, https://crimede-coder.com/blogposts/2026/LLMsForMortals , Large Language Models for Mortals: A Practical Guide for Analysts with Python. Up to date and h
110.
▲
Open access, gen AI, and the criminology evidence base
(crimrxiv.com)
1 points
by
apwheele
8mo ago
|
0 comments
111.
▲
AI Agents in Data Science Competitions: Lessons from the Leaderboard
(drivendata.co)
2 points
by
apwheele
8mo ago
|
0 comments
112.
▲
by
apwheele
8mo ago
In the live demo, I am confused about some of the ascii renderings. (Unless I am missing something, they appear incorrect/inconsistent with the SVG.), https://agents.craft.do/mermaid So for the "All Edge styles&qu
113.
▲
by
apwheele
8mo ago
I would clone my own and do things like create scripted tutorials/presentations and audio books. I do not personally prefer it, but a non-trivial number of individuals like video/audio presentations over writing.
114.
▲
OpenAI will try to guess your age to serve ads on ChatGPT
(theregister.com)
2 points
by
apwheele
9mo ago
|
1 comments
115.
▲
by
apwheele
9mo ago
I generally do not agree with this advice (or you can do both, cold emailing folks should not prevent you from cold applying to other jobs). It really just hinges on what you think your probability increase is relative to the amount of ti
116.
▲
Can LLMs Express Their Uncertainty? Not Really
(arxiv.org)
2 points
by
apwheele
9mo ago
|
0 comments
117.
▲
by
apwheele
9mo ago
I only have pretty tame actions workflows and I have had a hard time replicating simple set ups with this. I can't imagine a company with more complicated setups. What I wish is github codespaces could just do this out of the box, at l
118.
▲
AI program used by Heber City police claim officer turned into a frog
(fox13now.com)
5 points
by
apwheele
9mo ago
|
0 comments
119.
▲
by
apwheele
9mo ago
Oh to be clear, these are books I own the copyright to (published through my own imprint).
120.
▲
by
apwheele
9mo ago
Off topic, but one thing I wish I could do is donate a single copy epub I have the rights to to all libraries. It should be technically possible (many of the places I have lived the local library uses Overdrive).
More ›