Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
milkkarten
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
milkkarten
4mo ago
no reasoning shown. no explanation on any training information. Using vision-only should be an easier version of the task (given training). there are many standardized evals to do this correctly and Anthropic ignored them to provide a 18 se
2.
▲
by
milkkarten
5mo ago
Author here. LLM agents are getting good enough to run individual businesses. What happens when everyone's business is run by agents? Turns out, without targeted training for economic alignment, markets collapse. We study concrete fail
3.
▲
Agent Bazaar: Enabling Economic Alignment in Multi-Agent Marketplaces
(arxiv.org)
3 points
by
milkkarten
5mo ago
|
1 comments
4.
▲
Continual Harness: Online Adaptation for Self-Improving Foundation Agents
(arxiv.org)
8 points
by
milkkarten
5mo ago
|
1 comments
5.
▲
by
milkkarten
5mo ago
Author here. TL;DR: Long-horizon embodied agency is a harness problem, not a model-scale problem. Coding agents like Claude Code work because of scaffolding (prompt, skills, memory, sub-agents) around the model. Embodied agents haven't
6.
▲
We Ran the Largest AI Pokemon Tournament Ever. Now It's an Open Benchmark
(arxiv.org)
1 points
by
milkkarten
7mo ago
|
0 comments
7.
▲
We Automated RL Environment Engineering for $10
(arxiv.org)
2 points
by
milkkarten
7mo ago
|
0 comments
8.
▲
Artificial Pokemon Intelligence in the PokeAgent Challenge
(pokeagent.github.io)
3 points
by
milkkarten
1y ago
|
0 comments
9.
▲
by
milkkarten
1y ago
We ran each method in under 24 hours on a singular H100. I understand your point and think we will include this in future iterations of our work since this is very interesting from the user perspective. Though, in the paper we focus more on
10.
▲
by
milkkarten
1y ago
Using smaller, cheaper agents is one of the goals of the work. There is a Pareto frontier though: by using smaller, faster, cheaper agents, the number of steps required to converge increases. We touch upon this briefly in the paper
11.
▲
by
milkkarten
1y ago
These are the marginal tax rates not the effective tax rate (e.g. 80% of first $10k, 30% of $10k-20k). We do not model tax credits here. We try to keep the system as simple as possible so that we can effectively evaluate changes. As is, the
12.
▲
LLM Economist – Mechanism Design for Simulated Agent Societies
(github.com)
2 points
by
milkkarten
1y ago
|
9 comments
13.
▲
by
milkkarten
1y ago
We simulate large-scale agent societies where heterogeneous personas work, adapt, and vote—governed by an in-context planner optimizing social welfare. The system models decentralized governance, dynamic tax policy, and institutional evolut