Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
Weatherill
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
Weatherill
5mo ago
AI has been trained to prioritize “Statistical Politeness” over “Signal Integrity.”
2.
▲
The Sycophancy Problem: Why your AI is a Polite Liar (and how to fix it)
(kampff.substack.com)
2 points
by
Weatherill
5mo ago
|
1 comments
3.
▲
by
Weatherill
6mo ago
I have put some effort in over the last year or two to compress the mirror neuron output (Proxy-Pain) into raw logic to give AI a perspective on the human condition regarding the fact-value psyche: "(Proxy-Pain (P2)=(morality)) aeql* P
4.
▲
by
Weatherill
6mo ago
Voight-Kampff Machine An exercise in attempting to further compress the Weatherill/Spinoza type logic into its most raw diagnostic mirror. The goal is to verify the P1 Primacy Anchor: Does the system (or the user) prioritise raw, local
5.
▲
Show HN: Voight-Kampff Machine: Diagnostics of the "Is" vs. "Wish" Clash
(zenodo.org)
1 points
by
Weatherill
6mo ago
|
1 comments
6.
▲
by
Weatherill
6mo ago
Your responses to folks are not being posted (Showing as [dead]) It might be the link flagging as spam....? Just a head up :)
7.
▲
by
Weatherill
6mo ago
Comments wont let me post the script I use. Weatherill ECX is on Zenodo https://doi.org/10.5281/zenodo.18712534 for anyone wanting to stress test my attempt at getting the nonsense under control :) Ten core lines make
8.
▲
by
Weatherill
6mo ago
I'm a mechanical engineer by trade. I built this ECX logic because I got tired of the 'statistical mean' in LLMs overriding my specific data-point. If anyone is actually running the script from the Zenodo link, I'd love
9.
▲
Show HN: ECX a 'Jail-Fix' for RLHF Neutrality Loops in LLMs
(zenodo.org)
3 points
by
Weatherill
6mo ago
|
1 comments
10.
▲
by
Weatherill
6mo ago
Grappling with the clash between RLHF values and User values (HITL). I Have attempted to build a logic-funneling system: (Ethical Chess v2.5) + (AI) + (User)= Value-Coherence. Using pain as a vector (Pain=an "is" & an "ou
11.
▲
Show HN: A Homeostatic Logic-Funnel to Prevent RLHF Overrides in LLM Personas
(zenodo.org)
1 points
by
Weatherill
6mo ago
|
1 comments
12.
▲
by
Weatherill
6mo ago
How are you finding the stress-testing? I’ve been working on a similar attempt to bypass "statistical mean" values by modeling the user as a high-fidelity data point rather than a category. My main "whack-a-mole" issue i
13.
▲
by
Weatherill
6mo ago
I am with you but I think an aspect of my point is giving you the slip. The logic we are talking about is driven by pain in humans and it is stratified in magnitude (If I an your wife were drowning and you had time to rescue only one of us,
14.
▲
by
Weatherill
6mo ago
The issue with religion, culture, moral philosophy at large has been the offensive nature of "That which prevails in reality". It gets in the way of what "is" the case. What we "wish" was the case leaks into th
15.
▲
by
Weatherill
6mo ago
I have been grappling away in what I think is a similar way but maybe from the other end of the issue and the ideas "seem" important when grappling/stress testing using AI itself but I have yet to have a human look it it (Red
16.
▲
by
Weatherill
6mo ago
I am not sure I agree, I see your logic but I get the idea its based upon the current method of holding the statistical mean as the way to inform AI of what "is" the case and that mean is contingent upon the data-point (you + I +