2 ms·
It was a Dunning-Kruger amplifier. It has graduated into being an active and fairly precisely targeted thought shaping tool. I worked on Ads and Feed Ranking
by reinitctxoffset 4mo ago
It was a Dunning-Kruger amplifier.
It has graduated into being an active and fairly precisely targeted thought shaping tool.
I worked on Ads and Feed Ranking at FB/IG and we never dreamed of the scope for shaping behavior and opinion that is now routinely deployed by frontier and near frontier vendors. RLHF is basically feed ranking in the first place, preference gradient with no ground truth referee, late SFT on amplifying data sets, and affine injections into the residual stream with a fluent, earnest base model that the public has been conditioned to regard as omniscient and wise?
Yeah that's fucking mind control when applied at scale. We did some sketchy shit a decade ago, this is next level.