4 ms·
Can an expert explain how this protects against adversarial actors? At a glance it looks like something akin to a computing a checksum that's locality sensitiv
by xmasotto 1y ago
Can an expert explain how this protects against adversarial actors?
At a glance it looks like something akin to a computing a checksum that's locality sensitive, so it's robust to floating point errors, etc.
What's to stop someone from sending bad data + a matching bad checksum?
- yorwba 1y agoThe validation procedure is described on page 8 of the TOPLOC paper: https://arxiv.org/abs/2501.16007 https://arxiv.org/abs/2501.16007 The checksum is validated by redoing the computation, but making use of the fact that you already have the entire response to enable greater parallelism than when generating it one token at a time.
- DoctorOetker 1y agoTOPLOC attempts to detect model substitution, i.e. responses being generated by a different model than requested, it comes with certain caveats, as far as I can tell the TOPLOC paper considers verifiable learning / training as out of scope.