3 ms·
Isn’t this basically the Swiss cheese model? If your two input AIs hallucinate, or your consensus AI misunderstands the input, you will still have confabulation
by stereo 1y ago
Isn’t this basically the Swiss cheese model? If your two input AIs hallucinate, or your consensus AI misunderstands the input, you will still have confabulations in the output?
- TheKelsbee 1y agoI have this same thought, and have tried similar approaches. OP: Have you trained or fine tuned a model that specifically reasons the worker model inputs against the user input? Or is this basically just taking a model and turning the temperature down to near 0?
- kuberwastaken 1y agoLow temperature, heavy prompting to answer in a structured way. Sadly can't fine train models since this is API based but the approach does work!
- kuberwastaken 1y agoFrom all my testing, this never really happened even once honestly, plus the judge model (that I've kept strictly a reasoning model) also evaluates individually before "judging" the consensus.