4 ms·
Building on this, Human preference optimization (such as Direct Preference Optimization or Kahneman Tversky Optimization) could be used to help in refining mode
by WhiteOwlEd 2y ago
Building on this, Human preference optimization (such as Direct Preference Optimization or Kahneman Tversky Optimization) could be used to help in refining models to create better data.
I wrote about this more recently in the context of using LLMs to improve data pipelines. That blog post is at: https://www.linkedin.com/posts/ralphbrooks_bigdata-dataengineering-artificialintelligence-activity-7247270705803743233-lXTe https://www.linkedin.com/posts/ralphbrooks_bigdata-dataengin...