Question about your preference training process
#1
by AustinAligned - opened
Hi, I came across the Persona Maker series and your earlier work targeting positivity bias while researching how people run preference training on open-weight models. I'm curious how you build your preference data for those goals, and how you figure out whether a run actually changed the model the way you wanted.
Would you be open to a quick chat? Happy to do email if easier, I'm at [email protected].
Just trying to learn, not selling anything.
Thanks either way,
Austin
SicariusSicariiStuff changed discussion status to closed