Hacker News
new
past
comments
ask
show
jobs
points
by
lossolo
21 hours ago
|
comments
by
altcognito
20 hours ago
|
next
[-]
Well, yes, but remember there is the reinforcement learning that is applied after, and the system prompts that will bend the results.
reply
by
lossolo
20 hours ago
|
parent
|
[-]
Yeah, agree on both points. You can embed any bias you want using RL, regardless of training data.
reply
by
slibhb
20 hours ago
|
prev
|
[-]
No, they're very clearly trained to be truthful.
reply