From the source
OpenAI announces a new research direction called weak-to-strong generalization, which explores using the generalization properties of deep learning to enable weak supervisors to control strong models, with initial results.
From the source
From the source
OpenAI announces a new research direction called weak-to-strong generalization, which explores using the generalization properties of deep learning to enable weak supervisors to control strong models, with initial results.
From the source
We present a new research direction for superalignment, together with promising initial results: can we leverage the generalization properties of deep learning to control strong models with weak supervisors?
openai.com