← Flash Papers
Weak-to-Strong Generalization: Eliciting Strong Capabilities with Weak Supervision
Share
Share this Flash Paper
Weak-to-Strong Generalization: Eliciting Strong Capabilities with Weak Supervision
https://flashpapers.ai/p/hrarug52he
Copy link
Post to X
Post to Bluesky
Share on LinkedIn
Submit to Hacker News
Submit to Reddit
Email a link
Source PDF ↗
Open standalone ↗
Make your own
Make your own paper
Collin Burns, Pavel Izmailov +10 more
cs.CL
2023
arXiv:2312.09390
Code
· 2.6k★
· updated 2y ago
Details ▸
Details ▴
Performance Gap Recovered (PGR):
nearly 80%
· NLP Benchmarks
Performance Gap Recovered (PGR):
roughly 10%
· ChatGPT Reward Modeling
Performance Gap Recovered (PGR):
nearly 80%
· NLP Benchmarks
Performance Gap Recovered (PGR):
roughly 10%
· ChatGPT Reward Modeling