arxiv:2412.14093
Jared Kaplan
FrizzleFried
AI & ML interests
None yet
Recent Activity
authored
a paper
about 24 hours ago
Alignment faking in large language models
Organizations
None yet
models
None public yet
datasets
None public yet