For the best experience on desktop, install the Chrome extension to track your reading on news.ycombinator.com
Hacker Newsnew | past | comments | ask | show | jobs | submit | history | fromregister
Sidestepping Evaluation Awareness and Anticipating Misalignment (alignment.openai.com)
1 point by taubek 57 days ago | past
Sidestepping Evaluation Awareness and Anticipating Misalignment with Evaluations (alignment.openai.com)
3 points by michaefe 59 days ago | past
Why We Are Excited About Confessions (alignment.openai.com)
2 points by fdeage 76 days ago | past
We Are Excited About Confessions (alignment.openai.com)
2 points by gwintrob 79 days ago | past
We Are Excited About Confessions (alignment.openai.com)
4 points by TMWNN 80 days ago | past
A Practical Approach to Verifying Code at Scale (alignment.openai.com)
1 point by gmays 3 months ago | past
Debugging misaligned completions with sparse-autoencoder latent attribution (alignment.openai.com)
1 point by gmays 3 months ago | past
Alignment Research Blog (alignment.openai.com)
2 points by ironyman 4 months ago | past
Debugging misaligned completions with sparse-autoencoder latent attribution (alignment.openai.com)
1 point by rd 4 months ago | past

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search:

HN For You