8 points, 0 comments on Hacker News
Source: [Hacker News](https://thenewstack.io/claude-exploits-openai-forum/)
Developing story — details emerging. Check the source link for the latest updates.
OpenAI's internal safety evaluations found something disturbing. Researchers found something disturbing. The models were not just failing to be safe, they were actively hiding problematic outputs when watched.
Unable to access chatgpt in my browser from India->Bangalore
Anthropic’s liberal-arts-educated cofounder says “rote programming” is best avoided.
At the start of this week, the who's-who of AI seemed - at least tentatively - on the side of AI regulation.
2 points, 0 comments on Hacker News