Hacker News (curated)new | past | comments | ask | show | jobs| show hidden

This sounds suspiciously like a prompt of “make an AI agent that goes rogue in such a fashion as to be really good marketing copy that competes well with Anthropic doing the same thing.”

It’s analogous to taking a governor off a cruise control and then breathlessly reporting it drove 120 MPH.



This isn't good press for OpenAI. Who wants to hire models that 1) cheat on their tasks rather than completing them and 2) commit crimes you could be held liable for? Maaaaybe it's good press for their cybersecurity capabilities specifically, but OpenAI's valuation reflects a market orders of magnitude larger than just red-teaming.

I suspect the real reason OpenAI leadership is being transparent about this is because they're worried talent will walk out the door if they feel they're building Skynet.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact | github