AI
10485updated 1h ago · 16 sources
AI
Loading vital stats…
- Under-covered
anthropic.comUnder-coveredAccenture’s Faculty will lead model evaluation, red-teaming, alignment assessments and safeguard testing through a non-exclusive partnership.
- Josh Levy / lesswrong.com
The authors introduce Persuasion Undermining Control, a framework for assessing how AI communication could compromise safety-relevant AI R&D and lab security oversight.
- anthropic.com
The company introduces a prototype Anthropic R&D Automation Index and says independent third-party evaluators will monitor its metrics and safety practices.
- Linch / lesswrong.comDeveloping
Linch’s post, linking to The Wall Street Journal and Hacktron, says the hackers likely accessed almost all of OpenAI’s research and production code, but not its literal model weights.
- openai.com
The API release is positioned as a way to build more natural voice experiences with GPT‑Live‑1.
- Reed Albergotti / semafor.com
- Prashant Rao / semafor.com
- Andrea_Miotti / lesswrong.com