AI

10487
updated 1h ago · 16 sources
Loading vital stats…
  1. Under-covered
    anthropic.com

    Accenture’s Faculty will lead model evaluation, red-teaming, alignment assessments and safeguard testing through a non-exclusive partnership.

  2. Josh Levy / lesswrong.com

    The authors introduce Persuasion Undermining Control, a framework for assessing how AI communication could compromise safety-relevant AI R&D and lab security oversight.

  3. Reed Albergotti / semafor.com

    Reed Albergotti says an apparent OpenAI-Anthropic agreement to slow AI research left staff confused, while the labs still need time to build compute capacity and address safety concerns.

  4. Andrea_Miotti / lesswrong.com

    The authors argue that no company, government, or individual knows how to keep a superintelligent AI system under human control and call for immediate action.

  5. Ashley Gold / semafor.com

    Nvidia’s Jensen Huang and Apple’s Tim Cook are reportedly attending, while officials discuss a possible adjacent White House meeting with AI CEOs and cooperation on AI safety.

  6. Adam Chlipala / lesswrong.com
  7. Venkat T / lesswrong.com