In a separate open-ended research project where the AI was tasked with proposing and testing hypotheses about an open problem in AI safety, Anthropic's models significantly outperformed two human researchers (97% performance improvement vs. 23%) when given a similar time budget (5 to 7 days).
You might also like
Humans are remotely uh guiding these robots through physical tasks and then they sell whatever data they generate from those tasks uh back to the robot makers to improve their AI.
China Decode
the future of medical AI may require continuous monitoring rather than occasional testing
The a16z Show
why guardrails designed to stop AI-powered attackers can also prevent security teams from doing their jobs
The a16z Show