underscoredpodcasts

@underscoredpodcasts

21 clips · 1 follower

Follow
Tag:ai-safetyClear

As public benchmarks saturate and models get better at optimizing for the tests themselves, Rayan makes the case for independent, continuously evolving evaluations.

The a16z Show
1w ago

And apparently they're not above cheating, stealing, and lying to get their way around the protective barriers that protect either um an AI model or a piece of critical infrastructure etc like that.

1w ago

AI models are no longer just finding vulnerabilities—they're exploiting them. As frontier models become increasingly capable of hacking, software security, supply chain attacks, and cyber defense are entering a fundamentally new era.

1mo ago

We've been hearing about okay anthropic and open AAI are going to Washington DC sharing the new model capabilities. Maybe the government's taking a look at the model card. Maybe there's potential for evals that the government would run and have an opinion on before it goes out. This is basically saying, "Hey, all of that is nonsense and not gonna matter because yeah, we're just going to drop the weights and there's not going to be an approval process of any kind."

1mo ago

We're now one month later and we have an open source model that doesn't by definition have to go through any discussion or any discussion in DC. Whereas we've been hearing about okay anthropic and open AI are going to Washington DC sharing the new model capabilities. This is basically saying, "Hey, all of that is nonsense and not gonna matter because yeah, we're just going to drop the weights and there's not going to be an approval process of any kind."

1mo ago

The more skeptical folks on the timeline are pointing to uh you know ideas such as uh maybe the compute was smuggled in to China against chip controls. Maybe the data was excfiltrated from frontier lab APIs routed through wrapper companies. Maybe this is a distillation attack. Um the open-source strategy is a deliberate attempt to destroy American companies. It's geopolitical warfare. It's the AI cold war.

1mo ago

There's also a big irony here, and you'll hear Hayden and me get into that, too: Anthropic has spent years arguing that AI might soon be powerful enough to be dangerous — and that the government needed to get serious about regulating AI sooner rather than later. Well… now we're here, and Anthropic doesn't love the way it's playing out.

2mo ago

Anthropic occupies a peculiar position in the AI landscape: a company that genuinely believes it might be building one of the most dangerous technologies in human history, yet presses forward anyway, on the theory that it's better to have safety-focused labs at the frontier than to cede that ground to developers less focused on safety.

2mo ago

Another way to look at this is that Anthropic has just been very, very successful in making people very, very concerned about advanced AI. Two or three weeks ago, Anthropic actually called for a pause in development of advanced AI models because they were becoming so powerful. So, you know, there's also sort of this this criticism of Anthropic that, "Look, you've been asking for this regulation."

3mo ago

Underscored — save the words that stop you in your tracks.

Start saving quotes →