
7 segments available
Ian Webster (Promptfoo) on DeepSeek’s Security Vulnerabilities Ian Webster, founder of Promptfoo, joins a16z partner Joel de la Garza to break down the security risks embedded within DeepSeek’s reasoning model. As generative AI systems become more powerful, they also become more susceptible to attack. Ian explains how vulnerabilities like jailbreaks, backdoors, and model censorship can be exploited—and what developers and enterprises can do to defend against them. He also shares insights into how AI security testing is evolving, why transparency in model training matters, and what lessons companies can take from past security breaches to safeguard the next wave of AI applications. Learn more: What Are the Security Risks of Deploying DeepSeek-R1? - https://www.promptfoo.dev/blog/deepseek-redteam/ Follow everybody on social media: Ian Webster - https://x.com/iwebst Joel de la Garza - https://www.linkedin.com/in/3448827723723234/ Check out everything a16z is doing with artificial intelligence, including articles, projects, and more podcasts, here: https://a16z.com/ai/ 01:11 - DeepSeek: The Golden Age of AI or an existential threat? 02:18 - Red team testing, prompt injections, jail brakes - adversarial techniques 02:48 - Speech limitations 04:14 - Maturity and complexity of DeepSeek vs. other models 05:36 - Anything you build on top of DeepSeek will be subject to its insecurities 06:12 - Hosted model from China vs. open source/running locally 07:46 - DeepSeek benchmark on politically sensitive topics 08:54 - Western censorship vs. DeepSeek censorship 12:38 - How can we use it safely? Protecting infrastructure 14:09 - Wait for a more trusted source to run locally?
Exploring the implications and considerations of deploying China's Deep Seek in enterprise settings.
"the excitement around it is well warranted but I think in an Enterprise or infrastructure context I would probably wait for something that is more stable and then doesn't have these questions hanging ..."
Exploring the limits of Deep Seek's defenses against adversarial techniques and political sensitivities.
"limits of the model in terms of uh just red teaming it and seeing what sorts of adversarial techniques it responds to or doesn't respond to and by adversarial techniques what do you what do you mean e..."
This segment analyzes the weaknesses and susceptibility of Deep Seek to jailbreaks compared to other models like GPT.
"the US and you know you you you know for your for your company and I guess probably as a side project as well you you spent a lot of time breaking these things I'm curious your estimation of sort of t..."
This segment explores the security implications of using Deep Seek in both hosted and local environments, highlighting the misconceptions around censorship.
"all sorts of hyperventilation and maybe maybe it helps for folks to understand sort of like the the way most people were interacting with deep seek was through a hosted model that was in China but the..."
Exploring the hidden complexities and uncertainties surrounding Deep Seek's censorship and controls compared to Western models.
"version of Deep seek that that you see out there I think the the part that's really interesting to me is not the obvious stuff that that we measured the interesting part is the is is like the addition..."
This segment explores the varying levels of censorship among US AI models when handling sensitive Chinese political topics, revealing noteworthy discrepancies.
"anyway we we were going to run it on we were going to run benchmarks on like sensitive us political topics but as a baseline I was like let me you know let's let's do this and just run the all the fla..."
This segment highlights recommended strategies for enterprises considering the use of Deep Seek, emphasizing the importance of utilizing trusted models and being aware of potential risks.
"I mean that's that's to me is amazing that you know the a lot of American commentators were deriding the Chinese model for censoring things and CH sensitive Chinese topics and then kind of look in you..."