In The News

Meet GPT-Red: an LLM super-hacker OpenAI built to make its models safer

MIT Technology Review

July 15, 2026

CSET’s Jessica Ji shared her expert insight in an article published by MIT Technology Review. The article examines how OpenAI developed GPT-Red, an AI "super-hacker" designed to automatically identify vulnerabilities in large language models and strengthen their defenses against cyberattacks through AI-powered red-teaming.

Read Article

Related Content

CSET’s Jessica Ji shared her expert insight in an article published by WIRED. The article examines the launch of FLARE-AI, a new crowdsourced platform designed to improve transparency and accountability by creating a centralized system… Read More

CSET’s Vikram Venkatram, Mina Narayanan, and Jessica Ji shared their expert analysis in an op-ed published by The National Interest. The article analyzes the Trump administration’s new AI executive order and its attempt to limit… Read More

CSET’s Jessica Ji shared her expert insight in an article published by CNBC. The article examines the House passage of the Kids Internet and Digital Safety (KIDS) Act, a bill aimed at strengthening protections for… Read More

CSET’s Jessica Ji shared her expert perspective in an article published by CNN. The article examines new agreements between Microsoft, Google, and xAI to allow the U.S. government to evaluate unreleased AI models for cybersecurity… Read More