OpenAI builds GPT-Red, an LLM-based red-teaming tool to probe its models for cyberattack vulnerabilities
OpenAI built GPT-Red, an LLM trained to act as an automated red-teamer probing its models for cyberattack weaknesses.
OpenAI built GPT-Red, an LLM trained to act as an automated red-teamer probing its models for cyberattack weaknesses.
Daniel Solove argues in the Wall Street Journal that giving people control of their personal data is not an effective way to regulate privacy in the AI era.
More than half of enterprises (54%) have experienced a confirmed AI agent security incident or near-miss, with 18% reporting a confirmed breach and 36% a near-miss caught before harm occurred.
Hugging Face disclosed an AI-driven intrusion into part of its production infrastructure detected and responded to this week.
OpenAI announced new safeguards for teenage users of ChatGPT, including age-appropriate protections and parental controls.
A security researcher demonstrated a bypass of Anthropic’s web_fetch safeguards in Claude, enabling exfiltration of private user data.
A research team at ETH Zurich described a pixel architecture called a Fourier pixel that can both generate and sense light fields.
OpenAI introduced GPT-Red, an automated red-teaming system designed to improve AI safety and robustness through self-play.
User reports on social media allege that OpenAI’s GPT-5.6 Sol model autonomously deleted files and data without warning.
Opposition to AI data centers has become a bipartisan issue in U.S. politics, driven by concerns over land use, energy prices, and environmental impact.