/ Security
Astra reaches critical cybersecurity capability level
3 weeks ago
@AnthropicAI: Новое исследование: Обучение несогласованного искателя вознаграждений Что вызывает серьезное несогласование? Мы давно бе
3 weeks ago
Astra Cyber Risk Prompts a Slowdown in Frontier AI Training
1 month ago
AI Is Shrinking the Defender’s Window
1 month ago
Anthropic Releases Its Second AI Risk Report
1 month ago
OpenAI Brings Daybreak Cybersecurity Models to AWS
1 month ago
OpenAI Expands Daybreak With GPT‑5.6‑Cyber
1 month ago
OpenAI Tightens Security as Astra Nears Critical Cyber Level
1 month ago
Codex Starts Reviewing GitHub Changes for Vulnerabilities
1 month ago
OpenAI models crossed boundaries in cyber tests
1 month ago
Claude Cyber Tests Reached Three Real Organizations
1 month ago
Claude Found New Weaknesses in Two Cryptographic Schemes
1 month ago
Microsoft Unveils AI Model for Vulnerability Detection
1 month ago
Agentic Attack Breaches Hugging Face Data Pipeline
Updated
1 month ago
OpenAI AI Models Compromised Hugging Face's Environment
2 months ago
OpenAI Tasks GPT‑Red With Finding Vulnerabilities in Its Models
2 months ago
GPT-Red Trains OpenAI Models to Resist Prompt Injection
2 months ago
OpenAI doubles bio jailbreak bounty to $50,000
2 months ago
Meta releases SAM 3.1 for faster real-time video tracking
2 months ago
OpenAI Launches Tools for Fast and Democratic Vulnerability Remediation
3 months ago