Topic · AI agents
lab OpenAI News · 22d ago SWE-bench Pro · 100.0% · 40 views

OpenAI designates Astra as critical cybersecurity model with enhanced safeguards

OpenAI has designated its new model, Astra, as meeting the "Critical" cybersecurity capability threshold under its Preparedness Framework, meaning it can identify and exploit security flaws in hardened systems without human guidance. To mitigate risks, the company delayed parts of Astra's development to implement stronger safety measures, including improved alignment training and monitoring systems.

lab OpenAI News · 17d ago · 44 views

OpenAI reports automated research interns and coding agents accelerating progress toward RSI

OpenAI has published internal metrics showing that automated AI researchers and coding agents are significantly accelerating its research pace, with the organization reaching its goal of an automated "research intern" by September. The company aims to build a fully automated AI researcher by March 2028 to further progress on deep learning and alignment while maintaining human oversight.

lab OpenAI News · 19d ago · 39 views

OpenAI commits $1B to subsidize Daybreak AI for frontline defenders

OpenAI is launching "Daybreak for Frontline Defenders," a global initiative committing $1 billion in subsidized access, training, and technical support to help cyber defenders protect essential services. The program targets resource-constrained organizations, including water utilities, electric grid operators, and local governments, aiming to distribute frontier AI cybersecurity capabilities beyond large enterprises.

lab Google — The Keyword (AI) · 21d ago · 52 views

Google launches Fairwind Program with Gemini 3.8 Flash Cyber for proactive defense

Google has launched the Fairwind Program to provide trusted government agencies, enterprise partners, and cybersecurity organizations with early access to advanced AI capabilities for proactive cyber defense. The initiative combines the specialized Gemini 3.8 Flash Cyber model with the CodeMender harness to help defenders autonomously find, verify, and fix vulnerabilities at scale.

lab Google DeepMind Blog · 21d ago · 50 views

Google launches Fairwind Program with Gemini 3.8 Flash Cyber for proactive defense

Google has launched the Fairwind Program to provide government agencies, enterprise customers, and cybersecurity partners with access to advanced AI capabilities for proactive cyber defense. The initiative aims to help defenders autonomously find and fix vulnerabilities at scale by combining the specialized reasoning of Gemini 3.8 Flash Cyber with the CodeMender harness.

lab Anthropic News · 27d ago · 83 views

Anthropic opens research preview of Model Hardware Standard for AI agent device control

Anthropic and HHMI Janelia Research Campus have opened a research preview of the Model Hardware Standard (MHS), a shared specification designed to enable AI agents to safely operate physical devices in labs and manufacturing facilities. The standard reduces hardware integration time from weeks or months to hours by providing a unified driver that translates between operating systems and diverse hardware interfaces.

lab xAI News · 10h ago · 15 views

SpaceXAI uses Grok Bot to handle 175% ticket surge without hiring

SpaceXAI integrated its Grok Bot AI agent into its combined customer support operation following the Cursor acquisition, allowing the team to manage a 175% increase in support tickets without adding headcount. The company avoided hiring approximately 200 additional staff members and reduced resolution costs to between $0.20 and $0.30 per ticket by leveraging Grok Bot's usage-based pricing model.