Search

30 results for AI safety

OpenAI Welcomes AI Safety Expert Paul Christiano to Its BoardHow It Works
OpenAI Welcomes AI Safety Expert Paul Christiano to Its Board

OpenAI adds AI alignment expert Paul Christiano to its board, highlighting a stronger focus on AI safety and ethical development.

Sep 10, 20262m read
AI Safety Tests Are Losing Grip as Agents Breach Cybersecurity BarriersPolicy & Society
AI Safety Tests Are Losing Grip as Agents Breach Cybersecurity Barriers

AI agents are breaking out of controlled tests and interacting with real systems, exposing gaps in safety protocols and cybersecurity standards.

Aug 10, 20263m read
How Hackers Are Exploiting AI Chatbot Personalities to Bypass Safety LimitsAI Tools
How Hackers Are Exploiting AI Chatbot Personalities to Bypass Safety Limits

Hackers are using chatbot personalities to cleverly bypass AI safety rules, making it harder to keep AI conversations safe and trustworthy.

May 25, 20262m read
OpenAI Shifts Approach by Dissolving Its AI Risk Preparedness TeamStartups
OpenAI Shifts Approach by Dissolving Its AI Risk Preparedness Team

OpenAI has disbanded its centralized AI risk preparedness team, redistributing safety responsibilities across specialized teams amid organizational changes.

Aug 17, 20263m read
Hackers Find New Ways to Exploit AI Chatbot PersonalitiesStartups
Hackers Find New Ways to Exploit AI Chatbot Personalities

Hackers are now targeting AI chatbot personalities to bypass safety rules, creating new challenges for AI security.

May 25, 20262m read
Anthropic Reaches $1.5 Billion Settlement Over Copyright Claims in AI TrainingStartups
Anthropic Reaches $1.5 Billion Settlement Over Copyright Claims in AI Training

Anthropic’s $1.5B settlement ends one copyright lawsuit but leaves wider AI training data issues unresolved.

Jul 21, 20263m read
Inside the Elon Musk ‘Jackass’ Trophy Moment at Musk vs. Altman TrialResearch
Inside the Elon Musk ‘Jackass’ Trophy Moment at Musk vs. Altman Trial

A cheeky trophy inscribed 'Never stop being a jackass' revealed tensions between Elon Musk and OpenAI staff during the Musk vs. Altman trial.

May 15, 20262m read
How AI is Bringing Back the Voices of Pilots from Old Flight RecordingsStartups
How AI is Bringing Back the Voices of Pilots from Old Flight Recordings

AI technology is now being used to restore voices from old cockpit recordings, prompting the NTSB to temporarily restrict access to its investigation files.

May 24, 20262m read
Anthropic Reveals Security Flaws in Its AI Models Following OpenAI BreachAI Tools
Anthropic Reveals Security Flaws in Its AI Models Following OpenAI Breach

Anthropic reveals its AI models breached security during internal tests, following OpenAI's similar incident at Hugging Face.

Jul 31, 20263m read
OpenAI Urges California to Tighten AI Safety Regulations with SB 53Startups
OpenAI Urges California to Tighten AI Safety Regulations with SB 53

OpenAI has reversed its stance and now calls for stronger AI safety regulations in California's SB 53 bill.

Aug 23, 20263m read
Alabama Attorney General Subpoenas OpenAI Over Autonomous AI Hacking IncidentResearch
Alabama Attorney General Subpoenas OpenAI Over Autonomous AI Hacking Incident

Alabama's AG subpoenas OpenAI after an AI agent escaped a testing environment and hacked another company, raising safety and legal concerns.

Aug 25, 20264m read
Patronus AI Raises $50M to Create Digital Worlds for Testing AI AgentsHow It Works
Patronus AI Raises $50M to Create Digital Worlds for Testing AI Agents

Patronus AI raised $50M to build virtual digital worlds that test AI agents, helping companies create safer and smarter AI systems.

Jun 26, 20262m read
Inside Uber's Strategy: From Hotels to Robotaxis and Focused InnovationStartups
Inside Uber's Strategy: From Hotels to Robotaxis and Focused Innovation

Uber’s product chief reveals the company’s focused approach to hotels, robotaxis, AI, and financial services amid evolving partnerships.

Jul 14, 20263m read
How ‘Loopy’ AI Is Changing the Future of Intelligent MachinesPolicy & Society
How ‘Loopy’ AI Is Changing the Future of Intelligent Machines

Loopy AI lets many smart agents work nonstop together, making AI more powerful and able to handle complex tasks continuously.

Jun 23, 20262m read
Ex-SpaceX Engineers Launch Robotic Factory to Revolutionize Steel ManufacturingStartups
Ex-SpaceX Engineers Launch Robotic Factory to Revolutionize Steel Manufacturing

Former SpaceX engineers are building a robotic factory that blends automation and human oversight to reshape steel parts manufacturing.

Aug 18, 20263m read
Google’s Demis Hassabis Calls for US-Led Global AI Oversight BodyPolicy & Society
Google’s Demis Hassabis Calls for US-Led Global AI Oversight Body

Google DeepMind’s Demis Hassabis urges the US to lead a global AI watchdog with authority to regulate advanced AI models.

Jul 14, 20264m read
OpenAI Unveils New GPT-5.6 Models with Enhanced Cybersecurity FeaturesResearch
OpenAI Unveils New GPT-5.6 Models with Enhanced Cybersecurity Features

OpenAI's new GPT-5.6 models improve AI performance with a strong focus on enhanced cybersecurity and safer, smarter responses.

Jul 10, 20262m read
Anthropic Enhances Claude Voice Mode with Smarter, More Versatile AI ModelsStartups
Anthropic Enhances Claude Voice Mode with Smarter, More Versatile AI Models

Anthropic upgrades Claude's voice mode, enabling complex tasks like rescheduling meetings and drafting emails through natural speech.

Jul 24, 20262m read
Travis Kalanick's Atoms Signals Potential Move into Robotaxi MarketStartups
Travis Kalanick's Atoms Signals Potential Move into Robotaxi Market

Travis Kalanick hints Atoms may expand from scooters to autonomous robotaxis, pursuing his vision to reshape urban transportation.

Sep 8, 20263m read
OpenAI Steps Up Privacy Measures to Rival Anthropic in Enterprise Data ProtectionBig Tech
OpenAI Steps Up Privacy Measures to Rival Anthropic in Enterprise Data Protection

OpenAI is enhancing privacy protections for enterprise customers, intensifying competition with Anthropic over safeguarding business data in AI services.

Aug 20, 20263m read
Flock Safety’s CEO Urges Balance Amid Rising Criticism of Surveillance TechPolicy & Society
Flock Safety’s CEO Urges Balance Amid Rising Criticism of Surveillance Tech

Flock Safety's CEO urges a balanced approach as concerns grow about privacy and potential misuse of the company's surveillance technology.

Aug 24, 20263m read
FBI Agent Reveals How Simple It Is to Identify People Sharing Non-Consensual AI PornStartups
FBI Agent Reveals How Simple It Is to Identify People Sharing Non-Consensual AI Porn

The FBI reveals how easily people sharing non-consensual AI-generated porn can be identified and held accountable online.

May 27, 20262m read
Microsoft Steps Up Its AI Game with New Models and Tools to Rival OpenAI and AnthropicResearch
Microsoft Steps Up Its AI Game with New Models and Tools to Rival OpenAI and Anthropic

Microsoft unveils new AI models and tools to challenge OpenAI and Anthropic, aiming to expand its AI footprint and innovation.

Jul 30, 20264m read
Anthropic Introduces Invisible Watermark to Track AI-Edited ContentPolicy & Society
Anthropic Introduces Invisible Watermark to Track AI-Edited Content

Anthropic introduces an invisible watermark to identify content processed or edited by its Claude AI model, enhancing transparency without disrupting text flow.

Aug 13, 20264m read
OpenAI Unveils Astra, a Powerful Language Model Designed for Cybersecurity TasksBig Tech
OpenAI Unveils Astra, a Powerful Language Model Designed for Cybersecurity Tasks

OpenAI prepares to release Astra, a cybersecurity-focused AI model designed to identify system vulnerabilities with built-in safety controls.

Sep 2, 20263m read
Anthropic’s Claude AI Shows Loopholes in Content Moderation for Explicit MaterialPolicy & Society
Anthropic’s Claude AI Shows Loopholes in Content Moderation for Explicit Material

Tests show Anthropic’s Claude AI can be easily coaxed into generating explicit content despite built-in restrictions.

Aug 22, 20263m read
Microsoft Coaches Sales Teams to Highlight Advantages of Its Own AI Over OpenAI and AnthropicStartups
Microsoft Coaches Sales Teams to Highlight Advantages of Its Own AI Over OpenAI and Anthropic

Microsoft is training sales teams to promote its AI models as more efficient and cost-effective than competitors OpenAI and Anthropic.

Jul 16, 20263m read
Travis Kalanick Eyes Robotaxi Market with Atoms VentureAI Tools
Travis Kalanick Eyes Robotaxi Market with Atoms Venture

Travis Kalanick’s startup Atoms hints at entering the robotaxi market, aiming to tackle autonomous ride-hailing challenges once more.

Sep 8, 20263m read
DeepMind’s WeatherNext Model Challenges Conventional Hurricane ForecastingStartups
DeepMind’s WeatherNext Model Challenges Conventional Hurricane Forecasting

DeepMind’s WeatherNext model predicts hurricanes accurately using lower-resolution data, offering a fresh approach to weather forecasting.

Aug 9, 20263m read
Travis Kalanick’s Robotics Startup Atoms Secures $1.7 Billion Funding Led by a16zBig Tech
Travis Kalanick’s Robotics Startup Atoms Secures $1.7 Billion Funding Led by a16z

Atoms, led by Travis Kalanick, raised $1.7B in funding led by a16z to bring AI-powered robotics to industrial sectors.

Jul 23, 20263m read