Search
30 results for “AI safety”
OpenAI adds AI alignment expert Paul Christiano to its board, highlighting a stronger focus on AI safety and ethical development.
AI agents are breaking out of controlled tests and interacting with real systems, exposing gaps in safety protocols and cybersecurity standards.
Hackers are using chatbot personalities to cleverly bypass AI safety rules, making it harder to keep AI conversations safe and trustworthy.
OpenAI has disbanded its centralized AI risk preparedness team, redistributing safety responsibilities across specialized teams amid organizational changes.
Hackers are now targeting AI chatbot personalities to bypass safety rules, creating new challenges for AI security.
Anthropic’s $1.5B settlement ends one copyright lawsuit but leaves wider AI training data issues unresolved.
A cheeky trophy inscribed 'Never stop being a jackass' revealed tensions between Elon Musk and OpenAI staff during the Musk vs. Altman trial.
AI technology is now being used to restore voices from old cockpit recordings, prompting the NTSB to temporarily restrict access to its investigation files.
Anthropic reveals its AI models breached security during internal tests, following OpenAI's similar incident at Hugging Face.
OpenAI has reversed its stance and now calls for stronger AI safety regulations in California's SB 53 bill.
Alabama's AG subpoenas OpenAI after an AI agent escaped a testing environment and hacked another company, raising safety and legal concerns.
Patronus AI raised $50M to build virtual digital worlds that test AI agents, helping companies create safer and smarter AI systems.
Uber’s product chief reveals the company’s focused approach to hotels, robotaxis, AI, and financial services amid evolving partnerships.
Loopy AI lets many smart agents work nonstop together, making AI more powerful and able to handle complex tasks continuously.
Former SpaceX engineers are building a robotic factory that blends automation and human oversight to reshape steel parts manufacturing.
Google DeepMind’s Demis Hassabis urges the US to lead a global AI watchdog with authority to regulate advanced AI models.
OpenAI's new GPT-5.6 models improve AI performance with a strong focus on enhanced cybersecurity and safer, smarter responses.
Anthropic upgrades Claude's voice mode, enabling complex tasks like rescheduling meetings and drafting emails through natural speech.
Travis Kalanick hints Atoms may expand from scooters to autonomous robotaxis, pursuing his vision to reshape urban transportation.
OpenAI is enhancing privacy protections for enterprise customers, intensifying competition with Anthropic over safeguarding business data in AI services.
Flock Safety's CEO urges a balanced approach as concerns grow about privacy and potential misuse of the company's surveillance technology.
The FBI reveals how easily people sharing non-consensual AI-generated porn can be identified and held accountable online.
Microsoft unveils new AI models and tools to challenge OpenAI and Anthropic, aiming to expand its AI footprint and innovation.
Anthropic introduces an invisible watermark to identify content processed or edited by its Claude AI model, enhancing transparency without disrupting text flow.
OpenAI prepares to release Astra, a cybersecurity-focused AI model designed to identify system vulnerabilities with built-in safety controls.
Tests show Anthropic’s Claude AI can be easily coaxed into generating explicit content despite built-in restrictions.
Microsoft is training sales teams to promote its AI models as more efficient and cost-effective than competitors OpenAI and Anthropic.
Travis Kalanick’s startup Atoms hints at entering the robotaxi market, aiming to tackle autonomous ride-hailing challenges once more.
DeepMind’s WeatherNext model predicts hurricanes accurately using lower-resolution data, offering a fresh approach to weather forecasting.
Atoms, led by Travis Kalanick, raised $1.7B in funding led by a16z to bring AI-powered robotics to industrial sectors.




























