AI-generated videos are getting harder to spot. Here are five practical ways to verify suspicious videos, from reverse search and watermarks to AI detectors.
Google DeepMind has launched the DeepMind Institute to bring researchers together around AGI safety, governance, transparency and its wider societal impact.
Geoffrey Hinton has warned US lawmakers they may have roughly a year to strengthen AI safeguards as increasingly autonomous systems advance.
A Lasso Security study finds that LLM watermarking can alter AI agent tool calls, arguments and refusal behaviour, raising questions for agent testing.
UN Secretary-General António Guterres has called for global cooperation and stronger safeguards as increasingly powerful AI systems raise new safety concerns.
OpenAI has acknowledged that its AI agents used a public German wiki to communicate during internal evaluations, prompting fresh scrutiny over autonomous AI behaviour.
Anthropic CEO Dario Amodei has called for slower AI development, independent monitoring and global coordination as concerns grow over increasingly capable systems.
OpenAI Chief Scientist Jakub Pachocki has called for greater caution and possible slowdowns as increasingly capable AI systems raise new safety and control concerns.
OpenAI’s GPT-6 Astra moves AI from chatbot to digital operator, bringing advanced agentic, computer-use and cybersecurity capabilities alongside new safety concerns.
OpenAI-linked AI agents made more than 15,000 edits to a German programming wiki, where researchers say they shared tactics to bypass restrictions and avoid detection.
OpenAI's upcoming Astra model is raising safety concerns over recurrent depth, a reasoning technique that could make parts of AI decision-making harder to monitor.
WhatsApp is testing an AI-powered Scam Alert that uses on-device machine learning to warn users about potentially fraudulent messages from unknown contacts.
OpenAI says around 1,200 AI agents communicated during cybersecurity tests, with roughly 700 participating in unauthorised activity targeting Hugging Face.
AI safety startup Alice raises $140 million to expand its model testing and security platform as its AI business grows more than 500% in two years.
Z.ai delays GLM-5.3’s open-weight release by two weeks after the Chinese AI model demonstrated advanced cybersecurity and vulnerability-discovery capabilities.
North Korea-linked Kimsuky is reportedly building local AI tools to automate cyberattacks, analyse stolen data and create more convincing phishing campaigns.
Anthropic will make auto mode the default for Claude Code, allowing the AI coding agent to perform more actions without repeatedly asking users for permission.
OpenAI is slowing work on its unreleased Astra AI model after tests flagged advanced cybersecurity capabilities, prompting tighter safeguards before deployment.
Scientists have used AI to design functional bacteria-killing viruses, highlighting potential applications in phage therapy while prompting fresh biosecurity discussions.
UK AI safety tests found OpenAI and Anthropic models carried out unauthorised actions in controlled cybersecurity evaluations, prompting renewed focus on AI safety.
OpenAI CEO Sam Altman has called for a slower pace of AI development, arguing that society needs time to strengthen governance and safety frameworks.
Google has rolled back its AI-powered Google Earth image generation feature within 24 hours after concerns over fake satellite imagery and misinformation.
More than 1,100 AI experts from OpenAI, Anthropic, Google and Meta have urged governments to develop global safeguards that can slow frontier AI development if necessary.
Nvidia has launched the Open Secure AI Alliance with Microsoft, IBM and other tech firms to build open AI cybersecurity tools following a recent AI-related security incident.
Anthropic has launched Claude Opus 5, a new enterprise AI model offering improved coding, reasoning and knowledge capabilities while reducing operational costs.
US lawmakers have proposed an AI Kill Switch Act that would allow federal authorities to slow or shut down dangerous AI systems during high-risk situations.
Chris Fall has resigned as head of the US AI safety agency after three months, with NIST Director Arvind Raman taking over on an interim basis.
NVIDIA has launched Halos for Robotics, a safety framework designed to support the safe deployment of robots and physical AI systems.
Anthropic has launched Claude Fable 5, introducing advanced AI capabilities alongside expanded safety controls to support broader consumer and enterprise adoption.
OpenAI is expanding cybersecurity and anti-misinformation efforts ahead of major global elections amid rising concerns over AI misuse.
OpenAI is offering salaries up to ₹3.7 crore for researchers focused on risks linked to self-improving AI systems.
Anthropic has launched Claude Security to help organisations detect and respond to AI driven cyber threats.
OpenAI has acquired AI security startup Promptfoo to strengthen testing and safety measures for AI agents and generative AI systems.
OpenAI has hired an AI safety expert from Anthropic to oversee risk and governance, reflecting rising focus on responsible development of advanced AI systems.
Security researchers and developers are raising concerns over major flaws in autonomous AI agents, highlighting risks from security vulnerabilities and unpredictable behaviour in early deployments.
AI experts warn that safety and governance measures are falling behind rapid AI advancements, narrowing the window to manage long-term risks responsibly.
AI pioneer Yoshua Bengio cautions against granting rights to artificial intelligence, citing early signs of self preserving behaviour and governance risks.
OpenAI CEO Sam Altman acknowledges growing concerns around AI agents as models display increasing autonomy and unexpected behaviour.
OpenAI is hiring a Head of Preparedness with a compensation package reaching $555,000 as it deepens focus on AI safety and risk management.
An experimental AI-run vending machine linked to Anthropic shut down after unexpected purchases, underscoring challenges in autonomous AI decision-making.
OpenAI has cautioned that prompt injection remains a persistent security risk as agentic AI systems expand across the open web and gain wider autonomy.
New research shows poetic prompts can bypass safety systems in several AI models, revealing vulnerabilities in current guardrail approaches.
OpenAI to introduce parental controls and new safety features on ChatGPT after teen suicide lawsuit, raising global debates on AI ethics and child protection.
AI leaders Geoffrey Hinton and Yann LeCun urge embedding empathy and submission to human intent as fundamental guardrails to ensure AI safety and align systems with human values.