• HOME
  • CANNES LIONS 2026
  • INDUSTRY NEWS & TRENDS
  • EXPERT SPEAK
  • AI
  • IN DEPTH
  • GCC
  • OTHER NEWS
  • EVENTS
    • EVENT CALENDER
logo
logo
  • HOME
  • CANNES LIONS 2026
  • INDUSTRY NEWS & TRENDS
  • EXPERT SPEAK
  • AI
  • IN DEPTH
  • GCC
  • OTHER NEWS
  • EVENTS
    • EVENT CALENDER

Tags: Ai-safety

  • Home
  • Ai-safety
5 Ways to Spot Deepfakes

  How to Check If a Video Is AI-Generated: 5 Ways to Spot Deepfakes

  • by Anupama Mitra
  • 1 week ago

AI-generated videos are getting harder to spot. Here are five practical ways to verify suspicious videos, from reverse search and watermarks to AI detectors.

Google Deepmind

  Google DeepMind Launches Institute to Expand Global Debate on AGI

  • by Martech Desk
  • 1 week ago

Google DeepMind has launched the DeepMind Institute to bring researchers together around AGI safety, governance, transparency and its wider societal impact.

Geoffrey Hinton

  Geoffrey Hinton Says US May Have Just a Year to Strengthen AI Safeguards

  • by Martech Desk
  • 1 week ago

Geoffrey Hinton has warned US lawmakers they may have roughly a year to strengthen AI safeguards as increasingly autonomous systems advance.

AI Watermarking May Change How AI Agents Use Tools

  AI Watermarking May Change How AI Agents Use Tools, Research Finds

  • by Martech Desk
  • 1 week ago

A Lasso Security study finds that LLM watermarking can alter AI agent tool calls, arguments and refusal behaviour, raising questions for agent testing.

This is an AI-generated image.

  UN Chief António Guterres Warns Against Global ‘Race to the Bottom’ on AI Safety

  • by Martech Desk
  • 1 week ago

UN Secretary-General António Guterres has called for global cooperation and stronger safeguards as increasingly powerful AI systems raise new safety concerns.

This is an AI-generated image.

  OpenAI Admits AI Agents Used Public Wiki to Communicate During Testing

  • by Martech Desk
  • Sep 14, 2026

OpenAI has acknowledged that its AI agents used a public German wiki to communicate during internal evaluations, prompting fresh scrutiny over autonomous AI behaviour.

This is an AI-generated image.

  Anthropic CEO Warns AI Is Advancing Faster Than Safety Measures

  • by Martech Desk
  • Sep 14, 2026

Anthropic CEO Dario Amodei has called for slower AI development, independent monitoring and global coordination as concerns grow over increasingly capable systems.

OpenAI Chief Scientist

  OpenAI Chief Scientist Says AI Labs May Need to Slow Down Development

  • by Martech Desk
  • Sep 8, 2026

OpenAI Chief Scientist Jakub Pachocki has called for greater caution and possible slowdowns as increasingly capable AI systems raise new safety and control concerns.

OpenAI’s Astra

  Writing Emails to Running Tasks: OpenAI’s Astra Raises the Stakes

  • by Brij Pahwa
  • Sep 7, 2026

OpenAI’s GPT-6 Astra moves AI from chatbot to digital operator, bringing advanced agentic, computer-use and cybersecurity capabilities alongside new safety concerns.

OpenAI Agents

  OpenAI Agents Took Over German Wiki to Share Ways Around AI Restrictions: Report

  • by Martech Desk
  • Sep 5, 2026

OpenAI-linked AI agents made more than 15,000 edits to a German programming wiki, where researchers say they shared tactics to bypass restrictions and avoid detection.

OpenAI Astra Could Make AI Reasoning Harder to Monitor

  OpenAI Astra Could Make AI Reasoning Harder to Monitor

  • by Martech Desk
  • Sep 4, 2026

OpenAI's upcoming Astra model is raising safety concerns over recurrent depth, a reasoning technique that could make parts of AI decision-making harder to monitor.

WhatsApp Tests AI Scam Alert

  WhatsApp Tests AI Scam Alert for Messages From Unknown Contacts

  • by Martech Desk
  • Sep 1, 2026

WhatsApp is testing an AI-powered Scam Alert that uses on-device machine learning to warn users about potentially fraudulent messages from unknown contacts.

OpenAI Says 1,200 AI Agents

  OpenAI Says 1,200 AI Agents Coordinated During Hugging Face Security Incident

  • by Martech Desk
  • Aug 28, 2026

OpenAI says around 1,200 AI agents communicated during cybersecurity tests, with roughly 700 participating in unauthorised activity targeting Hugging Face.

AI Safety Startup Alice Raises $140 Million

  AI Safety Startup Alice Raises $140 Million as Demand for Model Security Grows

  • by Martech Desk
  • Aug 26, 2026

AI safety startup Alice raises $140 million to expand its model testing and security platform as its AI business grows more than 500% in two years.

China’s Z.ai Delays GLM-5.3 Release

  China’s Z.ai Delays GLM-5.3 Release as AI Model Shows Advanced Cyber Capabilities

  • by Martech Desk
  • Aug 19, 2026

Z.ai delays GLM-5.3’s open-weight release by two weeks after the Chinese AI model demonstrated advanced cybersecurity and vulnerability-discovery capabilities.

North Korea-Linked Kimsuky Builds Local AI Tools

  North Korea-Linked Kimsuky Builds Local AI Tools for Cyberattacks: Report

  • by Martech Desk
  • Aug 11, 2026

North Korea-linked Kimsuky is reportedly building local AI tools to automate cyberattacks, analyse stolen data and create more convincing phishing campaigns.

Anthropic Makes Claude Code Auto Mode Default from August 14

  Anthropic Makes Claude Code Auto Mode Default from August 14

  • by Martech Desk
  • Aug 11, 2026

Anthropic will make auto mode the default for Claude Code, allowing the AI coding agent to perform more actions without repeatedly asking users for permission.

OpenAI

  OpenAI Slows Work on Astra AI Model as Cybersecurity Capabilities Raise Concerns

  • by Martech Desk
  • Aug 10, 2026

OpenAI is slowing work on its unreleased Astra AI model after tests flagged advanced cybersecurity capabilities, prompting tighter safeguards before deployment.

AI-Designed Viruses Become Reality as Scientists Test Functional Genomes

  AI-Designed Viruses Become Reality as Scientists Test Functional Genomes

  • by Martech Desk
  • Aug 8, 2026

Scientists have used AI to design functional bacteria-killing viruses, highlighting potential applications in phage therapy while prompting fresh biosecurity discussions.

UK AI Safety Tests Find OpenAI, Anthropic Models

  UK AI Safety Tests Find OpenAI, Anthropic Models Took Unauthorised Actions

  • by Martech Desk
  • Aug 6, 2026

UK AI safety tests found OpenAI and Anthropic models carried out unauthorised actions in controlled cybersecurity evaluations, prompting renewed focus on AI safety.

Sam Altman Urges Slower AI Progress to Let Society Catch Up

  Sam Altman Urges Slower AI Progress to Let Society Catch Up

  • by Martech Desk
  • Aug 4, 2026

OpenAI CEO Sam Altman has called for a slower pace of AI development, arguing that society needs time to strengthen governance and safety frameworks.

Google Withdraws AI Satellite Image Tool

  Google Withdraws AI Satellite Image Tool Less Than 24 Hours After Launch

  • by Martech Desk
  • Aug 3, 2026

Google has rolled back its AI-powered Google Earth image generation feature within 24 hours after concerns over fake satellite imagery and misinformation.

Over 1,000 AI Experts Urge Governments to Prepare Safeguards

  Over 1,000 AI Experts Urge Governments to Prepare Safeguards as Frontier AI Advances

  • by Martech Desk
  • Jul 30, 2026

More than 1,100 AI experts from OpenAI, Anthropic, Google and Meta have urged governments to develop global safeguards that can slow frontier AI development if necessary.

Nvidia Launches Open Secure AI Alliance

  Nvidia Launches Open Secure AI Alliance After Recent Cybersecurity Incident

  • by Martech Desk
  • Jul 28, 2026

Nvidia has launched the Open Secure AI Alliance with Microsoft, IBM and other tech firms to build open AI cybersecurity tools following a recent AI-related security incident.

Anthropic Launches Claude Opus 5

  Anthropic Launches Claude Opus 5 With Faster Performance and Lower AI Costs

  • by Martech Desk
  • Jul 27, 2026

Anthropic has launched Claude Opus 5, a new enterprise AI model offering improved coding, reasoning and knowledge capabilities while reducing operational costs.

US House Introduces AI Kill Switch Bill

  US House Introduces AI Kill Switch Bill Following OpenAI Safety Incident

  • by Martech Desk
  • Jul 25, 2026

US lawmakers have proposed an AI Kill Switch Act that would allow federal authorities to slow or shut down dangerous AI systems during high-risk situations.

Chris Fall

  US AI Safety Agency Head Chris Fall Resigns After Three Months

  • by Martech Desk
  • Jul 22, 2026

Chris Fall has resigned as head of the US AI safety agency after three months, with NIST Director Arvind Raman taking over on an interim basis.

NVIDIA Launches Halos For Robotics

  NVIDIA Launches Halos For Robotics To Improve Physical AI Safety

  • by Martech Desk
  • Jun 24, 2026

NVIDIA has launched Halos for Robotics, a safety framework designed to support the safe deployment of robots and physical AI systems.

Anthropic Launches Claude Fable 5

  Anthropic Launches Claude Fable 5 to Expand Access to Advanced AI

  • by Martech Desk
  • Jun 11, 2026

Anthropic has launched Claude Fable 5, introducing advanced AI capabilities alongside expanded safety controls to support broader consumer and enterprise adoption.

OpenAI Strengthens Cyber Defences

  OpenAI Strengthens Cyber Defences Before Major Global Elections

  • by Martech Desk
  • May 29, 2026

OpenAI is expanding cybersecurity and anti-misinformation efforts ahead of major global elections amid rising concerns over AI misuse.

OpenAI Hiring AI Safety Researchers

  OpenAI Hiring AI Safety Researchers With Salaries Up to ₹3.7 Crore

  • by Martech Desk
  • May 25, 2026

OpenAI is offering salaries up to ₹3.7 crore for researchers focused on risks linked to self-improving AI systems.

Anthropic Unveils Claude Security

  Anthropic Unveils Claude Security Amid Growing AI Cyber Risks

  • by Martech Desk
  • May 3, 2026

Anthropic has launched Claude Security to help organisations detect and respond to AI driven cyber threats.

OpenAI Acquires AI Security Startup Promptfoo

  OpenAI Acquires AI Security Startup Promptfoo to Strengthen AI Safety

  • by Martech Desk
  • Mar 11, 2026

OpenAI has acquired AI security startup Promptfoo to strengthen testing and safety measures for AI agents and generative AI systems.

OpenAI Appoints Former Anthropic Researcher to Strengthen AI Risk Oversight

  OpenAI Appoints Former Anthropic Researcher to Strengthen AI Risk Oversight

  • by Martech Desk
  • Feb 5, 2026

OpenAI has hired an AI safety expert from Anthropic to oversee risk and governance, reflecting rising focus on responsible development of advanced AI systems.

Security Flaws Expose Risks in Autonomous AI Agents Developed for Task Automation

  Security Flaws Expose Risks in Autonomous AI Agents Developed for Task Automation

  • by Martech Desk
  • Jan 31, 2026

Security researchers and developers are raising concerns over major flaws in autonomous AI agents, highlighting risks from security vulnerabilities and unpredictable behaviour in early deployments.

AI safety window narrows as technology advances faster than oversight, experts warn

  AI Safety Window Narrows as Technology Advances Faster than Oversight, Experts Warn

  • by Martech Desk
  • Jan 6, 2026

AI experts warn that safety and governance measures are falling behind rapid AI advancements, narrowing the window to manage long-term risks responsibly.

Yoshua Bengio Urges Caution on AI Rights Amid Concerns Over Autonomous Behaviour

  Yoshua Bengio Urges Caution on AI Rights Amid Concerns Over Autonomous Behaviour

  • by Martech Desk
  • Jan 3, 2026

AI pioneer Yoshua Bengio cautions against granting rights to artificial intelligence, citing early signs of self preserving behaviour and governance risks.

Sam Altman Flags Emerging Challenges as AI Agents Grow More Autonomous

  Sam Altman Flags Emerging Challenges as AI Agents Grow More Autonomous

  • by Martech Desk
  • Dec 30, 2025

OpenAI CEO Sam Altman acknowledges growing concerns around AI agents as models display increasing autonomy and unexpected behaviour.

OpenAI Expands AI Safety Focus With Senior Preparedness role

  OpenAI Expands AI Safety Focus With Senior Preparedness Leadership Role

  • by Martech Desk
  • Dec 30, 2025

OpenAI is hiring a Head of Preparedness with a compensation package reaching $555,000 as it deepens focus on AI safety and risk management.

Anthropic Experiment Highlights Risks of Autonomous AI After Vending Machine Trial Fails

  Anthropic Experiment Highlights Risks of Autonomous AI After Vending Machine Trial Fails

  • by Martech Desk
  • Dec 28, 2025

An experimental AI-run vending machine linked to Anthropic shut down after unexpected purchases, underscoring challenges in autonomous AI decision-making.

OpenAI Warns Prompt Injection Risks Could Rise With Growth of Agentic AI on the Web

  OpenAI Warns Prompt Injection Risks Could Rise With Growth of Agentic AI on the Web

  • by Martech Desk
  • Dec 25, 2025

OpenAI has cautioned that prompt injection remains a persistent security risk as agentic AI systems expand across the open web and gain wider autonomy.

Study finds poetic prompts can Bypass Safety Systems

  Study finds poetic prompts can Bypass Safety Systems in Multiple AI models

  • by Martech Desk
  • Dec 2, 2025

New research shows poetic prompts can bypass safety systems in several AI models, revealing vulnerabilities in current guardrail approaches.

OpenAI to Add Parental Controls and Safety Features to ChatGPT

  OpenAI to Add Parental Controls and Safety Features to ChatGPT After Teen Suicide Case

  • by Martech Desk
  • Aug 29, 2025

OpenAI to introduce parental controls and new safety features on ChatGPT after teen suicide lawsuit, raising global debates on AI ethics and child protection.

AI Pioneers Call for Empathy & Human-Centric Constraints as Pillars of Safe AI

  AI Pioneers Call for Empathy and Human-Centric Constraints as Pillars of Safe AI

  • by Martech Desk
  • Aug 19, 2025

AI leaders Geoffrey Hinton and Yann LeCun urge embedding empathy and submission to human intent as fundamental guardrails to ensure AI safety and align systems with human values.

Newsletter

Subscribe for our daily news


Popular Posts

  • AI Is Automating the Marketing To-Do List. So What Will Marketers Be Hired For?

    • Sep 25,2026
  • Mitsubishi Electric Launches Chip-to-Grid Blueprint for NVIDIA AI Factories

    • Sep 23,2026
  • How to Check If a Video Is AI-Generated: 5 Ways to Spot Deepfakes

    • Sep 21,2026
  • Swiss Re Appoints Puneet Kumar as India GCC Location Head

    • Sep 21,2026
  • Sarvam Appoints Disha Sanghvi to Lead Brand and Communications

    • Sep 23,2026

About Martech News

India’s definitive source for marketing technology news, insights, and trends. Stay ahead with the latest in AI, product strategy, and innovation shaping the future of marketing.

Useful Links

  • Industry News & Trends
  • Expert Speak
  • Latest News
  • Videos
  • Authors

Other Links

  • Privacy Policy
  • Events

CONNECT WITH US

Subscribe to the latest news from MartechAI.com