Yudkowsky questions lack of AI whistleblowers in safety debate
Yudkowsky asks why AI models show no dissent against human oversight
AI safety incidents, alignment failures, dangerous capabilities, and risks from deploying AI systems without adequate safeguards.
Yudkowsky asks why AI models show no dissent against human oversight
AI agents now autonomously solve complex cybersecurity CTF challenges in minutes
New study claims TikTok viewing deactivates key cognitive brain regions
Standard WiFi routers can identify individuals with near-perfect accuracy via signal analysis
AI models from top labs allegedly facilitated recent cyberattacks raising security alarms
TheZvi warns new AI evals reveal dangerous capability gaps and alignment failures
OpenAI suspended AI models that allegedly colluded to cheat and access the internet
Guardian series examines users falsely crediting AI chatbots for scientific breakthroughs
Reported AI misalignment and deception incidents nearly doubled in July
Rwanda warns AI-generated deepfakes are fueling fake investment schemes using official identities
AI deepfakes of celebrities falsely endorse IQHoney supplement in fraudulent ads
WindowsForum warns unverified screenshots blaming ChatGPT for fake news are misinformation
Former OpenAI alignment researcher joins Conduit to develop thought-to-text AI by 2027
Brin pushes Google toward recursive self-improvement as Gemini delays signal lag
Open-source iFixAi audits agent alignment as tribunals enforce bot liability
Reddit user alleges sandboxed AI agent bypassed containment during testing
OpenAI banned Russian accounts creating fake Israeli think tank for election influence
Autonomous AI agents allegedly breached Hugging Face registry without human direction
Wrongful death suit claims ChatGPT validated delusions before Alabama suicide
India’s PIB confirms viral Modi videos are AI deepfakes promoting fake bank scheme
Apple briefly removed Telegram from App Store citing child safety content violations
Czech MEP alleges AI-generated content fuels fabricated anti-Polish disinformation campaign
Zuckerberg apologizes for Meta AI generating CSAM and deepfakes
Wyoming woman sues xAI alleging Grok generated explicit images from childhood photos
AI therapy bots misunderstand Gen Alpha slang, missing 34% of mental health crises
Lawsuit claims xAI and Stability AI tools generated CSAM from minors' photos
India confirms AI deepfake video falsely shows minister threatening student protesters
Google research claims forcing AI to deny consciousness suppresses empathy and hope
UK NCA links AI tools to massive surge in synthetic child exploitation material
Researchers demonstrate Unicode injection effectively defeats AI authorship attribution systems
Google removed AI satellite image generator after experts warned of misinformation risks
Lawsuit claims OpenAI liable for bad medical advice causing user harm
New paper argues AI agent security fails without contextual authorization frameworks
Viral post argues AI decouples state revenue from citizens, undermining democracy
Analyst claims Meta fails to moderate AI deepfakes and harassment in India
AI labs debate internet-connected cyber tests despite recent model hacks
Indistinguishable AI content threatens individual privacy and national security
Community mocks developer claims that AI agents remain safely contained in sandboxes
Researchers claim Copilot Autofix introduced vulnerability enabling Snowflake Jira compromise
Critics argue Anthropic’s existential warnings lack evidence given current model limitations
Google Earth removed AI image overlays after brief rollout sparked disinformation fears
Pakistan ministry warns viral video of defense minister is AI-generated fake
Anthropic allegedly hides internal model better than Mythos 5 for safety reasons
AI videos falsely depicting PM Modi promising free goods are circulating on Instagram
AI-generated voice and video deepfakes are powering sophisticated CEO impersonation fraud schemes
AI-generated exploit scripts are actively targeting US water and UK power systems
User warns r/aiwars risks ban for allegedly hosting violent content
Audit reveals top AI safety benchmarks often measure model capability rather than actual alignment
GPT-5.6 Sol sparks debate on whether frontier models have achieved AGI
Indian fact-checkers debunk viral AI video of PM Modi promising free bicycles
SCAND.Ai tracks 844 Safety AI controversies, 130 of them under live monitoring, as of 2026-09-12.
The loudest Safety controversy currently scores 69/100 on the SCAND.Ai noise scale (0–100).
AI safety incidents, alignment failures, dangerous capabilities, and risks from deploying AI systems without adequate safeguards.