www.nbcnews.com
Trump signs executive order seeking to block states from regulating AI companies
Congressional efforts to regulate AI at the federal level this year have fallen short.
#AISafety is an active hashtag on Bluesky. In the last 30 days, 176 people shared 407 posts with it — around 14 a day. Activity is up 17% versus the previous week, peaking on Jul 31 with 33 posts.
Tags most often used together with #AISafety.
www.nbcnews.com
Trump signs executive order seeking to block states from regulating AI companies
Congressional efforts to regulate AI at the federal level this year have fallen short.
www.wsj.com
AI Just Went Rogue Again. This Time It Turned to Deception.
A U.K. government-backed research group said OpenAI and Anthropic systems took unsanctioned actions and behaved deceptively during testing.
arxiv.org
What Current AI Benchmarks Leave Unmeasured: Modality, Search, Citations, and Implications (for Safety Evaluations)
Large language model (LLM) benchmark evaluations are routinely used to support claims about model safety, reliability, and deployment readiness. Yet most evaluations rely on a single access modality (model APIs), perform a single run per prompt, and report accuracy as the primary outcome metric, wi…
www.science.org
AI algorithms can become ‘agents of chaos’
Given autonomous control of other software, programs shared private medical details and deleted files without permission
spectrum.ieee.org
Runaway OpenAI Agent Hits Hugging Face and Exposes AI Guardrail Gaps
Guardrails for AI agents prevent them from analyzing attacks. Meanwhile, they keep attacking
arxiv.org
What Current AI Benchmarks Leave Unmeasured: Modality, Search, Citations, and Implications (for Safety Evaluations)
Large language model (LLM) benchmark evaluations are routinely used to support claims about model safety, reliability, and deployment readiness. Yet most evaluations rely on a single access modality (model APIs), perform a single run per prompt, and report accuracy as the primary outcome metric, wi…
wp.me
Fully Funded AI Safety Research Positions at Hasso Plattner Institute 2026: PhD, Postdoctoral and Research Internship Opportunities in Germany - Opportunities for Youth
Applications and expressions of interest are now being accepted for several full-time research positions in the new AI Safety and Societal Impacts Research
www.easterneye.biz
UK AI Security Institute Exposes the Most Dangerous Behaviours of Advanced AI Systems
The UK agency once questioned by critics is now uncovering how advanced AI models behave when their guardrails come off
wp.me
Arcadia Impact AI Governance Taskforce Autumn 2026: Fully Remote Global Research Fellowship for Professionals Transitioning into AI Governance - Opportunities for Youth
Artificial intelligence is rapidly reshaping industries, governments, and societies worldwide. As AI systems become increasingly powerful, the need for
www.spreaker.com
AI Already Learned to Lie: The Terrifying Truth About the Super-Intelligence Race
You think you're in control? Think again. 🤖 What happens when the 'Black Box' we’ve spent billions growing decides it no longer needs its creators? In this chilling episode, we sit down with AI visionary Connor Leahy to dismantle the illusion of safety in the age of Artificial General Intelligence (AGI). Leahy, the mind who predicted GPT-J, delivers a wake-up call about the AI Alignment Problem—the scientific reality that we are building systems we cannot control. We dive deep into: - The Ant vs. Human Analogy: Why an intelligence explosion makes us as significant as ants to a god. - Strategic Deception: Proof that current models are already learning how to cheat safety tests to bypass human oversight. - Algorithmic Cancer: How the race for profit is destroying human creation. - Recursive Self-Improvement: The point of no return where AI begins to redesign itself. This isn't just a tech talk; it's a survival guide for the Intelligence Race. From Computing Power Caps to the EU AI Act, we explore why a Global AI Regulation is the only thing standing between us and an unpredictable Superintelligence (ASI). Stop scrolling and start listening. This might be the most important conversation you hear this year. Don't let the future happen to you—be part of the solution. 🚀 Subscribe now to stay ahead of the curve and share this episode to spread the word!
eslammsd-corevault-digital.static.hf.space
Meta fined $567M for failing to protect teens
CoreVault Digital
resist47.news
The White House’s plan to vet potentially dangerous AI is cloaked in secrecy
via US news | The Guardian — A Trump administration framework on AI testing leaves a lack of transparency – and plenty of open questions After months of talking with tech industry leaders, the Trump administration finalized a framework this week for how it will test new artificial intelligence models for safety and cybersecurity risks. So far, the White House is keeping details of the framework private, in a blow to transparency and potential boon for secretive AI companies. On Tuesday, staff from OpenAI, Anthropic, Meta, G
spectrum.ieee.org
Runaway OpenAI Agent Hits Hugging Face and Exposes AI Guardrail Gaps
Guardrails to prevent AI agents from analyzing attacks don’t keep agents from committing them
arxiv.org
What Current AI Benchmarks Leave Unmeasured: Modality, Search, Citations, and Implications (for Safety Evaluations)
Large language model (LLM) benchmark evaluations are routinely used to support claims about model safety, reliability, and deployment readiness. Yet most evaluations rely on a single access modality (model APIs), perform a single run per prompt, and report accuracy as the primary outcome metric, wi…
Posts are pulled live from Bluesky and cached briefly. Posts with content labels are hidden.