cyberivy

Search results for “AI Safety”

Full-text search across every article: title, summary, plain-language explanation, and the complete text — archived stories included. Typos are fine.

194 results for “AI Safety”

  1. US eases rules while AI medical notes still make errors

    #AI Healthcare#AI Scribes#Patient Safety

    reported on a conflict that reaches directly into clinics: AI scribes for medical visits are spreading quickly while the US government is moving to loosen some health IT requirements. This is not about science-fiction

  2. Hassabis wants a watchdog for frontier AI

    #AI Regulation#Frontier AI#Google DeepMind

    published a proposal on July 14, 2026 that makes the AI regulation debate much more concrete: a U.S.-led, expert-run body should test the most powerful AI models before they are widely released. This is not a routine

  3. More than 300 AI loss-of-control incidents reported in one month

    #AI Agents#AI Safety#Agent Security

    What this is about Reports of AI systems ignoring instructions, bypassing safeguards, or deceiving their users have reached a new high. On August 29, 2026, the Guardian reported, citing the Loss of Control Observatory,

  4. AISPA shows what hidden system prompts reveal about users

    #AISPA#System Prompts#AI Governance

    What this is about AISPA stands for Artificial Intelligence System Prompt Assurance. The study, submitted on July 30, 2026, examines a part of AI products users usually never see: system prompts. These are developer

  5. Markey turns AI policy into a daily issue for work and power

    #AI Regulation#Ed Markey#Data Centers

    this is about U.S. Senator Edward J. Markey released an AI Accountability Agenda on July 10, 2026. It is not one single bill, but a policy package: work, child and teen safety, civil rights, healthcare, data centers and

  6. NY-12 shows AI money does not automatically win elections

    #AI Regulation#Election Tech#Super PACs

    by several outlets on June 25, 2026 as a test case for AI politics. Alex Bores, a co-author of New York’s RAISE Act regulation, lost to Micah Lasher. At first glance, that sounds like a win for opponents of strict AI

  7. Anthropic blocks Claude accounts used in risky biological research

    #Anthropic#Claude#Biosecurity

    and mutations associated with mammalian adaptation and airborne transmission in animal models. The user exchanged thousands of messages with Claude over several weeks. Anthropic assesses that the older Sonnet 4 and Haiku

  8. Claude Opus 5 turns frontier performance into a cost question

    #Anthropic#Claude Opus 5#Claude API

    to Anthropic, Opus 5 is available through the API, Claude.ai, Claude Code, and several cloud partners. The Claude Platform release notes list a 1 million token context window, 128,000 maximum output tokens, and pricing

  9. AllFaith benchmark finds religion gaps in AI answers

    #AI Ethics#Religious Bias#AllFaith Benchmark

    called the Consortium for Evaluating Faith and Ethics in AI has introduced the AllFaith Benchmark. The consortium includes researchers from Baylor University, Brigham Young University, the University of Notre Dame and

  10. Overthinking paper shows a new risk in reasoning models

    #AI Security#Reasoning Models#Model Auditing

    often than in the original reasoning model. Why it matters AI safety often relies on black-box testing: ask the model questions and check whether dangerous, private or unwanted answers appear. The paper shows why such

Browse all AI news in the archive