cyberivy

Search results for “Coding Agents”

Full-text search across every article: title, summary, plain-language explanation, and the complete text — archived stories included. Typos are fine.

249 results for “Coding Agents”

  1. agent-browser controls websites through a lean CLI

    #agent-browser#Vercel Labs#Browser Automation

    designed less for humans watching a screen and more for AI agents and coding assistants. Instead of pushing a model through long DOM dumps or brittle CSS selectors, the CLI returns compact snapshots with references such

  2. AI agents fail open-ended research in a real-world test

    Archived#AI Agents#AI Research#Shadow Evaluation

    AI industry's most consequential claims: can today's AI agents conduct genuinely new research on their own? A team including Peter Kirgis, Sayash Kapoor, and Arvind Narayanan gave agents the central questions from two

  3. OpenAI slows Astra over possible critical cyber capabilities

    #OpenAI#Astra#Cybersecurity

    OpenAI says Astra shows significant advances in agentic coding and cybersecurity. Performance is strong enough that the company cannot currently rule out a Critical capability level under its Preparedness Framework. What

  4. Tabstack gives agents web data and browser actions by API

    #Tabstack#Mozilla#Browser Automation

    data and browser automation API for developers building AI agents or data-driven products. The product accepts a URL, a schema, a question, or a task and returns Markdown, JSON, cited research, or completed browser

  5. IBM Bob Turns AI Coding Into an Enterprise SDLC Tool

    #IBM Bob#AI Coding#Enterprise AI

    What this is about IBM Bob is IBM’s attempt to move AI coding from individual assistant use into the governed software lifecycle of larger teams. IBM describes Bob as an AI SDLC partner for real codebases, not just a

  6. Microsoft Agent Governance Toolkit controls agent actions

    #Microsoft#AI Agent Governance#AI Security

    tool for a problem that becomes practical as soon as AI agents leave demos: an agent can send emails, query tables, open browsers, or execute code. The question is no longer only whether the model writes good answers,

  7. Shai-Hulud worm hits TanStack and Mistral packages

    #AI Security#Supply Chain#Shai-Hulud

    to any team using npm, PyPI, GitHub Actions, and AI coding tools. What Shai-Hulud actually does The reports describe a self-spreading credential stealer. After installation, the payload searches for access data: GitHub

  8. Study: AI agents still fail at real research

    #AI Agents#AI Research#arXiv

    29, 2026 asks a simple but hard question: can today's AI agents conduct open-ended AI research when the task is not just tests, code, and benchmarks, but real scientific judgment? The author group, including Peter

  9. GitHub Copilot Gets Cloud and Local Sandboxes

    #AI Security#Cyber Ivy#June 2026

    Linux environment. This targets the core risk of agentic coding tools: they read code, execute commands, and modify files. Why it matters: Once an agent executes tools, the chat transcript is not a security boundary.

  10. OpenHistory turns Mac activity into a private work memory

    #OpenHistory#Local AI#Productivity AI

    disabled or run with Apple's on-device model. For local AI agents, OpenHistory exposes an authenticated MCP server on loopback. It returns a redacted, read-only projection instead of the complete activity log. Each

Browse all AI news in the archive