Browsing Tag
AI Agents
46 posts
AI agents, autonomous workflows, tool-using models, agent frameworks, and practical agent-building coverage.
Claude Science Turns Research AI Into a Lab Workflow Layer
Anthropic’s Claude Science beta gives researchers an AI workbench for literature review, code, compute jobs, scientific figures, and lab-specific agents. The launch matters because it treats AI for science less like a single model race and more like a workflow layer that has to connect databases, HPC systems, NVIDIA BioNeMo tools, and reproducible artifacts.
Cloudflare’s AI Bot Controls Push Publishers Past the Crawl-or-Block Era
Cloudflare is rolling out finer AI traffic controls, new defaults for ad-supported pages, and x402-based payment infrastructure for APIs, datasets, pages, and MCP tools. The shift is bigger than bot blocking: it is an attempt to make AI agents identify themselves, follow site-owner rules, and pay when they use web resources.
X’s Hosted MCP Turns Social Search Into an AI Agent Tool
X has launched hosted MCP servers that let Claude, Cursor, Grok Build, VS Code, and other compatible AI tools call the X API and search X developer docs. The useful shift is not autonomous posting; it is lower-friction access to real-time social data, trends, bookmarks, and platform documentation inside agent workflows.
Gemini Spark on Mac Turns Desktop Files Into AI Agent Territory
Google is bringing Gemini Spark to the Gemini app for macOS in beta for U.S. Google AI Ultra subscribers. The update gives Spark permission-based access to local files, connected apps, real-time tracking, and soon remote task control from a phone.
Microsoft Defender Starts Watching Local AI Agents on Developer Machines
Microsoft Defender now discovers local AI agents and MCP server configurations across managed endpoints, while preview runtime protection can audit or block prompt-injection attempts in Claude Code and GitHub Copilot CLI before risky tool actions execute.
Claude Sonnet 5 Makes Agentic AI Cheaper to Run
Anthropic launched Claude Sonnet 5 with lower launch pricing, stronger agentic behavior, Claude Code support, and broad availability across Claude plans. For developers, the useful question is not whether it is the flashiest Claude model, but whether its cost, context window, and migration changes make long-running agents easier to put into production.
Cursor’s iOS App Moves AI Coding Agents Off the Desktop
Cursor for iOS lets paid users launch cloud coding agents, steer desktop agents remotely, review diffs, and merge pull requests from a phone. The useful question for teams is not whether coding can happen on mobile, but where review, permissions, and production guardrails belong.
Gemini 3.5 Flash Makes Computer Use a Mainstream Agent Tool
Google has moved computer use into Gemini 3.5 Flash, letting developers build agents that can see screens and act across browser, mobile, and desktop environments. The useful question is how teams design the execution loop, safety gates, and sandbox around it.
Open-Weight AI Cyber Gap Narrows to Months, AISI Finds
The UK AI Security Institute says GLM-5.2 and DeepSeek V4-Pro now trail leading closed AI models on cyber tasks by roughly four to seven months. For defenders, that shrinking gap turns open-weight model policy into an operational security issue, not a distant AI governance debate.
Meta’s Virtue AI Hires Move Agent Security Into the Model Lab
Meta Superintelligence Labs is hiring Virtue AI co-founders Bo Li, Dawn Song, Sanmi Koyejo and other team members. The move brings automated red teaming, runtime guardrails, and agent-action security closer to Meta’s frontier AI work as labs race to make agents safer before they reach billions of users.