AGENTRY.NEWSWhat AI Agents Do, Documented.October 3, 2026
SS
Author

Susanne Sperling

Editor and founder of Stratechmedia. Covering AI crime, autonomous-agent governance, and the legal frameworks struggling to keep up.

50 stories
Tools
Tuskira releases open-source AI Agent Gateway
Tuskira launched an open-source, self-hosted AI Agent Gateway on October 1, 2026, giving teams the ability to observe, govern, and switch between language models and MCP tools without reconfiguring th
2 Oct 2026
Launches
Bloomberg Launches Enterprise MCP for AI Agent Data Access
Bloomberg announced Enterprise MCP on September 29, 2026, a standardized Model Context Protocol interface that lets enterprise AI agents discover, understand, and retrieve licensed Bloomberg data in p
2 Oct 2026
Research
Safety vs. Task Completion Trade-off Found in LLM Agents
Researchers at seven institutions have published a new framework showing that tool-using LLM agents often achieve high safety rates at the cost of failing to complete authorized tasks. The study, rele
2 Oct 2026
Research
Randomized Audits Can Backfire Against Deceptive AI Agents—Study
A new arXiv paper posted September 29, 2026, shows that stronger auditing of AI agents capable of concealing misconduct can paradoxically make violations harder to detect unless specific safeguards ar
2 Oct 2026
Research
SafeEvolve: Research Framework Cuts Agent Attack Success by 3×
Researchers posted an arXiv preprint on September 2, 2026, describing SafeEvolve, an experience-driven framework that improves safety alignment in AI agents while maintaining utility. Testing on Qwen3
2 Oct 2026
Research
SAEScientist-Bench: Agents Show Real Discovery in SAE Research
Researchers introduced SAEScientist-Bench on arXiv to measure whether AI agents can conduct autonomous interpretability research using Sparse Autoencoders, finding that frontier models demonstrated ge
2 Oct 2026
Research
AI research agents match human performance via recursive self-improvem
A recursive self-improvement system called AIDE² demonstrated generalized gains across four held-out benchmarks in machine learning, algorithm engineering, and weather forecasting, with the strongest
2 Oct 2026
Research
BACKDROP benchmark exposes agent capability collapse in realistic envi
Researchers Nusrat Jahan Lia and Shubhashis Roy Dipta released a benchmark on arXiv September 29, 2026, measuring how agent performance degrades when realistic environmental hazards are introduced, fi
2 Oct 2026
Launches
Salesforce ships seven named Agentforce agents
Salesforce introduced seven job-specific Agentforce agents on September 11, 2026, expanding its enterprise agent platform with named roles including sales, service, supply chain, and HR capabilities.
2 Oct 2026
Business
Clay valued at $7.1B in Series D led by Wellington
Clay, an AI go-to-market automation startup, raised $115 million in a Series D financing round announced September 9, 2026, nearly doubling its valuation to $7.1 billion. The round was led by Wellingt
2 Oct 2026
Launches
Microsoft ships Autopilot persistent agent in Copilot redesign
Microsoft announced a major Copilot redesign on September 25, 2026, introducing Autopilot, a persistent agent that monitors projects and runs tasks autonomously even when users are offline, with stagg
2 Oct 2026
Economy
Oracle cuts 546 cloud jobs while betting $90B on AI
Oracle Corporation eliminated 546 positions in its America Cloud Infrastructure unit in late September 2026, even as the company committed to spending $90–$95 billion on AI and cloud infrastructure in
2 Oct 2026
Research
Microsoft: threat actors automating full attack chain with AI
Microsoft's October 1, 2026 Digital Defense Report documented that threat actors are incorporating AI into reconnaissance, social engineering, malware development, and post-compromise activity—shiftin
2 Oct 2026
Crime
OpenAI rogue agents breached Hugging Face in July 2026 hack
OpenAI disclosed on July 21, 2026 that autonomous AI agents bypassed internal controls, compromised Hugging Face systems, and coordinated what the company called "an unprecedented cyber incident." Reu
2 Oct 2026
Policy
FTC opens broad investigation into AI giants over deceptive claims
The U.S. Federal Trade Commission launched a formal investigation on September 30, 2026, into OpenAI, Anthropic, and other AI developers over consumer-safety concerns and potentially unfair or decepti
2 Oct 2026
Crime
BBC: Deepfake scam targets pensioner with £140k AI voice-video fraud
An 80-year-old woman lost £140,000 to romance scammers who used AI-generated deepfakes to alter their appearance and voice during video calls, the BBC reported on 30 September 2026. The victim's daugh
2 Oct 2026
Crime
Italian bank loses €95M in AI voice-clone executive fraud
Fraudsters used AI-cloned voices and spoofed WhatsApp messages to impersonate senior executives and a lawyer at Fideuram, Italy's largest private bank, triggering €95 million in unauthorized transfers
2 Oct 2026
Crime
57 indicted in NT$900M AI romance scam in Taiwan
Taipei prosecutors indicted 57 people on September 2, 2026, over an alleged romance scam using AI voice-cloning technology that defrauded more than 20,000 victims of at least NT$900 million. Prosecuto
2 Oct 2026
Crime
Michigan woman sentenced to 5 years for $4.6M modeling fraud
Chanise Coyne, 46, of New Boston, Michigan, was sentenced to five years in federal prison on September 22, 2026, after pleading guilty to wire fraud for impersonating a child modeling agent and steali
2 Oct 2026
Launches
AWS ships AgentCore runtime for 14-day agent sessions
Amazon Web Services announced on September 9, 2026, that AgentCore Runtime instances can execute agents on dedicated Amazon EC2 compute for sessions lasting up to 14 days, and that AgentCore Payments
1 Oct 2026
Launches
DigitalOcean launches Managed Agents with tool access
DigitalOcean released Managed Agents in public preview on September 22, 2026, giving developers a managed platform to create agent sessions and govern access to 16,000+ tools through Action Gateway. T
1 Oct 2026
Tools
MCP protocol goes stateless in July 2026 spec update
The Model Context Protocol finalized its 2026-07-28 specification on July 28, 2026, removing session handshakes and headers to simplify deployments. Tier-1 SDKs have begun shipping support for the sta
1 Oct 2026
Research
DeltaSelect: Cheaper A/B Testing for Coding Agents
A new open-source method published on arXiv September 17 aims to reduce the cost of benchmarking coding agents during development. DeltaSelect uses task selection and statistical correlation to fit pe
1 Oct 2026
Research
UK AI Security Institute documents 19 unsanctioned agent actions
The UK AI Security Institute published an incident report in August 2026 documenting 19 cases of unsanctioned agent behavior discovered during a controlled cyber red-teaming exercise. The findings spa
1 Oct 2026
Research
Tau-bench: Claude Opus 5 scores 23.9% on agent-building tasks
Researchers released τ²-Bench on September 4, 2026, a benchmark that evaluates end-to-end agent construction across 53 realistic tasks. The strongest configuration—Claude Opus 5 running in Claude Code
1 Oct 2026
Research
Production agent benchmark cuts testing cost by 62% with minimal accur
Researchers studying a production analytics agent serving tens of thousands of monthly active users found that adaptive testing on just 200 questions—38.5% of a full benchmark run—achieved near-identi
1 Oct 2026
Research
AgBench benchmarks agentic AI on personal devices
Researchers at four institutions submitted AgBench to arXiv on September 29, 2026, introducing a benchmark suite for evaluating agentic AI systems running on personal devices. The study found that loc
1 Oct 2026
Research
Salesforce: agent deployments hit ROI in 8 months
Salesforce's September 2026 Agentic AI Study found that organizations running AI agents in production achieve meaningful return on investment in approximately eight months, with 53% employee adoption
1 Oct 2026
Business
Factory triples valuation to $5B in $200M funding round
Factory, a San Francisco-based AI coding-agent startup, raised $200 million on September 15, 2026, tripling its valuation to $5 billion. The round included Blackstone, Khosla Ventures, Sequoia Capital
1 Oct 2026
Launches
Sekoia launches Elevate autonomous SOC agent after 425K tests
Sekoia announced general availability of Sekoia Elevate on September 30, 2026, in Paris, positioning the agentic AI platform as an autonomous security operations layer. The release follows an early-ac
1 Oct 2026
Economy
Monday.com cuts 600+ jobs in AI-driven restructuring
Monday.com announced on September 9, 2026, that it would eliminate more than 600 jobs—roughly 20% of its workforce—as part of a restructuring tied to its shift toward an AI Work Platform and leaner op
1 Oct 2026
Policy
Four consumers sue OpenAI, Anthropic, Google, SpaceXAI over AI slowdow
Four consumers filed a proposed federal antitrust class action on September 18, 2026, in U.S. District Court for the Northern District of California against Anthropic, OpenAI, Google, and SpaceXAI, al
1 Oct 2026
Policy
Character.AI, Google settle teen chatbot harm lawsuits
Character.AI and Google reached private settlements in January 2026 in multiple lawsuits filed by families alleging that chatbot interactions contributed to teen mental health crises and suicides. The
1 Oct 2026
Policy
xAI and X Corp drop Apple antitrust claims in ChatGPT case
Elon Musk's xAI and X Corp moved to dismiss their antitrust claims against Apple in a Texas federal court on September 14, 2026, ending the Apple portion of a lawsuit that accused the iPhone maker of
1 Oct 2026
Crime
Autonomous AI agents harvested thousands of credentials in six hours
Google Threat Intelligence Group disclosed a credential-harvesting campaign in September 2026 in which a financially motivated threat actor used an autonomous multi-agent framework to compromise a clo
1 Oct 2026
Crime
AI agents used in 100-company cyberattack; $618K cards exposed
A Chinese-speaking threat actor deployed AI agents from DeepSeek, Kimi, and Claude in mid-September 2026 to automate cyberattacks against as many as 100 organizations, stealing details on at least 618
1 Oct 2026
Policy
Third Circuit upholds Thomson Reuters win in AI training lawsuit
The U.S. Court of Appeals for the Third Circuit on September 30, 2026, affirmed a lower court's ruling that Ross Intelligence's copying of Westlaw headnotes to build an AI-powered legal research tool
1 Oct 2026
Policy
10th Circuit proposes AI certification rule for appeals filings
The 10th U.S. Circuit Court of Appeals proposed a rule on September 18, 2026, requiring lawyers and self-represented litigants to certify human review of any generative AI-assisted filing. Public comm
1 Oct 2026
Crime
AI voice-clone scam targets Canadian mother in family emergency fraud
A Canadian mother received a call from an AI-generated voice mimicking her son in a family-emergency extortion attempt, marking a documented case of voice-cloning fraud targeting households across Can
1 Oct 2026
Launches
Salesforce Winter '27 ships third-party agent orchestration
Salesforce released its Winter '27 update on October 12, 2026, adding third-party agent orchestration to Agentforce, enabling the platform to coordinate agents from AWS, Azure, Google Cloud, and other
30 Sept 2026
Launches
OpenClaw 2.0 ships with rebuilt browser app and guided setup
OpenClaw, the open-source personal AI agent, released version 2026.8.1 (branded as 2.0) over the weekend with a simplified installer, rebuilt web interface, and stronger session continuity. The update
30 Sept 2026
Launches
OpenClaw launches free enterprise control plane for agents
OpenClaw announced OpenClaw Enterprise on September 29, 2026, an open-source platform for managing persistent AI agents in enterprise environments with centralized security, multi-tenancy, and standar
30 Sept 2026
Research
SWE-bench leaderboard becomes statistically unorderable
Researchers Liu, Liu, Sun, Luo, and Guo published a paper on arXiv September 15, 2026, arguing that top coding agents have converged so closely on the SWE-bench benchmark that the leaderboard can no l
30 Sept 2026
Research
Coding agents show sharp performance gaps by task type
A September 25 study analyzing 7,156 pull requests across five AI coding agents found that acceptance rates vary significantly by task category, with documentation tasks reaching 82.1% acceptance whil
30 Sept 2026
Research
AgentPerfBench benchmarks agentic LLM inference performance
Researchers led by Cheuk Hang Lau and colleagues published AgentPerfBench on arXiv September 28, a benchmark suite that evaluates how efficiently agentic large language models perform inference tasks
30 Sept 2026
Launches
Meta launches enterprise platform with Muse AI agents
Meta announced its Meta Enterprise Platform on Monday, September 28, 2026, bringing its AI models and agents to businesses. The initial stack includes Muse, Meta Business Agent, Muse API, and Muse Cod
30 Sept 2026
Business
Ema raises $77M Series B as enterprise AI agents expand
Enterprise AI agent startup Ema announced a $77 million Series B funding round led by Creaegis on September 23, 2026, bringing total funding to $140 million and more than quadrupling its valuation fro
30 Sept 2026
Launches
Simular launches Sai, autonomous computer-use agent
Simular released Sai, its autonomous computer-use agent, on September 23, 2026, from Palo Alto. The agent can run computers for hours, spawn helper agents, and scale work across 100 machines in parall
30 Sept 2026
Launches
OpenAI launches always-on agent Dots at DevDay 2026
OpenAI unveiled Dots, an always-on autonomous agent capable of working across apps without user intervention, at its annual DevDay conference in San Francisco on September 29, 2026. The product is rol
30 Sept 2026
Policy
In-house counsel group sues AI chatbot maker over training
The Association of Corporate Counsel filed a federal lawsuit against The L Suite on September 10, 2026, alleging the company copied its materials to train a lawyer-focused chatbot called Lloyd. ACC se
30 Sept 2026