ACM

Non classé

Four of five enterprises that secured AI agent identities still can’t contain one that goes rogue

Visa’s president of technology, Rajat Taneja, walked the VB Transform 2026 audience through aiming Anthropic’s Mythos at Visa’s own payment network. The model stitched minor weaknesses into working exploit chains, and Visa open-sourced the harness that governed the hunt. That’s what it looks like when an enterprise has the engineering depth to act on what …

Four of five enterprises that secured AI agent identities still can’t contain one that goes rogue Read More »

SpaceXAI debuts Grok 4.6, overtaking Kimi K3’s performance and matching GPT-5.6 Sol for world’s third best on Artificial Analysis

Elon Musk’s company SpaceXAI, formerly known as xAI, has released Grok 4.6, its latest frontier AI model, with a focus on long-running agents, coding and knowledge work — and a pricing strategy designed to make those workloads cheaper to run. The model scores 61 on the third-party Artificial Analysis Intelligence Index, surpassing the popular open …

SpaceXAI debuts Grok 4.6, overtaking Kimi K3’s performance and matching GPT-5.6 Sol for world’s third best on Artificial Analysis Read More »

Skan AI raises $63 million betting that watching how employees actually work is the missing layer of enterprise AI

Skan AI, a startup that builds what it calls a “context graph of work” by observing how employees actually perform their jobs across enterprise software, has raised $63 million in Series C funding co-led by Cathay Innovation and Dell Technologies Capital, the company announced Wednesday. Citi Ventures, Bloomberg Beta, State Farm Ventures, and Wipro Ventures …

Skan AI raises $63 million betting that watching how employees actually work is the missing layer of enterprise AI Read More »

Agentic security: Enterprises enforce agent permissions two-thirds of the time — and isolate high-risk agents less than one in five

Across 116 enterprises, agents are in production and so are the incidents: A majority have already had a confirmed agent security event or a near-miss. Two-thirds of enterprises enforce scoped permissions at runtime. Barely one in five isolates its highest-risk agents, making containment the weakest layer in the stack precisely as autonomy scales. Credential sharing …

Agentic security: Enterprises enforce agent permissions two-thirds of the time — and isolate high-risk agents less than one in five Read More »

Agentic reliability and evaluations : Enterprises that got burned by a bad eval are the most likely to remove humans from the loop, not the least

Across 108 enterprises, trust in automated agent evaluation rose sharply in July — and the failure rate it is supposed to predict did not move at all. The share of organizations that fully trust automated evaluation nearly tripled, from 5% in June to 13%, and the complaint that evaluations don’t match real-world outcomes fell 10 …

Agentic reliability and evaluations : Enterprises that got burned by a bad eval are the most likely to remove humans from the loop, not the least Read More »

Agent context layers: Enterprises governing their AI data are catching twice as many bad answers as the ones who aren’t

Across 101 enterprises, the context feeding AI agents is failing often and repeatedly. Sixty-eight percent have traced a confident but wrong agent answer to missing or inconsistent business context in the past six months, and the single most common answer is not “once” but “more than once.” The counterintuitive part is which companies report it. …

Agent context layers: Enterprises governing their AI data are catching twice as many bad answers as the ones who aren’t Read More »

Agentic orchestration: Enterprise AI organizations know how to govern agents but still can’t meter what they cost

Across 107 enterprises, agentic orchestration is not a choice of a single platform. The typical enterprise runs three orchestration platforms at once, and selects them for flexibility across models rather than affinity to any single one. Microsoft leads primary usage while Anthropic leads forward consideration by a wide margin.  The AI control plane enterprises expect …

Agentic orchestration: Enterprise AI organizations know how to govern agents but still can’t meter what they cost Read More »

Infrastructure and compute: Enterprises are buying AI compute for speed while flying blind on what it costs

Across 170 enterprises, AI infrastructure has moved decisively into production — two-thirds now run AI workloads live and three in 10 run them at scale — while the ability to account for what that infrastructure costs has not kept pace. Enterprises have quietly demoted cost in the buying decision: performance and GPU availability now outrank …

Infrastructure and compute: Enterprises are buying AI compute for speed while flying blind on what it costs Read More »

SpaceXAI’s Grok Bot turns agents into persistent digital coworkers that can operate your apps for $120-per-month

SpaceXAI, the division of SpaceX formerly known as xAI, is launching an early beta version of Grok Bot, a new agent designed to move AI assistants beyond answering prompts and toward continuously executing work across the software employees already use. The central idea is straightforward: instead of opening an AI assistant whenever a task arises, …

SpaceXAI’s Grok Bot turns agents into persistent digital coworkers that can operate your apps for $120-per-month Read More »

Why AI-driven purchase intent so rarely becomes a completed sale

Presented by Rezolve Ai When an AI assistant recommends a product or brand, it generates something valuable: a purchase-ready consumer with high intent and low friction in their decision. That consumer has already compared options, asked follow-up questions, and arrived at a conclusion. They want to buy. What they encounter next is a commerce infrastructure …

Why AI-driven purchase intent so rarely becomes a completed sale Read More »