HomeAIAI Agents

AI Agents - Page 7

Stories from AI Agents come first, newest first, then stories from connected topics (AI Agents, OpenClaw, MCP, Open Source, coding agents), each labelled with where it comes from. This view is not indexed by search engines.

Latest

Anthropic Launches the Anthropic Cyber Mission: Critical-Infrastructure Programme and Free Open-Source Scanner

Anthropic has launched the Anthropic Cyber Mission: a Critical Infrastructure Defense Program with 11 founding partners and OSS Scanner, a free opt-in service that sends open-source projects unreviewed, model-written bug reports. What is confirmed, what is Anthropic's expectation, and what is not yet proven.

An AI-Built Map of the UV Sky, and Two Warnings About Agent Scores

Anthropic says a team of Claude Science agents helped build the first complete UV map of the sky, about a third of it predicted rather than observed. A preprint finds agent benchmark scores can move because of the grader, not the agent, and Exa previews a search benchmark it designed and runs itself. What each one shows, and what it doesn't.

Apple Reportedly Trims iPhone 18 Pro Parts Orders, and a $99 Ring for AI Agents

Two device stories from 8–9 October 2026. Nikkei Asia reports that Apple told some suppliers to cut production of iPhone 18 Pro components after price rises dampened demand; Apple has not commented. And start-up Natura has shown Interface, a $99 smart ring for sending voice requests to AI agents, which cannot be ordered yet.

An Open-Source AI Hacking Tool Was Used Against South Korean Banks, CrowdStrike Says

CrowdStrike says an unnamed, likely Chinese-speaking attacker used ARTEX, an open-source AI penetration-testing agent, alongside several AI models to steal data from South Korean financial firms. What CrowdStrike found, what Korean investigators say, and what it does not prove.

AI Rules This Week: Anthropic Rewrites Its Rulebook, and the UK Regulator Turns to AI Agents

Anthropic has published a new Usage Policy, taking effect on 12 November, with a section on deceptive campaigns, tighter wording on weapons and surveillance, and a ban on extreme abuse of its models. The UK's ICO says ten AI developers have made or committed to data protection changes, and has opened a call for evidence on AI agents. What changes, and for whom.

Two Open Models From 8 October: a Fast Coding Agent and a Gene Finder

JetBrains has released Mellum2.1, a small open model for coding agents, and Hugging Face has released Carbon-A, a 1.2B-parameter model that predicts protein-coding regions in DNA, plus a database of 566 million predicted loci. What each does, and which numbers are the developers' own.

Gemini for Mac Tests a Hidden “Full Access” Setting

An unreleased "Additional sandbox options" setting found in Google's Gemini desktop app for Mac would let the AI reach files, apps and network connections beyond the folders users connect today. It is hidden, unconfirmed by Google and arrives just as Apple is tightening Full Disk Access.

Samsung’s Record Quarter and AWS’s Linux Foundation Seat: Two Signs of How AI Is Reshaping Big Tech

Samsung's third-quarter guidance points to about ₩107.4 trillion in operating profit on ₩195 trillion of sales, up from ₩12.17 trillion a year ago, with AI data-centre memory demand credited. AWS becomes a Platinum member of the Linux Foundation, with a board seat. What the numbers say and what they leave out.

AI Money and Mood: Agents and Voices Draw Cash While the Public Wants Tighter Rules

Manus raises more than $500m after Beijing blocked Meta's takeover. Tab emerges at a $300m valuation. Nvidia reportedly discusses another $1bn for Figure. ElevenLabs says it has passed a $600m run rate, and Synthesia launches Syren in beta. A Reuters/Ipsos poll finds most Americans think government is not taking AI risks seriously enough.

AI Research Tools: TasteVal Measures “Research Taste”, Karotte Targets Reward Hacking and Zephon Fixes Data Order

A new benchmark, TasteVal, says the best AI model now reaches expert-level results on eight AI research tasks with 2.3 times less compute, though its authors list reasons that may overstate progress. Preference Model has open-sourced Karotte, a framework for training environments that are harder to game, and DatologyAI has released Zephon, a data loader that keeps experiments repeatable when the GPU count changes. What each claims, and who measured it.

Latest on TSN

Free TSN tools: crypto calculator, Flux dashboard and more.