
MLCommons Agent Reliability Profile Reaches Final Round of C:>DIR Agentic Regulator Hackathon
MLCommons announced its Agent Reliability Profile, a flagship initiative of its Financial Services Working Group, has been named a finalist in the C:>DIR Global "Agentic Regulator" Hackathon. The profile advances to the final round, with winners scheduled to be announced on September 18.
Sources and evidence
Summary last validated Sep 15, 2026
Reader actions
Report an issue
Use this for an incorrect summary, wrong source, duplicate story, or wrong category. Submissions are private and do not change the story automatically.
Related coverage
- agents
NVIDIA Highlights Developers Using Frontier AI Agents With Omniverse Libraries
NVIDIA published a blog post describing how developers combine frontier AI models with NVIDIA Omniverse libraries to build simulation applications. The post says developers direct AI agents to assemble assets, connect physics and rendering, and verify scene behavior, supporting work such as exploring scenarios, investigating failures and improving designs.
- agents
Claude Code v2.1.295 adds hook failure blocking and gateway model controls
Anthropic released Claude Code v2.1.295, adding onFailure:"block" so hooks that fail to start, time out, or exit unexpectedly block the action. It also adds Program Status Protocol (OSC 7501) terminal support, optional per-upstream models lists with wildcards, upstream_ttfb_ms stream timeouts, and gateway login settings.
- agents
BlockRun and Incarna use Amazon Bedrock AgentCore payments for per-inference agent billing
Amazon Bedrock AgentCore payments lets AI agents pay for services on demand with infrastructure-enforced spending limits. Incarna's agents pay BlockRun for model inference one request at a time over x402, reducing the work of adding x402 payment support from months to days, according to an AWS Machine Learning Blog post.
- agents
NVIDIA KGMON Team Places Second in KDD Cup 2026 Data Agents Competition
NVIDIA's KGMON team placed second in the KDD Cup 2026 Data Agents competition. The team built a system around making an agent's harness smaller, clearer, and easier to verify. The competition required agents to answer natural-language questions across heterogeneous sources including databases, CSV and JSON files, prose documents, PDFs, and briefing videos.
- industry
MLCommons publishes business guide to measuring and managing AI risk
MLCommons released a guide titled "Assessing & Managing AI Risk for Business and Commercial Deployments," aimed at business owners. It argues teams managing AI risk face tension between protecting the business and adopting AI faster, and that independent measurement can reconcile the two. The guide frames AI risk as a cost that can be measured, priced, and managed.