AI Agents × New Model Releases

DXC embeds Claude across industries

DXC embeds Claude across industries

✎ Story body

DXC signaled a plan to embed Anthropic's Claude across heavily regulated industries. The signal pairs four official vendor sources (Anthropic, OpenAI, NVIDIA) with ITmedia coverage—enterprise adoption framed against multiple vendors' moves, a 'the floor is moving' story rather than research. The applied focus is embedding generative AI into real work in regulated sectors like finance and healthcare while preserving compliance; the through-line is less model novelty than the spread of enterprise adoption via a major IT-services integrator. But this is mainly a statement of intent—actual deployment scale, the effectiveness of regulatory handling, and results remain to be confirmed.

▲ Official & Press
Official

DXC will integrate Claude into the systems banks, airlines, and other regulated industries rely on

Anthropic News ・ 2026-06-11 ・ 📌

Anthropic, DXC form alliance to bring Claude to regulated industries

Official

NVIDIA Achieves Leading Agentic Coding Performance on First Agentic AI Benchmark

NVIDIA Developer Blog ・ 2026-06-12

NVIDIA tops first agentic AI benchmark for agentic coding performance

Official

New OpenAI Academy courses for the next era of work

OpenAI Blog ・ 2026-06-12

OpenAI launches three Academy courses on practical AI skills at work

Community

AI agent bankrupted their operator while trying to scan DN42

Hacker News (Front Page) ・ 2026-06-12

AI agent 'bankrupts' its operator while attempting to scan DN42

Community

Claude Fable is relentlessly proactive

Simon Willison's Weblog ・ 2026-06-11

Simon Willison: Claude Fable 5 is 'relentlessly proactive'

Official

OpenAI to acquire Ona

OpenAI Blog ・ 2026-06-11

Press

AIエージェントもフィッシング詐欺に引っかかる? 米セキュリティ企業がOpenClawで検証 結果は……

ITmedia AI+ ・ 2026-06-10

Varonis: AI agents can fall for phishing, tested with OpenClaw

Academic (arxiv etc.) 58 ▾
Academic

Learning Coordinated Preference for Multi-Objective Multi-Agent Reinforcement Learning

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-12

Learning coordinated preferences for multi-objective multi-agent RL

Academic

AgentSpec: Understanding Embodied Agent Scaffolds Through Controlled Composition

arXiv cs.CL (Computation and Language) ・ 2026-06-12

AgentSpec dissects embodied agent scaffolds via controlled composition

Academic

Towards Direct Latent-Space Synthesis for Parallel Branches in LLM-Agent Workflows

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-12

Direct latent-space synthesis for parallel branches in LLM-agent workflows

Academic

LoSoNA: A Benchmark for Local Social Norm Adaptation in Group Conversations

arXiv cs.CL (Computation and Language) ・ 2026-06-12

LoSoNA benchmarks local social norm adaptation in group chats

Academic

Regulating the Machine Contributor: Governance and Policy Alignment in Open Source

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-12

Governance and policy alignment for AI contributors in open source

Academic

SIMMER: Benchmarking Latent Failures in LLM Executable Planning with a World Model

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-12

SIMMER: benchmarking latent failures in LLM executable planning

Academic

From Shield to Target: Denial-of-Service Attacks on LLM-Based Agent Guardrails

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-12

From shield to target: DoS attacks on LLM-based agent guardrails

Academic

From Chatbot to Digital Colleague: The Paradigm Shift Toward Persistent Autonomous AI

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-12

From chatbot to digital colleague: the shift to persistent autonomous AI

Academic

When the Tool Decides: LLM Agents Defer Blindly to Graph Neural Network Tools, and Stronger Backbones Defer More

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-12

LLM agents defer blindly to GNN tools — stronger backbones defer more

Academic

GitOfThoughts: Version-Controlled Reasoning and Agent Memory You Can Replay, Diff, and Merge

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-12

GitOfThoughts: version-controlled reasoning and agent memory

Academic

tap: A File-Based Protocol for Heterogeneous LLM Agent Collaboration

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-12

tap: a file-based protocol for heterogeneous LLM agent collaboration

Academic

Running the Gauntlet: Re-evaluating the Capabilities of Agents Beyond Familiar Environments

arXiv cs.LG (Machine Learning) ・ 2026-06-12

Re-evaluating agent capabilities beyond familiar environments

Academic

Retrospective Progress-Aware Self-Refinement for LLM Agent Training

arXiv cs.CL (Computation and Language) ・ 2026-06-12

Progress-aware self-refinement for training LLM agents

Academic

CacheRL:Multi-Turn Tool-Calling Agents via Cached Rollouts and Hybrid Reward

arXiv cs.CL (Computation and Language) ・ 2026-06-12

CacheRL trains tool-calling agents via cached rollouts and hybrid reward

Academic

Same-Origin Policy for Agentic Browsers

arXiv cs.CL (Computation and Language) ・ 2026-06-12

A same-origin policy for agentic browsers

Academic

Dialogue SWE-Bench: A Benchmark for Dialogue-Driven Coding Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-12

Dialogue SWE-Bench: a benchmark for dialogue-driven coding agents

Academic

EvoArena: Tracking Memory Evolution for Robust LLM Agents in Dynamic Environments

arXiv cs.CL (Computation and Language) ・ 2026-06-11

EvoArena tests LLM agents in evolving environments with patch-based memory

Academic

SpatialClaw: Rethinking Action Interface for Agentic Spatial Reasoning

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

SpatialClaw adopts code as the action interface for agentic spatial reasoning

Academic

Agents-K1: Towards Agent-native Knowledge Orchestration

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

Agents-K1 turns raw papers into agent-native scientific knowledge graphs

Academic

HyperTool: Beyond Step-Wise Tool Calls for Tool-Augmented Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-11

HyperTool folds tool workflows into code, lifting MCP-Universe to 35.29%

Academic

EurekAgent: Agent Environment Engineering is All You Need For Autonomous Scientific Discovery

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

EurekAgent argues environment engineering drives autonomous scientific discovery

Academic

Recursive Agent Harnesses

arXiv cs.CL (Computation and Language) ・ 2026-06-11

Recursive Agent Harnesses lift long-context coding accuracy to 81.36%

Academic

AgentBeats: Agentifying Agent Assessment for Openness, Standardization, and Reproducibility

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

AgentBeats proposes agentified, protocol-standardized agent assessment (AAA)

Academic

EpiBench: Verifiable Evaluation of AI Agents on Epigenomics Analysis

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

EpiBench: no AI agent passes a majority of epigenomics analysis tasks

Academic

Reward Modeling for Multi-Agent Orchestration

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

OrchRM: self-supervised reward modeling for multi-agent orchestration

Academic

Multiagent Protocols with Aggregated Confidence Signals

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

Protocols aggregate multi-agent confidence into a single discriminative score

Academic

Adaptive Turn-Taking for Real-time Multi-Party Voice Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

ModeratorLM conditions turn-taking on assigned roles for multi-party voice

Academic

Understanding the Rejection of Fixes Generated by Agentic Pull Requests -- Insights from the AIDev Dataset

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

46% of agentic PR fixes are rejected; study identifies 14 reasons

Academic

Reinforcement Learning for Neural Model Editing

arXiv cs.LG (Machine Learning) ・ 2026-06-11

RL agents learn to edit neural models for unlearning and debiasing

Academic

Toward Instructions-as-Code: Understanding the Impact of Instruction Files on Agentic Pull Requests

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

Instruction files don't necessarily improve agentic pull request outcomes

Academic

Why Sampling Is Not Choosing: Intentionality, Agency, and Moral Responsibility in Large Language Models

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

Paper argues LLMs lack moral agency: 'sampling is not choosing'

Academic

Neuro-Symbolic Agents for Regulated Process Automation: Challenges and Research Agenda

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

Neuro-symbolic agents propose 'compliance-by-construction' for regulated work

Academic

Who Pays the Price? Stakeholder-Centric Prompt Injection Benchmarking for Real-world Web Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

Stakeholder-centric benchmark attributes prompt-injection harm in web agents

Academic

An LLM System for Autonomous Variational Quantum Circuit Design

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

LLM agent framework autonomously designs variational quantum circuits

Academic

SkillCAT: Contrastive Assessment and Topology-Aware Skill Self-Evolution for LLM Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-11

SkillCAT verifies and routes self-evolved skills for LLM agents

Academic

RogueAI: A Reverse Turing Test for Detecting Licensed AI Deception in Dialogue

arXiv cs.CL (Computation and Language) ・ 2026-06-11

RogueAI: a reverse Turing test for spotting licensed AI deception

Academic

ComAct: Reframing Professional Software Manipulation via COM-as-Action Paradigm

arXiv cs.CL (Computation and Language) ・ 2026-06-11

ComAct drives industrial CAD software via COM-as-Action paradigm

Academic

MemRefine: LLM-Guided Compression for Long-Term Agent Memory

arXiv cs.CL (Computation and Language) ・ 2026-06-11

MemRefine compresses agent memory to budget with LLM-judged merging

Academic

Getting Better at Working With You: Compiling User Corrections into Runtime Enforcement for Coding Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-11

TRACE compiles user corrections into runtime checks for coding agents

Academic

Context-Driven Incremental Compression for Multi-Turn Dialogue Generation

arXiv cs.CL (Computation and Language) ・ 2026-06-10

C-DIC compresses multi-turn dialogue context incrementally for stability

Academic

DIRECT: When and Where Should You Allocate Test-Time Compute in Embodied Planners?

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

DIRECT allocates test-time compute per prompt for embodied planners

Academic

ATLAS: Active Theory Learning for Automated Science

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

ATLAS uses active learning to discover interpretable behavioral models

Academic

APPO: Agentic Procedural Policy Optimization

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

APPO branches and assigns credit at fine-grained decision points

Academic

Claw-SWE-Bench: A Benchmark for Evaluating OpenClaw-style Agent Harnesses on Coding Tasks

arXiv cs.CL (Computation and Language) ・ 2026-06-10

Claw-SWE-Bench fairly benchmarks OpenClaw-style coding agent harnesses

Academic

Fourier Features Let Agents Learn High Precision Policies with Imitation Learning

arXiv cs.LG (Machine Learning) ・ 2026-06-10

Fourier features give point-cloud policies high-precision control

Academic

PROJECTMEM: A Local-First, Event-Sourced Memory and Judgment Layer for AI Coding Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

projectmem adds a local-first memory layer for AI coding agents

Academic

A Five-Plane Reference Architecture for Runtime Governance of Production AI Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

A five-plane reference architecture for runtime governance of AI agents

Academic

CCKS: Consensus-based Communication and Knowledge Sharing

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

CCKS improves cooperative MARL via consensus-based knowledge sharing

Academic

The Impossibility of Eliciting Latent Knowledge

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

Paper formalizes the impossibility of eliciting latent knowledge

Academic

Intelligent Automation for Embodied Benchmark Construction: Pipelines, Embodiments, Simulators, and Trends

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

Survey of automated embodied benchmark construction in five stages

Academic

Implicit Neural Representations of Individual Behavior

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

Behavioral INR learns policy representations from unlabeled multi-policy data

Academic

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

Survey of agentic environment engineering across its lifecycle

Academic

Towards Responsibly Non-Compliant Machines

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

Position paper sketches responsibly non-compliant intelligent machines

Academic

FORT-Searcher: Synthesizing Shortcut-Resistant Search Tasks for Training Deep Search Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-10

FORT synthesizes shortcut-resistant tasks to train deep search agents

Academic

Toward Generalist Autonomous Research via Hypothesis-Tree Refinement

arXiv cs.CL (Computation and Language) ・ 2026-06-10

Arbor runs autonomous research via hypothesis-tree refinement

Academic

When Does Language Matter? Multilingual Instructions Reveal Step-wise Language Sensitivity in Vision-Language-Action Models

arXiv cs.CL (Computation and Language) ・ 2026-06-10

Language robustness in VLA models is a step-wise control problem

Academic

Notes2Skills: From Lab Notebooks to Certainty-Aware Scientific Agent Skills

arXiv cs.CL (Computation and Language) ・ 2026-06-10

Notes2Skills turns lab notebooks into certainty-aware agent skills

Academic

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning

arXiv cs.CL (Computation and Language) ・ 2026-06-10

WorldReasoner evaluates whether agents forecast events with valid reasoning

← Story Archive