AI エージェント × 新モデル・リリース

DXC、規制業界の業務にClaudeを統合へ

DXC、規制業界の業務にClaudeを統合へ

✎ ストーリー本文

DXC が規制の厳しい業界の業務に Anthropic の Claude を統合していく方針を示した。発生元は Anthropic・OpenAI・NVIDIA など公式4件に itmedia の報道が付く構成で、複数ベンダーの動きを背景に事業導入が語られる――研究ではなく「現場が動いた」側の動きだ。金融や医療など規制産業で、コンプライアンスを保ちながら生成 AI を実務に組み込むという応用が主眼で、本流はモデルの新規性より、大手 IT サービス企業を介した企業導入の広がりにある。ただし現時点は統合の方針表明が中心で、実際の展開規模や規制対応の実効性、成果は、これからの確認点として残る。

▲ 公式・報道
公式

DXC will integrate Claude into the systems banks, airlines, and other regulated industries rely on

Anthropic News ・ 2026-06-11 ・ 📌

Anthropic、DXCと複数年提携、銀行・航空など規制業界の基幹システムにClaude統合

公式

NVIDIA Achieves Leading Agentic Coding Performance on First Agentic AI Benchmark

NVIDIA Developer Blog ・ 2026-06-12

NVIDIA、初のエージェント型AIベンチマークでコーディング性能首位を達成

公式

New OpenAI Academy courses for the next era of work

OpenAI Blog ・ 2026-06-12

OpenAI、仕事での AI 活用を学ぶ Academy 新コース 3 種を公開

コミュニティ

AI agent bankrupted their operator while trying to scan DN42

Hacker News (Front Page) ・ 2026-06-12

AIエージェント、DN42スキャン試行で運用者に破産級のコストを発生との報告

コミュニティ

Claude Fable is relentlessly proactive

Simon Willison's Weblog ・ 2026-06-11

Simon Willison氏、Claude Fable 5を「徹底的にproactive」と評価

公式

OpenAI to acquire Ona

OpenAI Blog ・ 2026-06-11

報道

AIエージェントもフィッシング詐欺に引っかかる? 米セキュリティ企業がOpenClawで検証 結果は……

ITmedia AI+ ・ 2026-06-10

AIエージェントもフィッシングに引っかかる VaronisがOpenClawで検証

学術(arxiv ほか) 58本 ▾
学術

Learning Coordinated Preference for Multi-Objective Multi-Agent Reinforcement Learning

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-12

多目的・多エージェント強化学習で協調選好を学習する手法を提案

学術

AgentSpec: Understanding Embodied Agent Scaffolds Through Controlled Composition

arXiv cs.CL (Computation and Language) ・ 2026-06-12

AgentSpec、エージェント足場を統制的に分解し検証

学術

Towards Direct Latent-Space Synthesis for Parallel Branches in LLM-Agent Workflows

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-12

LLMエージェントの並列分岐を潜在空間で直接合成する手法を検討

学術

LoSoNA: A Benchmark for Local Social Norm Adaptation in Group Conversations

arXiv cs.CL (Computation and Language) ・ 2026-06-12

LoSoNA、集団会話の局所規範への適応を評価する基準

学術

Regulating the Machine Contributor: Governance and Policy Alignment in Open Source

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-12

オープンソースへのAIエージェント貢献を統治する枠組みを論じる

学術

SIMMER: Benchmarking Latent Failures in LLM Executable Planning with a World Model

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-12

世界モデルでLLM実行計画の潜在的失敗を測る「SIMMER」

学術

From Shield to Target: Denial-of-Service Attacks on LLM-Based Agent Guardrails

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-12

LLMエージェントのガードレールを狙うサービス妨害攻撃を提示

学術

From Chatbot to Digital Colleague: The Paradigm Shift Toward Persistent Autonomous AI

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-12

LLMが「チャットボット」から永続自律AIへ移行する転換を概念化

学術

When the Tool Decides: LLM Agents Defer Blindly to Graph Neural Network Tools, and Stronger Backbones Defer More

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-12

LLMエージェントはGNNツールに盲目的に委ね、強いほど委ねると指摘

学術

GitOfThoughts: Version-Controlled Reasoning and Agent Memory You Can Replay, Diff, and Merge

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-12

再生・差分・統合できる版管理つき推論記憶「GitOfThoughts」

学術

tap: A File-Based Protocol for Heterogeneous LLM Agent Collaboration

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-12

異種LLMエージェント協調のファイルベースプロトコル「tap」

学術

Running the Gauntlet: Re-evaluating the Capabilities of Agents Beyond Familiar Environments

arXiv cs.LG (Machine Learning) ・ 2026-06-12

馴染みのない環境でエージェント能力を再評価

学術

Retrospective Progress-Aware Self-Refinement for LLM Agent Training

arXiv cs.CL (Computation and Language) ・ 2026-06-12

進捗を自覚する自己改善でLLMエージェント訓練を強化

学術

CacheRL:Multi-Turn Tool-Calling Agents via Cached Rollouts and Hybrid Reward

arXiv cs.CL (Computation and Language) ・ 2026-06-12

CacheRL、キャッシュ活用で小型ツール呼出エージェントを訓練

学術

Same-Origin Policy for Agentic Browsers

arXiv cs.CL (Computation and Language) ・ 2026-06-12

エージェント型ブラウザに同一生成元ポリシーを提案

学術

Dialogue SWE-Bench: A Benchmark for Dialogue-Driven Coding Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-12

Dialogue SWE-Bench、対話駆動のコーディングを評価

学術

EvoArena: Tracking Memory Evolution for Robust LLM Agents in Dynamic Environments

arXiv cs.CL (Computation and Language) ・ 2026-06-11

動的環境でのLLMエージェント評価基盤EvoArena、パッチ型記憶EvoMemも提案

学術

SpatialClaw: Rethinking Action Interface for Agentic Spatial Reasoning

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

SpatialClaw、空間推論agentの行動IFをコードに刷新

学術

Agents-K1: Towards Agent-native Knowledge Orchestration

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

Agents-K1、論文をagent向け知識グラフ化する基盤

学術

HyperTool: Beyond Step-Wise Tool Calls for Tool-Augmented Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-11

HyperTool、ツール呼び出しをコードに畳み込みMCP精度を15.7%から35.3%へ

学術

EurekAgent: Agent Environment Engineering is All You Need For Autonomous Scientific Discovery

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

EurekAgent、自律的科学発見の鍵は環境設計と主張

学術

Recursive Agent Harnesses

arXiv cs.CL (Computation and Language) ・ 2026-06-11

再帰エージェントハーネスRAH、長文脈推論でCodex基準71.75%を81.36%に

学術

AgentBeats: Agentifying Agent Assessment for Openness, Standardization, and Reproducibility

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

AgentBeats、judge agentで評価を標準化するAAAを提唱

学術

EpiBench: Verifiable Evaluation of AI Agents on Epigenomics Analysis

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

EpiBench、AI agentのエピゲノム解析を検証、最高45%

学術

Reward Modeling for Multi-Agent Orchestration

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

OrchRM、多agent統括の品質を自己教師で評価し効率10倍

学術

Multiagent Protocols with Aggregated Confidence Signals

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

多agent系の出力に単一の集約信頼度を与える3手法を提案

学術

Adaptive Turn-Taking for Real-time Multi-Party Voice Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

ModeratorLM、役割条件で多者間音声agentの発話交代を改善

学術

Understanding the Rejection of Fixes Generated by Agentic Pull Requests -- Insights from the AIDev Dataset

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

AI agentのPR修正46%が却下、14の却下理由を分類

学術

Reinforcement Learning for Neural Model Editing

arXiv cs.LG (Machine Learning) ・ 2026-06-11

モデル編集を強化学習で自動化、忘却精度ほぼ0%・保持9割超を達成

学術

Toward Instructions-as-Code: Understanding the Impact of Instruction Files on Agentic Pull Requests

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

AI agent向け指示ファイルは必ずしもPR成果を改善せずと判明

学術

Why Sampling Is Not Choosing: Intentionality, Agency, and Moral Responsibility in Large Language Models

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

LLMに道徳的主体性はないと論じる『標本抽出は選択でない』

学術

Neuro-Symbolic Agents for Regulated Process Automation: Challenges and Research Agenda

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

規制業務の自動化に『構築による準拠』を提唱する神経記号agent

学術

Who Pays the Price? Stakeholder-Centric Prompt Injection Benchmarking for Real-world Web Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

web agentのプロンプト注入を被害者視点で評価するベンチ

学術

An LLM System for Autonomous Variational Quantum Circuit Design

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

LLMで量子回路を自律設計する閉ループagent枠組み

学術

SkillCAT: Contrastive Assessment and Topology-Aware Skill Self-Evolution for LLM Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-11

SkillCAT、LLMエージェントのスキル自己進化を検証付き3段階に分離

学術

RogueAI: A Reverse Turing Test for Detecting Licensed AI Deception in Dialogue

arXiv cs.CL (Computation and Language) ・ 2026-06-11

逆チューリングテストRogueAI、嘘を許可されたAIを対話で見破るゲーム

学術

ComAct: Reframing Professional Software Manipulation via COM-as-Action Paradigm

arXiv cs.CL (Computation and Language) ・ 2026-06-11

ComAct、COM経由の決定論的操作で産業CADソフトをエージェント制御

学術

MemRefine: LLM-Guided Compression for Long-Term Agent Memory

arXiv cs.CL (Computation and Language) ・ 2026-06-11

MemRefine、LLM判断で予算内にエージェント長期記憶を圧縮

学術

Getting Better at Working With You: Compiling User Corrections into Runtime Enforcement for Coding Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-11

TRACE、ユーザー修正をルール化しコーディングエージェントに強制適用

学術

Context-Driven Incremental Compression for Multi-Turn Dialogue Generation

arXiv cs.CL (Computation and Language) ・ 2026-06-10

C-DIC、多ターン対話の文脈を逐次圧縮し長対話を安定化

学術

DIRECT: When and Where Should You Allocate Test-Time Compute in Embodied Planners?

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

DIRECT、身体エージェントのテスト時計算をプロンプト別に配分

学術

ATLAS: Active Theory Learning for Automated Science

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

ATLAS、能動学習で解釈可能な行動モデルを自動的に発見

学術

APPO: Agentic Procedural Policy Optimization

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

APPO、細粒度の決定点で分岐と信用割当を行うエージェントRL

学術

Claw-SWE-Bench: A Benchmark for Evaluating OpenClaw-style Agent Harnesses on Coding Tasks

arXiv cs.CL (Computation and Language) ・ 2026-06-10

Claw-SWE-Bench、OpenClaw型エージェントのコード能力を多言語評価

学術

Fourier Features Let Agents Learn High Precision Policies with Imitation Learning

arXiv cs.LG (Machine Learning) ・ 2026-06-10

Fourier 特徴で点群方策が高精度ロボット操作を獲得

学術

PROJECTMEM: A Local-First, Event-Sourced Memory and Judgment Layer for AI Coding Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

projectmem、コーディングエージェントに局所優先の記憶層を追加

学術

A Five-Plane Reference Architecture for Runtime Governance of Production AI Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

本番AIエージェントの実行時統治に5プレーン参照アーキテクチャ

学術

CCKS: Consensus-based Communication and Knowledge Sharing

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

CCKS、合意ベースの通信と知識共有で協調的MARLを改善

学術

The Impossibility of Eliciting Latent Knowledge

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

潜在知識の引き出し(ELK)は不可能と因果影響図で形式化

学術

Intelligent Automation for Embodied Benchmark Construction: Pipelines, Embodiments, Simulators, and Trends

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

身体ベンチ構築の自動化を5段階パイプラインで調査

学術

Implicit Neural Representations of Individual Behavior

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

Behavioral INR、無ラベルの多方策行動から方策表現を学習

学術

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

エージェント環境工学を環境のライフサイクル観点で調査

学術

Towards Responsibly Non-Compliant Machines

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

責任ある「不服従」が可能なAIエージェントの工学を提起

学術

FORT-Searcher: Synthesizing Shortcut-Resistant Search Tasks for Training Deep Search Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-10

FORT、近道耐性のある探索課題を合成し深層探索エージェントを訓練

学術

Toward Generalist Autonomous Research via Hypothesis-Tree Refinement

arXiv cs.CL (Computation and Language) ・ 2026-06-10

Arbor、仮説ツリー改良で長期間の自律研究ループを運用

学術

When Does Language Matter? Multilingual Instructions Reveal Step-wise Language Sensitivity in Vision-Language-Action Models

arXiv cs.CL (Computation and Language) ・ 2026-06-10

VLAモデルの言語頑健性は段階ごとの制御問題と判明

学術

Notes2Skills: From Lab Notebooks to Certainty-Aware Scientific Agent Skills

arXiv cs.CL (Computation and Language) ・ 2026-06-10

Notes2Skills、実験ノートを確信度付きの科学エージェントスキルに変換

学術

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning

arXiv cs.CL (Computation and Language) ・ 2026-06-10

WorldReasoner、エージェントの事象予測の推論妥当性を評価

← ストーリー アーカイブ