NVIDIA × 新モデル・リリース

NVIDIA、AI工場のエージェント統制手法を提示

NVIDIA、AI工場のエージェント統制手法を提示

✎ ストーリー本文

NVIDIA が「Enterprise AI Factories でエージェントをどう統制するか」の設計指針を提示。話題は公式ドキュメントを核に、Microsoft Research の Memora(抽象と特異性を両立するメモリ表現)や SkillOpt(skill を訓練可能パラメータ化)など研究系が厚く並走した。エージェント個体の性能より、複数エージェントの権限・記憶・スキル管理という運用側の設計が主戦場になった週。HF の ScarfBench による Java 移行ベンチも「実務で agent を回すと何が問題になるか」を測る動き。実装が発表を追い越し始めた兆候。

▲ 公式・報道
公式

How to Govern Autonomous Agents in Enterprise AI Factories

NVIDIA Developer Blog ・ 2026-06-29 ・ 📌

NVIDIA、企業の自律AIエージェント統治の手引きを公開

公式

ScarfBench: Benchmarking AI Agents for Enterprise Java Framework Migration

Hugging Face Blog ・ 2026-06-30

IBM、Java移行AIエージェントの評価基盤「ScarfBench」を公開

コミュニティ

Have your agent record video demos of its work with shot-scraper video

Simon Willison's Weblog ・ 2026-06-30

Simon Willison氏、shot-scraperでエージェントに作業動画を記録させる手法

公式

SkillOpt: Agent skills as trainable parameters

Microsoft Research Blog ・ 2026-06-30

Microsoft、エージェントのスキル編集を学習化する「SkillOpt」を提案

報道

「従来のモデルルーティングはクソ」 コスト35%減で最先端モデルの性能を維持する「Devin Fusion」発表

ITmedia AI+ ・ 2026-06-30

コスト35%減で最先端性能を保つ「Devin Fusion」が発表

報道

【徹底入門】AIエージェントで注目の「AX」とは何か 人間だけじゃなく“AI視点の使いやすさ”も重要に?

ITmedia AI+ ・ 2026-06-30

ITmedia解説、AIエージェント時代のキーワード「AX」を徹底入門

報道

社長もAIが代わる時代に 社員の相談にいつでも答えるエージェント「AI社長」が登場

ITmedia AI+ ・ 2026-06-29

社員の相談に答えるAIエージェント「AI社長」が登場

報道

自社の業務に合わせたAIエージェントを「10分で作成」 freeeが「AI戦略」を強化

ITmedia AI+ ・ 2026-06-29

freee、自社業務向けAIエージェントを10分で作れる機能で「AI戦略」強化

報道

メール、Teams、Slack――バラバラな連絡ツールの「見落とし」 ChatGPT Agentsで解決する方法

ITmedia AI+ ・ 2026-06-29

ChatGPT Agentsで、メール・Teams・Slackの連絡見落としを防ぐ

公式

Memora: A Harmonic Memory Representation Balancing Abstraction and Specificity

Microsoft Research Blog ・ 2026-06-29

Microsoft、エージェント向け記憶システム「Memora」を提案

報道

「AI活用が単発質問の企業は大敗」 楽天にコストと遅延の30%低下も達成させた、AIエージェント運用の勝ち筋

ITmedia AI+ ・ 2026-06-29

Anthropic、企業向けAIエージェント活用ガイドを公開、楽天は30%改善

報道

製造現場のトラブル解消を「AI工場長」が支援? 「エージェント型工場」とは

ITmedia AI+ ・ 2026-06-28

Accenture・Avanade、Microsoftと「エージェント型工場」を開発

コミュニティ

Quoting Jon Udell

Simon Willison's Weblog ・ 2026-06-28

ジョン・ユーデル、「人間がループ内」を反転しAIを輪に招く発想

学術(arxiv ほか) 48本 ▾
学術

QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-30

学術

Generative Skill Composition for LLM Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-30

学術

Scalable Behaviour Cloning on Browser Using via Skill Distillation

arXiv cs.CL (Computation and Language) ・ 2026-06-30

学術

DigitalCoach: Communication and Grounding Gaps in Human and Agentic Computer Use Coaching

arXiv cs.CL (Computation and Language) ・ 2026-06-30

学術

MECoBench: A Systematic Study of Multimodal Agent Collaboration in Embodied Environments

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-30

学術

MVP-Nav: Multi-layer Value Map Planner Navigator

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-30

学術

Theory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and Action

arXiv cs.CL (Computation and Language) ・ 2026-06-30

学術

Better Understanding, Understanding Better

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-30

学術

Bridging Local Observation and Global Simulation in Closed-Loop Traffic Modeling

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-30

学術

An Agentic AI Framework to Accelerate Scientific Discovery in Plant Phenotyping

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-30

学術

ShopX: A Foundation Model for Intent-to-Item Fulfillment in Agentic Shopping

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-30

学術

ECHO: Prune to act, trace to learn with selective turn memory in agentic RL

arXiv cs.LG (Machine Learning) ・ 2026-06-30

学術

Think in English, Answer in Korean: Efficient Adaptation of Multilingual Tool-Using Agents

arXiv cs.LG (Machine Learning) ・ 2026-06-30

学術

AutoTrainess: Teaching Language Models to Improve Language Models Autonomously

arXiv cs.CL (Computation and Language) ・ 2026-06-30

学術

FinPersona-Bench: A Benchmark for Longitudinal Psychometric Stability of Autonomous Financial Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-30

学術

Calibrating the Evaluator: Does Probability Calibration Mitigate Preference Coupling in LLM Agent Feedback Loops?

arXiv cs.CL (Computation and Language) ・ 2026-06-30

学術

When the Database Fails: Prompting LLM Dialogue Agents for Safe Recovery in Task-Oriented Dialogue

arXiv cs.CL (Computation and Language) ・ 2026-06-30

学術

The Decomposition Is the Fingerprint: Per-Component Identity for Agent Skills

arXiv cs.CL (Computation and Language) ・ 2026-06-30

学術

Learning from Failure: Inference-Time Self-Improvement for Computer-Use Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-30

学術

Can LLMs Imagine Moral Alternatives Beyond Binary Dilemmas?

arXiv cs.CL (Computation and Language) ・ 2026-06-30

学術

Self-Evolving World Models for LLM Agent Planning

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

学術

Scaling the Horizon, Not the Parameters: Reaching Trillion-Parameter Performance with a 35B Agent

arXiv cs.CL (Computation and Language) ・ 2026-06-29

学術

SWE-INTERACT: Reimagining SWE Benchmarks as User-Driven Long-Horizon Coding Sessions

arXiv cs.LG (Machine Learning) ・ 2026-06-29

学術

Attractor States Emerge in Multi-Turn LLM Conversations

arXiv cs.CL (Computation and Language) ・ 2026-06-29

学術

Forensic Trajectory Signatures for Agent Memory Poisoning Detection

arXiv cs.LG (Machine Learning) ・ 2026-06-29

学術

TraceLab: Characterizing Coding Agent Workloads for LLM Serving

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

学術

Linguistic Firewall: Geometry as Defense in Multi-Agent Systems Routing

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

学術

TRACE: Temporal Relationship-Aware Conversational Entrainment Detection in Dyadic Speech

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

学術

Entity Binding Failures in Tool-Augmented Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

学術

Field Order Should Not Matter: Permutation-Invariant Embedding Model Fine-Tuning for Structured Metadata Retrieval

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

学術

Collective cooperation without individual fidelity in LLM agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

学術

Whose Side Is Your Agent On? Multi-Party Principal Loyalty in LLM Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

学術

BayesEvolve: Explicit Belief States for Autonomous Scientific Discovery

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

学術

Always-OnAgents:A Survey of Persistent Memory, State, and Governance in LLMAgents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

学術

ManimAgent: Self-Evolving Multimodal Agents for Visual Education

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

学術

Rehearsed Multi-Agent Live Product Demonstrations with Real-Time Voice Question Answering

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

学術

Towards Continual Motion-Language Agents: LoRA Variants for Incremental Motion Understanding and Generation

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

学術

Grounding LLM Reasoning under Incomplete Graph Evidence

arXiv cs.CL (Computation and Language) ・ 2026-06-29

学術

Clarus: Coordinating Autonomous Research Agents toward Web-Scale Scientific Collaboration

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

学術

DAIN: Dynamic Agent-Based Interaction Network for Efficient and Collaborative Multimodal Reasoning

arXiv cs.CL (Computation and Language) ・ 2026-06-29

学術

Dynamo: Dynamic Skill-Tool Evolution for Vision-Language Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

学術

MirrorCode: AI can rebuild entire programs from behavior alone

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

学術

Automating the Design of Embodied AgentArchitectures

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

学術

ACPO: Agent-Chained Policy Optimization for Multi-Agent Reinforcement Learning

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

学術

LLM Agents Are Latent Context Managers: Eliciting Self-Managed Context via a Proprioceptive Dashboard

arXiv cs.CL (Computation and Language) ・ 2026-06-29

学術

Neural Procedural Memory: Empowering LLM Agents with Implicit Activation Steering

arXiv cs.CL (Computation and Language) ・ 2026-06-29

学術

Mandol: An Agglomerative Agent Memory System for Long-Term Conversations

arXiv cs.CL (Computation and Language) ・ 2026-06-29

学術

A Diagnostic Framework and Multi-Evaluator Audit of Evaluator-Driven Preference Dynamics in Self-Adapting LLM Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-29

← ストーリー アーカイブ