NVIDIA × New Model Releases

NVIDIA outlines agent governance in AI

NVIDIA outlines agent governance in AI

✎ Story body

NVIDIA published guidance on governing autonomous agents in enterprise AI factories. The theme was operational, not the next model release. Microsoft Research parallel work — Memora on memory representation, SkillOpt on treating agent skills as trainable parameters — landed alongside, and Hugging Face's ScarfBench measured real Java migration workloads. Focus is shifting from raw agent capability to permissions, memory, and skill governance across many agents. Implementation is starting to outpace the announcement cycle.

▲ Official & Press
Official

How to Govern Autonomous Agents in Enterprise AI Factories

NVIDIA Developer Blog ・ 2026-06-29 ・ 📌

NVIDIA outlines governing autonomous agents in enterprise AI factories

Official

ScarfBench: Benchmarking AI Agents for Enterprise Java Framework Migration

Hugging Face Blog ・ 2026-06-30

IBM's ScarfBench benchmarks AI agents on enterprise Java framework migration

Community

Have your agent record video demos of its work with shot-scraper video

Simon Willison's Weblog ・ 2026-06-30

Record agent work as video demos with shot-scraper video

Official

SkillOpt: Agent skills as trainable parameters

Microsoft Research Blog ・ 2026-06-30

Microsoft's SkillOpt turns agent skill editing into a training process

Press

「従来のモデルルーティングはクソ」 コスト35%減で最先端モデルの性能を維持する「Devin Fusion」発表

ITmedia AI+ ・ 2026-06-30

'Devin Fusion' claims a 35% cost cut while keeping frontier performance

Press

【徹底入門】AIエージェントで注目の「AX」とは何か 人間だけじゃなく“AI視点の使いやすさ”も重要に?

ITmedia AI+ ・ 2026-06-30

ITmedia primer: what is 'AX' in the age of AI agents?

Press

社長もAIが代わる時代に 社員の相談にいつでも答えるエージェント「AI社長」が登場

ITmedia AI+ ・ 2026-06-29

'AI President' agent answers employees' questions anytime

Press

自社の業務に合わせたAIエージェントを「10分で作成」 freeeが「AI戦略」を強化

ITmedia AI+ ・ 2026-06-29

freee lets users build custom AI agents in 10 minutes

Press

メール、Teams、Slack――バラバラな連絡ツールの「見落とし」 ChatGPT Agentsで解決する方法

ITmedia AI+ ・ 2026-06-29

Using ChatGPT Agents to catch missed messages across tools

Official

Memora: A Harmonic Memory Representation Balancing Abstraction and Specificity

Microsoft Research Blog ・ 2026-06-29

Microsoft's Memora is a scalable memory system for AI agents

Press

「AI活用が単発質問の企業は大敗」 楽天にコストと遅延の30%低下も達成させた、AIエージェント運用の勝ち筋

ITmedia AI+ ・ 2026-06-29

Anthropic publishes enterprise AI-agent playbook; Rakuten cuts cost 30%

Press

製造現場のトラブル解消を「AI工場長」が支援? 「エージェント型工場」とは

ITmedia AI+ ・ 2026-06-28

Accenture and Avanade build an "agentic factory" system with Microsoft

Community

Quoting Jon Udell

Simon Willison's Weblog ・ 2026-06-28

Jon Udell: flip "human in the loop"—invite agents into our loop instead

Academic (arxiv etc.) 48 ▾
Academic

QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-30

Academic

Generative Skill Composition for LLM Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-30

Academic

Scalable Behaviour Cloning on Browser Using via Skill Distillation

arXiv cs.CL (Computation and Language) ・ 2026-06-30

Academic

DigitalCoach: Communication and Grounding Gaps in Human and Agentic Computer Use Coaching

arXiv cs.CL (Computation and Language) ・ 2026-06-30

Academic

MECoBench: A Systematic Study of Multimodal Agent Collaboration in Embodied Environments

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-30

Academic

MVP-Nav: Multi-layer Value Map Planner Navigator

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-30

Academic

Theory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and Action

arXiv cs.CL (Computation and Language) ・ 2026-06-30

Academic

Better Understanding, Understanding Better

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-30

Academic

Bridging Local Observation and Global Simulation in Closed-Loop Traffic Modeling

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-30

Academic

An Agentic AI Framework to Accelerate Scientific Discovery in Plant Phenotyping

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-30

Academic

ShopX: A Foundation Model for Intent-to-Item Fulfillment in Agentic Shopping

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-30

Academic

ECHO: Prune to act, trace to learn with selective turn memory in agentic RL

arXiv cs.LG (Machine Learning) ・ 2026-06-30

Academic

Think in English, Answer in Korean: Efficient Adaptation of Multilingual Tool-Using Agents

arXiv cs.LG (Machine Learning) ・ 2026-06-30

Academic

AutoTrainess: Teaching Language Models to Improve Language Models Autonomously

arXiv cs.CL (Computation and Language) ・ 2026-06-30

Academic

FinPersona-Bench: A Benchmark for Longitudinal Psychometric Stability of Autonomous Financial Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-30

Academic

Calibrating the Evaluator: Does Probability Calibration Mitigate Preference Coupling in LLM Agent Feedback Loops?

arXiv cs.CL (Computation and Language) ・ 2026-06-30

Academic

When the Database Fails: Prompting LLM Dialogue Agents for Safe Recovery in Task-Oriented Dialogue

arXiv cs.CL (Computation and Language) ・ 2026-06-30

Academic

The Decomposition Is the Fingerprint: Per-Component Identity for Agent Skills

arXiv cs.CL (Computation and Language) ・ 2026-06-30

Academic

Learning from Failure: Inference-Time Self-Improvement for Computer-Use Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-30

Academic

Can LLMs Imagine Moral Alternatives Beyond Binary Dilemmas?

arXiv cs.CL (Computation and Language) ・ 2026-06-30

Academic

Self-Evolving World Models for LLM Agent Planning

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

Academic

Scaling the Horizon, Not the Parameters: Reaching Trillion-Parameter Performance with a 35B Agent

arXiv cs.CL (Computation and Language) ・ 2026-06-29

Academic

SWE-INTERACT: Reimagining SWE Benchmarks as User-Driven Long-Horizon Coding Sessions

arXiv cs.LG (Machine Learning) ・ 2026-06-29

Academic

Attractor States Emerge in Multi-Turn LLM Conversations

arXiv cs.CL (Computation and Language) ・ 2026-06-29

Academic

Forensic Trajectory Signatures for Agent Memory Poisoning Detection

arXiv cs.LG (Machine Learning) ・ 2026-06-29

Academic

TraceLab: Characterizing Coding Agent Workloads for LLM Serving

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

Academic

Linguistic Firewall: Geometry as Defense in Multi-Agent Systems Routing

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

Academic

TRACE: Temporal Relationship-Aware Conversational Entrainment Detection in Dyadic Speech

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

Academic

Entity Binding Failures in Tool-Augmented Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

Academic

Field Order Should Not Matter: Permutation-Invariant Embedding Model Fine-Tuning for Structured Metadata Retrieval

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

Academic

Collective cooperation without individual fidelity in LLM agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

Academic

Whose Side Is Your Agent On? Multi-Party Principal Loyalty in LLM Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

Academic

BayesEvolve: Explicit Belief States for Autonomous Scientific Discovery

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

Academic

Always-OnAgents:A Survey of Persistent Memory, State, and Governance in LLMAgents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

Academic

ManimAgent: Self-Evolving Multimodal Agents for Visual Education

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

Academic

Rehearsed Multi-Agent Live Product Demonstrations with Real-Time Voice Question Answering

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

Academic

Towards Continual Motion-Language Agents: LoRA Variants for Incremental Motion Understanding and Generation

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

Academic

Grounding LLM Reasoning under Incomplete Graph Evidence

arXiv cs.CL (Computation and Language) ・ 2026-06-29

Academic

Clarus: Coordinating Autonomous Research Agents toward Web-Scale Scientific Collaboration

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

Academic

DAIN: Dynamic Agent-Based Interaction Network for Efficient and Collaborative Multimodal Reasoning

arXiv cs.CL (Computation and Language) ・ 2026-06-29

Academic

Dynamo: Dynamic Skill-Tool Evolution for Vision-Language Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

Academic

MirrorCode: AI can rebuild entire programs from behavior alone

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

Academic

Automating the Design of Embodied AgentArchitectures

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

Academic

ACPO: Agent-Chained Policy Optimization for Multi-Agent Reinforcement Learning

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-29

Academic

LLM Agents Are Latent Context Managers: Eliciting Self-Managed Context via a Proprioceptive Dashboard

arXiv cs.CL (Computation and Language) ・ 2026-06-29

Academic

Neural Procedural Memory: Empowering LLM Agents with Implicit Activation Steering

arXiv cs.CL (Computation and Language) ・ 2026-06-29

Academic

Mandol: An Agglomerative Agent Memory System for Long-Term Conversations

arXiv cs.CL (Computation and Language) ・ 2026-06-29

Academic

A Diagnostic Framework and Multi-Evaluator Audit of Evaluator-Driven Preference Dynamics in Self-Adapting LLM Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-29

← Story Archive