NVIDIA × Inference & Efficiency

NVIDIA speeds clinical ASR evaluation

NVIDIA speeds clinical ASR evaluation

✎ Story body

NVIDIA showed that its 'Nemotron Speech' model speeds up clinical speech-recognition (ASR) evaluation. All five sources are NVIDIA official (with some Hugging Face and DeepMind links)—a single-vendor technical presentation, announcement-led but stepping into healthcare as an applied domain. It targets clinical ASR—jargon-heavy, accuracy-critical tasks like transcribing medical records—to speed evaluation and processing; the through-line is a shift from general model performance toward practical accuracy and efficiency in a specific domain. It stands as an application to lighten documentation work in care settings. But this is mainly a technical presentation—clinical accuracy validation, regulatory handling, and adoption breadth are to confirm.

▲ Official & Press
Official

Evaluate Clinical ASR Models Faster with Agent Skills and NVIDIA Nemotron Speech

NVIDIA Developer Blog ・ 2026-06-09 ・ 📌

NVIDIA speeds clinical ASR evaluation via Agent Skills, Nemotron Speech

Press

AIエージェントもフィッシング詐欺に引っかかる? 米セキュリティ企業がOpenClawで検証 結果は……

ITmedia AI+ ・ 2026-06-10

Varonis: AI agents can fall for phishing, tested with OpenClaw

Community

Building a persistent cognitive architecture for LLM agents using Elixir and OTP

Lobste.rs (AI tagged) ・ 2026-06-09

Persistent cognitive architecture for LLM agents via Elixir/OTP

Official

Can Voice Agents Handle Bilingual Customers? Benchmarking Frontier ASR on Code-Switched Speech

Hugging Face Blog ・ 2026-06-09

ServiceNow AI benchmarks frontier ASR on code-switched speech

Official

Delivering Lifecycle Control for AI Infrastructure at Scale with NVIDIA DGX Spark Enterprise Manageability

NVIDIA Developer Blog ・ 2026-06-09

NVIDIA adds enterprise lifecycle control for AI infra via DGX Spark

Official

Model Quantization: Turn FP8 Checkpoints into High-Performance Inference Engines with NVIDIA TensorRT

NVIDIA Developer Blog ・ 2026-06-09

NVIDIA converts FP8 checkpoints into fast inference engines via TensorRT

Official

Accelerating Federated Learning Research with AI Agents and NVIDIA FLARE Auto-FL

NVIDIA Developer Blog ・ 2026-06-09

NVIDIA FLARE Auto-FL and AI agents accelerate FL research

Official

Fluid, natural voice translation with Gemini 3.5 Live Translate

Google DeepMind Blog ・ 2026-06-09

Google debuts Gemini 3.5 Live Translate for real-time speech

Official

Introducing North Mini Code: Cohere’s first model for developers

Cohere Blog ・ 2026-06-09

Cohere open-sources North Mini Code, its first agentic coding model

Press

Ubuntu、サンドボックス化された開発環境をコマンド一発で構築。新機能「Workshop」リリース

Publickey ・ 2026-06-08

Canonical launches Workshop, one-command sandboxed dev environments

Academic (arxiv etc.) 99 ▾
Academic

Operadic consistency: a label-free signal for compositional reasoning failures in LLMs

arXiv cs.CL (Computation and Language) ・ 2026-06-11

Operadic consistency flags LLM compositional reasoning errors label-free

Academic

Valid Inference with Synthetic Data via Task Exchangeability

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

Task exchangeability enables valid inference from synthetic data with guarantees

Academic

Beyond Uniform Tokens: Adaptive Compression for Time Series Language Models

arXiv cs.CL (Computation and Language) ・ 2026-06-11

Adaptive token compression streamlines time series language models

Academic

Beyond the Commitment Boundary: Probing Epiphenomenal Chain-of-Thought in Large Reasoning Models

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

Chain-of-thought reasoning crosses a 'commitment boundary,' study shows

Academic

Simplex-Constrained Sparse Bagging: Transitioning from Uniform Priors to Sparse Posteriors in Ensemble Learning

arXiv cs.LG (Machine Learning) ・ 2026-06-11

SCSB prunes bagging ensembles up to 96% while improving calibration

Academic

Existence Precedes Value: Joint Modeling of Observational Existence and Evolving States in Time Series Forecasting

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

Timeflies jointly models future observation existence and values in forecasting

Academic

A2D2: Fine-Tuning Any-Length Discrete Diffusion for Adaptive Decoding

arXiv cs.LG (Machine Learning) ・ 2026-06-11

A2D2 unifies reward-guided fine-tuning for any-length discrete diffusion

Academic

Is It You or Your Environment? A Bayesian Inference Framework for Genomically-Anchored Personalized Physiological Interpretation

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

Genomic priors solve the cold-start problem in personalized health AI

Academic

NetCause: Counterfactual Learning for Root Cause Analysis in Large-Scale Networks

arXiv cs.LG (Machine Learning) ・ 2026-06-11

NetCause ranks network incident root causes via counterfactuals

Academic

Graphical Causal Reasoning for Root Cause Analysis in Cloud Networks

arXiv cs.LG (Machine Learning) ・ 2026-06-11

Causal graph traversal recalls 85.7% of cloud incident root causes

Academic

Heterogeneous LiDAR Early Fusion and Learned Re-Ranking Strategy for Robust Long-Term Place Recognition in Unstructured Environments

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

Heterogeneous LiDAR fusion and re-ranking boost place recognition in fields

Academic

GF-DiT: Scheduling Parallelism for Diffusion Transformer Serving

arXiv cs.LG (Machine Learning) ・ 2026-06-11

GF-DiT makes GPU parallelism schedulable for diffusion transformers

Academic

Optical Implementation of Equilibrium Propagation Using Spatial Photonic Ising Machines

arXiv cs.LG (Machine Learning) ・ 2026-06-11

Equilibrium propagation realized on spatial photonic Ising machines

Academic

Accelerating Speculative Diffusions via Block Verification

arXiv cs.LG (Machine Learning) ・ 2026-06-11

Block verification speeds up speculative sampling for diffusion models

Academic

PolyFlow: Safe and Efficient Polytope-Constrained Flow Matching with Constraint Embedding and Projection-free Update

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

PolyFlow embeds polytope constraints into flow matching, projection-free

Academic

MiniMax Sparse Attention

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

MiniMax Sparse Attention enables efficient ultra-long-context LLMs

Academic

SmartFont: Dynamic Condition Allocation for Few-Shot Font Generation

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-11

SmartFont allocates global and local conditions for few-shot font generation

Academic

Hölder++: Improving the Quality-Coherence Trade-off in Multimodal VAEs

arXiv cs.LG (Machine Learning) ・ 2026-06-11

Hölder++ improves quality-coherence trade-off in multimodal VAEs

Academic

VideoMDM: Towards 3D Human Motion Generation From 2D Supervision

arXiv cs.LG (Machine Learning) ・ 2026-06-11

VideoMDM learns 3D human motion priors from 2D video supervision

Academic

SkillCAT: Contrastive Assessment and Topology-Aware Skill Self-Evolution for LLM Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-11

SkillCAT verifies and routes self-evolved skills for LLM agents

Academic

SICI: A Semantic-Pragmatic Complexity Index Reveals Regime Shifts in LLM Stance Detection

arXiv cs.CL (Computation and Language) ・ 2026-06-11

SICI complexity index reveals regime shifts in LLM stance detection

Academic

MiniPIC: Flexible Position-Independent Caching in <100LOC

arXiv cs.CL (Computation and Language) ・ 2026-06-11

MiniPIC adds position-independent KV caching to vLLM in <100 LOC

Academic

Reroute, Don't Remove: Recoverable Visual Token Routing for Vision-Language Models

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

Reroute replaces visual-token removal with recoverable routing in VLMs

Academic

Context-Driven Incremental Compression for Multi-Turn Dialogue Generation

arXiv cs.CL (Computation and Language) ・ 2026-06-10

C-DIC compresses multi-turn dialogue context incrementally for stability

Academic

DIRECT: When and Where Should You Allocate Test-Time Compute in Embodied Planners?

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

DIRECT allocates test-time compute per prompt for embodied planners

Academic

Doc-to-Atom: Learning to Compile and Compose Memory Atoms

arXiv cs.CL (Computation and Language) ・ 2026-06-10

Doc2Atom decomposes documents into composable micro-LoRA memory atoms

Academic

System Report for CCL25-Eval Task 5: New Dataset and LoRA-Fine-Tuned Qwen2.5

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

PoetryQwen specializes classical Chinese poetry appreciation via CCPoetry-49K

Academic

TAHOE: Text-to-SQL with Automated Hint Optimization from Experience

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

Tahoe learns hints from experience to optimize production Text-to-SQL

Academic

ATLAS: Active Theory Learning for Automated Science

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

ATLAS uses active learning to discover interpretable behavioral models

Academic

APPO: Agentic Procedural Policy Optimization

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

APPO branches and assigns credit at fine-grained decision points

Academic

Breaking Entropy Bounds: Accelerating RL Training via MTP with Rejection Sampling

arXiv cs.CL (Computation and Language) ・ 2026-06-10

Bebop boosts MTP acceptance via rejection sampling to speed RL training

Academic

Latent World Recovery for Multimodal Learning with Missing Modalities

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

LWR recovers a latent world for multimodal learning with missing modalities

Academic

CHORUS: Decentralized Multi-Embodiment Collaboration with One VLA Policy

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

CHORUS controls multi-robot teams with one decentralized VLA policy

Academic

Claw-SWE-Bench: A Benchmark for Evaluating OpenClaw-style Agent Harnesses on Coding Tasks

arXiv cs.CL (Computation and Language) ・ 2026-06-10

Claw-SWE-Bench fairly benchmarks OpenClaw-style coding agent harnesses

Academic

ALIGNBEAM : Inference-Time Alignment Transfer via Cross-Vocabulary Logit Mixing

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

ALIGNBEAM transfers safety logits across model vocabularies

Academic

Fourier Features Let Agents Learn High Precision Policies with Imitation Learning

arXiv cs.LG (Machine Learning) ・ 2026-06-10

Fourier features give point-cloud policies high-precision control

Academic

Measuring Semantic Progress in Multi-turn Dialogue via Information Gain

arXiv cs.CL (Computation and Language) ・ 2026-06-10

An information-gain metric measures semantic progress in multi-turn dialogue

Academic

PROJECTMEM: A Local-First, Event-Sourced Memory and Judgment Layer for AI Coding Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

projectmem adds a local-first memory layer for AI coding agents

Academic

A Five-Plane Reference Architecture for Runtime Governance of Production AI Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

A five-plane reference architecture for runtime governance of AI agents

Academic

Harness In-Context Operator Learning with Chain of Operators

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

CHOP chains operators to generalize a frozen ICON to OOD tasks

Academic

CCKS: Consensus-based Communication and Knowledge Sharing

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

CCKS improves cooperative MARL via consensus-based knowledge sharing

Academic

Holding the FP8 Quality Ceiling at 8-Bit Weights and Activations: INT8 and GGUF Post-Training Quantization of Ideogram 4.0 for Consumer GPUs

arXiv cs.LG (Machine Learning) ・ 2026-06-10

INT8 quantization of Ideogram 4.0 holds FP8 quality on consumer GPUs

Academic

Mathematical perspective on genetic algorithms with optimization guided operators

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

A mathematical model of genetic algorithms with optimization-guided operators

Academic

The Impossibility of Eliciting Latent Knowledge

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

Paper formalizes the impossibility of eliciting latent knowledge

Academic

VIA-SD: Verification via Intra-Model Routing for Speculative Decoding

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

VIA-SD speeds speculative decoding via intra-model verifier routing

Academic

Re-evaluating Confidence Remasking in Masked Diffusion Language Models

arXiv cs.LG (Machine Learning) ・ 2026-06-10

Re-evaluation: WINO remasking adds little in masked diffusion LLMs

Academic

Can News Predict the Market? Limits of Zero-Shot Financial NLP and the Role of Explainable AI

arXiv cs.CL (Computation and Language) ・ 2026-06-10

Zero-shot financial NLP fails to beat baselines at predicting markets

Academic

Intelligent Automation for Embodied Benchmark Construction: Pipelines, Embodiments, Simulators, and Trends

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

Survey of automated embodied benchmark construction in five stages

Academic

Adaptive Multi-Resolution Procedural Knowledge Compression for Large Language Models

arXiv cs.CL (Computation and Language) ・ 2026-06-10

SKIM compresses procedural LLM skills via adaptive multi-resolution tokens

Academic

Implicit Neural Representations of Individual Behavior

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

Behavioral INR learns policy representations from unlabeled multi-policy data

Academic

Which Speech Representation Better Matches Text-Native Reasoning? A Study of Speech-Text Alignment on Frame Rate and Representation

arXiv cs.CL (Computation and Language) ・ 2026-06-10

Study finds best speech-text alignment regime at 4.17Hz frame rate

Academic

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

Survey of agentic environment engineering across its lifecycle

Academic

A Resource for Enthymeme Detection in Controversial Political Discourse

arXiv cs.CL (Computation and Language) ・ 2026-06-10

New dataset enables enthymeme detection in political discourse

Academic

Towards Responsibly Non-Compliant Machines

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

Position paper sketches responsibly non-compliant intelligent machines

Academic

FORT-Searcher: Synthesizing Shortcut-Resistant Search Tasks for Training Deep Search Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-10

FORT synthesizes shortcut-resistant tasks to train deep search agents

Academic

"That's AI Slop, You Bot!" Studying Accusations, Evidence, and Credibility in Online Discourse Towards LLM-Generated Comments

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-10

Analysis of 25M comments tracks the surge of 'AI slop' accusations

Academic

Toward Generalist Autonomous Research via Hypothesis-Tree Refinement

arXiv cs.CL (Computation and Language) ・ 2026-06-10

Arbor runs autonomous research via hypothesis-tree refinement

Academic

When Does Language Matter? Multilingual Instructions Reveal Step-wise Language Sensitivity in Vision-Language-Action Models

arXiv cs.CL (Computation and Language) ・ 2026-06-10

Language robustness in VLA models is a step-wise control problem

Academic

Notes2Skills: From Lab Notebooks to Certainty-Aware Scientific Agent Skills

arXiv cs.CL (Computation and Language) ・ 2026-06-10

Notes2Skills turns lab notebooks into certainty-aware agent skills

Academic

Beyond representational alignment with brain-guided language models for robust reasoning

arXiv cs.CL (Computation and Language) ・ 2026-06-10

Brain-guided language models strengthen robust deductive reasoning

Academic

Fine-tuning Multi-modal LLMs with ART: Art-based Reinforcement Training

arXiv cs.CL (Computation and Language) ・ 2026-06-10

ART fine-tunes frozen MLLMs by optimizing only the visual input

Academic

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning

arXiv cs.CL (Computation and Language) ・ 2026-06-10

WorldReasoner evaluates whether agents forecast events with valid reasoning

Academic

MultiToP: Learning to Patch Visual Tokens to Mitigate Hallucinations in Video Large Multimodal Models

arXiv cs.CL (Computation and Language) ・ 2026-06-10

MultiToP patches visual tokens to cut video-LMM hallucinations

Academic

Fast Speech Foundation Model Distillation Using Interleaved Stacking

arXiv cs.CL (Computation and Language) ・ 2026-06-10

Interleaved stacking accelerates speech foundation model distillation

Academic

EEVEE: Towards Test-time Prompt Learning in the Real World for Self-Improving Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-09

EEVEE: multi-dataset test-time prompt learning for self-improving agents

Academic

Data Journalist Agent: Transforming Data into Verifiable Multimodal Stories

arXiv cs.CL (Computation and Language) ・ 2026-06-09

Data2Story: a multi-agent framework for verifiable, multimodal data stories

Academic

Multi-Faceted Interactivity Alignment in Full-Duplex Speech Models

arXiv cs.CL (Computation and Language) ・ 2026-06-09

RL post-training aligns full-duplex speech models across four interactivity axes

Academic

ReasonAlloc: Hierarchical Decoding-Time KV Cache Budget Allocation for Reasoning Models

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-09

ReasonAlloc: hierarchical KV cache budget allocation for reasoning models

Academic

Itô maps for any-step SDEs

arXiv cs.LG (Machine Learning) ・ 2026-06-09

Ito maps: any-step stochastic flow maps for posterior sampling and control

Academic

ABC-Bench: An Agentic Bio-Capabilities Benchmark for Biosecurity

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-09

ABC-Bench evaluates LLM agents' biosecurity-relevant capabilities

Academic

FADA: Accessible fetal ultrasound interpretation and annotation with a selectively distilled unified vision-language model

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-09

FADA: a unified vision-language model for fetal ultrasound interpretation

Academic

RoboNaldo: Accurate, Stable and Powerful Humanoid Soccer Shooting via Motion-Guided Curriculum Reinforcement Learning

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-09

RoboNaldo: motion-guided curriculum RL for powerful humanoid soccer shots

Academic

A History-Aware Visually Grounded Critic for Computer Use Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-09

HiViG: a history-aware, visually grounded critic for computer-use agents

Academic

T1-Bench: Benchmarking Multi-Scenario Agents in Real-World Domains

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-09

T1-Bench: high-fidelity evaluation of multi-domain agents across 25 domains

Academic

What Fits (Into Few Tokens) Doesn't Overfit: Compression and Generalization in ML Research Agents

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-09

Compressibility explains why ML research agents rarely overfit

Academic

Workflow-GYM: Towards Long-Horizon Evaluation of Computer-use Agentic tasks in Real-World Professional Fields

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-09

Workflow-GYM: long-horizon GUI tasks in professional software

Academic

AuRA: Internalizing Audio Understanding into LLMs as LoRA

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-09

AuRA: internalizing audio understanding into LLMs via LoRA

Academic

Diffusion Forcing Planner: History-Annealed Planning with Time-Dependent Guidance for Autonomous Driving

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-09

Diffusion Forcing Planner: history-guided diffusion planning for driving

Academic

Understanding and mitigating the risks of OpenClaw for non-technical users: A practical guide with Skill

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-09

A plain-language guide to OpenClaw risks for non-technical users

Academic

Mind the Gap: Can Frontier LLMs Pass a Standardized Office Proficiency Exam?

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-09

Frontier LLMs score at most 36.6% on a standardized Office proficiency exam

Academic

CLP: Collocation-Length Prediction for Zero-Loss Adaptive Multi-Token Inference

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-09

CLP: backbone-as-architect for zero-loss adaptive multi-token inference

Academic

Frontier Coding Agents Use Metaprogramming to Adapt to Unfamiliar Programming Languages

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-09

Frontier coding agents use metaprogramming for esoteric languages

Academic

Role-Agent: Bootstrapping LLM Agents via Dual-Role Evolution

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-09

Role-Agent: a single LLM as both agent and environment for co-evolution

Academic

Range Penalization: Theoretical Insights with Applications in Federated Learning

arXiv cs.LG (Machine Learning) ・ 2026-06-09

Range penalization for quantization-friendly federated learning

Academic

What Do Deepfake Speech Detectors Actually Hear?

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-09

What deepfake speech detectors actually hear, localized over time

Academic

Ethical and Technical Limits of Deepfake Speech Datasets

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-09

Auditing 39 deepfake speech datasets reveals fairness gaps and overlap

Academic

RAT: Reference-Augmented Training for ASV Anti-Spoofing

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-09

RAT: reference-augmented training sets a new ASV anti-spoofing SOTA

Academic

Pushing the Limits of LLM Tool Calling via Experiential Knowledge Integration and Activation

arXiv cs.CL (Computation and Language) ・ 2026-06-09

Boosting LLM tool calling via experiential knowledge integration

Academic

ConvMemory v2: A Recall-Preserving Top-10 Evidence Reranker for Conversational Memory Retrieval

arXiv cs.CL (Computation and Language) ・ 2026-06-09

ConvMemory v2: a recall-preserving Top-10 reranker for conversational memory

Academic

Attention-Discounted Adaptive Sampler for Masked Diffusion Language Models

arXiv cs.CL (Computation and Language) ・ 2026-06-09

ADAS: attention-discounted reranking for masked diffusion decoding

Academic

K-Forcing: Joint Next-K-Token Decoding via Push-Forward Language Modeling

arXiv cs.CL (Computation and Language) ・ 2026-06-09

K-Forcing: joint next-k-token decoding via push-forward language modeling

Academic

Recovering the Zipfian Distribution in Unsupervised Term Discovery

arXiv cs.CL (Computation and Language) ・ 2026-06-09

Graph clustering beats K-means in unsupervised term discovery

Academic

Continual LLM Upcycling: A Predictor-Gated Bank-Wise Sparsity Training Recipe for Dense-to-Sparse LLMs

arXiv cs.CL (Computation and Language) ・ 2026-06-09

Predictor-gated sparsity recipe upcycles dense LLMs to sparse

Academic

Attention Expansion: Enhancing Keyphrase Extraction from Long Documents with Attention-Augmented Contextualized Embeddings

arXiv cs.CL (Computation and Language) ・ 2026-06-09

Attention expansion boosts keyphrase extraction from long docs

Academic

REAL: A Reasoning-Enhanced Graph Framework for Long-Term Memory Management of LLMs

arXiv cs.CL (Computation and Language) ・ 2026-06-09

REAL manages LLM long-term memory with reasoning-enhanced graphs

Academic

Infini Memory: Maintainable Topic Documents for Long-Term LLM Agent Memory

arXiv cs.CL (Computation and Language) ・ 2026-06-09

Infini Memory: topic documents for long-term LLM agent memory

Academic

Multilingual Word-Level Forced Alignment with Self-Supervised Representations and Learned Dynamic Programming

arXiv cs.CL (Computation and Language) ・ 2026-06-09

Learned DP decoder advances multilingual forced alignment

Academic

Speaker Group Encoding in Self-supervised Speech Recognition Models

arXiv cs.CL (Computation and Language) ・ 2026-06-09

How self-supervised speech models encode speaker group traits

Academic

ParaBridge: Bridging Paralinguistic Perception and Dialogue Behavior in Speech Language Models

arXiv cs.CL (Computation and Language) ・ 2026-06-09

ParaBridge links paralinguistic cues to dialogue behavior

← Story Archive