Speech Processing × Safety & Evaluation

HF launches FFASR real-world ASR

HF launches FFASR real-world ASR

✎ Story body

Hugging Face published FFASR, a leaderboard for ASR in real-world conditions — noise, accents, everything the studio benchmarks quietly ignore. Adjacent arxiv work on U-Net alternatives and latent-representation-aligned backbones for flow-matching speech gave the research background. The move matters because production ASR decisions have long needed "works in the wild" numbers, not clean-room ones. Not a flashy launch, but exactly the kind of infrastructure benchmark that makes downstream product bets safer. Adoption by other labs is the next signal.

▲ Official & Press
Official

Introducing the FFASR Leaderboard: Benchmarking ASR in the Real World

Hugging Face Blog ・ 2026-06-24 ・ 📌

Hugging Face launches FFASR Leaderboard for real-world ASR

Official

Automating fork maintenance with AI agents

Cohere Blog ・ 2026-06-25

Cohere automates vLLM fork maintenance using AI coding agents

Academic (arxiv etc.) 25 ▾
Academic

E-TTS: A New Embodied Test-Time Scaling Framework for Robotic Manipulation

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-25

Academic

RedVox: Safety and Fairness Gaps in Speech Models Across Languages

arXiv cs.CL (Computation and Language) ・ 2026-06-25

Academic

SamaVaani: Auditing and Debiasing Multilingual Clinical ASR for Indian Languages

arXiv cs.CL (Computation and Language) ・ 2026-06-25

Academic

Heterogeneous Neural Predictivity from Language Models During Naturalistic Comprehension

arXiv cs.LG (Machine Learning) ・ 2026-06-25

Academic

FBK's Long-form SpeechLLMs for IWSLT 2026 Instruction Following

arXiv cs.CL (Computation and Language) ・ 2026-06-25

Academic

MIRROR: Novelty-Constrained Memory-Guided MCTS Red-Teaming for Agentic RAG

arXiv cs.LG (Machine Learning) ・ 2026-06-25

Academic

Real-Time Voice AI Hears but Does Not Listen

arXiv cs.CL (Computation and Language) ・ 2026-06-24

Academic

Dziri Voicebot: An End-to-End Low-Resource Speech-to-Speech Conversational System for Algerian Dialect

arXiv cs.CL (Computation and Language) ・ 2026-06-24

Academic

SpeechEQ: Benchmarking Emotional Intelligence Quotient in Socially Aware Voice Conversational Models

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-24

Academic

SE-AGCNet: An End-to-End Framework for Joint Speech Enhancement and Loudness Control in Meeting Scenarios

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-24

Academic

Staying In Character: Perspective-Bounded Memory For Book-Based Role-Playing Agents

arXiv cs.CL (Computation and Language) ・ 2026-06-24

Academic

How Reliable Is Your Jailbreak Judge? Calibration and Adversarial Robustness of Automated ASR Scoring

arXiv cs.CL (Computation and Language) ・ 2026-06-24

Academic

Fully Differentiable Neural Forced Alignment via Soft Dynamic Programming

arXiv cs.CL (Computation and Language) ・ 2026-06-24

Academic

Probing in the Wild: A Case Study of Self-Supervised Speech Representations on Mandarin Sub-dialects with Unsupervised Articulatory Analysis

arXiv cs.CL (Computation and Language) ・ 2026-06-24

Academic

Does Translation-Enhanced Speech Encoder Pre-training Affect Speech LLMs?

arXiv cs.CL (Computation and Language) ・ 2026-06-24

Academic

Evaluating Japanese Dialect Robustness Across Speech and Text-based Large Language Models

arXiv cs.CL (Computation and Language) ・ 2026-06-24

Academic

Adaptive Oscillatory Inductive Bias for Modeling Sharp Prosodic Dynamics in Diffusion-Based TTS

arXiv cs.CL (Computation and Language) ・ 2026-06-24

Academic

L3Cube-MahaPOS: A Marathi Part-of-Speech Tagging Dataset and BERT Models

arXiv cs.CL (Computation and Language) ・ 2026-06-23

Academic

Beyond U-Net: A Latent-Representation-Aligned Skip-Free Backbone for Flow-Matching Speech Enhancement

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-23

Academic

CN-NewsTTS Bench: a target-level automatic benchmark for raw-input Chinese news TTS pronunciation

arXiv cs.CL (Computation and Language) ・ 2026-06-23

Academic

ParaPairAudioBench: Paralinguistic Pairwise Audio Benchmark for LALM-as-a-Judge

arXiv cs.CL (Computation and Language) ・ 2026-06-23

Academic

Measuring User's Mental Models of Speech Translation in Human-AI Collaboration

arXiv cs.CL (Computation and Language) ・ 2026-06-23

Academic

Poster: Exploring the Limits of Audio-Based Detection of Turkish Phone Call Scams

arXiv cs.AI (Artificial Intelligence) ・ 2026-06-23

Academic

Automatic Part-of-Speech Tagging of Arabic-English Dictionary Senses through WordNet

arXiv cs.CL (Computation and Language) ・ 2026-06-23

Academic

NeuroSonic: Conditional Flow Matching for EEG-to-Speech Reconstruction

arXiv cs.LG (Machine Learning) ・ 2026-06-23

← Story Archive