Safety & Evaluation A
Showing 31–60 of 94
-
EgoGenesis: Egocentric World-Action Modeling with Online Anchored Projective Memory and Action-3D RoPE
-
AI and Authenticity in Islamic Research: A Critical Evaluation of Generative AI Reliability, Hallucination, and Source Fidelity in Quranic, Hadith, and Fiqh Knowledge
-
Security of World-Model-Based Embodied AI: A Lifecycle of Threats, Defenses, and Evaluation
-
Fidelity Is Not Safety: Gently-Compressed LLMs Pass Every Data-Free Quality Guard Yet Invent Procedure Steps in Agentic Execution
-
Integrating AI into Requirements Quality Learning in Software Engineering Education: A TPACK-Guided Empirical Study
-
OPLD: On-Policy Latent Distillation for Multimodal Reasoning
-
Can Agents Deceive? Evaluating Reasoning and Deception in ParliamentBench using a Social Deduction Game
-
Asymmetric Communication: Large Language Models and Language Games
-
An Instrument to Evaluate Governance Proposals: AI Policy Analysis at Scale
-
Diversifying Personalized Research Ideation against AI-Induced Homogenization
-
ClawTrack: Towards Trace-Level Evaluation and Improvement of Real-World Autonomous Agents
-
Measuring Alignment With Reader Highlights Net of Position and Length
-
DualAnchor: Preserving Language Priors and Improving Lexical Fidelity in Gloss-Free Sign Language Translation
-
Cost-Sensitive Conformal Prediction and Human-in-the-Loop Abstention for Imbalanced High-Stakes Decision Support: A Multi-Domain Benchmark
-
How to Self-Host a Validated AI Coding Assistant with NVIDIA NeMo GuardrailsNVIDIA: self-host a validated AI coding assistant via NeMo GuardrailsAn NVIDIA developer-blog post on self-hosting a validated AI coding assistant using NeMo Guardrails, framed around agent operation, infrastructure and safety. Note: the raw excerpt was blocked by a content guard, so specific components, supported models and guardrail rules are inferred from the title and URL and remain unverified from the body.
-
SciFigQual-Bench: A Benchmark for Scientific Figure Quality Assessment with Full-Manuscript Context
-
On-Policy Distillation for LLM Safety: A Routing Approach to Template-Robust Realignment
-
Visual Credit Audit for Multimodal Spatial Reasoning
-
SciFigAlign: Scoring Scientific Figures by Fine-tuned Alignment of Visuals with Manuscript Evidence
-
OptimismBench: Forecasting Bias and the Alignment Effect in Language Model Judgment
-
TREK: A Travel Reasoning and Evaluation Kit for LLM Agents in Complex Trip Planning
-
Progressive Multimodal Alignment for Continual Instruction Tuning
-
Belief-Guided Decision Making with Uncertainty Gating in the Game of Go
-
Defending Against Backdoor Attacks via Alignment Checking in Model-Contrastive Federated Learning
-
BioVLN: A Simulation Platform for Visual Language Navigation in Biomedical Laboratories
-
Dual-Path LLM Reasoning for Multimodal Few-Shot Knowledge Graph Completion
-
From Found to Designed: Concepts as a Design Axis for Large Language Models
-
FedTopo: Relation-Level Topology Sharing for Model-Heterogeneous Federated Learning
-
Dual Inversion for Text-to-Image Diffusion Models: From Both Prompt and Noise Perspectives
-
MPEcho: A Melody and Phoneme-Aware Generative Framework for Controllable Cover Song Generation