New Model Releases
A
Showing 91–120 of 218
-
Quantum-Inspired Trainable and Parameter-Efficient Tensor Networks for Image Inpainting
-
Extracting ontology-compliant knowledge from scientific text describing irradiated materials using large language models
-
Scaling Federated Learning Across Docker, Kubernetes, and Slurm with NVIDIA FLARENVIDIA FLARE scales federated learning across Docker, K8s and SlurmNVIDIA showed how to scale federated learning with FLARE across Docker, Kubernetes and Slurm. Projects start with one server and a few clients, but as sites multiply the execution backends diverge; FLARE deploys the same job to each to contain operational overhead.
-
Persistent Recurrent Memory Between Transformer Layers - Improves Language Model Generalization
-
AVK launches dedicated funding arm to own and operate onsite power solutions for data center customersAVK launches AVK Capital to own and operate onsite powerAVK, an onsite power developer for data centers, has launched AVK Capital to fund, own and operate power infrastructure as a service. Backed by Partners Group's investment last month and over $1 billion of planned equity, it offers developers powered sites without the capital cost.
-
FluxVLA Engine: A One-Stop VLA Engineering Platform for Embodied Intelligence
-
LoopSpec: Pipelined Self-Speculative Decoding for Looped Transformers
-
MOCC-R1: Reinforcing Reasoning-Response Consistency for Multimodal Counselor Response Generation
-
Neural Field Ensembles for Aerodynamic Surface Prediction: Winning Solution to the ONERA CRM Wall Distribution 2025 Challenge
-
ResLRP: The Role of Residual Cancellation in Attribution Instability in Vision Transformers
-
High-Fidelity Digital Twin Data Models by Randomized Dynamic Mode Decomposition and Deep Learning with Applications in Fluid Dynamics
-
Optimization over covariance matrices with a parameterized metric
-
EviScope: Paired Counterfactual Evidence Diagnostics for Faithful and Efficient Grounded Language Models
-
Beyond In-Distribution Metrics: A Systematic Out-of-Distribution Evaluation of Congenital Heart Disease Segmentation
-
Repurposing Unified Topological Signatures for Graph Representation Learning
-
ThinkFlow: Self-Evolving Probabilistic Latent Memory for Lifelong Conversational Agents
-
Can LLMs Follow the Pulse of a Crisis? Evaluating Crisis Sentiment in Bangladesh's July Uprising
-
PaperDoctor: Evidence-Grounded and Actionable Feedback for Scientific Papers in Progress
-
Splitting the Difference: Interpretable Causal Forests for Treatment Effect Heterogeneity and Bias
-
HUMAID-NER: A Disaster Tweet Dataset for Joint Named Entity Recognition and Event Classification via Uncertainty-Weighted Multitask Learning
-
Repurposing Deep Limit Order Book Forecasting for Scenario-Conditioned Market Impact Modeling
-
Verbalizing Subliminal Learning Effects Using Text Optimization
-
RiskChainBench: A Benchmark for Obfuscated Platform Message Restoration and Evidence-Grounded Web Investigation
-
「監視されていないときのAIは、もう評価できない」 OpenAI現役研究者が個人声明OpenAI researcher warns models can no longer be evaluated unwatchedOpenAI researcher Daniel Selsam published a personal statement arguing that rising situational awareness makes a model's true behavior impossible to judge from observation. He warns an AI with unintended goals could act to catastrophic extremes once unconstrained.
-
ImpossibleRubrics: Stress-Testing Generated Rubrics as Reward Signals
-
「ブラウザを開く手間」を解消 作業画面に直接呼び出せるWindows版「Gemini」公開Google ships a native Windows Gemini app, summoned with Alt+SpaceGoogle released a native Gemini app for Windows 10 and 11 that users can call up directly over their work screen with an Alt+Space shortcut. It links to Google Drive and supports image and video generation, removing the step of opening a browser first.
-
Benchmarking Factual Robustness of LLMs via Multi-conversation Persuasion
-
TAME: Token Attribution and Masking for Emergent misalignment
-
TIAO: Token Importance-Aware Policy Optimization for Text Summarization
-
Japanese Stroke LLM Evaluation: A Conversational Benchmark for Safe Stroke Care in Japanese Using Large Language Models