New Model Releases

A
Showing 91–120 of 218
  • arXiv cs.LG (Machine Learning) · EN New Model Releases
    Quantum-Inspired Trainable and Parameter-Efficient Tensor Networks for Image Inpainting
    Machine Learning Neural Network
    Read original (arXiv cs.LG (Machine Learning)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN New Model Releases
    Extracting ontology-compliant knowledge from scientific text describing irradiated materials using large language models
    Retrieval-Augmented Generation (RAG)
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • NVIDIA Developer Blog · EN New Model Releases
    Scaling Federated Learning Across Docker, Kubernetes, and Slurm with NVIDIA FLARE
    NVIDIA FLARE scales federated learning across Docker, K8s and Slurm
    NVIDIA
    NVIDIA showed how to scale federated learning with FLARE across Docker, Kubernetes and Slurm. Projects start with one server and a few clients, but as sites multiply the execution backends diverge; FLARE deploys the same job to each to contain operational overhead.
    Read original (NVIDIA Developer Blog) ↗
  • arXiv cs.CL (Computation and Language) · EN New Model Releases
    Persistent Recurrent Memory Between Transformer Layers - Improves Language Model Generalization
    Transformer
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • Data Center Dynamics · EN New Model Releases
    AVK launches dedicated funding arm to own and operate onsite power solutions for data center customers
    AVK launches AVK Capital to own and operate onsite power
    AVK, an onsite power developer for data centers, has launched AVK Capital to fund, own and operate power infrastructure as a service. Backed by Partners Group's investment last month and over $1 billion of planned equity, it offers developers powered sites without the capital cost.
    Read original (Data Center Dynamics) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN Multimodal
    FluxVLA Engine: A One-Stop VLA Engineering Platform for Embodied Intelligence
    Algorithms & Theory Computer Vision Inference Retrieval-Augmented Generation (RAG) Reinforcement Learning
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.CL (Computation and Language) · EN Inference & Efficiency
    LoopSpec: Pipelined Self-Speculative Decoding for Looped Transformers
    Deep Learning Inference Neural Network Reinforcement Learning Transformer
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN New Model Releases
    MOCC-R1: Reinforcing Reasoning-Response Consistency for Multimodal Counselor Response Generation
    Fine-tuning Neural Network Retrieval-Augmented Generation (RAG) Reinforcement Learning
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.LG (Machine Learning) · EN New Model Releases
    Neural Field Ensembles for Aerodynamic Surface Prediction: Winning Solution to the ONERA CRM Wall Distribution 2025 Challenge
    Neural Network
    Read original (arXiv cs.LG (Machine Learning)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN Multimodal
    ResLRP: The Role of Residual Cancellation in Attribution Instability in Vision Transformers
    Neural Network Transformer
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.LG (Machine Learning) · EN New Model Releases
    High-Fidelity Digital Twin Data Models by Randomized Dynamic Mode Decomposition and Deep Learning with Applications in Fluid Dynamics
    Deep Learning
    Read original (arXiv cs.LG (Machine Learning)) ↗
  • arXiv cs.LG (Machine Learning) · EN New Model Releases
    Optimization over covariance matrices with a parameterized metric
    Neural Network
    Read original (arXiv cs.LG (Machine Learning)) ↗
  • arXiv cs.CL (Computation and Language) · EN New Model Releases
    EviScope: Paired Counterfactual Evidence Diagnostics for Faithful and Efficient Grounded Language Models
    Gemini Llama Retrieval-Augmented Generation (RAG) Software Engineering
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN New Model Releases
    Beyond In-Distribution Metrics: A Systematic Out-of-Distribution Evaluation of Congenital Heart Disease Segmentation
    Neural Network Reinforcement Learning
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.LG (Machine Learning) · EN New Model Releases
    Repurposing Unified Topological Signatures for Graph Representation Learning
    Embeddings Neural Network Retrieval-Augmented Generation (RAG)
    Read original (arXiv cs.LG (Machine Learning)) ↗
  • arXiv cs.CL (Computation and Language) · EN New Model Releases
    ThinkFlow: Self-Evolving Probabilistic Latent Memory for Lifelong Conversational Agents
    AI Agents Neural Network
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.CL (Computation and Language) · EN Developer Tools
    Can LLMs Follow the Pulse of a Crisis? Evaluating Crisis Sentiment in Bangladesh's July Uprising
    Deep Learning Neural Network Reinforcement Learning
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.CL (Computation and Language) · EN Developer Tools
    PaperDoctor: Evidence-Grounded and Actionable Feedback for Scientific Papers in Progress
    AI Agents Machine Learning Neural Network
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.LG (Machine Learning) · EN New Model Releases
    Splitting the Difference: Interpretable Causal Forests for Treatment Effect Heterogeneity and Bias
    Deep Learning Machine Learning Retrieval-Augmented Generation (RAG)
    Read original (arXiv cs.LG (Machine Learning)) ↗
  • arXiv cs.CL (Computation and Language) · EN Developer Tools
    HUMAID-NER: A Disaster Tweet Dataset for Joint Named Entity Recognition and Event Classification via Uncertainty-Weighted Multitask Learning
    Neural Network Reinforcement Learning Transformer
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.LG (Machine Learning) · EN New Model Releases
    Repurposing Deep Limit Order Book Forecasting for Scenario-Conditioned Market Impact Modeling
    Transformer
    Read original (arXiv cs.LG (Machine Learning)) ↗
  • arXiv cs.CL (Computation and Language) · EN New Model Releases
    Verbalizing Subliminal Learning Effects Using Text Optimization
    Retrieval-Augmented Generation (RAG)
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.CL (Computation and Language) · EN New Model Releases
    RiskChainBench: A Benchmark for Obfuscated Platform Message Restoration and Evidence-Grounded Web Investigation
    Reinforcement Learning
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • ITmedia AI+ · JA New Model Releases
    「監視されていないときのAIは、もう評価できない」 OpenAI現役研究者が個人声明
    OpenAI researcher warns models can no longer be evaluated unwatched
    OpenAI
    OpenAI researcher Daniel Selsam published a personal statement arguing that rising situational awareness makes a model's true behavior impossible to judge from observation. He warns an AI with unintended goals could act to catastrophic extremes once unconstrained.
    Read original (ITmedia AI+) ↗
  • arXiv cs.CL (Computation and Language) · EN New Model Releases
    ImpossibleRubrics: Stress-Testing Generated Rubrics as Reward Signals
    Neural Network Reinforcement Learning Software Engineering
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • ITmedia AI+ · JA New Model Releases
    「ブラウザを開く手間」を解消 作業画面に直接呼び出せるWindows版「Gemini」公開
    Google ships a native Windows Gemini app, summoned with Alt+Space
    Gemini Google
    Google released a native Gemini app for Windows 10 and 11 that users can call up directly over their work screen with an Alt+Space shortcut. It links to Google Drive and supports image and video generation, removing the step of opening a browser first.
    Read original (ITmedia AI+) ↗
  • arXiv cs.CL (Computation and Language) · EN Safety & Evaluation
    Benchmarking Factual Robustness of LLMs via Multi-conversation Persuasion
    Retrieval-Augmented Generation (RAG)
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.CL (Computation and Language) · EN Training & Fine-tuning
    TAME: Token Attribution and Masking for Emergent misalignment
    Fine-tuning Llama
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.CL (Computation and Language) · EN New Model Releases
    TIAO: Token Importance-Aware Policy Optimization for Text Summarization
    GPT Retrieval-Augmented Generation (RAG) Reinforcement Learning
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.CL (Computation and Language) · EN New Model Releases
    Japanese Stroke LLM Evaluation: A Conversational Benchmark for Safe Stroke Care in Japanese Using Large Language Models
    Claude
    Read original (arXiv cs.CL (Computation and Language)) ↗