Developer Tools

B
Showing 31–60 of 326
  • arXiv cs.AI (Artificial Intelligence) · EN Developer Tools
    Beyond Outcomes: Dual-View Relational Learning for Efficient Agent Benchmarking
    AI Agents Software Engineering
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.CL (Computation and Language) · EN New Model Releases
    How Much is a Human Right Worth? ECtHR-NPD: A Benchmark for Predicting Non-Pecuniary Damage Awards
    AI Agents
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.CL (Computation and Language) · EN New Model Releases
    Structured Claim-Level Discourse Representations for Dense Health Narratives
    Inference Neural Network Retrieval-Augmented Generation (RAG)
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.LG (Machine Learning) · EN Developer Tools
    Fast Learning Rates for Physics-Informed Kernel Methods
    Machine Learning
    Read original (arXiv cs.LG (Machine Learning)) ↗
  • NVIDIA Developer Blog · EN Infrastructure & Hardware
    Translating CUDA Tile Operations from Python to Rust Using Agentic AI
    NVIDIA uses an AI agent skill to port TileGym kernels to cuTile Rust
    AI Agents Generative AI Mixture of Experts (MoE) NVIDIA
    NVIDIA unveiled cuTile Rust, extending Rust's ownership model to tile-based GPU kernels, plus an AI agent skill that translates cuTile Python and Triton-TileIR kernels into it. A bounded multi-agent pipeline ported all 24 public TileGym operators at 99.5% of cuTile Python performance.
    Read original (NVIDIA Developer Blog) ↗
  • arXiv cs.LG (Machine Learning) · EN Developer Tools
    Learning Lyapunov Operators for Nonlinear Systems
    Machine Learning
    Read original (arXiv cs.LG (Machine Learning)) ↗
  • arXiv cs.LG (Machine Learning) · EN Developer Tools
    Preventing Model Collapse: A Fisher-Rao Perspective on the Dynamics of Training with Synthetic Data
    Retrieval-Augmented Generation (RAG) Reinforcement Learning
    Read original (arXiv cs.LG (Machine Learning)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN New Model Releases
    ASLEval: Measuring Privacy Exposure Displacement in LLM Agent Sessions
    AI Agents
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.LG (Machine Learning) · EN Developer Tools
    Physics-based prediction, uncertainty quantification and decision-making for IN718 crystallographic texture intensity across LPBF defocus regimes
    Neural Network Retrieval-Augmented Generation (RAG) Reinforcement Learning
    Read original (arXiv cs.LG (Machine Learning)) ↗
  • Simon Willison's Weblog · EN Safety & Evaluation
    Quoting Mustafa Suleyman
    Willison quotes Suleyman's warning against granting models rights
    Microsoft
    Simon Willison highlights an essay by Microsoft AI CEO Mustafa Suleyman warning against treating models as if they had feelings, preferences or rights. Suleyman argues consciousness underpins our ethical, legal and political systems, that extending such rights is not justified by the evidence, and that doing so would make AI containment and alignment harder.
    Read original (Simon Willison's Weblog) ↗
  • arXiv cs.CL (Computation and Language) · EN New Model Releases
    PersonaPath: Towards Knowledge-Centric Personalized Learning Path Planning
    Neural Network
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN New Model Releases
    GrainSpeech: Less Context, More Detail for Compact Speech Synthesis
    Reinforcement Learning Speech Processing
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.CL (Computation and Language) · EN Developer Tools
    EviGen: Predictive Evidence Scaffolding for Verifiable Clinical Rationale Generation
    Retrieval-Augmented Generation (RAG)
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN Inference & Efficiency
    Ask the Tool, Don't Guess: Agent Tool Calls Hold Their Progress, and the Serving System Should Read It
    Neural Network Software Engineering
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.CL (Computation and Language) · EN Developer Tools
    ReFigBench: Benchmarking Scientific Figure Reconstruction as Editable PowerPoint Artifacts
    AI Agents Neural Network Reinforcement Learning
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.LG (Machine Learning) · EN Developer Tools
    Interpretable Multi-Instance Learning Enables Early Prediction of Key Molecular Alterations from Routine Flow Cytometry in Acute Myeloid Leukemia
    Deep Learning Machine Learning Neural Network Reinforcement Learning
    Read original (arXiv cs.LG (Machine Learning)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN Infrastructure & Hardware
    Using OCR Heads to Verbalize Image Semantics
    Reinforcement Learning
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN New Model Releases
    ProgramDistill: From Interactive Web Apps to Verifiable Reference-Guided SWE Tasks
    AI Agents Claude GPT Software Engineering
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.CL (Computation and Language) · EN Developer Tools
    Beyond frequency measures: Can contextual embeddings capture meaning change in scientific texts?
    Embeddings Neural Network Natural Language Processing (NLP) Retrieval-Augmented Generation (RAG) Reinforcement Learning
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • Sakana AI Blog (ja) · EN Developer Tools
    Inside Sakana AI's Product Team
    Sakana AI opens up its Product Team: culture, workflow, and hiring FAQ
    Retrieval-Augmented Generation (RAG) Reinforcement Learning Software Engineering
    Sakana AI published a look inside its Product Team, which turns research into shipped products. Based on member interviews, it covers team composition, flexible hours without core time, a typical day for applied research and software engineers, and the hiring process.
    Read original (Sakana AI Blog (ja)) ↗
  • arXiv cs.CL (Computation and Language) · EN Developer Tools
    Zero-Shot Cross-Lingual Recognition of Sign Language Handshapes
    Deep Learning Machine Learning Neural Network Retrieval-Augmented Generation (RAG) Reinforcement Learning
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN Industry Adoption
    Version- and Scope-Aware Question Answering over Normative Documents: A Deployed System and an End-to-End Evaluation at Production Scale
    Software Engineering
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.CL (Computation and Language) · EN Developer Tools
    TeleAntiFraud 2.0: A Refreshable, Profile-Grounded, and Audio-Based Benchmark for Telecom Fraud Detection
    Deep Learning Retrieval-Augmented Generation (RAG) Speech Processing
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.LG (Machine Learning) · EN New Model Releases
    When Edit Flows are Edit Jumps: replicating Edit Flows and EvoFlows
    Reinforcement Learning
    Read original (arXiv cs.LG (Machine Learning)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN Inference & Efficiency
    A Scalable Framework for Automated NER Annotation Correction in Low-Resource Languages
    Inference Neural Network Natural Language Processing (NLP) Retrieval-Augmented Generation (RAG)
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN New Model Releases
    Which LLM is Best for Translating Natural Language Goals to PDDL
    Deep Learning Gemini GPT Neural Network Reinforcement Learning
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.CL (Computation and Language) · EN Developer Tools
    "If I Had to Buy Just ONE: Galaxy S26 Ultra": Auditing AI-Generated Product Recommendations
    Gemini Google GPT OpenAI Retrieval-Augmented Generation (RAG)
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.CL (Computation and Language) · EN Training & Fine-tuning
    LocQE: Principled Domain Adaptation for Localisation Quality Estimation by Leveraging Post-Edits
    Deep Learning Fine-tuning Retrieval-Augmented Generation (RAG) Reinforcement Learning
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN New Model Releases
    Rethinking Critic Learning in PPO: Understanding and Mitigating Value Flattening
    Reinforcement Learning
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN Developer Tools
    Echo: Learning-based Matching Decompilation using Trusted Back Translation
    GPT Retrieval-Augmented Generation (RAG)
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗