Developer Tools
B
Showing 31–60 of 326
-
Beyond Outcomes: Dual-View Relational Learning for Efficient Agent Benchmarking
-
How Much is a Human Right Worth? ECtHR-NPD: A Benchmark for Predicting Non-Pecuniary Damage Awards
-
Structured Claim-Level Discourse Representations for Dense Health Narratives
-
Fast Learning Rates for Physics-Informed Kernel Methods
-
Translating CUDA Tile Operations from Python to Rust Using Agentic AINVIDIA uses an AI agent skill to port TileGym kernels to cuTile RustNVIDIA unveiled cuTile Rust, extending Rust's ownership model to tile-based GPU kernels, plus an AI agent skill that translates cuTile Python and Triton-TileIR kernels into it. A bounded multi-agent pipeline ported all 24 public TileGym operators at 99.5% of cuTile Python performance.
-
Learning Lyapunov Operators for Nonlinear Systems
-
Preventing Model Collapse: A Fisher-Rao Perspective on the Dynamics of Training with Synthetic Data
-
ASLEval: Measuring Privacy Exposure Displacement in LLM Agent Sessions
-
Physics-based prediction, uncertainty quantification and decision-making for IN718 crystallographic texture intensity across LPBF defocus regimes
-
Quoting Mustafa SuleymanWillison quotes Suleyman's warning against granting models rightsSimon Willison highlights an essay by Microsoft AI CEO Mustafa Suleyman warning against treating models as if they had feelings, preferences or rights. Suleyman argues consciousness underpins our ethical, legal and political systems, that extending such rights is not justified by the evidence, and that doing so would make AI containment and alignment harder.
-
PersonaPath: Towards Knowledge-Centric Personalized Learning Path Planning
-
GrainSpeech: Less Context, More Detail for Compact Speech Synthesis
-
EviGen: Predictive Evidence Scaffolding for Verifiable Clinical Rationale Generation
-
Ask the Tool, Don't Guess: Agent Tool Calls Hold Their Progress, and the Serving System Should Read It
-
ReFigBench: Benchmarking Scientific Figure Reconstruction as Editable PowerPoint Artifacts
-
Interpretable Multi-Instance Learning Enables Early Prediction of Key Molecular Alterations from Routine Flow Cytometry in Acute Myeloid Leukemia
-
Using OCR Heads to Verbalize Image Semantics
-
ProgramDistill: From Interactive Web Apps to Verifiable Reference-Guided SWE Tasks
-
Beyond frequency measures: Can contextual embeddings capture meaning change in scientific texts?
-
Inside Sakana AI's Product TeamSakana AI opens up its Product Team: culture, workflow, and hiring FAQSakana AI published a look inside its Product Team, which turns research into shipped products. Based on member interviews, it covers team composition, flexible hours without core time, a typical day for applied research and software engineers, and the hiring process.
-
Zero-Shot Cross-Lingual Recognition of Sign Language Handshapes
-
Version- and Scope-Aware Question Answering over Normative Documents: A Deployed System and an End-to-End Evaluation at Production Scale
-
TeleAntiFraud 2.0: A Refreshable, Profile-Grounded, and Audio-Based Benchmark for Telecom Fraud Detection
-
When Edit Flows are Edit Jumps: replicating Edit Flows and EvoFlows
-
A Scalable Framework for Automated NER Annotation Correction in Low-Resource Languages
-
Which LLM is Best for Translating Natural Language Goals to PDDL
-
"If I Had to Buy Just ONE: Galaxy S26 Ultra": Auditing AI-Generated Product Recommendations
-
LocQE: Principled Domain Adaptation for Localisation Quality Estimation by Leveraging Post-Edits
-
Rethinking Critic Learning in PPO: Understanding and Mitigating Value Flattening
-
Echo: Learning-based Matching Decompilation using Trusted Back Translation