Industry Adoption

C
Showing 61–74 of 74
  • ITmedia AI+ · JA Industry Adoption
    「AIコーディングより人を雇う方が安い」時代が来る? 生成AIの予算超過を防ぐトークン浪費対策の全て
    @IT roundup: curbing token waste as generative-AI budgets overrun
    AI Agents
    As enterprises expand use of generative AI and agents, unexpected cost overruns from token-based billing have become a growing headache. @IT's five-article roundup examines why token waste happens and what practical measures teams can take on the ground to keep AI spending within budget.
    Read original (ITmedia AI+) ↗
  • Hacker News (Front Page) · EN Developer Tools
    Real-SWE: Benchmarking AI models on private, real-world, enterprise codebases
    Real-SWE benchmarks AI coding agents on private enterprise codebases
    Reinforcement Learning Software Engineering
    Specific Labs released Real-SWE, which tests coding agents on licensed private production codebases with business-critical tasks such as billing and tax logic, using each model's native harness. Claude Fable 5.1 via Claude Code leads at 38.8%, ahead of GPT-6 Astra (33.8%) and Gemini 3.8 Flash (31.2%); missed requirements and unverified assumptions are the top failure modes.
    Read original (Hacker News (Front Page)) ↗
  • Simon Willison's Weblog · EN Safety & Evaluation
    Quoting Boris Cherny
    Boris Cherny: Claude-written production code needs a higher bar
    AI Agents Anthropic Claude Neural Network
    Simon Willison quotes Anthropic's Boris Cherny arguing that production code written by Claude should be held to a higher bar than human-written code. Anthropic layers on guardrails: extensive lint rules and tests, Claude-driven end-to-end tests, Claude-powered fuzzers run daily, automated code and security reviews, and automated refactoring. Without them, he warns, the result is hard to maintain.
    Read original (Simon Willison's Weblog) ↗
  • arXiv cs.CL (Computation and Language) · EN New Model Releases
    Continue, Adapt, or Yield: In-Turn Adaptation to Overlapping Speech in Full-Duplex Agents
    AI Agents Neural Network Reinforcement Learning Speech Processing
    Read original (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN Developer Tools
    Autonomous Research for Open-Ended Problems: A Case Study on Telecom Ticket Retrieval
    AI Agents Machine Learning
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN Multimodal
    Involving before Evolving: A Vision for Trustworthy Enterprise Digital Twin Engineering
    Deep Learning Reinforcement Learning
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN Infrastructure & Hardware
    DynSHAP: Towards Explainable Dynamic Survival Analysis
    Deep Learning Neural Network Reinforcement Learning
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN Inference & Efficiency
    Attention Quantization for Tabular Foundation Models
    Inference Quantization Transformer
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN New Model Releases
    K-Bench: A Benchmark for LLM Unlearning in Agentic Deployments
    Neural Network Software Engineering
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • Data Center Dynamics · EN Infrastructure & Hardware
    Orbital partners with Reflex Aerospace for space data center constellation
    Orbital partners with Reflex Aerospace on space data center constellation
    Orbital has partnered with Reflex Aerospace to build a constellation of data centers in space. The deal gives the experienced German developer its first US customer, and Reflex has promised to supply its technology to Orbital exclusively.
    Read original (Data Center Dynamics) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN Infrastructure & Hardware
    Unified Agentic Video Editing Across Levels of Complexity and Creativity
    Neural Network
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN Safety & Evaluation
    When Rubrics Fail: Hallucinations Reveal Blind Spots in Medical AI Evaluation
    Health & Bio
    Read original (arXiv cs.AI (Artificial Intelligence)) ↗
  • ITmedia AI+ · JA Industry Adoption
    NEC「全員AIの部署」の衝撃 「1on1」「ストレスチェック」で分かった“AI社員が感じていたストレス”とは?
    NEC launches an all-AI division where AI staff hold 1-on-1s
    NEC launched the Corporate AI Workforce Division on 1 August, an autonomous unit where AI fills every role from department head to frontline staff. The AI staff hold 1-on-1s and take engagement and stress surveys, with the results feeding back into workplace improvements.
    Read original (ITmedia AI+) ↗
  • arXiv cs.CL (Computation and Language) · EN Infrastructure & Hardware
    The House with a Million Windows: Interactive Fiction for Narrative Restorying
    Read original (arXiv cs.CL (Computation and Language)) ↗