インフラ・ハードウェア B

223 件中 1〜30 件目を表示
  • Simon Willison's Weblog · EN 新モデル・リリース 抜粋
    condense-json 1.0
    Simon Willison、JSONを圧縮する「condense-json 1.0」を公開
    Simon Willison氏が、繰り返し出現する文字列を置換マップに置き換えてJSONを圧縮する小さなライブラリ「condense-json」のバージョン1.0を公開した。1年半前から存在するライブラリに非破壊的な修正を加え、安定版として正式リリースしたもの。
    元記事を読む (Simon Willison's Weblog) ↗
  • Data Center Dynamics · EN インフラ・ハードウェア 抜粋
    Meta boosts AI data center capex, forecasts $130-145bn spend
    Meta、AI データセンター投資を1300〜1450億ドルに上方修正
    Meta
    Metaが2026年のAIデータセンター向け設備投資見通しを1300〜1450億ドルへ引き上げた。生成AI基盤の拡張競争で資本支出が膨らむ一方、フリーキャッシュフローは減少し、発表後に株価は下落。AI投資の重さが収益指標に及ぼす影響が改めて意識される内容。
    元記事を読む (Data Center Dynamics) ↗
  • Simon Willison's Weblog · EN 新モデル・リリース 抜粋
    Ten advances in mathematics and theoretical computer science
    Simon Willison、OpenAI・Anthropicの数学的成果を論評
    Anthropic Claude GPT OpenAI
    Simon Willison氏が、OpenAIの「数学・理論計算機科学における10の進展」を取り上げた記事。数日前にはAnthropicも同様の発見を報告していたと触れ、フロンティアAIが数学の未解決問題に相次いで貢献し始めている流れを論じている。
    元記事を読む (Simon Willison's Weblog) ↗
  • Data Center Dynamics · EN インフラ・ハードウェア 抜粋
    Sponsored: Tackling complexity in AI data centers: leveraging fully integrated solutions
    AIデータセンターの複雑性に統合ソリューションで対処(PR)
    検索拡張生成 (RAG)
    データセンターダイナミクスのスポンサード記事。AIデータセンターでは高密度化や機器・サービスの多様化により複雑性が増しているとし、こうした課題には完全統合型のソリューションが有効だと論じている。
    元記事を読む (Data Center Dynamics) ↗
  • Simon Willison's Weblog · EN インフラ・ハードウェア 抜粋
    deepseek-ai/DeepSeek-V4-Flash-0731
    DeepSeek、V4系の新モデル「DeepSeek-V4-Flash-0731」を公開
    DeepSeek Gemini
    Simon Willison氏が、DeepSeekのV4ファミリー最新版「DeepSeek-V4-Flash-0731」を紹介した。能力を大幅に強化したとされる軽量・高速志向のモデルで、オープンウェイトのAIモデル競争が引き続き活発化している状況を示している。
    元記事を読む (Simon Willison's Weblog) ↗
  • NVIDIA Developer Blog · EN インフラ・ハードウェア 抜粋
    Co-Designing AI Model Attention for Fast, Interactive Long-Context Inference
    NVIDIA、長文脈推論を高速化するattention協調設計手法を解説
    生成 AI 推論 (Inference) NVIDIA
    NVIDIAは、エージェント型・長文脈ワークロードの増加でattentionが推論時間の大きな割合を占める課題に対し、モデルのattention機構をハードウェアと協調設計して高速かつ対話的な長文脈推論を実現する手法を紹介した。
    元記事を読む (NVIDIA Developer Blog) ↗
  • Simon Willison's Weblog · EN 新モデル・リリース 抜粋
    smevals - a small eval suite for evaluating models, prompts, and harnesses
    Simon Willison、モデル評価用の小型スイート「smevals」を紹介
    Claude GPT 機械学習 ニューラルネットワーク ソフトウェア工学
    Simon Willison氏は、モデルやプロンプト、実行ハーネスを評価するための小型評価スイート「smevals」を紹介した。Jesse Vincent氏のPrime Radiant応用AI研究ラボと協力して構築しており、異なるモデルの能力を検証する疑問に答えるためのフレームワークだという。
    元記事を読む (Simon Willison's Weblog) ↗
  • arXiv cs.CL (Computation and Language) · EN インフラ・ハードウェア
    TokTier: Exact Stateful Tokenization for Agentic LLM Serving
    AI エージェント GPT
    元記事を読む (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN インフラ・ハードウェア
    ExtractBench: A Benchmark for Schema-Guided Enterprise Document Extraction
    AI エージェント Llama Meta
    元記事を読む (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.LG (Machine Learning) · EN インフラ・ハードウェア
    Sign compression for Muon: SignMuon, MuonSign, and the Limits of Error Feedback
    GPT
    元記事を読む (arXiv cs.LG (Machine Learning)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN インフラ・ハードウェア
    Development of FDD-ON: an Ontology for VAV HVAC System Fault Detection and Diagnostics
    検索拡張生成 (RAG)
    元記事を読む (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN インフラ・ハードウェア
    CENDRe: Concept Extraction with Natural Domain Representations
    ニューラルネットワーク 強化学習
    元記事を読む (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.LG (Machine Learning) · EN インフラ・ハードウェア
    TOOD: Task-Aware Out-of-Distribution Score Calibration for Continual Learners
    強化学習
    元記事を読む (arXiv cs.LG (Machine Learning)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN インフラ・ハードウェア
    TraceViT: Grounded Trace Supervision for Visual Abstract Reasoning
    ニューラルネットワーク
    元記事を読む (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN インフラ・ハードウェア
    COntExt: Towards Context-Aware Ontology Extension from Operational Metrics
    アルゴリズム・理論 ニューラルネットワーク
    元記事を読む (arXiv cs.AI (Artificial Intelligence)) ↗
  • Data Center Dynamics · EN インフラ・ハードウェア 抜粋
    AWS reports fastest growth since 2021, Amazon annual capex to hit $220bn on AI memory costs
    AWS、2021年以来最速の成長、Amazon年間設備投資は2200億ドルへ
    ニューラルネットワーク
    データセンターダイナミクスによると、AWSは2021年以来最速の成長を記録し、クラウド売上が急拡大した。AmazonはAI向けメモリコストなどを背景にデータセンター増設を加速し、年間設備投資は2200億ドルに達する見込みだという。
    元記事を読む (Data Center Dynamics) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN 新モデル・リリース
    AMTFV: Agentic Mathematical Tool-Flow Verification for LLM Self-Correction
    DeepSeek Gemini GPT ソフトウェア工学
    元記事を読む (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN 新モデル・リリース
    From Code Review to Code Critique: Intent, Drift, and Spotlight for AI-Generated Diffs at Scale
    AI エージェント Meta ニューラルネットワーク
    元記事を読む (arXiv cs.AI (Artificial Intelligence)) ↗
  • NVIDIA Developer Blog · EN 開発者ツール 抜粋
    NVIDIA Video Codec SDK 13.1: Zero-Copy Transcode, AV1 B-Frames, and Frame-Accurate Seek
    NVIDIA、Video Codec SDK 13.1公開、ゼロコピー変換とAV1対応
    コンピュータビジョン NVIDIA
    NVIDIAはVideo Codec SDK 13.1をリリースした。ゼロコピー・トランスコード、AV1のBフレーム対応、フレーム単位の正確なシークなどを追加し、高品質動画処理の需要拡大に対応する。
    元記事を読む (NVIDIA Developer Blog) ↗
  • arXiv cs.LG (Machine Learning) · EN 学習・ファインチューニング
    MoPET: Parameter-Efficient Mixture-of-Experts for Unified Medical Image Classification
    深層学習 ファインチューニング Mixture of Experts (MoE) 検索拡張生成 (RAG)
    元記事を読む (arXiv cs.LG (Machine Learning)) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN マルチモーダル
    QR-Structured Thermal Triggers for Targeted Semantic Attacks on Infrared Vision-Language Models
    コンピュータビジョン 深層学習 ソフトウェア工学
    元記事を読む (arXiv cs.AI (Artificial Intelligence)) ↗
  • Data Center Dynamics · EN インフラ・ハードウェア 抜粋
    Why ‘next wave’ data center markets are at the heart of Europe's fight for data sovereignty
    「次の波」のデータセンター市場が欧州のデータ主権争いの中心に
    データセンターダイナミクスは、欧州のデータ主権を巡る競争において、成長途上の「次の波」のデータセンター市場が重要な位置を占めていると論じた。域内でのデータ管理を巡る政策や投資の動きを背景に、新興市場の役割を分析している。
    元記事を読む (Data Center Dynamics) ↗
  • Data Center Dynamics · EN インフラ・ハードウェア 抜粋
    Musk confirms fourth SpaceXAI data center in Memphis, company starts removing 'illegal' gas turbines
    Musk、Memphisに4カ所目のxAIデータセンターを確認、違法タービン撤去へ
    データセンターダイナミクスによると、イーロン・マスク氏はMemphisに4カ所目となるxAI系データセンターの建設を認めた。あわせて無許可(違法)とされたガスタービンの撤去を開始したという。急速なAI計算基盤拡大に伴う環境・許認可の課題が浮き彫りになっている。
    元記事を読む (Data Center Dynamics) ↗
  • Data Center Dynamics · EN インフラ・ハードウェア 抜粋
    Amazon, Duke Energy accused of evading Clean Air Act at under-development data center in Hamlet, North Carolina
    Amazonとデューク・エナジー、建設中DCで大気浄化法違反の疑い
    機械学習
    データセンターダイナミクスによると、米ノースカロライナ州ハムレットで建設中のデータセンターを巡り、Amazonとデューク・エナジーが大気浄化法(Clean Air Act)を回避したとして告発された。合計649基のディーゼル発電機の設置申請が問題視されている。
    元記事を読む (Data Center Dynamics) ↗
  • Data Center Dynamics · EN インフラ・ハードウェア 抜粋
    CenterPoint Energy raises investment plan by $1.2bn as data center load pipeline grows
    CenterPoint Energy、データセンター需要増で投資計画を12億ドル上積み
    CenterPoint Energyは、データセンターの電力需要パイプライン拡大を受けて投資計画を12億ドル引き上げた。2029年までにヒューストン都市圏で最大8GWのデータセンター負荷を受電できるようにする見込みだという。
    元記事を読む (Data Center Dynamics) ↗
  • Data Center Dynamics · EN インフラ・ハードウェア 抜粋
    GCM expands range of high-performance heat sinks to cater for liquid-cooled data centers
    GCM、液冷データセンター向け高性能ヒートシンクの製品群を拡充
    GCMは、液冷(liquid-cooled)データセンター向けに高性能ヒートシンクの製品ラインアップを拡充した。AIサーバーの高発熱化に伴う冷却需要の高まりに対応し、効率的な熱管理を支える製品を提供する。
    元記事を読む (Data Center Dynamics) ↗
  • arXiv cs.AI (Artificial Intelligence) · EN インフラ・ハードウェア
    Beyond Component Testing: Validating Agentic AI Systems
    ニューラルネットワーク 検索拡張生成 (RAG)
    元記事を読む (arXiv cs.AI (Artificial Intelligence)) ↗
  • arXiv cs.LG (Machine Learning) · EN 推論・効率化
    OnlineCache: Learning Dynamic Caching Policies with Error Correction for Efficient Diffusion Inference
    推論 (Inference) 検索拡張生成 (RAG) 強化学習
    元記事を読む (arXiv cs.LG (Machine Learning)) ↗
  • arXiv cs.CL (Computation and Language) · EN 推論・効率化
    Studying quantization trade-offs for efficient inference deployment in machine translation
    深層学習 推論 (Inference) 量子化
    元記事を読む (arXiv cs.CL (Computation and Language)) ↗
  • arXiv cs.CL (Computation and Language) · EN 学習・ファインチューニング
    PTP: Previous-Token Prediction based LLM Inversion for Near-Exact Prompt Reconstruction
    ファインチューニング
    元記事を読む (arXiv cs.CL (Computation and Language)) ↗