New Model Releases A
Showing 1–30 of 330
-
Rust製のフルスタックWebアプリフレームワーク「Topcoat」登場。非同期ランタイム「Tokio」上でサーバサイドレンダリング、ルーティング、コンポーネントライブラリなど
-
クラウドインフラのシェア、AWSが28%と変わらず、Google Cloudは1ポイント上昇して15%に。市場全体が年42%と過去最高の成長率に。2026年第2四半期、Synergy Research
-
「Qwen3.8-Max」登場、オープン化は「来週」 一部「Fable 5」「GPT-5.6 Sol」超えの性能うたうAlibaba Cloud releases Qwen3.8-Max; open weights due next weekAlibaba Cloud, part of China's Alibaba, released its large AI model Qwen3.8-Max on Aug 3, claiming it beats Fable 5 and GPT-5.6 Sol on some benchmarks. The company says it will publish the model weights next week, continuing its open-weight strategy.
-
condense-json 1.0Simon Willison ships condense-json 1.0 for compact JSONSimon Willison released condense-json 1.0, a small Python library that shrinks JSON by replacing repeated strings with a short replacements map (e.g. mapping a key to a recurring phrase). Now a year and a half old, it graduates to a stable 1.0 with sensible, non-disruptive fixes.
-
OpenAI、次期主力モデル「Astra」の存在を明らかに――未解決の数学問題10件を「解決」と発表OpenAI unveils 'Astra,' says it solved 10 open math problemsOpenAI disclosed 'Astra,' its next flagship model, saying an internal version produced new results on 10 unsolved problems in mathematics and theoretical computer science—the first public use of the name. Compute cost was held to about $2,000 in 'Sol' terms, and formal proofs via the Lean proof assistant were published on GitHub, signaling potential for advanced math research.
-
アトラシアン、AI時代の仕事用ブラウザ「Diaブラウザ」のWindows版リリースへ。ウェイトリストへの登録開始Atlassian brings its AI work browser Dia to Windows this fallAtlassian announced that Dia, its AI-era productivity browser currently available only on Mac, will get a Windows version this fall. A waitlist is now open, extending availability to Windows users.
-
日本におけるクラウドネイティブコミュニティの開発者数が約100万人に、CNCFが調査結果を発表CNCF: Japan's cloud-native developer community nears 1 millionThe Cloud Native Computing Foundation and analyst firm SlashData released survey findings estimating that Japan's cloud-native developer community has reached roughly 1 million, underscoring the growing domestic adoption of Kubernetes and related cloud-native technologies.
-
Sakana AI、日本語特化のLLM API「Sakana Namazu」を提供開始Sakana AI launches Namazu, a Japanese-focused OpenAI-compatible LLM APISakana AI released Namazu, an LLM API tuned for Japanese and local business use. Built on Moonshot AI's open Kimi K2.6 and refined with in-house data, it adds built-in web search and code execution. Being OpenAI-compatible, existing code works by swapping the base_url, filling the gap between costly frontier models and raw open ones.
-
July 2026 newsletterSimon Willison publishes his latest monthly newsletterDeveloper Simon Willison released the latest edition of his sponsors-only monthly newsletter. It rounds up recent developments across AI models and tooling—spanning GPT, Claude, DeepSeek, Anthropic, and MCP—offering an individual's closely watched view of the fast-moving AI landscape.
-
Google、パーソナルAI「Gemini Spark」を日本でも利用可能に Chrome統合は米国からGoogle expands Gemini Spark personal AI to 160+ countries incl. JapanGoogle extended its Gemini Spark personal AI agent to more than 160 countries, including Japan. Running on Google's cloud, it can act even when a PC is off or a phone is locked, handling tasks based on triggers. Chrome integration will roll out first in the US.
-
OpenAI、アクティブユーザー10億人超に 導入企業は200万社超OpenAI passes 1 billion active users and 2 million business customersOpenAI said it surpassed one billion active users and two million business customers. It cited efficiency gains from retained reasoning, better context management, and production optimization that cut costs and improved token throughput, alongside price cuts on some GPT-5.6 models.
-
datasette-apps 0.2a0Simon Willison releases datasette-apps 0.2a0Simon Willison released datasette-apps 0.2a0, an update to the Datasette extension. The release includes changes that improve building and editing Datasette Apps directly within Datasette, continuing steady development of the open-source data-exploration ecosystem.
-
Ten advances in mathematics and theoretical computer scienceSimon Willison weighs in on OpenAI and Anthropic's math resultsSimon Willison discussed OpenAI's 'ten advances in mathematics and theoretical computer science,' noting that just days earlier Anthropic had reported similar discoveries. The post reflects on a growing trend of frontier AI models contributing to open problems in mathematics.
-
Stateless MCP has recaptured my interest (and inspired mcp-explorer and datasette-mcp)Simon Willison: stateless MCP (MCP 2.0) has recaptured my interestSimon Willison wrote that the rollout of stateless MCP—the MCP 2.0 or 2026-07-28 Model Context Protocol specification—has renewed his interest in the protocol. He says it inspired him to build tools such as mcp-explorer and datasette-mcp on top of the new stateless design.
-
llm-mcp-client 0.1a0Simon Willison releases llm-mcp-client 0.1a0Simon Willison released llm-mcp-client 0.1a0, a tool for connecting his LLM utility to Model Context Protocol (MCP) servers. Detailed in an accompanying blog post, the release adds to the growing set of tooling built around the MCP ecosystem.
-
smevals - a small eval suite for evaluating models, prompts, and harnessesSimon Willison introduces 'smevals,' a small eval suite for modelsSimon Willison introduced smevals, a small evaluation suite for testing models, prompts, and harnesses. Built in collaboration with Jesse Vincent's Prime Radiant applied AI research lab, the framework aims to help answer questions about the capabilities of different AI models.
-
GQ-FSL: Green Quantized Federated Split Learning
-
Evolving language compositionality in a frequency-structured meaning space
-
AgentHPOBench: A Benchmark For Evaluating LLM Agents as Sequential Hyperparameter Optimizers
-
The Theoretical Foundation of Socratic Tests: Dynamic, Multimodal, Conversational Examinations
-
When Does On-Policy Interaction Help? Representational Tradeoffs in Value-Based Imitation Learning
-
QASP: Query-Adaptive Robust Vector Search Policy
-
FriendBench: Benchmarking Dyadic Familiarity Inference in Humans and Multimodal Large Language Models
-
TraceViT: Grounded Trace Supervision for Visual Abstract Reasoning
-
DungeonBench: A Benchmark for Rules-Rich Tactical Reasoning in Dungeons & Dragons Combat
-
AMTFV: Agentic Mathematical Tool-Flow Verification for LLM Self-Correction
-
ARB: A Matched Authorship-Rewriting Benchmark Dataset for AI-Text Detector Evaluation
-
A Neurosymbolic Approach for Explainable Early Diagnosis of Alzheimer's Disease
-
TerraNova: A Foundation Model for the Anthropocene
-
From Code Review to Code Critique: Intent, Drift, and Spotlight for AI-Generated Diffs at Scale