A study comparing whether ChatGPT or Google search yields better learning outcomes drew attention. The signal mixes ITmedia coverage with community reaction on lobste.rs and HN and related arXiv work—no official announcement, driven by press and field discussion, so verification and reception led rather than a launch. It measures comprehension and retention when generative AI is used as a study tool, in a controlled contrast with conventional search; the through-line is evaluating AI in education by evidence rather than impression. But it's a single study—variation by subject and conditions, and whether it generalizes, await replication.
Study compares ChatGPT vs Google
Study compares ChatGPT vs Google
ChatGPT vs. Google検索──どっちで調べるのが学習効果が高い? 8日間の実験で検証した研究
Study: does ChatGPT or Google search aid learning more? An 8-day test
GPT‑NL: a sovereign language model for the Netherlands
GPT-NL: a sovereign language model for the Netherlands
OpenAIの高度AIでソフトバンクの脆弱性を1万件発見 孫正義氏「大変な危機」 日本の重要インフラ企業へ診断サービス提供
SoftBank unveils OpenAI-powered Patching-as-a-Service security offering
CrankGPT — Local Human-powered AI
CrankGPT pitches a tongue-in-cheek local, human-powered 'AI'
Academic (arxiv etc.) 16 ▾
ReproRepo: Scaling Reproducibility Audits with GitHub Repository Issues
ReproRepo scales reproducibility audits using GitHub repo issues
RubricsTree: scalable open-ended evaluation of personal health agents
An agentic benchmark for implicit animal welfare in frontier AI
Structural role injection in Handlebars-templated LLM prompts
Querying an astronomical database using large language models: the ALeRCE text-to-SQL system
A text-to-SQL system for querying the ALeRCE astronomical database
Security and Privacy Prompts in the Wild: What Users Ask LLMs and How LLMs Respond
Security and privacy prompts in the wild: what users ask LLMs
Synthetic lived experience in AI peer-like caregiver support
Toward Accessible Psychotherapy Training Using AI-Driven Interactive Patient Avatars
AI-driven patient avatars for more accessible psychotherapy training
From Trainee to Trainer: LLM-Designed Training Environment for RL with Multi-Agent Reasoning
From trainee to trainer: LLM-designed RL training environments
Binary Tracking for Spatial QA and Navigation with Open Vision-Language Models
Binary Tracking: open vision-language models for spatial QA and navigation
Reasoning hop-count predicts clinical AI failure in EHR QA
Paper: framework measures LLM search-agent endorsement risk
LLM-based Visual Code Completion for Aerospace Geometric Design
Paper: LLM visual-programming copilot for aerospace design
Multimodal Evaluator Preference Collapse: Cross-Modal Contagion in Self-Evolving Agents
Paper on evaluator preference collapse in self-evolving agents
Paper frames LLM sycophancy as material failure (title only)
BD-LSC: a new benchmark dataset for lexical semantic change detection