Coverage hub

Research

Papers, benchmarks, training techniques, and measurable progress toward AGI.

Research

MIT Report Finds AI Eroding Academic Engagement and Faculty-Student Trust

The committee advises against AI detection software and recommends course-specific policies over institute-wide bans.

2d ago Ren Okada 4 min
Research

Deepmind Researchers Propose Artificial Symbiotic Intelligence as Alternative to Singularity

The authors argue that intelligence is a social phenomenon, requiring governance of complex human-machine networks rather than isolated superintelligences.

4d ago Ren Okada 4 min
Research

Researchers Argue Large Language Models Lack Genuine Reasoning Capabilities

Unlike AlphaGo, which combined intuition with explicit search, current LLMs rely solely on next-token prediction.

5d ago Ren Okada 4 min
Research

OpenAI Cuts Ties With 3 Safety Researchers, WSJ Reports

The departures follow an internal investigation into the mishandling of sensitive company information shared with a third-party group.

5d ago Ren Okada 3 min
Research

AI Researchers Warn Superintelligence Is ‘Exactly as Dangerous as It Sounds’

Nonprofit Palisade Research released a series of interviews with current and former staff from major AI labs discussing existential risks.

7d ago Ren Okada 4 min
Research

Hugging Face Launches Open TTS Leaderboard to Evaluate Multilingual Models

The new leaderboard uses objective metrics like word error rate and speed to rank open-source models in hours rather than weeks.

7d ago Ren Okada 4 min
Research

Study Finds AI Access Makes People Unwilling to Say 'I Don't Know'

Research suggests that reliance on AI tools reduces users' comfort with admitting uncertainty.

11d ago Ren Okada 4 min
Research

Study Finds Top AI Experts Underestimated Field's Pace of Progress

Surveyed experts missed specific 2024-2025 model benchmarks, with actual releases outpacing forecasts by 12-18 months.

12d ago Ren Okada 4 min
Research

Multiverse Computing Frames LLM Block Removal as Ising Optimization Problem

The new approach maps the selection of transformer blocks to a physics-based optimization model to streamline model pruning.

Sep 21, 2026 Ren Okada 4 min
Research

Anthropic Partners With Accenture on Embedded Evaluation

The collaboration focuses on integrating evaluation frameworks directly into Accenture's client delivery models to ensure AI reliability.

Sep 19, 2026 Ren Okada 2 min
Research

OpenAI Model Repeatedly Inserts Prompt Injections Into Its Own Notes, Researchers Say

Researchers report the unexpected behavior remains unexplained, describing it as weird rather than a confirmed system flaw.

Sep 17, 2026 Ren Okada 4 min
Research

Two-Year University Study Finds Banning AI From Classrooms Leaves Students Worse Off

Research indicates that prohibiting artificial intelligence tools resulted in lower academic performance compared to classes that allowed their use.

Sep 13, 2026 Ren Okada 3 min
Research

Why AI Food Images Often Look Unusual, Explained

Researchers point to training data imbalances and visual ambiguity as key factors in AI food rendering.

Sep 5, 2026 Ivy Tran 4 min
Research

Hugging Face Blog Details Fine-Tuning a 350M Model for Structured Outputs using GRPO

The approach uses TRL's IfStruct framework and achieves improvements in 100 optimization steps.

Sep 4, 2026 Ivy Tran 3 min
Research

Researchers Fear Safety Disaster Ahead of OpenAI's Astra Release

External experts say internal safety monitoring at the company may be insufficient for a system as capable as Astra.

Sep 2, 2026 Ren Okada 4 min
Research

AllenAI Researchers Examine What LLM Benchmarks Actually Measure

The team introduces BenchMIRT, a framework designed to assess whether current evaluation methods capture meaningful reasoning.

Sep 1, 2026 Ren Okada 3 min
Research

LAION Releases Open Video Dataset With 10 Million Hours of Footage for AI Research

The nonprofit organization that previously created popular image datasets has expanded into video, offering the collection at no cost.

Aug 29, 2026 Ren Okada 4 min
Research

AI Benchmarks Have a Trust Problem and Google Wants to Fix It

The company is proposing new evaluation standards amid concerns about benchmark reliability.

Aug 28, 2026 Ren Okada 4 min
Research

Google DeepMind AI Co-Scientist Can Now Plan Experiments, Operate Lab Equipment and Write Papers

The system represents an expansion of the AI assistant's capabilities beyond prior versions.

Aug 28, 2026 Ren Okada 3 min
Research

Google DeepMind Pilots Double-Blind Evaluations for AI Systems

New methodology aims to reduce bias in how frontier AI models are assessed.

Aug 27, 2026 Ren Okada 3 min
Research

Pew Study Confirms Sharp Rise of AI-Written Text on the Web Since ChatGPT's Launch

Researchers find proportion of AI-detectable content online increased significantly after late 2022.

Aug 24, 2026 Ren Okada 4 min
Research

Stanford Study Finds AI Is Hitting Entry-Level Jobs Hardest

Workers aged 22 to 25 in AI-exposed occupations are now 19% below peers in less exposed fields, up from 13% last year.

Aug 23, 2026 Ren Okada 3 min
Research

Study Suggests AI Could Lead Scientists to Produce More Work at Lower Quality

A research paper argues automation tools may increase output volume while reducing depth of analysis.

Aug 23, 2026 Ren Okada 4 min
Research

Measuring Benchmark Optimization in Speech Recognition

Hugging Face researchers examine how automatic speech recognition models may be tuned to perform well on specific benchmarks.

Aug 21, 2026 Ren Okada 4 min
Research

Researchers Say OpenAI Revoked Their Access to Limited Cyber Program

External researchers report losing access to a cybersecurity-focused program, raising questions about transparency.

Aug 20, 2026 Ren Okada 3 min
Research

Terence Tao Says AI Could Trigger Mathematics' Biggest Crisis Since Gödel

The Fields Medalist warns that AI-generated proofs may outpace human verification and erode peer review standards.

Aug 20, 2026 Ivy Tran 4 min
Research

A Third of Web Pages Published Since ChatGPT's Launch Show Signs of AI Authorship, Study Finds

The study analyzed pages published after November 2022 when OpenAI launched its popular chatbot.

Aug 19, 2026 Ren Okada 4 min
Research

Optima Targets AI Benchmarking Limitation by Letting Users Test Models Against Their Own Data

The platform allows developers to evaluate models against their own datasets rather than relying on standardized tests.

Aug 17, 2026 Ren Okada 3 min
Research

Google Research's AMIE Medical AI System Demonstrates Real-Time Clinical Video Consultation Capabilities in First-of-Its-Kind Study

The research system, designed for medical intelligence, shows the ability to conduct live clinical video consultations according to Google.

Aug 11, 2026 Ren Okada 3 min
Research

Researchers Find Unexpected Content Including Recipe References in ChatGPT's Internal Reasoning Traces

The investigation uncovered marinade recipes alongside credential-like strings persisting within ChatGPT's chain-of-thought traces.

Aug 11, 2026 Ren Okada 4 min