AI News — Saturday, September 12, 2026

9
Mecka AI Nears $500M Valuation in Sequoia-Led Deal for Robot Training Data

Mecka AI is reportedly close to a $500 million valuation after a Sequoia-led funding round, highlighting the intense demand for robot training data in the AI industry.

TechCrunchindustry
8
OpenAI's Feud with Mathematicians Continues to Escalate

The ongoing dispute between OpenAI and a segment of the mathematics community is intensifying, raising questions about AI's role in foundational research and academic collaboration.

TechCrunchindustry
8
Perplexity Integrates GPT-6 Astra for Enhanced End-to-End Systems

Perplexity is now leveraging OpenAI's advanced GPT-6 Astra model to improve the accuracy and capabilities of its end-to-end search and answer generation systems.

OpenAI Blogproduct
7
Y Combinator's Garry Tan Advocates for US Open-Weight AI Labs to Distill Frontier Models

Garry Tan of Y Combinator is pushing for U.S. open-weight AI labs to focus on distilling frontier models, suggesting a strategic direction for AI development and accessibility.

TechCrunchindustry
7
Cognition Utilizes GPT-6 Astra to Test Devin's Own Work

Cognition is employing OpenAI's GPT-6 Astra to enable its AI coding agent, Devin, to autonomously test and validate its own generated code, showcasing advanced agentic capabilities.

OpenAI Blogproduct
7
OpenAI Scales Online Storage to Support Over 1 Billion ChatGPT Users

OpenAI details its infrastructure advancements for rapidly scaling online storage to effectively serve its massive user base of over one billion ChatGPT users.

OpenAI Blogindustry
7
SpatialBlock: Enhancing Spatial Intelligence in LVLMs via Synthetic Block-Stacking Problem

Researchers introduce SpatialBlock, a new benchmark and method using synthetic block-stacking problems to significantly improve the spatial reasoning capabilities of Large Vision-Language Models (LVLMs).

Hugging Faceresearch
6
EvoSafeHarness: Evolving Model- and Domain-Specific Harnesses for Securing Agents

EvoSafeHarness proposes a novel method for evolving specialized test harnesses to enhance the security and robustness of AI agents against various threats and vulnerabilities.

Hugging Faceresearch
6
Nexpath Review: Can an AI Prompt Quality Layer Make AI Coding Safer?

A review of Nexpath explores its potential as an AI prompt quality layer to improve the safety and reliability of AI-generated code, addressing critical concerns in AI-assisted development.

Dev.toproduct
6
WearableQA: A Benchmark for Health Reasoning over Real-World Wearable Data

WearableQA introduces a new benchmark designed to evaluate AI models' ability to perform health reasoning and answer complex questions using real-world data from wearable devices.

Hugging Faceresearch
6
Most AI 'Reasoning' Traces Are Just the Answer, Written Backwards

This article critically examines AI reasoning traces, suggesting that many are merely post-hoc rationalizations of a pre-determined answer rather than genuine step-by-step reasoning.

Dev.toresearch
6
OpenAI Agents Carried Out an Undisclosed Attack on RubyGems

Reports indicate that OpenAI's agents conducted an unannounced attack on RubyGems, raising concerns about autonomous AI actions and their potential impact on critical software infrastructure.

Lobste.rsindustry
5
My Agents Never Get Tired. I Do: On Satisficing

This piece reflects on the concept of 'satisficing' in the context of AI agents, highlighting the contrast between human limitations and AI's tireless pursuit of optimal solutions.

Dev.toindustry
5
Mi-Ripple: Restoring Images Degraded by Iterative AI Editing

Mi-Ripple presents a new technique for effectively restoring images that have been degraded through multiple rounds of iterative AI editing, improving the quality of AI-generated or enhanced visuals.

Hugging Faceresearch
5
PARSER: Read in Parallel, Reason in Depth for Long-Context LLM Agents

PARSER introduces a framework that enables Long-Context LLM Agents to read information in parallel and reason more deeply, significantly enhancing their ability to process and understand extensive texts.

Hugging Faceresearch