AI News — Monday, September 21, 2026
This research introduces a novel method for agentic reinforcement learning that allows on-policy distillation to self-retire, improving efficiency and performance.
OpenAI leveraged its own large language models to assist in the design process of a new hardware chip, showcasing a practical application of AI in engineering.
This article discusses the increasing trend of secrecy among leading AI model development companies, raising concerns about transparency and open research.
An analysis questions the industry's willingness to decelerate AI development amidst calls for caution and regulation, highlighting the ongoing race for innovation.
New research presents WeVisDoc, a method aimed at enhancing the robustness and capability of end-to-end document parsing systems.
VABench is introduced as a new benchmark designed to evaluate embodied spatial intelligence in AI agents using visual demonstrations and active perception.
Laya is announced as a new multilingual decision engine boasting an impressive 33-millisecond response time, indicating advancements in real-time AI processing.
This paper explores CodeMidas, a framework for scaling reinforcement learning environments for agentic coding directly from existing codebases.
Researchers propose EvoOntology, a self-evolving ontology layer designed to enhance the knowledge representation and adaptability of data agents.
This research introduces a method for hybrid reasoning models to learn difficulty-aware length control, optimizing efficiency without sacrificing accuracy.
UFO presents a novel chain-of-evaluation approach to achieve omni-condition alignment, significantly improving multi-modal image generation.
A study reveals that the candidate-generation strategy, not just sample count, critically influences the energy consumption and performance of LLM scaling during testing.
This paper investigates how observation supervision fundamentally alters exploration strategies for agents in reinforcement learning environments.
OpenAI has unveiled the Australian Youth Safety Blueprint, outlining strategies and commitments to ensure safer AI interactions for young people.
This article provides guidance on building robust and secure DevSecOps pipelines specifically tailored for the deployment and management of enterprise AI agents.