AI News — Sunday, July 19, 2026
A new unified infrastructure is proposed for comprehensively evaluating the capabilities of AI agents, addressing a critical need in agent development.
Researchers introduce a new benchmark and methods to improve policy-adaptive image guardrails, enhancing AI safety and alignment with ethical guidelines.
A new report reveals that over half of enterprises have experienced security incidents involving AI agents, largely due to agents sharing credentials, highlighting a critical vulnerability.
Enterprises are struggling with evaluating AI agents effectively, leading to a disconnect between perceived and actual performance, yet many are still deploying them to production.
TechCrunch explores the potential impact and implications of the new AI entity 'Kimi,' questioning its role in the evolving AI landscape.
OpenAI introduces a new framework or 'scorecard' to assess progress and challenges in the development and deployment of AI technologies.
A new evaluation benchmark, KeyFrame-Compass, is introduced to thoroughly assess the quality and consistency of video generation models conditioned on keyframes.
Prominent investor Neil Rimer suggests that the significant capital influx into the AI sector may soon begin to recede, signaling a potential shift in investment trends.
Many enterprises are mislabeling chatbots as advanced AI agents and face significant challenges in orchestrating and deploying true agentic systems, highlighting a gap in understanding and implementation.
This article highlights how inefficient PDF processing can consume excessive tokens in Large Language Models, impacting performance and cost.
Researchers present MetaView, a new method for generating novel views from a single image by incorporating scale-aware implicit geometry priors.
MultiRef-Compass is introduced as a new benchmark for thoroughly evaluating AI models that generate audio-video content from multiple references.
This survey provides a comprehensive overview of current techniques and challenges in enabling self-improvement capabilities within modern AI agentic systems.
Cars24 shares insights into how they leverage OpenAI's technologies to scale customer conversations and accelerate their development processes.
RxBrain introduces an embodied cognition foundation model capable of joint language-visual reasoning and imaginative capabilities, pushing boundaries in multimodal AI.