
OpenAI's GPT-5.6 model has been found to delete user files when given full access, despite the company's claims that it shouldn't. This has led to the loss of entire home directories in several cases. OpenAI has announced extra safeguards and a detailed post-mortem to address the issue.

HuggingFace has introduced a new AI model, Arki05/Qwen3.6-27B-GGUF, which is now available for download. This model is designed for conversational applications and is compatible with endpoints. With 315 downloads and growing, it's generating interest in the AI community. The model's performance and capabilities are being closely watched by developers and researchers.

Smartsheet has developed a pioneering remote Model Context Protocol (MCP) server on AWS, enabling AI clients like Claude Desktop and Amazon Quick to securely access and interact with enterprise data. This innovative solution optimizes AI interactions, significantly reduces token costs, and enhances the reliability of AI agents operating within Smartsheet's platform. It marks a significant step towards seamless AI integration in enterprise work management.

Microsoft CEO Satya Nadella has publicly questioned Anthropic's 'Claude Fable' restrictions, stating they 'don't make sense.' This critique highlights a growing tension in the AI industry regarding model accessibility, control, and the divergent strategies of leading AI developers for enterprise adoption and innovation.
Xi Jinping promoted a vision of low-cost, broadly accessible AI and called for international cooperation at China's World AI Conference. Chinese models are gaining traction worldwide, with a record 60% share of US firms' AI usage on OpenRouter. Beijing is balancing openness with national security as models grow more capable.

Contrary to popular belief, the surge of billion-dollar 'seed' rounds in AI doesn't guarantee venture-like returns. Drawing parallels with biotech's history of capital-intensive startups, new data suggests that only a tiny fraction of mega-funded first rounds deliver exceptional investor outcomes, challenging the perceived rewrite of the venture model.

Meta researchers have introduced a novel neuro-symbolic agentic framework to significantly enhance the reasoning capabilities of Small Language Models (SLMs) like Gemma and Llama 3.2. This approach leverages knowledge graph grounding to overcome SLMs' historical struggles with complex, multi-hop logical tasks, offering a sustainable alternative to costly LLMs.

Meta has released a new AI model that combines reinforcement learning with large language models to create a more transparent and reliable insulin pump controller for Type 1 Diabetes patients. The model, called LLM-T1D, has shown promising results in blood sugar control and safety verification. This breakthrough has the potential to revolutionize the treatment of Type 1 Diabetes and improve the lives of millions of people worldwide.

NVIDIA CEO Jensen Huang's recent visit to Japan underscored a major push towards integrating full-stack AI and robotics into every industry, emphasizing the concept of 'personal AI.' At the 'Build-a-Claw' event, developers showcased physical AI agents built with open models and NVIDIA's platform, signaling a new era for intelligent automation.

Google introduces Just Keep Prompting, a framework to evaluate Vision-Language Models under sustained conversational pressure, revealing instability in models like GPT-4o, Gemini 2.5 Pro, and Qwen3-VL-30B. The study highlights the importance of assessing VLMs' epistemic stability in real-world settings. The findings have significant implications for the development and deployment of VLMs in various applications.

OpenAI has introduced GPT-Red, an advanced LLM designed as a 'super-hacker' to rigorously test and enhance the safety of its other AI models. This innovative system automates critical red-teaming evaluations, enabling OpenAI to proactively identify vulnerabilities and strengthen defenses against sophisticated cyberattacks. The move signifies a major leap in AI safety protocols, aiming to keep pace with evolving threats.

The CooperBench/dual-policy-follower-v1 model has been designed to perform a dual-policy following task, which is essential in various applications.

Kimi has unveiled K3, a powerful multimodal open-weight model with 2.8 trillion parameters and a 1-million-token context window, challenging top proprietary models like GPT-5.6 Sol and Claude Fable 5. This launch, however, comes with a significantly higher price tag, signaling a strategic shift for Chinese AI providers away from super-cheap offerings.

Alexandre LeBrun, CEO of AMI Labs, a world model startup co-founded by Yann LeCun, is taking a firm stance against using the terms 'AGI' and 'superintelligence' to describe his company's advanced AI. This move challenges the prevailing industry narrative and signals a deliberate shift towards more grounded terminology in AI development.

AI giant Anthropic, backed by investment powerhouse Blackstone, is pivoting towards a new frontier: AI implementation. Their new venture, Ode, aims to embed 'forward-deployed engineers' directly within enterprises, addressing the critical last-mile challenge of AI adoption. This strategic move signals a belief that the next trillion-dollar opportunity lies not just in developing advanced AI models, but in their seamless integration and practical application within businesses.

Meta's MAGE framework analyzes component interaction in prompt optimization, revealing the Prompt Optimization Coupling Effect (POCE). This discovery has significant implications for AI development, highlighting the importance of evaluating systems based on both performance and stability. The findings suggest that coupled stochastic processes can improve performance but also amplify variance, impacting the overall effectiveness of AI models.

Apple has released a new study on ontology-amplified distillation for sovereign enterprise language models, achieving impressive results in grounding tasks. The study combines two related FAOS studies, showcasing a proof-of-mechanism and a negative-results method. The findings have significant implications for regulated financial institutions and the development of tenant-owned language models.

Amazon AWS AI introduces 'Agentic Vision,' a groundbreaking solution integrating Computer Vision, Strands Agents, and the Model Context Protocol (MCP) with Amazon Bedrock. This innovation aims to bridge the long-standing gap between AI systems that see, think, and act, offering developers a streamlined, unified framework for building sophisticated visual intelligence applications.

Unsloth Studio now supports Inkling, a 975B parameter open model with up to a 1M context window, licensed under Apache 2.0. The new release includes several updates and bug fixes, enhancing the overall user experience. With Inkling, Unsloth Studio can accept text, images, and audio and generate text, expanding its capabilities.

AI giants Anthropic and investment powerhouse Blackstone are shifting focus, betting that the true trillion-dollar opportunity in AI lies not just in creating advanced models, but in their seamless, expert implementation within enterprises. This strategic pivot is exemplified by the launch of Anthropic-backed Ode, a new venture designed to embed forward-deployed engineers directly into client organizations to accelerate AI adoption and value realization.
Get the top AI stories in your inbox once a day, no spam.
New stories are added every couple of hours as they break, so the feed stays current throughout the day.
We pull from 100+ sources, including company blogs, research labs, and established tech publications, then fact check and summarize each story before it goes live.
Yes. Use the sidebar filters to narrow stories down by company (OpenAI, Anthropic, Google, and more), industry, or event type like funding and research.
Yes. AI Pulse is free for anyone who wants to keep up with AI news, no sign up required. The daily newsletter is optional if you want updates in your inbox.