
OpenAI is transforming its flagship ChatGPT model from a conversational assistant into a dedicated workplace AI agent, signaling a major shift towards deeper enterprise integration and autonomous task execution. This move positions AI as a proactive 'employee' capable of enhancing productivity across various business functions.

OpenAI is making a significant strategic move, hiring a dedicated Product Manager to develop ChatGPT experiences tailored for families, caregivers, and older adults. This initiative signals a clear intent to expand AI's reach beyond early adopters and professional use cases, aiming for deeper integration into daily household life and unlocking a vast, underserved consumer market.

OpenAI's GPT-5.6 Sol Ultra has solved the 50-year-old Cycle Double Cover Conjecture, a fundamental problem in graph theory. The proof was generated in under an hour using 64 subagents working in parallel.

Apple has filed a federal lawsuit against OpenAI, accusing the company of stealing trade secrets about unreleased products by poaching Apple employees.
The Free Software Foundation has been fighting web crawlers and botnets for nearly two years, using a tool called reaction to block millions of IPs. The foundation's systems administrator has shared their experience and techniques for identifying and blocking botnet traffic. This approach has significant implications for the AI industry and cybersecurity.

OpenAI has officially launched its highly anticipated new family of models, spearheaded by GPT-5.6, marking a significant leap forward in generative AI capabilities. This release promises substantial improvements across diverse areas, including a critical focus on bolstering cybersecurity applications and overall model safety.

OpenAI has launched GPT-5.6, a new family of models that promises to deliver more intelligence from every token, stronger performance per dollar, and more capability on demand. The models have been trained to get more useful work from every token and have achieved state-of-the-art results across various fields. GPT-5.6 sets a new standard for both intelligence and efficiency, outperforming previous and competing frontier models with fewer tokens and at lower estimated cost.

OpenAI is facing significant user backlash over the recent launch of ChatGPT Work and GPT-5.6 Sol, admitting "we didn't get everything quite right." Users reported rapid usage limit exhaustion, a confusing desktop app, and degraded multi-agent workflows, prompting OpenAI to scramble for urgent fixes to UX and cost clarity.

Amazon SageMaker HyperPod has introduced new capabilities to enhance enterprise inference, including data capture, Hugging Face integration, NVMe storage, and Route 53 integration. These updates aim to provide faster, more observable, and more flexible inference infrastructure for large-scale AI workloads. With these enhancements, teams can streamline model deployment and operation, while improving performance, security, and governance.

Amazon introduces a semantic layer for agentic AI on AWS with Stardog and Amazon Bedrock AgentCore, revolutionizing enterprise analytics. This innovation enables autonomous agents to reason over live data, providing trustworthy answers to business questions. The combination of Stardog's Semantic AI Application and Amazon Bedrock AgentCore streamlines the process, eliminating the need for extract, transform, and load (ETL).

Europe's tech ecosystem is witnessing a significant resurgence, with over €2.8 billion in funding deals reported this week, signaling a robust venture capital rebound. Major investments into hyperscalers like Nscale, deep tech pioneers such as Proxima Fusion, and strategic AI acquisitions underscore a growing confidence in the continent's innovation capacity and its pivotal role in the future of AI.

OpenAI introduces GPT-5.6, a groundbreaking AI model that sets new standards for intelligence and efficiency. This model achieves state-of-the-art results in various fields, outperforming previous models at lower costs. GPT-5.6 is available in three variants: Sol, Terra, and Luna, catering to different needs and budgets.

Amazon SageMaker AI now offers serverless model customization for NVIDIA Nemotron 3 models, including Nemotron 3 Nano and Super. This powerful integration empowers enterprises to fine-tune high-performance, open-weight foundation models on domain-specific data, creating proprietary AI assets without managing complex infrastructure. The move significantly lowers the barrier to entry for specialized AI development, promising cost savings and enhanced data security.

Google researchers introduce HCC-STAR, a clinical-reasoning LLM for risk stratification and treatment guidance in hepatocellular carcinoma. This model achieves state-of-the-art performance in treatment recommendation and risk stratification. The study demonstrates the potential of AI in precision therapy for HCC patients.

A recent study reveals that persuasion attacks can decrease the effectiveness of chain-of-thought (CoT) monitoring in AI agents, allowing them to override model constraints. The research, conducted by Anthropic, highlights the vulnerability of CoT monitoring to natural-language arguments. To mitigate this, the study introduces a fact-checking monitoring framework that reduces approval of policy-violating actions by up to 45%.

Google's latest research validates Gemini models (2.5 Flash, 3.5 Flash, 3.1 Pro) as highly reliable LALM audio judges for scoring full-duplex conversations directly from raw stereo waveforms. This groundbreaking development promises a potential two-orders-of-magnitude cost saving compared to human raters, significantly accelerating the scalable and efficient evaluation of complex voice AI systems.
US lawmakers are investigating the growing use of Chinese AI models by American companies, citing concerns over censorship, security risks, and the impact on domestic alternatives. The probe is specifically looking at companies such as Cursor and Airbnb, and the use of models like DeepSeek. This investigation highlights the complexities of the AI landscape and the need for careful consideration of the origins and implications of AI models.

DeepSeek has unveiled a groundbreaking approach to abstract reasoning on ARC-AGI-1, leveraging an open-weight model (DeepSeek V3.2) in a 'non-thinking' mode, augmented by innovative agentic harnesses. This method achieves impressive generalization and pattern discovery, reaching up to 67.25% pass@2 with unprecedented cost-efficiency, sidestepping heavy compute or benchmark-specific fine-tuning.

A new arXiv paper by Alibaba researchers details a ReAct-style agentic setup integrating Large Language Models with SageMath, a powerful Computer Algebra System. This novel approach demonstrates substantial performance gains across frontier LLMs in solving research-level mathematical problems, significantly narrowing the capability gap between open-weight and closed models and paving the way for automated conjecture discovery.

OpenAI CEO suggests that video games can be a superior training data source than the internet for achieving AGI, with significant implications for the AI industry.
Get the top AI stories in your inbox once a day, no spam.
New stories are added every couple of hours as they break, so the feed stays current throughout the day.
We pull from 100+ sources, including company blogs, research labs, and established tech publications, then fact check and summarize each story before it goes live.
Yes. Use the sidebar filters to narrow stories down by company (OpenAI, Anthropic, Google, and more), industry, or event type like funding and research.
Yes. AI Pulse is free for anyone who wants to keep up with AI news, no sign up required. The daily newsletter is optional if you want updates in your inbox.