
Alibaba's Qwen team has released Qwen3.8-Max, a 2.4 trillion parameter mixture-of-experts model, and confirmed open weights will ship next week. This model accepts text, image, and video inputs and returns text, making it a powerful tool for various industries. With its impressive capabilities, Qwen3.8-Max is set to revolutionize software engineering, legal and financial document review, media, and e-commerce operations.

OpenAI has announced GPT-5.6, an incremental yet significant model release focused on advancing the price-performance frontier for large language models. This update promises enhanced efficiency, lower operational costs, and improved accessibility, setting a new benchmark for practical AI deployment across industries.

OpenAI has released a new study on improving the faithfulness of podcasts generated from documents using large language models. The study introduces a framework for evaluating faithfulness and proposes a model-agnostic approach to detect and rewrite unfaithful conversational turns. This breakthrough has significant implications for the AI industry, developers, and businesses.

OpenAI introduces a new method for detecting model distillation in large language models, raising questions about fairness and policy violations. The approach uses reference-based membership inference to identify teacher models. This breakthrough has significant implications for the AI industry, developers, and businesses.

OpenAI has officially launched its highly anticipated new family of models, spearheaded by GPT-5.6, marking a significant leap forward in generative AI capabilities. This release promises substantial improvements across diverse areas, including a critical focus on bolstering cybersecurity applications and overall model safety.

OpenAI's latest research demonstrates how large language models can automate training data labeling for entity matching, reducing manual effort by 99% and slashing costs. This breakthrough enables faster, cheaper AI deployment for businesses.

A new arXiv study shows OpenEvidence's specialized clinical tool beats top general‑purpose models (Claude Opus 4.8, Gemini 3.1 Pro, GPT‑5.5) on 620 real‑world point‑of‑care questions. Physicians across 30 specialties rated the specialized tool higher on accuracy, utility, source quality, verifiability and completeness.
OpenAI introduces IMCBench, a benchmark for multimodal large language models in image-grounded medical conversations, evaluating safety, accuracy, and uncertainty in diagnosis. The benchmark tests eight models, with Claude Opus 4.6 achieving the highest overall score. The results highlight the need for multi-dimensional evaluation frameworks in medical AI.
Get the top AI stories in your inbox once a day, no spam.
New stories are added every couple of hours as they break, so the feed stays current throughout the day.
We pull from 100+ sources, including company blogs, research labs, and established tech publications, then fact check and summarize each story before it goes live.
Yes. Use the sidebar filters to narrow stories down by company (OpenAI, Anthropic, Google, and more), industry, or event type like funding and research.
Yes. AI Pulse is free for anyone who wants to keep up with AI news, no sign up required. The daily newsletter is optional if you want updates in your inbox.