
Alibaba's Qwen introduces Chain-of-Models, an automated audit pipeline to mitigate cognitive biases in large language models. This approach uses a second model to inspect the first model's reasoning trace, reducing bias and improving judgment. The study reveals that auditor identity and bias type significantly impact audit effectiveness.

Anthropic researchers have introduced LLM-SoccerArena, a groundbreaking prospective live benchmark designed to evaluate how well Large Language Models (LLMs) forecast real-world events before outcomes are known. This open-source platform moves beyond static, retrospective evaluations, offering a dynamic environment to test LLMs' ability to synthesize information and predict uncertain future sports outcomes like the FIFA World Cup.

Groundbreaking Arxiv research reveals how large language model safety mechanisms are encoded and can be bypassed, introducing novel 'Activation-Guided' adversarial attacks. The study finds safety representations are distributed across model layers, not localized, and proposes a 33x faster attack method, Soft-GCG, offering critical insights for designing more robust AI alignment strategies.

OpenAI has officially launched its highly anticipated new family of models, spearheaded by GPT-5.6, marking a significant leap forward in generative AI capabilities. This release promises substantial improvements across diverse areas, including a critical focus on bolstering cybersecurity applications and overall model safety.

Meta AI introduces ReContext, a groundbreaking training-free inference method that significantly boosts Large Language Model (LLM) performance on long contexts. By recursively replaying relevant evidence, ReContext enhances effective context utilization, bridging the gap between vast context windows and accurate reasoning without requiring retraining or external memory. This innovation promises to unlock more reliable and powerful LLM applications across industries.
Get the top AI stories in your inbox once a day, no spam.
New stories are added every couple of hours as they break, so the feed stays current throughout the day.
We pull from 100+ sources, including company blogs, research labs, and established tech publications, then fact check and summarize each story before it goes live.
Yes. Use the sidebar filters to narrow stories down by company (OpenAI, Anthropic, Google, and more), industry, or event type like funding and research.
Yes. AI Pulse is free for anyone who wants to keep up with AI news, no sign up required. The daily newsletter is optional if you want updates in your inbox.