
A recent study examines the use of Large Language Models (LLMs) for specialised terminology, evaluating four proprietary models in two domains. The results highlight the potential of LLMs as useful tools for specialised translators, but also note their limitations. The study paves the way for future work on the practical usefulness of LLMs in work and educational contexts.

New research reveals a critical safety blind spot in leading Large Language Models (LLMs), including ChatGPT-4o, DeepSeek, and Llama 3.1, when assessing multi-sensor physical hazard data. While excelling at single-sensor violations, these models consistently failed to issue warnings when multiple sensors collectively indicated danger below individual thresholds, posing significant risks for AI-powered safety systems.
Get the top AI stories in your inbox once a day, no spam.
New stories are added every couple of hours as they break, so the feed stays current throughout the day.
We pull from 100+ sources, including company blogs, research labs, and established tech publications, then fact check and summarize each story before it goes live.
Yes. Use the sidebar filters to narrow stories down by company (OpenAI, Anthropic, Google, and more), industry, or event type like funding and research.
Yes. AI Pulse is free for anyone who wants to keep up with AI news, no sign up required. The daily newsletter is optional if you want updates in your inbox.