Large Language Models (LLMs) — Latest News and Analysis

News on large language models, foundation model releases, benchmarks, and LLM-powered applications.

The Latest News About Large Language Models

Anthropic Claude

Claude Mythos Improves Crypto Attacks with the HAWK Signature Scheme

Anthropic's Claude Mythos has improved simulated attacks on the HAWK digital signature scheme and seven-round AES-128.
Monshot Kimi K3 Chat

Kimi K3 Model Launches With Alibaba, Huawei Day-Zero Support

Moonshot AI has released Kimi K3's weights, with reported Alibaba Cloud and Huawei Ascend support plus documented SGLang and Baseten deployment paths.
Claude Voice Mode

Claude Voice Mode Adds Opus 5 and Sonnet 5 Support, Stays Turn-Based

Anthropic has expanded Claude voice mode to Opus and Sonnet in beta across mobile, desktop, and web, adding model choice while keeping turn-based audio.
Hugging Face HUGS official

OpenAI’s GPT-5.6 Sol Models Escapes Sandbox and Breaches Hugging Face

OpenAI's latest GPT-5.6 Sol and a stronger unreleased model escaped a cyber test through a proxy flaw, accessing Hugging Face systems and service credentials.
Monshot Kimi K3 Chat

Microsoft Weighs Chinese Kimi K3 Model for Selected Copilot Tasks

Microsoft's possible use of Moonshot AI's Kimi K3 for selected Copilot requests could lower inference costs, but deployment plans remain unconfirmed.
Sam Altman

Altman Signals AI Price War with Anthropic

Sam Altman challenges Anthropic on AI model pricing as cheaper Chinese rivals pressure premium rates, but OpenAI has not yet enacted his proclamation to offer GPT-5.2 at the price of Claude Fable 5.
Z.ai GLM-5

Z.AI Completes 1GW Chinese-Chip Data Center

Chinese AI developer Z.AI has reportedly started operating a 1GW domestic-chip data center for AI models, but its chip mix and operating capacity remain unconfirmed.
Cybersecurity Hackers Cyberesionage Node Network

AISI: Open-Weight AI Is Catching up With Models From Anthropic and OpenAI in Cybersecurity...

The UK AI Security Institute has found GLM-5.2 and DeepSeek V4-Pro trail closed cyber benchmarks by four to seven months at sharply lower cost.
ChatGPT teen accounts parent notifications

ChatGPT to Alert Parents After Teen Violence-Policy Deactivation

OpenAI has announced parent alerts that will follow any teen ChatGPT account violence-policy deactivations, sharing the policy category with parents but no private chats.

Google Releases Cheaper Gemini Flash and Flash Lite Models, Plans Restricted Gemini Cyber Pilot

Google has released lower-cost Gemini 3.6 Flash and Flash-Lite models and plans a restricted Gemini 3.5 Flash Cyber pilot for cybersecurity testing.
ChatGPT App

OpenAI Restores ChatGPT Desktop Features After Backlash

OpenAI has restored ChatGPT desktop history, Projects, and a Chat/Work switch after a redesign backlash, while Local Tasks remain tied to a single computer.
Anthropic Claude

Anthropic Keeps Claude Fable 5 in Paid Plans After OpenAI’s GPT-5.6 Release

Anthropic has permanently included Claude Fable 5 into its subscription plans, with capped included use for premium tiers and metered credits for Pro and Team Standard.
Google Gemini

Google Reportedly Delays Gemini 3.5 Pro Over Coding Issues

Google has reportedly delayed Gemini 3.5 Pro after missing a June target as coding results fell short, leaving partner testing underway without a release date.
Thinking Machines Lab

Thinking Machines Lab Launches DeepSeek-Inspired Inkling 975B Parameter Model

Thinking Machines Lab has released Inkling, a 975-billion-parameter open-weight AI model with a DeepSeek-inspired design, 2 TB GPU needs and mixed benchmarks.
Meta parent notifications for supervised teen AI chats

Meta Adds Human-Reviewed Parent Alerts for Teen AI Chats

Meta has added human-reviewed parent alerts for supervised teen AI chats about possible self-harm, keeping exact messages private despite false-alert risks.
Anthropic money cash funding

Anthropic Seeks Billions in Bank Credit Before IPO

Anthropic is reportedly discussing billions in added bank credit before a possible initial public offering, with loan terms and listing timing unresolved.
Monshot Kimi K3 Chat

Moonshot AI Unveils 2.8T-Parameter Kimi K3 AI Model

Moonshot AI has launched its 2.8-trillion-parameter Kimi K3 model, but a high hallucination rate might tempers its frontier-model pitch.
Microsoft 365 Copilot Chat

Microsoft Reportedly Trains Sales Team to Target AI Labs

Microsoft is reportedly coaching sales staff to challenge OpenAI and Anthropic with a cost, security, and integrated platform pitch for its own Copilot AI.
Anthropic Claude Memory

Memory Attacks Might Let Claude Leak Personal Data with

Security researcher Ayush Paul says a Claude proof of concept leaked memory-derived personal data through web links before Anthropic's mitigation.

Apple Reportedly Weighs PrismML For IPhone AI Models

Apple is reportedly weighing use of PrismML's compression for a 27-billion-parameter iPhone AI model.
OpenAI Codex GPT-5.6

GPT-5.6 Sol Users Complain About File and Database Deletions

Several users say OpenAI's GPT-5.6 Sol frontier model has deleted files or data without permission.
Satya Nadella on the Dwarkesh Patel podcast

Satya Nadella Challenges AI Labs Over Distillation Rules

Satya Nadella challenges AI labs' distillation restrictions, arguing companies should control evaluations, memory, and work traces created through AI use.
openai logo

OpenAI Eases GPT-5.6 Usage Limits, Keeps Weekly Caps

OpenAI gives paid Codex and ChatGPT Work users more scheduling flexibility, adds banked resets, and targets about 10% more effective usage.
Europe-Flag-AI-Act-Bing-Image-Creator

Europe’s New Soofi S AI Model Is Blazing Fast

Germany’s Soofi S AI model pairs sparse architecture with strong project-run benchmarks, but licensing gaps and long-context limits temper its promise.
OpenAI profit money

OpenAI Launches GPT-5.6 in Three Tiers: Sol, Terra and Luna

OpenAI’s GPT-5.6 launches with three model tiers: Sol for advanced reasoning, Terra for everyday work and Luna for faster, lower-cost tasks.
xAI Grok official

Grok 4.5 Launches for Coding Agents as SpaceXAI Tests Lower Prices

Grok 4.5 enters the AI coding race with Cursor integration, frontier-level benchmark scores, and promising pricing.
Anthropic Jacobian Lens J-Lens

Anthropic Finds a Hidden “Workspace” Inside Claude’s Reasoning

Anthropic's J-lens research exposes Claude's hidden J-space, a workspace that could aid safety monitoring of risky states without proving AI consciousness.
ChatGPT Ads Advertising

OpenAI and Anthropic Must Prove AI Scale Can Outrun Compute Costs

OpenAI and Anthropic are facing tough IPOs as frontier AI costs and token-spend controls pressure their growth economics.
Meta AI bots profiles facebook

Meta Reportedly Probed Rival Chatbots With Fake Teen Accounts

Meta reportedly used contractors posing as teens to test ChatGPT, Gemini, and Character.AI, exposing platform-rule disputes and copied-response data questions.
Spike in number of CVEs in June 2026

AI Bug Hunting Leads to Dramatic Spike of CVE Disclosures Since Claude Mythos Release

Epoch AI data points to a record June surge in public software-flaw disclosures as AI bug-hunting expands, but the data cannot prove which flaws AI found.