LLM

Large Language Models

Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
LLMGoogle

Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

Google is introducing new Gemini models including Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber. These releases...

DeepMind Blog
Google launches a cheaper alternative to large AI security models like Mythos
LLMGoogle

Google launches a cheaper alternative to large AI security models like Mythos

Google launches Gemini 3.5 Flash Cyber, a cost-efficient AI security model designed to quickly identify and patch...

The Verge AI
Alibaba's Qwen Audio 3.0 TTS Plus tops the competition in the text-to-speech rankings
LLM

Alibaba's Qwen Audio 3.0 TTS Plus tops the competition in the text-to-speech rankings

Alibaba's Qwen Audio 3.0 TTS Plus has topped the Artificial Analysis Speech Arena leaderboard for text-to-speech...

The Decoder
America needs to stop getting shocked by Chinese AI
LLM

America needs to stop getting shocked by Chinese AI

Two Chinese AI companies unveiled large language models that can credibly compete with OpenAI and Anthropic's best...

The Verge AI
The Army Is Burning Through Its AI Tokens
LLM

The Army Is Burning Through Its AI Tokens

The U.S. Army has issued a notice to personnel that they are rapidly consuming their allocated AI tokens and must...

Wired AI
Google is working on a new AI chip designed to make Gemini more efficient
LLMGoogle

Google is working on a new AI chip designed to make Gemini more efficient

Alphabet is developing a new AI chip optimized to improve the efficiency of its Gemini large language models. The chip...

TechCrunch AI
Alibaba’s Tongyi Lab Releases Qwen-Audio-3.0-TTS, a Hosted Text-to-Speech Model in Flash and Plus Tiers Across 16 Languages
LLM

Alibaba’s Tongyi Lab Releases Qwen-Audio-3.0-TTS, a Hosted Text-to-Speech Model in Flash and Plus Tiers Across 16 Languages

Alibaba's Tongyi Lab released Qwen-Audio-3.0-TTS, a production-ready text-to-speech model available in two variants:...

MarkTechPost
Google's "Frozen v2" chip reportedly bakes Gemini's architecture directly into silicon for efficiency gains
LLMGoogle

Google's "Frozen v2" chip reportedly bakes Gemini's architecture directly into silicon for efficiency gains

Google is developing 'Frozen v2,' a specialized server chip designed to run Gemini's architecture with 6-10x greater...

The Decoder
District 9 director Neill Blomkamp releases first short film made entirely with AI video generation
LLM

District 9 director Neill Blomkamp releases first short film made entirely with AI video generation

Director Neill Blomkamp released "Nightborne," a 13-minute sci-fi horror short film created entirely using the Seedance...

The Decoder
Nvidia's grip on AI chips weakens as Microsoft turns to AMD and Anthropic may follow
LLMMicrosoft

Nvidia's grip on AI chips weakens as Microsoft turns to AMD and Anthropic may follow

Microsoft is expanding Azure's AI infrastructure by partnering with AMD's new Helios platform to challenge Nvidia's GPU...

The Decoder
As AI Spending Climbs, Enterprises Get Serious About Token Costs
LLM

As AI Spending Climbs, Enterprises Get Serious About Token Costs

Enterprises are increasingly concerned about the opaque pricing models and backward-looking billing practices...

AI Business
China delivers a one-two punch to America’s AI dominance 
LLM

China delivers a one-two punch to America’s AI dominance 

China's Moonshot AI and Alibaba have unveiled new AI models that claim to compete with OpenAI and Anthropic's systems...

The Verge AI
Kimi K3 open-weight model: China’s biggest AI is a bet on memory, not compute
LLM

Kimi K3 open-weight model: China’s biggest AI is a bet on memory, not compute

Moonshot AI released Kimi K3, an open-weight model with 2.8 trillion parameters, making it the largest open-weight...

AI News
Moonshot pauses new Kimi K3 subscriptions after GPU demand maxes out in 48 hours
LLM

Moonshot pauses new Kimi K3 subscriptions after GPU demand maxes out in 48 hours

Moonshot has temporarily paused new subscriptions for its Kimi K3 AI model after demand exhausted GPU capacity within...

The Decoder
LLM

Someone Fine-Tuned OpenBMB’s MiniCPM5-1B on Claude Fable 5 Traces to Ship a 657MB Local Thinking Model

A community developer fine-tuned OpenBMB's MiniCPM5-1B model on Claude Fable 5 traces to create a 657MB local thinking...

MarkTechPost
Best Local LLMs You Can Run on a Single 24GB GPU in 2026: Qwen, Gemma, Mistral, DeepSeek Compared
LLM

Best Local LLMs You Can Run on a Single 24GB GPU in 2026: Qwen, Gemma, Mistral, DeepSeek Compared

A comparative guide to six open-weight large language models that can run on a single 24GB GPU, including Qwen3.6,...

MarkTechPost
LLM

Feyn AI Releases SQRL, a Text-to-SQL Model Family That Inspects the Database Before Writing a Query

Feyn Labs released SQRL, a family of text-to-SQL models that inspect databases before generating queries. The flagship...

MarkTechPost
Alibaba Previews Qwen3.8-Max, a 2.4 Trillion-Parameter Multimodal Model, Days After Moonshot’s Kimi K3 Open-Weight Launch
LLM

Alibaba Previews Qwen3.8-Max, a 2.4 Trillion-Parameter Multimodal Model, Days After Moonshot’s Kimi K3 Open-Weight Launch

Alibaba previewed Qwen3.8-Max-Preview, a 2.4 trillion-parameter multimodal MoE model available at discounted pricing on...

MarkTechPost
Alibaba's Qwen takes on Kimi K3 with open-weight Qwen 3.8, says model is "second only to Fable 5"
LLM

Alibaba's Qwen takes on Kimi K3 with open-weight Qwen 3.8, says model is "second only to Fable 5"

Alibaba has released Qwen 3.8, a multimodal AI model with 2.4 trillion parameters that the company claims rivals...

The Decoder