LLM · Other Companies

Best Local LLMs You Can Run on a Single 24GB GPU in 2026: Qwen, Gemma, Mistral, DeepSeek Compared
LLM

Best Local LLMs You Can Run on a Single 24GB GPU in 2026: Qwen, Gemma, Mistral, DeepSeek Compared

A comparative guide to six open-weight large language models that can run on a single 24GB GPU, including Qwen3.6,...

MarkTechPost
LLM

Feyn AI Releases SQRL, a Text-to-SQL Model Family That Inspects the Database Before Writing a Query

Feyn Labs released SQRL, a family of text-to-SQL models that inspect databases before generating queries. The flagship...

MarkTechPost
Alibaba Previews Qwen3.8-Max, a 2.4 Trillion-Parameter Multimodal Model, Days After Moonshot’s Kimi K3 Open-Weight Launch
LLM

Alibaba Previews Qwen3.8-Max, a 2.4 Trillion-Parameter Multimodal Model, Days After Moonshot’s Kimi K3 Open-Weight Launch

Alibaba previewed Qwen3.8-Max-Preview, a 2.4 trillion-parameter multimodal MoE model available at discounted pricing on...

MarkTechPost
Alibaba's Qwen takes on Kimi K3 with open-weight Qwen 3.8, says model is "second only to Fable 5"
LLM

Alibaba's Qwen takes on Kimi K3 with open-weight Qwen 3.8, says model is "second only to Fable 5"

Alibaba has released Qwen 3.8, a multimodal AI model with 2.4 trillion parameters that the company claims rivals...

The Decoder
Moonshot's Kimi K3 outperforms Fable 5 in frontend code but lags far behind in complex math
LLM

Moonshot's Kimi K3 outperforms Fable 5 in frontend code but lags far behind in complex math

Moonshot's Kimi K3 achieved top rankings in frontend code benchmarks, outperforming Claude Fable 5 and GPT-5.6 Sol, but...

The Decoder
Kimi K3 vs DeepSeek V4 Pro vs GLM-5.2: Open Trillion-Scale MoE Models Compared on Benchmarks, License, and Serving Cost
LLM

Kimi K3 vs DeepSeek V4 Pro vs GLM-5.2: Open Trillion-Scale MoE Models Compared on Benchmarks, License, and Serving Cost

The article compares three open-source trillion-scale Mixture of Experts (MoE) language models: Kimi K3, DeepSeek V4...

MarkTechPost
Fine-Tuning Qwen3 with LoRA Using NVIDIA NeMo AutoModel: A Complete Single-GPU Google Colab Workflow Tutorial
LLM

Fine-Tuning Qwen3 with LoRA Using NVIDIA NeMo AutoModel: A Complete Single-GPU Google Colab Workflow Tutorial

A tutorial demonstrating how to fine-tune the Qwen3-0.6B language model using LoRA (Low-Rank Adaptation) with NVIDIA...

MarkTechPost
Kimi: Threat or menace?
LLM

Kimi: Threat or menace?

Chinese company Moonshot AI released a new version of its Kimi large language model this week. The release has sparked...

TechCrunch AI
Just like Deepseek, China's Kimi K3 is forcing Western AI labs to question their compute advantage
LLM

Just like Deepseek, China's Kimi K3 is forcing Western AI labs to question their compute advantage

Moonshot AI released Kimi K3, a large language model built by 300 people that matches Anthropic's Claude Opus 4.8,...

The Decoder
Chinese AI Startup Releases Massive Open Weight Model
LLM

Chinese AI Startup Releases Massive Open Weight Model

Chinese AI startup Kimi has released K3, an open-weight large language model with 2.8 trillion parameters. U.S.-based...

AI Business
Netflix's 300 AI productions show how fast the technology is spreading through entertainment
LLM

Netflix's 300 AI productions show how fast the technology is spreading through entertainment

Netflix is using AI in approximately 300 productions, primarily for post-production work. The company reports...

The Decoder
NVIDIA AI Releases Nemotron 3 Embed: An Open Embedding Collection Whose 8B Checkpoint Ranks #1 on RTEB
LLM

NVIDIA AI Releases Nemotron 3 Embed: An Open Embedding Collection Whose 8B Checkpoint Ranks #1 on RTEB

NVIDIA released Nemotron 3 Embed, a collection of open embedding models with three checkpoints, where the 8B model...

MarkTechPost
Moonshot AI Releases Kimi K3: A 2.8 Trillion Parameter Open MoE Model With Kimi Delta Attention and 1M Context
LLM

Moonshot AI Releases Kimi K3: A 2.8 Trillion Parameter Open MoE Model With Kimi Delta Attention and 1M Context

Moonshot AI released Kimi K3, a 2.8-trillion-parameter open MoE (Mixture of Experts) model featuring Kimi Delta...

MarkTechPost
Kimi's open model K3 nears GPT-5.6 Sol and Fable 5 while signaling the end of super cheap Chinese AI
LLM

Kimi's open model K3 nears GPT-5.6 Sol and Fable 5 while signaling the end of super cheap Chinese AI

Kimi has launched K3, a multimodal open-weight model with 2.8 trillion parameters and 1 million token context window...

The Decoder
Thinking Machines Rolls Out Broad but Efficient Model
LLM

Thinking Machines Rolls Out Broad but Efficient Model

Thinking Machines, founded by OpenAI's former CTO, has released Inkling, a general-purpose AI model designed with token...

AI Business
The AI compute gap: Enterprises are buying infrastructure faster than they can measure what it costs
LLM

The AI compute gap: Enterprises are buying infrastructure faster than they can measure what it costs

Enterprises are rapidly increasing AI infrastructure spending but lack visibility into costs and efficiency, with 83%...

VentureBeat AI
Moonshot’s upcoming Kimi 3 is expected to close the gap with Anthropic’s Opus 4.8
LLM

Moonshot’s upcoming Kimi 3 is expected to close the gap with Anthropic’s Opus 4.8

Moonshot's upcoming Kimi K3 is expected to be China's largest open AI model with 2-3 trillion parameters, positioning...

TechCrunch AI
Sakana AI's orchestrator adds Nvidia Nemotron to prove "collective intelligence" can rival single frontier models
LLM

Sakana AI's orchestrator adds Nvidia Nemotron to prove "collective intelligence" can rival single frontier models

Sakana AI is integrating Nvidia's Nemotron open-source models into its Fugu orchestrator to demonstrate that...

The Decoder
Ex-OpenAI CTO Murati's Thinking Machines drops Inkling, a 975B parameter model that leads US labs but trails China
LLM

Ex-OpenAI CTO Murati's Thinking Machines drops Inkling, a 975B parameter model that leads US labs but trails China

Thinking Machines Lab, founded by ex-OpenAI CTO Mira Murati, released Inkling, a 975B parameter multimodal open-weights...

The Decoder