LLM · Other Companies

The token bill comes due: Inside the industry scramble to manage AI’s runaway costs
LLM

The token bill comes due: Inside the industry scramble to manage AI’s runaway costs

The AI industry is shifting focus from rapid expansion and cost maximization toward implementing cost controls and...

TechCrunch AI
NVIDIA AI Releases Nemotron 3 Ultra: An Open 550B Mixture-of-Experts Hybrid Mamba-Transformer for Long-Running Agents
LLM

NVIDIA AI Releases Nemotron 3 Ultra: An Open 550B Mixture-of-Experts Hybrid Mamba-Transformer for Long-Running Agents

NVIDIA has released Nemotron 3 Ultra, a 550B parameter Mixture-of-Experts hybrid Mamba-Transformer model designed for...

MarkTechPost
Miso Labs Releases MisoTTS: An 8B Emotive Text-to-Speech Model with Open Weights
LLM

Miso Labs Releases MisoTTS: An 8B Emotive Text-to-Speech Model with Open Weights

Miso Labs has released MisoTTS, an open-weights 8 billion parameter text-to-speech model that uses residual vector...

MarkTechPost
Ideogram 4.0 drops as an open-weight model with native 2K resolution and improved text rendering
LLM

Ideogram 4.0 drops as an open-weight model with native 2K resolution and improved text rendering

Ideogram has released version 4.0 of its text-to-image model as an open-weight model featuring native 2K resolution,...

The Decoder
Perplexity announces hybrid AI system that decides what runs locally or in the cloud
LLM

Perplexity announces hybrid AI system that decides what runs locally or in the cloud

Perplexity has announced a hybrid AI orchestrator system that intelligently distributes tasks between local and...

The Decoder
How to Fine-Tune LFM2 Using QLoRA and DPO: A Complete Step-by-Step Coding Tutorial on Google Colab
LLM

How to Fine-Tune LFM2 Using QLoRA and DPO: A Complete Step-by-Step Coding Tutorial on Google Colab

A tutorial on fine-tuning the LFM2 language model using QLoRA and DPO techniques on Google Colab. The guide covers...

MarkTechPost
OpenAI vs. Anthropic vs. Google: But the Model Isn't the Point
LLM

OpenAI vs. Anthropic vs. Google: But the Model Isn't the Point

Enterprise customers are prioritizing practical AI solutions and business outcomes over specific AI model providers or...

AI Business
JetBrains Releases Mellum2: A 12B MoE Model for Fast, Specialized Tasks in Multi-Model AI Pipelines
LLM

JetBrains Releases Mellum2: A 12B MoE Model for Fast, Specialized Tasks in Multi-Model AI Pipelines

JetBrains has released Mellum2, a 12B parameter Mixture of Experts model trained on 10.6 trillion tokens, under an...

MarkTechPost
MiniMax Releases MiniMax M3 with MSA Architecture Supporting 1M-Token Context, Native Multimodality, and Agentic Coding
LLM

MiniMax Releases MiniMax M3 with MSA Architecture Supporting 1M-Token Context, Native Multimodality, and Agentic Coding

MiniMax released M3, a new large language model featuring MiniMax Sparse Attention architecture that supports 1 million...

MarkTechPost
MiniMax M3: Open-weight model with a million-token context challenges proprietary leaders
LLM

MiniMax M3: Open-weight model with a million-token context challenges proprietary leaders

Chinese AI company MiniMax has released its M3 model, an open-weight AI model featuring top-tier coding performance, a...

The Decoder
Nvidia's Nemotron 3 Ultra becomes the smartest open US model, but China still leads
LLM

Nvidia's Nemotron 3 Ultra becomes the smartest open US model, but China still leads

Nvidia's new Nemotron 3 Ultra has been ranked as the most capable open-source AI model from the US according to...

The Decoder
Best Text-to-Speech TTS Models in 2026: A Benchmark-Based Comparison
LLM

Best Text-to-Speech TTS Models in 2026: A Benchmark-Based Comparison

This article benchmarks and compares leading text-to-speech models in 2026, evaluating their quality, latency, cost,...

MarkTechPost
NVIDIA Introduces X-Token: Projection-Guided Cross-Tokenizer KD That Outperforms GOLD by +3.82 Average Points on Llama-3.2-1B
LLM

NVIDIA Introduces X-Token: Projection-Guided Cross-Tokenizer KD That Outperforms GOLD by +3.82 Average Points on Llama-3.2-1B

NVIDIA introduces X-Token, a projection-guided cross-tokenizer knowledge distillation technique that improves upon the...

MarkTechPost
StepFun Releases Step 3.7 Flash: A 198B MoE Vision-Language Model for Coding Agents and Search Workflows
LLM

StepFun Releases Step 3.7 Flash: A 198B MoE Vision-Language Model for Coding Agents and Search Workflows

StepFun has released Step 3.7 Flash, a 198 billion parameter mixture-of-experts vision-language model designed for...

MarkTechPost
Liquid AI Releases LFM2.5-8B-A1B: An On-Device MoE Model With 8.3B Total and 1.5B Active Parameters
LLM

Liquid AI Releases LFM2.5-8B-A1B: An On-Device MoE Model With 8.3B Total and 1.5B Active Parameters

Liquid AI released LFM2.5-8B-A1B, a mixture-of-experts (MoE) model with 8.3B total parameters but only 1.5B active...

MarkTechPost
Perplexity AI Open-Sources Unigram Tokenizer That Achieves 5x Lower p50 Latency Than Hugging Face tokenizers Crate
LLM

Perplexity AI Open-Sources Unigram Tokenizer That Achieves 5x Lower p50 Latency Than Hugging Face tokenizers Crate

Perplexity AI has open-sourced an optimized Unigram tokenizer that significantly improves performance compared to...

MarkTechPost
Amazon builds its own AI production platform and greenlights three AI animated series for Prime Video
LLM

Amazon builds its own AI production platform and greenlights three AI animated series for Prime Video

Amazon MGM Studios and AWS have launched a GenAI Creators' Fund providing filmmakers access to their proprietary AI...

The Decoder
ElevenLabs Music v2 promises opera-to-metal transitions without losing musical coherence
LLM

ElevenLabs Music v2 promises opera-to-metal transitions without losing musical coherence

ElevenLabs has released Music v2, an AI music generation model capable of seamlessly transitioning between different...

The Decoder
The AI boom drove Nvidia's yearly Taiwan spending from $15 billion to $150 billion
LLM

The AI boom drove Nvidia's yearly Taiwan spending from $15 billion to $150 billion

Nvidia's annual spending with Taiwan-based suppliers, primarily TSMC, has surged from $15 billion to $150 billion due...

The Decoder