LLM · Other Companies

Flux 3 generates videos with native audio up to 20 seconds long, a first for Black Forest Labs
Featured
LLM

Flux 3 generates videos with native audio up to 20 seconds long, a first for Black Forest Labs

Black Forest Labs has released Flux 3, a multimodal foundation model that learns from images, video, and audio and can generate video with native sound for the first time. BFL's own tests put it just...

The Decoder
Read more
Poolside's Laguna S 2.1 is a small open-weight coding model that punches well above its size
LLM

Poolside's Laguna S 2.1 is a small open-weight coding model that punches well above its size

Poolside released Laguna S 2.1, a compact coding model that outperforms larger competitors by using training techniques...

The Decoder
China’s Open AI Models Are Challenging Silicon Valley’s Playbook
LLM

China’s Open AI Models Are Challenging Silicon Valley’s Playbook

Chinese AI laboratories are offering open-source large language models as alternatives to restricted frontier models...

Wired AI
Cisco bets its small open cybersecurity models can outperform GPT-5.5 at vulnerability detection for a fraction of the cost
LLM

Cisco bets its small open cybersecurity models can outperform GPT-5.5 at vulnerability detection for a fraction of the cost

Cisco has released two small, open-source AI models specifically designed for cybersecurity that detect approximately...

The Decoder
Unsloth vs Axolotl vs TRL vs LLaMA-Factory: A Fine-Tuning Framework Comparison on Speed, VRAM, and Multi-GPU
LLM

Unsloth vs Axolotl vs TRL vs LLaMA-Factory: A Fine-Tuning Framework Comparison on Speed, VRAM, and Multi-GPU

This article compares four open-source LLM fine-tuning frameworks—Unsloth, Axolotl, TRL, and LLaMA-Factory—analyzing...

MarkTechPost
Cisco Foundation AI Releases Antares: 350M and 1B Open-Weight Models That Localize Known Vulnerabilities Inside Real Codebases
LLM

Cisco Foundation AI Releases Antares: 350M and 1B Open-Weight Models That Localize Known Vulnerabilities Inside Real Codebases

Cisco Foundation AI released Antares, a family of small language models (350M and 1B parameters) designed to identify...

MarkTechPost
LLM

Poolside Releases Laguna S 2.1, an Open-Weight Agentic Coding Model Punching Above Its Weight Class on SWE-Bench Multilingual

Poolside has released Laguna S 2.1, a 118B open-weight Mixture-of-Experts coding model that achieves performance...

MarkTechPost
Validating Distributed LLM Serving Benchmarks with NVIDIA srt-slurm, SLURM Recipes, Parameter Sweeps, and Pareto Analysis
LLM

Validating Distributed LLM Serving Benchmarks with NVIDIA srt-slurm, SLURM Recipes, Parameter Sweeps, and Pareto Analysis

This tutorial explores NVIDIA's srt-slurm framework for converting YAML configurations into reproducible SLURM...

MarkTechPost
Alibaba Qwen 3.8 Max Shows China Closing in on U.S. Models
LLM

Alibaba Qwen 3.8 Max Shows China Closing in on U.S. Models

Alibaba's Qwen 3.8 Max, a low-cost open-weight language model, demonstrates that Chinese AI providers are closing the...

AI Business
Alibaba's Qwen-Image-3.0 renders full infographic grids and readable ten-pixel text in a single pass
LLM

Alibaba's Qwen-Image-3.0 renders full infographic grids and readable ten-pixel text in a single pass

Alibaba's Qwen-Image-3.0 is an advanced image generation model that accepts prompts up to 4,500 tokens and can render...

The Decoder
Alibaba's Qwen Audio 3.0 TTS Plus tops the competition in the text-to-speech rankings
LLM

Alibaba's Qwen Audio 3.0 TTS Plus tops the competition in the text-to-speech rankings

Alibaba's Qwen Audio 3.0 TTS Plus has topped the Artificial Analysis Speech Arena leaderboard for text-to-speech...

The Decoder
America needs to stop getting shocked by Chinese AI
LLM

America needs to stop getting shocked by Chinese AI

Two Chinese AI companies unveiled large language models that can credibly compete with OpenAI and Anthropic's best...

The Verge AI
The Army Is Burning Through Its AI Tokens
LLM

The Army Is Burning Through Its AI Tokens

The U.S. Army has issued a notice to personnel that they are rapidly consuming their allocated AI tokens and must...

Wired AI
Alibaba’s Tongyi Lab Releases Qwen-Audio-3.0-TTS, a Hosted Text-to-Speech Model in Flash and Plus Tiers Across 16 Languages
LLM

Alibaba’s Tongyi Lab Releases Qwen-Audio-3.0-TTS, a Hosted Text-to-Speech Model in Flash and Plus Tiers Across 16 Languages

Alibaba's Tongyi Lab released Qwen-Audio-3.0-TTS, a production-ready text-to-speech model available in two variants:...

MarkTechPost
District 9 director Neill Blomkamp releases first short film made entirely with AI video generation
LLM

District 9 director Neill Blomkamp releases first short film made entirely with AI video generation

Director Neill Blomkamp released "Nightborne," a 13-minute sci-fi horror short film created entirely using the Seedance...

The Decoder
As AI Spending Climbs, Enterprises Get Serious About Token Costs
LLM

As AI Spending Climbs, Enterprises Get Serious About Token Costs

Enterprises are increasingly concerned about the opaque pricing models and backward-looking billing practices...

AI Business
China delivers a one-two punch to America’s AI dominance 
LLM

China delivers a one-two punch to America’s AI dominance 

China's Moonshot AI and Alibaba have unveiled new AI models that claim to compete with OpenAI and Anthropic's systems...

The Verge AI
Kimi K3 open-weight model: China’s biggest AI is a bet on memory, not compute
LLM

Kimi K3 open-weight model: China’s biggest AI is a bet on memory, not compute

Moonshot AI released Kimi K3, an open-weight model with 2.8 trillion parameters, making it the largest open-weight...

AI News
Moonshot pauses new Kimi K3 subscriptions after GPU demand maxes out in 48 hours
LLM

Moonshot pauses new Kimi K3 subscriptions after GPU demand maxes out in 48 hours

Moonshot has temporarily paused new subscriptions for its Kimi K3 AI model after demand exhausted GPU capacity within...

The Decoder
LLM

Someone Fine-Tuned OpenBMB’s MiniCPM5-1B on Claude Fable 5 Traces to Ship a 657MB Local Thinking Model

A community developer fine-tuned OpenBMB's MiniCPM5-1B model on Claude Fable 5 traces to create a 657MB local thinking...

MarkTechPost

Stay Updated

Get the latest AI news delivered to your inbox every morning. No spam, unsubscribe anytime.