LLM · Other Companies

Thinking Machines Lab Releases Inkling: A 975B-Parameter Open-Weights Multimodal MoE With 41B Active Parameters And Controllable Thinking Effort
LLM

Thinking Machines Lab Releases Inkling: A 975B-Parameter Open-Weights Multimodal MoE With 41B Active Parameters And Controllable Thinking Effort

Thinking Machines Lab released Inkling, a 975B-parameter open-weights multimodal mixture-of-experts model with 41B...

MarkTechPost
Sheetz is quitting VMware, migrating 11,000 virtual machines
LLM

Sheetz is quitting VMware, migrating 11,000 virtual machines

The convenience store chain will use StorMagic instead.

Ars Technica AI
Soofi Consortium Releases Soofi S 30B-A3B: An Open Hybrid Mamba-Transformer MoE Foundation Model For German And English
LLM

Soofi Consortium Releases Soofi S 30B-A3B: An Open Hybrid Mamba-Transformer MoE Foundation Model For German And English

Soofi Consortium has released Soofi S 30B-A3B, an open-source hybrid Mamba-Transformer mixture-of-experts foundation...

MarkTechPost
Windows 0-day drops the same day Microsoft releases record number of patches
LLM

Windows 0-day drops the same day Microsoft releases record number of patches

HiveLegacy is a "powerful primitive" that's likely capable of other nefarious actions.

Ars Technica AI
Thinking Machines Lab Drops Its First Model
LLM

Thinking Machines Lab Drops Its First Model

Thinking Machines Lab released Inkling, a 975-billion-parameter open source model trained to understand video and...

Wired AI
Thinking Machines amps up its bet against one-size-fits-all AI with its first open model, Inkling
LLM

Thinking Machines amps up its bet against one-size-fits-all AI with its first open model, Inkling

Thinking Machines has released Inkling, its first open AI model, marking the company's initial public demonstration...

TechCrunch AI
PrismML Releases Bonsai 27B: 1-bit and Ternary Builds of Qwen3.6-27B That Run on Laptops and Phones
LLM

PrismML Releases Bonsai 27B: 1-bit and Ternary Builds of Qwen3.6-27B That Run on Laptops and Phones

PrismML released Bonsai 27B, a quantized version of Qwen3.6-27B using 1-bit and ternary weight representations that...

MarkTechPost
The real AI race may no longer be at the frontier
LLM

The real AI race may no longer be at the frontier

Hugging Face CEO argues that the AI race is shifting from frontier models to open models, as enterprises increasingly...

TechCrunch AI
German AI consortium releases Soofi S, an open 30B model that tops benchmarks in both English and German
LLM

German AI consortium releases Soofi S, an open 30B model that tops benchmarks in both English and German

A German research consortium released Soofi S 30B-A3B, an open-source language model trained on Deutsche Telekom's...

The Decoder
Mira Murati’s Thinking Machines Lab Makes The Technical Case For Human-Centered AI Built On Customizable Model Weights
LLM

Mira Murati’s Thinking Machines Lab Makes The Technical Case For Human-Centered AI Built On Customizable Model Weights

Mira Murati's Thinking Machines Lab published an essay on human-centered AI that frames human participation, model...

MarkTechPost
Open source AI matters more than ever, according to Hugging Face’s Clem Delangue
LLM

Open source AI matters more than ever, according to Hugging Face’s Clem Delangue

Hugging Face CEO Clem Delangue highlights the growing importance of open source AI, noting that the platform has become...

TechCrunch Startups
Hugging Face’s CEO on why companies are done renting their AI
LLM

Hugging Face’s CEO on why companies are done renting their AI

Hugging Face CEO Clem Delangue discusses the growing trend of open source AI adoption, with the platform now used by...

TechCrunch Startups
Meet Nemotron Labs 3 Puzzle 75B A9B: A Compressed Hybrid MoE LLM Delivering 2.03x Server Throughput
LLM

Meet Nemotron Labs 3 Puzzle 75B A9B: A Compressed Hybrid MoE LLM Delivering 2.03x Server Throughput

NVIDIA released Nemotron-Labs-3-Puzzle-75B-A9B, a compressed hybrid MoE language model that reduces parameters from...

MarkTechPost
Fast token generation emerges as the key differentiator as heterogeneous inference takes hold
LLM

Fast token generation emerges as the key differentiator as heterogeneous inference takes hold

The article discusses how fast token generation has become the critical differentiator in AI inference, moving from...

SiliconANGLE
NVIDIA Releases Nemotron-Labs-3-Puzzle-75B-A9B: A Compressed Hybrid MoE LLM Delivering 2.03x Server Throughput at Matched User Throughput
LLM

NVIDIA Releases Nemotron-Labs-3-Puzzle-75B-A9B: A Compressed Hybrid MoE LLM Delivering 2.03x Server Throughput at Matched User Throughput

NVIDIA released Nemotron-Labs-3-Puzzle-75B-A9B, a compressed hybrid MoE language model that reduces parameters from...

MarkTechPost
Datalab Lift vs the Field: How a 9B Schema-First Extractor Compares with NuExtract3, LlamaExtract, Marker, and Docling
LLM

Datalab Lift vs the Field: How a 9B Schema-First Extractor Compares with NuExtract3, LlamaExtract, Marker, and Docling

Datalab's Lift is a 9B schema-first document extraction tool that converts PDFs and images directly to JSON without...

MarkTechPost
I Built a Self-Improving AI, and So Can You
LLM

I Built a Self-Improving AI, and So Can You

The article discusses experiments where AI is used to build and improve other AI systems, demonstrating that advanced...

Wired AI
Chinese AI startup MiniMax plans to open-source a 2.7 trillion parameter model later this year
LLM

Chinese AI startup MiniMax plans to open-source a 2.7 trillion parameter model later this year

Chinese AI startup MiniMax is developing a large language model with 2.7 trillion parameters and plans to release it as...

The Decoder
NVIDIA Releases Audex (Nemotron-Labs-Audex-30B-A3B): A Unified Audio-Text LLM That Preserves the Text Intelligence of Its Backbone
LLM

NVIDIA Releases Audex (Nemotron-Labs-Audex-30B-A3B): A Unified Audio-Text LLM That Preserves the Text Intelligence of Its Backbone

NVIDIA released Audex (Nemotron-Labs-Audex-30B-A3B), a unified audio-text large language model that combines audio...

MarkTechPost