LLM · Other Companies

Samsung and SK Hynix plan $590 billion chip investment as AI demand sends memory prices soaring
LLM

Samsung and SK Hynix plan $590 billion chip investment as AI demand sends memory prices soaring

Samsung and SK Hynix are investing $590 billion in chip manufacturing and packaging facilities to meet surging AI data...

The Decoder
China’s Z.ai claims it can match Mythos on cybersecurity
LLM

China’s Z.ai claims it can match Mythos on cybersecurity

China's Zhipu AI released GLM-5.2, an open-weight model that researchers claim matches Anthropic's Mythos in...

The Verge AI
Coinbase joins the rush to Chinese AI models as Western labs face a pricing stress test
LLM

Coinbase joins the rush to Chinese AI models as Western labs face a pricing stress test

Coinbase CEO Brian Armstrong is switching the company to Chinese AI models like GLM 5.2 and Kimi 2.7 with an automated...

The Decoder
Sina's open model VibeThinker-3B aims to show reasoning compresses well but factual knowledge doesn't
LLM

Sina's open model VibeThinker-3B aims to show reasoning compresses well but factual knowledge doesn't

Sina Weibo's VibeThinker-3B, a 3 billion parameter open model, achieves performance comparable to much larger models...

The Decoder
LLM

Liquid AI Ships LFM2.5-230M with llama.cpp, MLX, vLLM, SGLang, and ONNX Support for On-Device Inference

Liquid AI released LFM2.5-230M, a 230M-parameter open-weight model optimized for on-device inference, achieving 213...

MarkTechPost
ByteDance's "iLLaDA" is a diffusion language model that keeps up with Qwen2.5
LLM

ByteDance's "iLLaDA" is a diffusion language model that keeps up with Qwen2.5

ByteDance and Renmin University researchers released iLLaDA, an 8B diffusion-based language model that uses a different...

The Decoder
Building Supervised Fine-Tuning Data from NVIDIA Open-SWE-Traces: Trajectory Parsing, Patch Analysis, Token Budgets, and Tool-Use Metrics
LLM

Building Supervised Fine-Tuning Data from NVIDIA Open-SWE-Traces: Trajectory Parsing, Patch Analysis, Token Budgets, and Tool-Use Metrics

A tutorial on building supervised fine-tuning datasets from NVIDIA's Open-SWE-Traces dataset for training agentic...

MarkTechPost
DeepReinforce Releases Ornith-1.0: An Open-Source Coding Model Family That Learns Its Own RL Scaffolds
LLM

DeepReinforce Releases Ornith-1.0: An Open-Source Coding Model Family That Learns Its Own RL Scaffolds

DeepReinforce released Ornith-1.0, an open-source coding model family built on Gemma 4 and Qwen 3.5 that learns its own...

MarkTechPost
Baidu Releases Unlimited OCR, a 3B Model That Keeps the KV Cache Flat for Long-Document Parsing
LLM

Baidu Releases Unlimited OCR, a 3B Model That Keeps the KV Cache Flat for Long-Document Parsing

Baidu open-sourced Unlimited OCR, a 3B-parameter MoE model that efficiently parses long documents using Reference...

MarkTechPost
Companies are scrambling to stop employees from maxing out AI budgets with small tasks
LLM

Companies are scrambling to stop employees from maxing out AI budgets with small tasks

Companies are implementing token rationing and budget controls to prevent employees from depleting AI API quotas with...

TechCrunch AI
Gradium Launches stt-translate and s2s-translate, Real-Time Speech Translation Models Beating gpt-realtime-translate on Accuracy and Latency
LLM

Gradium Launches stt-translate and s2s-translate, Real-Time Speech Translation Models Beating gpt-realtime-translate on Accuracy and Latency

Gradium released stt-translate and s2s-translate, real-time speech translation models that support 20 language pairs...

MarkTechPost
Snowflake CEO finds GLM-5.2 competitive with Opus 4.7 at a fraction of the cost
LLM

Snowflake CEO finds GLM-5.2 competitive with Opus 4.7 at a fraction of the cost

Zhipu AI's GLM-5.2 model demonstrates competitive performance with Claude Opus 4.7 on coding benchmarks while costing...

The Decoder
Pangram CEO says language models give themselves away by making the same arguments
LLM

Pangram CEO says language models give themselves away by making the same arguments

Pangram CEO Max Spero argues that language models can be identified by their homogeneous reasoning patterns, as they...

The Decoder
Oracle’s 21,000 layoffs help drive its debt-fueled AI investments
LLM

Oracle’s 21,000 layoffs help drive its debt-fueled AI investments

Oracle is laying off 21,000 employees while investing billions in data center infrastructure to support AI...

Ars Technica AI
Datalab Releases lift: A 9B Open-Weights Vision Model That Extracts Structured JSON From PDFs Using Schemas
LLM

Datalab Releases lift: A 9B Open-Weights Vision Model That Extracts Structured JSON From PDFs Using Schemas

Datalab released lift, a 9B open-weights vision model that converts PDFs and images into structured JSON matching...

MarkTechPost
ByteDance's Seedance 2.5 breaks the 30-second barrier for AI video generation
LLM

ByteDance's Seedance 2.5 breaks the 30-second barrier for AI video generation

ByteDance introduced five new AI models at Volcano Engine's FORCE conference, with Seedance 2.5 as the flagship video...

The Decoder
GLM-5.2 OpenAI-Compatible API: A Hands-On Guide to Reasoning Effort, Function Calling, and Long-Context Retrieval
LLM

GLM-5.2 OpenAI-Compatible API: A Hands-On Guide to Reasoning Effort, Function Calling, and Long-Context Retrieval

A practical guide to using GLM-5.2's OpenAI-compatible API, demonstrating reasoning effort control, function calling,...

MarkTechPost
Sakana AI's Fugu orchestrates multiple LLMs to match Anthropic's Fable and Mythos benchmarks
LLM

Sakana AI's Fugu orchestrates multiple LLMs to match Anthropic's Fable and Mythos benchmarks

Sakana AI launched Fugu, a system that coordinates multiple AI models to compete with Anthropic's Fable 5 benchmarks....

The Decoder
Cisco AI Introduces FAPO: Pipeline-Aware Prompt Optimization With Step-Level Failure Attribution and Claude Code Orchestration
LLM

Cisco AI Introduces FAPO: Pipeline-Aware Prompt Optimization With Step-Level Failure Attribution and Claude Code Orchestration

Cisco Foundation AI has open-sourced FAPO, a Claude-powered system that autonomously optimizes multi-step LLM pipelines...

MarkTechPost