LLM

Large Language Models

How to Fine-Tune LFM2 Using QLoRA and DPO: A Complete Step-by-Step Coding Tutorial on Google Colab
LLM

How to Fine-Tune LFM2 Using QLoRA and DPO: A Complete Step-by-Step Coding Tutorial on Google Colab

A tutorial on fine-tuning the LFM2 language model using QLoRA and DPO techniques on Google Colab. The guide covers...

MarkTechPost
OpenAI vs. Anthropic vs. Google: But the Model Isn't the Point
LLM

OpenAI vs. Anthropic vs. Google: But the Model Isn't the Point

Enterprise customers are prioritizing practical AI solutions and business outcomes over specific AI model providers or...

AI Business
Microsoft’s first advanced reasoning AI is here
LLMMicrosoft

Microsoft’s first advanced reasoning AI is here

Microsoft announced MAI-Thinking-1, its new flagship in-house AI model trained from scratch on clean data without...

The Verge AI
JetBrains Releases Mellum2: A 12B MoE Model for Fast, Specialized Tasks in Multi-Model AI Pipelines
LLM

JetBrains Releases Mellum2: A 12B MoE Model for Fast, Specialized Tasks in Multi-Model AI Pipelines

JetBrains has released Mellum2, a 12B parameter Mixture of Experts model trained on 10.6 trillion tokens, under an...

MarkTechPost
OpenAI models now available on Amazon Web Services
LLMOpenAI

OpenAI models now available on Amazon Web Services

OpenAI has partnered with Amazon Web Services to make GPT-5.5, GPT-4.5, and Codex models available through Amazon...

The Decoder
LLMOpenAI

Codex is becoming a productivity tool for everyone

Codex, an AI-powered tool, is being positioned as a productivity solution for knowledge workers across various tasks....

OpenAI Blog
MiniMax Releases MiniMax M3 with MSA Architecture Supporting 1M-Token Context, Native Multimodality, and Agentic Coding
LLM

MiniMax Releases MiniMax M3 with MSA Architecture Supporting 1M-Token Context, Native Multimodality, and Agentic Coding

MiniMax released M3, a new large language model featuring MiniMax Sparse Attention architecture that supports 1 million...

MarkTechPost
MiniMax M3: Open-weight model with a million-token context challenges proprietary leaders
LLM

MiniMax M3: Open-weight model with a million-token context challenges proprietary leaders

Chinese AI company MiniMax has released its M3 model, an open-weight AI model featuring top-tier coding performance, a...

The Decoder
Nvidia's Nemotron 3 Ultra becomes the smartest open US model, but China still leads
LLM

Nvidia's Nemotron 3 Ultra becomes the smartest open US model, but China still leads

Nvidia's new Nemotron 3 Ultra has been ranked as the most capable open-source AI model from the US according to...

The Decoder
LLMOpenAI

OpenAI frontier models and Codex are now available on AWS

OpenAI's frontier models and Codex are now generally available on AWS, allowing enterprises to build with OpenAI...

OpenAI Blog
Best Text-to-Speech TTS Models in 2026: A Benchmark-Based Comparison
LLM

Best Text-to-Speech TTS Models in 2026: A Benchmark-Based Comparison

This article benchmarks and compares leading text-to-speech models in 2026, evaluating their quality, latency, cost,...

MarkTechPost
Meta's leaked memo reveals AI pendant, supersensing glasses, and enterprise wearables strategy
LLMMeta

Meta's leaked memo reveals AI pendant, supersensing glasses, and enterprise wearables strategy

Meta is pivoting its AI strategy toward hardware products including an AI pendant and supersensing glasses, after...

The Decoder
NVIDIA Introduces X-Token: Projection-Guided Cross-Tokenizer KD That Outperforms GOLD by +3.82 Average Points on Llama-3.2-1B
LLM

NVIDIA Introduces X-Token: Projection-Guided Cross-Tokenizer KD That Outperforms GOLD by +3.82 Average Points on Llama-3.2-1B

NVIDIA introduces X-Token, a projection-guided cross-tokenizer knowledge distillation technique that improves upon the...

MarkTechPost
StepFun Releases Step 3.7 Flash: A 198B MoE Vision-Language Model for Coding Agents and Search Workflows
LLM

StepFun Releases Step 3.7 Flash: A 198B MoE Vision-Language Model for Coding Agents and Search Workflows

StepFun has released Step 3.7 Flash, a 198 billion parameter mixture-of-experts vision-language model designed for...

MarkTechPost
OpenAI gives GPT-5.5 Instant a readability upgrade while phasing out two older models
LLMOpenAI

OpenAI gives GPT-5.5 Instant a readability upgrade while phasing out two older models

OpenAI is updating GPT-5.5 Instant with improved readability for more natural responses and phasing out the Canvas...

The Decoder
Google fixes several bugs in Gemini usage limits that burned through quotas too fast
LLMGoogle

Google fixes several bugs in Gemini usage limits that burned through quotas too fast

Google has fixed bugs in its Gemini app that caused video generations to rapidly consume user quotas. The fixes include...

The Decoder
One company reportedly spent $500 million on Claude in one month after failing to cap AI usage
LLMAnthropic

One company reportedly spent $500 million on Claude in one month after failing to cap AI usage

An unnamed company spent $500 million on Claude licenses in one month due to lack of usage limits and AI expertise. The...

The Decoder
Anthropic Opus 4.8 Shows the AI Lab is Paying Attention to Customers
LLMAnthropic

Anthropic Opus 4.8 Shows the AI Lab is Paying Attention to Customers

Anthropic has released Opus 4.8, a model designed to serve enterprise customers with complex workflows. The model...

AI Business
Anthropic releases Claude Opus 4.8
LLMAnthropic

Anthropic releases Claude Opus 4.8

Anthropic has released Claude Opus 4.8, an upgraded version of Claude Opus 4.7 with improvements in coding, agent work,...

AI News