Research & Papers · Other Companies

How to Build a Parsing Pipeline with Docling Parse for Layout-Aware Document Intelligence
Research & Papers

How to Build a Parsing Pipeline with Docling Parse for Layout-Aware Document Intelligence

This tutorial demonstrates how to use Docling Parse to analyze PDF documents with layout-aware document intelligence,...

MarkTechPost
Meet Flash-KMeans: An IO-Aware, Exact K-Means That Runs Over 200× Faster Than FAISS on GPUs
Research & Papers

Meet Flash-KMeans: An IO-Aware, Exact K-Means That Runs Over 200× Faster Than FAISS on GPUs

Flash-KMeans is an open-source, IO-aware GPU implementation of k-means clustering that achieves significant performance...

MarkTechPost
A Coding Hands-On on FineWeb for Streaming, Filtering, Deduplication, Tokenization, and Large-Scale Web Corpus Analytics
Research & Papers

A Coding Hands-On on FineWeb for Streaming, Filtering, Deduplication, Tokenization, and Large-Scale Web Corpus Analytics

This tutorial provides a hands-on guide to working with the FineWeb dataset, demonstrating how to stream and analyze...

MarkTechPost
New AI model called "Count Anything" does exactly what it says, and that's harder than it sounds
Research & Papers

New AI model called "Count Anything" does exactly what it says, and that's harder than it sounds

"Count Anything" is a new AI model that can count objects in any type of image using text prompts, cutting error rates...

The Decoder
A Coding Implementation on Spatial Graph Neural Networks for Urban Function Inference Using city2graph, OSMnx, and PyTorch Geometric
Research & Papers

A Coding Implementation on Spatial Graph Neural Networks for Urban Function Inference Using city2graph, OSMnx, and PyTorch Geometric

The article presents an end-to-end spatial graph learning pipeline using city2graph, OSMnx, and PyTorch Geometric to...

MarkTechPost
When it comes to predicting people’s preferences, it pays to consider “the power of three”
Research & Papers

When it comes to predicting people’s preferences, it pays to consider “the power of three”

MIT researchers have developed a significant advancement to random utility models, a theory nearly a century old that...

MIT News AI
How memory tools can make AI models worse
Research & Papers

How memory tools can make AI models worse

Recent research reveals that memory systems integrated into AI models can actually reduce performance and promote...

TechCrunch AI
ClawHub Security Signals: A Coding Guide to End-to-End Security Signal Analysis and Verdict Classification on the AI Skills Dataset
Research & Papers

ClawHub Security Signals: A Coding Guide to End-to-End Security Signal Analysis and Verdict Classification on the AI Skills Dataset

This tutorial explores the ClawHub Security Signals dataset to analyze how security scanners assess AI skills,...

MarkTechPost
Researchers pinpoint why larger language models pick up skills that small ones miss
Research & Papers

Researchers pinpoint why larger language models pick up skills that small ones miss

Researchers discovered that small language models fail at rare tasks because frequent tasks overwrite their learned...

The Decoder
A Hands-On Coding Tutorial on Qualcomm AI Hub Models for Classification, Object Detection, and Hardware-Aware Deployment
Research & Papers

A Hands-On Coding Tutorial on Qualcomm AI Hub Models for Classification, Object Detection, and Hardware-Aware Deployment

This tutorial demonstrates how to use Qualcomm AI Hub Models to run inference tasks including MobileNet-V2...

MarkTechPost
NVIDIA AI Releases Dynamo Snapshot: A CRIU-Based Fast Startup System for AI Inference on Kubernetes
Research & Papers

NVIDIA AI Releases Dynamo Snapshot: A CRIU-Based Fast Startup System for AI Inference on Kubernetes

NVIDIA has released Dynamo Snapshot, a system that uses CRIU and cuda-checkpoint tools to enable fast startup of vLLM...

MarkTechPost
Building a Semantic Search Engine and Open-Status Classifier over the ResearchMath-14k Dataset
Research & Papers

Building a Semantic Search Engine and Open-Status Classifier over the ResearchMath-14k Dataset

This tutorial demonstrates a complete NLP pipeline for research mathematics using the ResearchMath-14k dataset,...

MarkTechPost
NSF renews support for MIT-led AI and physics institute, expanding a new model for discovery
Research & Papers

NSF renews support for MIT-led AI and physics institute, expanding a new model for discovery

The National Science Foundation has renewed funding for IAIFI (Institute for Artificial Intelligence and Fundamental...

MIT News AI
NVIDIA Releases Cosmos 3: A Two-Tower Mixture-of-Transformers Foundation Model Unifying Physical Reasoning, World Generation, and Action Generation
Research & Papers

NVIDIA Releases Cosmos 3: A Two-Tower Mixture-of-Transformers Foundation Model Unifying Physical Reasoning, World Generation, and Action Generation

NVIDIA released Cosmos 3, an omnimodal world model featuring a two-tower mixture-of-transformers architecture that...

MarkTechPost
MIT researchers teach AI models to interpret charts
Research & Papers

MIT researchers teach AI models to interpret charts

MIT researchers have developed ChartNet, a new training dataset designed to improve how vision-language models...

MIT News AI
How to Speed Up Transformer Training Using NVIDIA Apex (FusedAdam, FusedLayerNorm) and Native torch.amp
Research & Papers

How to Speed Up Transformer Training Using NVIDIA Apex (FusedAdam, FusedLayerNorm) and Native torch.amp

This article discusses optimizing Transformer model training performance using NVIDIA Apex tools including FusedAdam...

MarkTechPost
Turing Award winner Richard Sutton says pure generative AI can't do real science
Research & Papers

Turing Award winner Richard Sutton says pure generative AI can't do real science

Turing Award winner Richard Sutton argues that pure generative AI lacks the ability to evaluate its own results,...

The Decoder
Nvidia bets big on physical AI at GTC Taipei with a new world model, driving brain, and open humanoid robot
Research & Papers

Nvidia bets big on physical AI at GTC Taipei with a new world model, driving brain, and open humanoid robot

Nvidia launched new AI models at GTC Taipei including Cosmos 3 world model, Alpamayo 2 Super driving model, and an open...

The Decoder
Parallax: A Parameterized Local Linear Attention That Keeps Softmax and Adds a Learned Covariance Correction Branch
Research & Papers

Parallax: A Parameterized Local Linear Attention That Keeps Softmax and Adds a Learned Covariance Correction Branch

Parallax is a new parameterized local linear attention mechanism that improves upon Local Linear Attention by replacing...

MarkTechPost