Skip to main content

Research & Benchmarks

Research, DemosWhitepapers

Technical deep-dives, product demonstrations, and research on enterprise AI infrastructure.

How we benchmark →
Autonomous Infrastructure Operations on Dell PowerEdge
BlogSeptember 2026

Autonomous Infrastructure Operations on Dell PowerEdge

Read article →
What happens when the KV cache skips host memory?
BlogSeptember 2026

What happens when the KV cache skips host memory?

What happens when compression moves to the SSD?
BlogSeptember 2026

What happens when compression moves to the SSD?

Scaling Vision Language Model Inference on Dell PowerEdge R770 with Intel Xeon 6
BlogJune 2026

Scaling Vision Language Model Inference on Dell PowerEdge R770 with Intel Xeon 6

Stop Paying Twice for the Same Tokens: Eliminating the KV Cache Recomputation Tax on the Dell PowerEdge XE9785 with AMD Instinct MI355X, Micron 9550 PRO, and LMCache
BlogJune 2026

Stop Paying Twice for the Same Tokens: Eliminating the KV Cache Recomputation Tax on the Dell PowerEdge XE9785 with AMD Instinct MI355X, Micron 9550 PRO, and LMCache

Secure Vector Ingestion at Enterprise Scale: Accelerating TLS-Protected Embedding Pipelines on Dell PowerEdge R770 with Intel QuickAssist Technology
BlogMay 2026

Secure Vector Ingestion at Enterprise Scale: Accelerating TLS-Protected Embedding Pipelines on Dell PowerEdge R770 with Intel QuickAssist Technology

Eliminate Redundant GPU Compute: How Solidigm D7-PS1010 NVMe KV Cache Offload Delivers 10x Faster AI Inference at 6.8x the Power Efficiency
BlogMarch 2026

Eliminate Redundant GPU Compute: How Solidigm D7-PS1010 NVMe KV Cache Offload Delivers 10x Faster AI Inference at 6.8x the Power Efficiency

The TCO Titan: MI355X vs. MI300X on Dell PowerEdge Servers with AMD Instinct Accelerators
BlogApril 2026

The TCO Titan: MI355X vs. MI300X on Dell PowerEdge Servers with AMD Instinct Accelerators

The FP4 Breakthrough: How MXFP4 Quantization Delivers Up to 6.1x Inference Throughput on Dell PowerEdge XE9785L with AMD Instinct MI355X Accelerators
BlogApril 2026

The FP4 Breakthrough: How MXFP4 Quantization Delivers Up to 6.1x Inference Throughput on Dell PowerEdge XE9785L with AMD Instinct MI355X Accelerators

The Free Upgrade: How ROCm 7 Unlocks Up to 2.2x More Throughput on Existing Dell PowerEdge XE9680 with AMD Instinct MI300X Hardware
BlogApril 2026

The Free Upgrade: How ROCm 7 Unlocks Up to 2.2x More Throughput on Existing Dell PowerEdge XE9680 with AMD Instinct MI300X Hardware

Day-0 Deployment: Running Kimi K2.5 at Enterprise Scale on a Single Dell PowerEdge XE9785L with AMD Instinct MI355X
BlogApril 2026

Day-0 Deployment: Running Kimi K2.5 at Enterprise Scale on a Single Dell PowerEdge XE9785L with AMD Instinct MI355X

Optimizing Computer Vision at Scale on Dell PowerEdge R770 with Intel Xeon 6
BlogJanuary 2026

Optimizing Computer Vision at Scale on Dell PowerEdge R770 with Intel Xeon 6

Accelerating Generative AI Image Workloads on Dell PowerEdge R770 with Intel Xeon 6
BlogJanuary 2026

Accelerating Generative AI Image Workloads on Dell PowerEdge R770 with Intel Xeon 6

Metrum Insights v3.7 - Side-by-Side Benchmarking, Smarter User Support Agent, and Support for Benchmarking Your Own Inference Endpoint
BlogDecember 2025

Metrum Insights v3.7 - Side-by-Side Benchmarking, Smarter User Support Agent, and Support for Benchmarking Your Own Inference Endpoint

Metrum Insights v3.6 - Upgraded Model Serving, Auto-Evaluation of Performance, and Support for Dynamic Prompt Generation and Reasoning Datasets
BlogNovember 2025

Metrum Insights v3.6 - Upgraded Model Serving, Auto-Evaluation of Performance, and Support for Dynamic Prompt Generation and Reasoning Datasets

Metrum Insights: A Comprehensive Methodology for LLM Inference Benchmarking
BlogNovember 2025

Metrum Insights: A Comprehensive Methodology for LLM Inference Benchmarking

Metrum AI Adds Support for NVIDIA DGX Spark in Metrum Insights
BlogOctober 2025

Metrum AI Adds Support for NVIDIA DGX Spark in Metrum Insights

Legislative & Fiscal Insights | Powered By Agentic RAG With AMD EPYC Processors on Dell PowerEdge Servers
BlogJune 2025

Legislative & Fiscal Insights | Powered By Agentic RAG With AMD EPYC Processors on Dell PowerEdge Servers

Accelerating Manufacturing Operations with Agentic RAG
BlogJune 2025

Accelerating Manufacturing Operations with Agentic RAG

Llama 4 Maverick on NVIDIA H200 versus B200 using vLLM: A Performance Analysis
BlogMay 2025

Llama 4 Maverick on NVIDIA H200 versus B200 using vLLM: A Performance Analysis

Accelerate Multi-Node Training with Dell PowerEdge XE9680 with Intel Gaudi 3 Accelerators and RoCE
BlogApril 2025

Accelerate Multi-Node Training with Dell PowerEdge XE9680 with Intel Gaudi 3 Accelerators and RoCE

Accelerating Distributed Fine-Tuning with RoCE: Performance Benchmarks on Dell PowerEdge XE9680 with Nvidia H200 Tensor Core GPUs
BlogApril 2025

Accelerating Distributed Fine-Tuning with RoCE: Performance Benchmarks on Dell PowerEdge XE9680 with Nvidia H200 Tensor Core GPUs

Deploy Next Generation Dell PowerEdge XE7745 Server for 5x AI Performance Gains
BlogMarch 2025

Deploy Next Generation Dell PowerEdge XE7745 Server for 5x AI Performance Gains

Autonomous Agents for IT Log Analysis: Transforming Infrastructure Management on Dell PowerEdge XE9680 with Nvidia H200 GPUs and Dell iDRAC
BlogDecember 2024

Autonomous Agents for IT Log Analysis: Transforming Infrastructure Management on Dell PowerEdge XE9680 with Nvidia H200 GPUs and Dell iDRAC

Enhancing Roof Damage Assessment and Reporting for Insurance Claims with RAG
BlogJanuary 2025

Enhancing Roof Damage Assessment and Reporting for Insurance Claims with RAG

Streamlining Technical Marketing with RAG-based Content Creation
BlogNovember 2024

Streamlining Technical Marketing with RAG-based Content Creation

Supercharge Multi-Node Training Performance with Dell PowerEdge XE9680 and RoCE
BlogOctober 2024

Supercharge Multi-Node Training Performance with Dell PowerEdge XE9680 and RoCE

Multimodal RAG-Based Healthcare Assistant on Dell PowerEdge XE9680 Rack Server with AMD Instinct MI300X Accelerators
BlogSeptember 2024

Multimodal RAG-Based Healthcare Assistant on Dell PowerEdge XE9680 Rack Server with AMD Instinct MI300X Accelerators

From RAGS to Riches | Industry First Multimodal RAG on 5th Gen Intel Xeon Processors
BlogAugust 2024

From RAGS to Riches | Industry First Multimodal RAG on 5th Gen Intel Xeon Processors

Putting AMD Instinct MI300X Accelerators to the Test on Dell PowerEdge XE9680 Rack Server With LORA Fine-tuning and vLLM Model Serving
BlogJune 2024

Putting AMD Instinct MI300X Accelerators to the Test on Dell PowerEdge XE9680 Rack Server With LORA Fine-tuning and vLLM Model Serving

Gen AI IT Log Analyzer | Enable IT Teams to Chat with Infrastructure on Dell PowerEdge R760xa Server with Nvidia H100 Data Center Tensor Core GPUs
BlogApril 2024

Gen AI IT Log Analyzer | Enable IT Teams to Chat with Infrastructure on Dell PowerEdge R760xa Server with Nvidia H100 Data Center Tensor Core GPUs

Enhancing Telco Quality of Service with Generative AI Agentic RAG
BlogFebruary 2025

Enhancing Telco Quality of Service with Generative AI Agentic RAG