Data Incoming

Know your hardware. Run your models.

Designed around tokens/sec benchmarks, cost-per-GB tracking, and compatibility matrices across NVIDIA, Apple Silicon, and AMD -- for people who run models, not rent them.

Memory intelligence Interface preview
One view for compatibility, value, and performance. DATA MODEL / 01

Illustrative launch-preview data. Live price tracking is not connected yet.

Live GPU Offers Tracked
Cloud Providers
GPU Models Priced
H100 SXM Median $/GPU-hr

Launching 2026. Inference benchmarks, cost tracking, compatibility data.

Illustrative price signals / launch preview

The numbers behind every inference decision

Which GPU runs your model, how fast, and at what cost. Benchmarks across platforms, new and used pricing, and the setup guides to get running.

Planned capability

Tokens/sec Benchmarks

Real inference speed on real models. We test every GPU across Llama, DeepSeek, Mistral, and more -- with llama.cpp, MLX, and vLLM backends. Quantized and full precision.

Planned capability

Cost/GB Tracker

New retail, eBay used, Apple refurb. We track cost-per-GB of usable memory across every inference-capable GPU and unified memory config. Updated daily.

Planned capability

Model Compatibility Matrix

Can your GPU run it? Enter a model name and size. We show every GPU that fits it, the expected tokens/sec at each quantization level, and what it costs.

Planned capability

Setup Guides by Platform

Mac + ollama. NVIDIA + llama.cpp. AMD + ROCm. Step-by-step from unboxing to first inference, with benchmarked settings for your exact hardware.

Run models locally. Make better hardware decisions.

This launch preview maps the inference benchmarks, cost trackers, and compatibility data planned for the full product.