RunInfraby RightNow
  • Pricing
  • Research
  • Contact
DashboardSign inGet started
Home/Methodology

Measurement methodology

The public method behind every measured claim

Method version one states the current public contract.

It changes when the contract changes. Measurement dates change only when the underlying evidence changes.

What we measure

A published row names its metric before it states a result.

Throughput

Tokens per second expresses how quickly a serving system produces tokens over an interval.

Time to first token

Time to first token measures elapsed time from request arrival until the first generated token becomes available.

Fiftieth and ninety-ninth percentile latency

The fiftieth and ninety-ninth percentiles mark different positions in an ordered latency distribution.

Quality scores against a baseline

Quantization quality recovery is the measured retention of task performance after a model moves to lower precision.

How a number is produced

  • Baseline and candidate runs use the same GPU, request protocol, and task definition.
  • Serving engine versions stay pinned beside the measurement, so later releases do not rewrite history.
  • Benchmark receipts record the run conditions and the date attached to each published claim.
  • Quality uses the same protocol and scoring method as its recorded baseline.
  • Derived cost stays labeled as derived from the recorded rate and measured throughput.
  • Comparison ratios are computed at render time from both absolutes and are never stored.

What we publish and refuse to publish

  • We publish measured rows, their conditions, dates, source status, and known absences.
  • An unmeasured combination renders an absence. We do not project it or fill it from a vendor table.
  • Vendor specifications appear only as labeled citations with a source and retrieval date.
  • A cited specification never becomes our measured data.
  • We do not present vendor-copied numbers as our measurements.

Versions and dates

  • Every measured claim carries an as-of or verified date from its owning evidence.
  • Update cadence follows measurement changes, not deploy timing.
  • A deploy alone does not advance a measurement date.
  • The visible method version changes when this public contract changes.

Reproduction trail

  • Every published row keeps a reproduction trail to its recorded conditions.
  • Each published package download includes its signed benchmark receipt.
  • The receipt records the conditions, engine version, date, and evidence shipped with that package.
  • Comparison rows preserve their source conditions verbatim, including workload, hardware, precision, concurrency, and statistic identity.
  • Reproduce a row under those conditions, compare both absolutes, then derive any ratio.
RunInfraby RightNow

© 2026 RunInfra. All rights reserved.

System status
Pipeline BuilderModel APIsCost CalculatorPricingStartupsBenchmarksDocsResearchNewsContact
Backed by
YCombinator
AICPA Type II
SOC 2
NVIDIA Inception ProgramNVIDIA Inception Program
Ask AI about RunInfra
Part of RightNow
SecurityDPAAUPCookiesTermsPrivacy