RunInfraby RightNow
  • CatalogNew
  • Pricing
  • Research
  • Contact
DashboardSign inGet started
Home/Catalog/Qwen3.6 27B
Model

Qwen3.6 27B measured packages and prices

RunInfra has published 1 measured package for Qwen/Qwen3.6-27B on H100. Every package below keeps its engine, price, and proof date together.

Parameters
27B
Modality
Text
Base model license
Apache-2.0
Hugging Face model
Qwen/Qwen3.6-27B

The published packages are available to buy

Published packages for Qwen3.6 27B, with GPU, engine, optimization, price, and proof date
PackageGPUEngineOptimizationPriceVerified
qwen3-6-27b-fp8cd-v3-mlponly-h100-vllmH100vLLM 0.25.1Channelwise FP8, measured selective-layer recipe$40 one-timeJul 25, 2026

What we measured

Selective channelwise FP8 leaves the sensitive paths at higher precision while converting the remaining eligible layers.

qwen3-6-27b-fp8cd-v3-mlponly-h100-vllm

Published Jul 25, 2026. Proof verified Jul 25, 2026.

1.28x faster

Median latency measured 2857 ms at baseline and 2215 ms optimized.

Quality evidence

gsm8k 99.87% recovery, passed.

Published package measurements and prices for Qwen3.6 27B, grouped from the public RunInfra catalog.

Measured by RunInfra on H100 with vLLM 0.25.1; proof dates are listed with each package.

We report package measurements here, and each package remains subject to its listed license.

Citation: RunInfra (2026). Measured open-model serving benchmarks. https://runinfra.ai/catalog.

What is not measured

Other GPU measurements for this model not published.

Other engine measurements for this model not published.

Measurements for context lengths not listed in these packages not published.

Published facts answer common questions

What does Qwen3.6 27B cost on RunInfra?

qwen3-6-27b-fp8cd-v3-mlponly-h100-vllm is listed at $40 one-time, published Jul 25, 2026.

Which GPU serves Qwen3.6 27B?

qwen3-6-27b-fp8cd-v3-mlponly-h100-vllm is measured on H100 with vLLM 0.25.1. Proof verified Jul 25, 2026.

Is quality measured for Qwen3.6 27B?

gsm8k 99.87% recovery, passed. Proof verified Jul 25, 2026.

If you need custom optimization for a specific model, describe what you need

Describe the model and hardware you want optimized...
ModelsAuto engineAuto GPU
End-to-end encryption
Isolated GPU infrastructure
No training on your data
SOC 2 Type II
RunInfraby RightNow

© 2026 RunInfra. All rights reserved.

System status
Pipeline BuilderModelsCost CalculatorPricingStartupsBenchmarksDocsResearchNewsContact
Backed by
YCombinator
AICPA Type II
SOC 2
NVIDIA Inception ProgramNVIDIA Inception Program
Ask AI about RunInfra
Part of RightNow
SecurityDPAAUPCookiesTermsPrivacy