OpenRating is an independent, open-source AI rating platform. Join the community

OOpenRating
All models

Alibaba · Released Mar 20, 2026

Qwen3.8 27B (xhigh)

A compact 27B parameter powerhouse capable of running locally or on single-GPU instances with near-frontier scores.

Open weights

OpenRating Index

52.1

0–100 composite

Reasoning Index

71.4

0–100 composite

Cost / Task

$0.26

benchmark run

Output speed

129 tok/s

blended tok/s

Time to first token

0.5s

median latency

Input price

$0.12

per 1M tokens

Output price

$0.48

per 1M tokens

Context window

128K

tokens

Strengths

  • Runs on single-GPU hardware with high inference efficiency

  • Rapid 129 tok/s output speed

  • Cost-effective alternative for private enterprise deployments

How we rate Qwen3.8 27B (xhigh)

OpenRating scores every model on the same axes. The OpenRating Index is a weighted composite of reasoning, knowledge, and agentic benchmarks, normalised to a 0–100 scale. Speed is blended output measured in tokens per second under standardised load, with time-to-first-token reported separately. Prices are list input and output prices per million tokens, and Task Cost measures end-to-end workload spending.

OpenRating Index

52.1

Reasoning Index

71.4

Token Pricing

Input per 1M tokens

$0.12

Output per 1M tokens

$0.48

Cost of 1M input + 1M output

$0.6

Task Cost Economics

Empirical cost breakdown for a standard OpenRating Index benchmark task ($0.26 total).

Reasoning (Chain of Thought)
$0.11
Answer generation
$0.030
Cache write overhead
$0.033
Cache hit queries
$0.076
Input payload
$0.004