Loading...
Loading...
Quality Index
45.0
20th of 440
Top 5%
Coding Index
41.3
24th of 350
Top 7%
Price/1M
$1.35
513th cheapest
335% above median
Top 76%
Speed
57 tok/s
Top 43%
TTFT
1.55s
Context Window
262K
61st largest
Top 25%
Input
$0.60
per 1M tokens
Output
$3.60
per 1M tokens
Blended
$1.35
per 1M tokens
Cheaper than 24% of models. Median price is $0.31/1M tokens.
Daily
$1.35
Monthly
$40.50
57
tokens/sec
Faster than 57% of models
1.55
seconds
Faster than 17% of models
Market Median
46 tok/s
24% faster
Median TTFT
0.43s
261% slower
Throughput/Dollar
42
tok/s per $/1M
Speed Comparison
Context Window
262K
tokens
Larger than 75% of models
Max Output
66K
tokens
25% of context
565.4K
Downloads (30d)
139
Likes
apache-2.0
Permissive license allowing commercial use, modification, and distribution.
Hardware Requirements
Quantization Available
Check HuggingFace for GGUF, GPTQ, and AWQ quantized versions that require significantly less VRAM.