vram.run Models Hardware Providers Cloud State of Inference
API provider data is live · Hardware & cloud pricing curated 2026-02-23

Featherless vs Novita

207 vs 67 models, 44 shared

Shared models

ModelFeatherless $/1M outFeatherless tok/sNovita $/1M outNovita tok/s
MiniMax-M286 tok/s
MiniMax-M2.172 tok/s
MiniMax-M2.567 tok/s
MiniMax-M2.740 tok/s
MiniMax-M339 tok/s
Qwen2.5-72B-Instruct27 tok/s
Qwen3-235B-A22B16 tok/s
Qwen3-235B-A22B-Thinking-250765 tok/s
Qwen3-Coder-480B-A35B-Instruct48 tok/s
Qwen3-Coder-Next120 tok/s
Qwen3-Next-80B-A3B-Instruct89 tok/s
Qwen3-VL-235B-A22B-Thinking43 tok/s
Qwen3-VL-30B-A3B-Instruct102 tok/s
Qwen3-VL-8B-Instruct73 tok/s
Qwen3.5-27B68 tok/s
Qwen3.5-397B-A17B73 tok/s
L3-8B-Lunaris-v171 tok/s
L3-8B-Stheno-v3.281 tok/s
DeepSeek-R1-052825 tok/s
DeepSeek-R1-Distill-Llama-70B60 tok/s
DeepSeek-V3-032437 tok/s
DeepSeek-V3.125 tok/s
DeepSeek-V3.1-Terminus27 tok/s
DeepSeek-V3.239 tok/s
DeepSeek-V4-Flash86 tok/s
DeepSeek-V4-Pro59 tok/s
gemma-4-26B-A4B-it34 tok/s
gemma-4-31B-it20 tok/s
Llama-3.1-8B-Instruct163 tok/s
Llama-3.3-70B-Instruct34 tok/s
Meta-Llama-3-70B-Instruct23 tok/s
Kimi-K2-Instruct25 tok/s
Kimi-K2-Instruct-090525 tok/s
Kimi-K2-Thinking62 tok/s
Kimi-K2.557 tok/s
Kimi-K2.684 tok/s
Kimi-K2.7-Code47 tok/s
gpt-oss-120b134 tok/s
gpt-oss-20b101 tok/s
GLM-4.657 tok/s
GLM-4.757 tok/s
GLM-4.7-Flash57 tok/s
GLM-535 tok/s
GLM-5.247 tok/s
Install CLI [email protected] Raw data · MIT · API data: live · HW/Cloud data: curated 2026-02-23 · v0.6.0