$ AICostSaverbeta
N

NVIDIA

synced

ModelTypeInputOutputContextLast sync
0A 01-ai/yi-large None LLM — — —
G3 Gemma 3 12B google LLM — — 131,072
G3 Gemma 3 4B google LLM — — 131,072
KK Kimi K2.6 moonshotai LLM — — 262,144
LG Llama Guard 4 12B meta-llama LLM — — 163,840
ML Mistral Large mistralai LLM — — 128,000
MG Muse Glimmer 30B meta LLM — — 131,072
A8 adept/fuyu-8b None LLM — — —
A1 ai21labs/jamba-1.5-large-instruct None LLM — — —
AL aisingapore/sea-lion-7b-instruct None LLM — — —
B1 bigcode/starcoder2-15b None LLM — — —
DI databricks/dbrx-instruct None LLM — — —
DA deepseek-ai/DeepSeek-V4.1-Flash None LLM — — —
DA deepseek-ai/deepseek-coder-6.7b-instruct None LLM — — —
G1 google/codegemma-1.1-7b None LLM — — —
G7 google/codegemma-7b None LLM — — —
G google/deplot None LLM — — —
G2 google/diffusiongemma-26b-a4b-it None LLM — — —
G2 google/gemma-2b None LLM — — —
G4 google/gemma-4-31B-it None LLM — — —
G2 google/recurrentgemma-2b None LLM — — —
GO gpt-oss-20b (batch) openai LLM — — 131,072
I3 ibm/granite-3.0-3b-a800m-instruct None LLM — — —
I3 ibm/granite-3.0-8b-instruct None LLM — — —
I3 ibm/granite-34b-code-instruct None LLM — — —
I8 ibm/granite-8b-code-instruct None LLM — — —
M7 meta/codellama-70b None LLM — — —
M3 meta/llama-3.2-11b-vision-instruct None LLM — — —
M3 meta/llama-3.2-90b-vision-instruct None LLM — — —
M7 meta/llama2-70b None LLM — — —
M2 microsoft/kosmos-2 None LLM — — —
M3 microsoft/phi-3-vision-128k-instruct None LLM — — —
M3 microsoft/phi-3.5-moe-instruct None LLM — — —
M2 mistralai/codestral-22b-instruct-v0.1 None LLM — — —
M7 mistralai/mistral-7b-instruct-v0.3 None LLM — — —
ML mistralai/mistral-large-2-instruct None LLM — — —
M8 mistralai/mixtral-8x22b-v0.1 None LLM — — —
MK moonshotai/Kimi-K3 None LLM — — —
NM nv-mistralai/mistral-nemo-12b-instruct None LLM — — —
NS nvidia/ai-synthetic-video-detector None Video — — —
NR nvidia/cosmos-reason2-8b None LLM — — —
NQ nvidia/embed-qa-4 None Embeddings — — —
NC nvidia/ising-calibration-1.5-31b None LLM — — —
N3 nvidia/llama-3.1-nemoguard-8b-content-safety None LLM — — —
N3 nvidia/llama-3.1-nemoguard-8b-topic-control None LLM — — —
N3 nvidia/llama-3.1-nemotron-51b-instruct None LLM — — —
N3 nvidia/llama-3.1-nemotron-70b-instruct None LLM — — —
N3 nvidia/llama-3.1-nemotron-safety-guard-8b-v3 None LLM — — —
N3 nvidia/llama-3.1-nemotron-ultra-253b-v1 None LLM — — —
N3 nvidia/llama-3.2-nemoretriever-1b-vlm-embed-v1 None Embeddings — — —
N3 nvidia/llama-3.2-nv-embedqa-1b-v1 None Embeddings — — —
NN nvidia/llama-nemotron-embed-vl-1b-v2 None Embeddings — — —
NC nvidia/llama3-chatqa-1.5-70b None LLM — — —
NN nvidia/mistral-nemo-minitron-8b-8k-instruct None LLM — — —
N3 nvidia/nemotron-3-embed-1b None Embeddings — — —
N3 nvidia/nemotron-3-nano-omni-30b-a3b-reasoning None LLM — — —
N3 nvidia/nemotron-3-super-120b-a12b None LLM — — —
N3 nvidia/nemotron-3-ultra-550b-a55b None LLM — — —
N3 nvidia/nemotron-3.5-content-safety None LLM — — —
N3 nvidia/nemotron-3.5-lightning-30b-a3b None LLM — — —
N4 nvidia/nemotron-4-340b-instruct None LLM — — —
N4 nvidia/nemotron-4-340b-reward None LLM — — —
NN nvidia/nemotron-nano-3-30b-a3b None LLM — — —
NP nvidia/nemotron-parse None LLM — — —
NP nvidia/nemotron-parse-2.0 None LLM — — —
N2 nvidia/neva-22b None LLM — — —
NE nvidia/nv-embedqa-mistral-7b-v2 None Embeddings — — —
N nvidia/nvclip None LLM — — —
NT nvidia/riva-translate-4b-instruct None LLM — — —
NT nvidia/riva-translate-4b-instruct-v1.1 None LLM — — —
NT nvidia/riva-translate-4b-instruct-v2 None LLM — — —
N nvidia/vila None LLM — — —
PX poolside/laguna-xs-2.1 None LLM — — —
SE snowflake/arctic-embed-l None Embeddings — — —
WC writer/palmyra-creative-122b None LLM — — —
WF writer/palmyra-fin-70b-32k None LLM — — —
WM writer/palmyra-med-70b None LLM — — —
WM writer/palmyra-med-70b-32k None LLM — — —
ZA z-ai/glm-5.3 None LLM — — —
ZA z-ai/glm-5.3-flash None LLM — — —
Z7 zyphra/zamba2-7b-instruct None LLM — — —