VRAM
calc
Calculator
All Models
Calculator
›
All Models
All Supported Models
29681 AI models · 198875 quantization formats tracked
A
Agga Min
1 models
llama-3-8b-Instruct-bnb-4bit-aiaustin-demo
8B
8K ctx
Q4_K_M
4.5GB file
6.63
GB
A
Aghorbani
1 models
h2o-danube2-1.8b-chat-gguf
1.8B
8K ctx
Q2_K
0.56GB file
2.59
GB
Q3_K_M
0.79GB file
2.82
GB
Q4_K_M
1.01GB file
3.04
GB
Q5_K_M
1.24GB file
3.27
GB
Q6_K
1.46GB file
3.49
GB
A
Agnuxo
93 models
Agente-Director-Qwen2-7B-Instruct_CODE_Python-GGUF_Spanish_English_8bit
7B
8K ctx
Q8_0
7.44GB file
9.55
GB
Agente-Director-Qwen2-7B-Instruct_CODE_Python-Spanish_English_GGUF_4bit
7B
8K ctx
Q4_K_M
3.94GB file
6.05
GB
Agente-Director-Qwen2-7B-Instruct_CODE_Python-Spanish_English_GGUF_q5_k
7B
8K ctx
Q5_K
4.81GB file
6.92
GB
Agente-Director-Qwen2-7B-Instruct_CODE_Python-Spanish_English_GGUF_q6_k
7B
8K ctx
Q6_K
5.69GB file
7.8
GB
Agente-Director-Qwen2-7Bron-Instruct_CODE_Python_English_GGUF_16bit
7B
8K ctx
FP16
14GB file
16.11
GB
Agente-GPT-Qwen-2.5-3B_GGUF_16bit
3B
8K ctx
FP16
6GB file
8.05
GB
Agente-GPT-Qwen-2.5-3B-Spanish_English_GGUF_32bit
3B
8K ctx
FP16
6GB file
8.05
GB
Agente-GPT-Qwen-2.5-7B_GGUF_16bit
7B
8K ctx
FP16
14GB file
16.11
GB
Agente-GPT-Qwen-2.5-7B-Spanish_8bit
7B
8K ctx
Q8_0
7.44GB file
9.55
GB
Agente-GPT-Qwen-2.5-7B-Spanish_English_GGUF_32bit
7B
8K ctx
FP16
14GB file
16.11
GB
Agente-GPT-Qwen-2.5-7B-Spanish_English_GGUF_4bit
7B
8K ctx
Q4_K_M
3.94GB file
6.05
GB
Agente-GPT-Qwen-2.5-7B-Spanish_English_GGUF_q5_k
7B
8K ctx
Q5_K
4.81GB file
6.92
GB
Agente-GPT-Qwen-2.5-7B-Spanish_English_GGUF_q6_k
7B
8K ctx
Q6_K
5.69GB file
7.8
GB
Agente-GPT-Qwen2.5-3B-Instruct-Spanish_8bit
3B
8K ctx
Q8_0
3.19GB file
5.24
GB
Agente-GPT-Qwen2.5-3B-Instruct-Spanish_English_GGUF_4bit
3B
8K ctx
Q4_K_M
1.69GB file
3.74
GB
Agente-GPT-Qwen2.5-3B-Instruct-Spanish_English_GGUF_q5_k
3B
8K ctx
Q5_K
2.06GB file
4.11
GB
Agente-GPT-Qwen2.5-3B-Instruct-Spanish_English_GGUF_q6_k
3B
8K ctx
Q6_K
2.44GB file
4.49
GB
Agente-Llama-3.1_GGUF_16bit
16B
8K ctx
FP16
32GB file
34.26
GB
Agente-Llama-3.1-Spanish_8bit
8B
8K ctx
Q8_0
8.5GB file
10.63
GB
Agente-Llama-3.1-Spanish_English_GGUF_32bit
32B
8K ctx
FP16
64GB file
66.52
GB
Agente-Llama-3.1-Spanish_English_GGUF_4bit
4B
8K ctx
Q4_K_M
2.25GB file
4.32
GB
Agente-Llama-3.1-Spanish_English_GGUF_q5_k
8B
8K ctx
Q5_K
5.5GB file
7.63
GB
BF16
16GB file
18.13
GB
Agente-Llama-3.1-Spanish_English_GGUF_q6_k
8B
8K ctx
Q6_K
6.5GB file
8.63
GB
BF16
16GB file
18.13
GB
CAJAL-4B
4B
8K ctx
Q4_0
2.25GB file
4.32
GB
Q8_0
4.25GB file
6.32
GB
cajal-9b-v2-f16-gguf
9B
8K ctx
FP16
18GB file
20.15
GB
cajal-9b-v2-q4_k_m
9B
8K ctx
Q4_0
5.06GB file
7.21
GB
FP16
18GB file
20.15
GB
cajal-9b-v2-q5_k_m
9B
8K ctx
Q5_0
6.19GB file
8.34
GB
FP16
18GB file
20.15
GB
cajal-9b-v2-q6_k
9B
8K ctx
Q6_K
7.31GB file
9.46
GB
cajal-9b-v2-q8_0
9B
8K ctx
Q8_0
9.56GB file
11.71
GB
gemma-2-2b-instruct-python_CODE_assistant-GGUF_16bit
2B
8K ctx
FP16
4GB file
6.03
GB
gemma-2-2b-instruct-python_CODE_assistant-GGUF_4bit
2B
8K ctx
Q4_K_M
1.13GB file
3.16
GB
BF16
4GB file
6.03
GB
gemma-2-2b-Python_CODE_assistant-GGUF_8bit
2B
8K ctx
Q8_0
2.13GB file
4.16
GB
Llama-3.1-Minitron-4B-Instruct_CODE_Python-GGUF_Spanish_English_8bit
4B
8K ctx
Q8_0
4.25GB file
6.32
GB
Llama-3.1-Minitron-4B-Instruct_CODE_Python-Spanish_English_GGUF_16bit
4B
8K ctx
FP16
8GB file
10.07
GB
Llama-3.1-Minitron-4B-Instruct_CODE_Python-Spanish_English_GGUF_4bit
4B
8K ctx
Q4_K_M
2.25GB file
4.32
GB
BF16
8GB file
10.07
GB
Llama-3.1-Minitron-4B-Instruct_CODE_Python-Spanish_English_GGUF_q5_k
4B
8K ctx
Q5_K
2.75GB file
4.82
GB
BF16
8GB file
10.07
GB
Mamba-Codestral-7B-Instruct_CODE_Python-GGUF_Spanish_English_8bit
7B
8K ctx
Q8_0
7.44GB file
9.55
GB
Mamba-Codestral-7B-Instruct_CODE_Python-Spanish_English_GGUF_16bit
7B
8K ctx
FP16
14GB file
16.11
GB
Mamba-Codestral-7B-Instruct_CODE_Python-Spanish_English_GGUF_4bit
7B
8K ctx
Q4_K_M
3.94GB file
6.05
GB
Mamba-Codestral-7B-v0.1-instruct-python_coding_assistant-GGUF_16bit
7B
8K ctx
FP16
14GB file
16.11
GB
Mamba-Codestral-7B-v0.1-instruct-python_coding_assistant-GGUF_4bit
7B
8K ctx
Q4_K_M
3.94GB file
6.05
GB
Mamba-Codestral-7B-v0.1-python_coding_assistant-GGUF_8bit
7B
8K ctx
Q8_0
7.44GB file
9.55
GB
Meta-Llama-3.1-8B-CODE-Alpaca-Python-8bit-GGUF
8B
8K ctx
Q8_0
8.5GB file
10.63
GB
Mistral-Nemo-CODE-Python_assistant-GGUF_16bit
16B
8K ctx
FP16
32GB file
34.26
GB
Mistral-Nemo-CODE-Python_assistant-GGUF_8bit
8B
8K ctx
Q8_0
8.5GB file
10.63
GB
Mistral-NeMo-Minitron-8B-Alpaca-CODE-Python-GGUF-16bit
8B
8K ctx
FP16
16GB file
18.13
GB
Mistral-NeMo-Minitron-8B-Alpaca-CODE-Python-GGUF-8bit
8B
8K ctx
Q8_0
8.5GB file
10.63
GB
Phi-3.5-mini-instruct-python_coding_assistant-GGUF_16bit
16B
8K ctx
FP16
32GB file
34.26
GB
Phi-3.5-mini-instruct-python_coding_assistant-GGUF_4bit
4B
8K ctx
Q4_K_M
2.25GB file
4.32
GB
Phi-3.5-mini-instruct-python_coding_assistant-GGUF_8bit
8B
8K ctx
Q8_0
8.5GB file
10.63
GB
Qwen2_0.5B-GGUF_Spanish_English_raspberry_pi_8bit
0.5B
8K ctx
Q8_0
0.53GB file
2.54
GB
Qwen2_0.5B-Spanish_English_raspberry_pi_GGUF_16bit
0.5B
8K ctx
FP16
1GB file
3.01
GB
Qwen2_0.5B-Spanish_English_raspberry_pi_GGUF_4bit
0.5B
8K ctx
Q4_K_M
0.28GB file
2.29
GB
Qwen2-1.5B-Instruct_MOE_assistant-GGUF_16bit
1.5B
8K ctx
FP16
3GB file
5.02
GB
Qwen2-1.5B-Instruct_MOE_assistant-GGUF_4bit
1.5B
8K ctx
Q4_K_M
0.84GB file
2.86
GB
BF16
3GB file
5.02
GB
Qwen2-1.5B-Instruct_MOE_assistant-GGUF_8bit
1.5B
8K ctx
Q8_0
1.59GB file
3.61
GB
Qwen2-1.5B-Instruct_MOE_BIOLOGY_assistant-GGUF_16bit
1.5B
8K ctx
FP16
3GB file
5.02
GB
Qwen2-1.5B-Instruct_MOE_BIOLOGY_assistant-GGUF_4bit
1.5B
8K ctx
Q4_K_M
0.84GB file
2.86
GB
BF16
3GB file
5.02
GB
Qwen2-1.5B-Instruct_MOE_BIOLOGY_assistant-GGUF_8bit
1.5B
8K ctx
Q8_0
1.59GB file
3.61
GB
Qwen2-1.5B-Instruct_MOE_CODE_assistant-GGUF_16bit
1.5B
8K ctx
FP16
3GB file
5.02
GB
Qwen2-1.5B-Instruct_MOE_CODE_assistant-GGUF_4bit
1.5B
8K ctx
Q4_K_M
0.84GB file
2.86
GB
BF16
3GB file
5.02
GB
Qwen2-1.5B-Instruct_MOE_CODE_assistant-GGUF_8bit
1.5B
8K ctx
Q8_0
1.59GB file
3.61
GB
Qwen2-1.5B-Instruct_MOE_Director-GGUF_16bit
1.5B
8K ctx
FP16
3GB file
5.02
GB
Qwen2-1.5B-Instruct_MOE_Director-GGUF_4bit
1.5B
8K ctx
Q4_K_M
0.84GB file
2.86
GB
BF16
3GB file
5.02
GB
Qwen2-1.5B-Instruct_MOE_Director-GGUF_8bit
1.5B
8K ctx
Q8_0
1.59GB file
3.61
GB
Qwen2-7B-Instruct_CODE_Python-GGUF_Spanish_English_8bit
7B
8K ctx
Q8_0
7.44GB file
9.55
GB
Qwen2-7B-Instruct_CODE_Python-Spanish_English_GGUF_16bit
7B
8K ctx
FP16
14GB file
16.11
GB
Qwen2-7B-Instruct_CODE_Python-Spanish_English_GGUF_4bit
7B
8K ctx
Q4_K_M
3.94GB file
6.05
GB
BF16
14GB file
16.11
GB
Qwen2-7B-v2-Instruct_CODE_Python-GGUF_Spanish_English_8bit
7B
8K ctx
Q8_0
7.44GB file
9.55
GB
Qwen2-7B-v2-Instruct_CODE_Python-Spanish_English_GGUF_16bit
7B
8K ctx
FP16
14GB file
16.11
GB
Qwen2-7B-v2-Instruct_CODE_Python-Spanish_English_GGUF_32bit
7B
8K ctx
FP16
14GB file
16.11
GB
Qwen2-7B-v2-Instruct_CODE_Python-Spanish_English_GGUF_4bit
7B
8K ctx
Q4_K_M
3.94GB file
6.05
GB
BF16
14GB file
16.11
GB
tiny-llama-GGUF_Spanish_English_raspberry_pi_8bit
8B
8K ctx
Q8_0
8.5GB file
10.63
GB
tiny-llama-Spanish_English_raspberry_pi_GGUF_16bit
16B
8K ctx
FP16
32GB file
34.26
GB
tiny-llama-Spanish_English_raspberry_pi_GGUF_4bit
4B
8K ctx
Q4_K_M
2.25GB file
4.32
GB
Tinytron-1B-TinyLlama-Instruct_CODE_Python-GGUF_Spanish_English_8bit
1B
8K ctx
Q8_0
1.06GB file
3.08
GB
Tinytron-1B-TinyLlama-Instruct_CODE_Python-Spanish_English_GGUF_16bit
1B
8K ctx
FP16
2GB file
4.02
GB
Tinytron-1B-TinyLlama-Instruct_CODE_Python-Spanish_English_GGUF_4bit
1B
8K ctx
Q4_K_M
0.56GB file
2.58
GB
Tinytron-1B-TinyLlama-Instruct_CODE_Python-Spanish_English_GGUF_q5_k
1B
8K ctx
Q5_K
0.69GB file
2.71
GB
BF16
2GB file
4.02
GB
Tinytron-1B-TinyLlama-Instruct_CODE_Python-Spanish_English_GGUF_q6_k
1B
8K ctx
Q6_K
0.81GB file
2.83
GB
BF16
2GB file
4.02
GB
Tinytron-ORCA-3B-Instruct_CODE_Python_English_Asistant-16bit-v2
3B
8K ctx
FP16
6GB file
8.05
GB
Tinytron-ORCA-3B-Instruct_CODE_Python-GGUF_Spanish_English_8bit
3B
8K ctx
Q8_0
3.19GB file
5.24
GB
Tinytron-ORCA-3B-Instruct_CODE_Python-Spanish_English_GGUF_4bit
3B
8K ctx
Q4_K_M
1.69GB file
3.74
GB
BF16
6GB file
8.05
GB
Tinytron-ORCA-3B-TinyLlama-Instruct_CODE_Python-extra_small_quantization_GGUF_3bit
3B
8K ctx
BF16
6GB file
8.05
GB
Tinytron-ORCA-7B-Instruct_CODE_Python_English_GGUF_16bit
7B
8K ctx
FP16
14GB file
16.11
GB
Tinytron-ORCA-7B-Instruct_CODE_Python-GGUF_Spanish_English_8bit
7B
8K ctx
Q8_0
7.44GB file
9.55
GB
Tinytron-ORCA-7B-Instruct_CODE_Python-Spanish_English_GGUF_4bit
7B
8K ctx
Q4_K_M
3.94GB file
6.05
GB
BF16
14GB file
16.11
GB
Tinytron-ORCA-7B-TinyLlama-Instruct_CODE_Python-extra_small_quantization_GGUF_3bit
7B
8K ctx
BF16
14GB file
16.11
GB
Tinytron-TinyLlama-Instruct_CODE_Python-GGUF_Spanish_English_8bit
8B
8K ctx
Q8_0
8.5GB file
10.63
GB
Tinytron-TinyLlama-Instruct_CODE_Python-Spanish_English_GGUF_16bit
16B
8K ctx
FP16
32GB file
34.26
GB
Tinytron-TinyLlama-Instruct_CODE_Python-Spanish_English_GGUF_4bit
4B
8K ctx
Q4_K_M
2.25GB file
4.32
GB
Tinytron-TinyLlama-Instruct_CODE_Python-Spanish_English_GGUF_q5_k
Unknown
8K ctx
BF16
0GB file
2
GB
Q5_K
0GB file
2
GB
Tinytron-TinyLlama-Instruct_CODE_Python-Spanish_English_GGUF_q6_k
Unknown
8K ctx
BF16
0GB file
2
GB
Q6_K
0GB file
2
GB
A
Agtian
1 models
llamas
Unknown
8K ctx
Q4_K_M
0GB file
2
GB
A
Ah Beyond
1 models
ERNIE-4.5-0.3B-Base-BF16-gguf
0.3B
8K ctx
BF16
0.6GB file
2.6
GB
A
Ahashan Habib Eatl
2 models
bloomz-1b1-Q4_K_M-GGUF
1B
8K ctx
Q4_0
0.56GB file
2.58
GB
FP16
2GB file
4.02
GB
Llama-3.1-8B-Instruct-Q2_K-GGUF
8B
8K ctx
Q2_0
2.5GB file
4.63
GB
FP16
16GB file
18.13
GB
A
Ahmad Abdo Shbat
1 models
TIA2.1-14B-GGUF
14B
8K ctx
Q8_0
14.88GB file
17.11
GB
A
Ahmadabdulnasir
1 models
NaijaPidgin-Qwen3-4B
4B
8K ctx
Q4_K_M
2.25GB file
4.32
GB
Q8_0
4.25GB file
6.32
GB
A
Ahmedx Saad
1 models
Qwen3Guard-Gen-4B-Q4_K_M-GGUF
4B
8K ctx
Q4_0
2.25GB file
4.32
GB
FP16
8GB file
10.07
GB
A
Ai A F
4 models
bluey-8B-Q4_K_M-GGUF
8B
8K ctx
Q4_0
4.5GB file
6.63
GB
FP16
16GB file
18.13
GB
Midnight-Miqu-70B-v1.5_GGUF
70B
8K ctx
Q2_K
21.88GB file
25.03
GB
Q3_K_L
30.63GB file
33.78
GB
Q3_K_M
30.63GB file
33.78
GB
Q3_K_S
30.63GB file
33.78
GB
Q4_0
39.38GB file
42.53
GB
Q4_1
39.38GB file
42.53
GB
Q4_K_M
39.38GB file
42.53
GB
Q4_K_S
39.38GB file
42.53
GB
Q5_0
48.13GB file
51.28
GB
Q5_1
48.13GB file
51.28
GB
Q5_K_M
48.13GB file
51.28
GB
Q5_K_S
48.13GB file
51.28
GB
Q6_K
56.88GB file
60.03
GB
Q8_0
74.38GB file
77.53
GB
FP16
140GB file
143.15
GB
Pretrained-Merged-QLoRA-r9kilo-V1.1_Ckpt-80
7B
8K ctx
Q2_K
2.19GB file
4.3
GB
Q3_K_L
3.06GB file
5.17
GB
Q3_K_M
3.06GB file
5.17
GB
Q3_K_S
3.06GB file
5.17
GB
Q4_0
3.94GB file
6.05
GB
Q4_K_M
3.94GB file
6.05
GB
Q4_K_S
3.94GB file
6.05
GB
Q5_0
4.81GB file
6.92
GB
Q5_K_M
4.81GB file
6.92
GB
Q5_K_S
4.81GB file
6.92
GB
Q6_K
5.69GB file
7.8
GB
Q8_0
7.44GB file
9.55
GB
rp-sft-merged_1000_GGUF
Unknown
8K ctx
Q2_K
0GB file
2
GB
Q3_K_L
0GB file
2
GB
Q3_K_M
0GB file
2
GB
Q3_K_S
0GB file
2
GB
Q4_0
0GB file
2
GB
Q4_1
0GB file
2
GB
Q4_K_M
0GB file
2
GB
Q4_K_S
0GB file
2
GB
Q5_0
0GB file
2
GB
Q5_1
0GB file
2
GB
Q5_K_M
0GB file
2
GB
Q5_K_S
0GB file
2
GB
Q6_K
0GB file
2
GB
Q8_0
0GB file
2
GB
Previous
Next
Showing
201
to
210
of
5981
results
1
2
...
18
19
20
21
22
23
24
...
598
599