MODEL DIRECTORY

Supported Local AI Models

Browse verified GGUF text and reasoning models ready for local execution with Ollama, LM Studio, or llama.cpp. Click any model to calculate required VRAM, RAM and tokens per second for your setup.

Showing 177 models Updated continuously from Hugging Face
unsloth GGUF Ready

Qwen3-Coder-30B-A3B-Instruct

unsloth/Qwen3-Coder-30B-A3B-Instruct-GGUF

12,834,876 1018
Calculate
LiquidAI GGUF Ready

LFM2.5-2.6B

LiquidAI/LFM2.5-2.6B-GGUF

1,149,670 342
Calculate
LiquidAI GGUF Ready

LFM2.5-230M

LiquidAI/LFM2.5-230M-GGUF

588,021 98
Calculate
unsloth GGUF Ready

gpt-oss-20b

unsloth/gpt-oss-20b-GGUF

552,424 820
Calculate
unsloth GGUF Ready

Ornith-1.0-9B

unsloth/Ornith-1.0-9B-GGUF

534,817 62
Calculate
LiquidAI GGUF Ready

LFM2.5-8B-A1B

LiquidAI/LFM2.5-8B-A1B-GGUF

495,233 297
Calculate
unsloth GGUF Ready

Qwen3-4B

unsloth/Qwen3-4B-GGUF

494,362 240
Calculate
unsloth GGUF Ready

GLM-5.3-Flash

unsloth/GLM-5.3-Flash-GGUF

457,926 408
Calculate
bartowski GGUF Ready

DeepSeek-Coder-V2-Lite-Instruct

bartowski/DeepSeek-Coder-V2-Lite-Instruct-GGUF

449,414 190
Calculate
unsloth GGUF Ready

Qwen-AgentWorld-35B-A3B

unsloth/Qwen-AgentWorld-35B-A3B-GGUF

434,894 244
Calculate
unsloth GGUF Ready

GLM-5.3

unsloth/GLM-5.3-GGUF

409,762 78
Calculate
bartowski GGUF Ready

Kwaipilot_KAT-Coder-V2.5-Dev

bartowski/Kwaipilot_KAT-Coder-V2.5-Dev-GGUF

395,181 131
Calculate
unsloth GGUF Ready

GLM-5.2

unsloth/GLM-5.2-GGUF

385,343 643
Calculate
Qwen GGUF Ready

Qwen3-8B

Qwen/Qwen3-8B-GGUF

368,660 276
Calculate
Qwen GGUF Ready

Qwen2.5-3B-Instruct

Qwen/Qwen2.5-3B-Instruct-GGUF

351,812 184
Calculate
LiquidAI GGUF Ready

LFM2.5-1.2B-Instruct

LiquidAI/LFM2.5-1.2B-Instruct-GGUF

337,254 218
Calculate
bartowski GGUF Ready

Meta-Llama-3.1-8B-Instruct

bartowski/Meta-Llama-3.1-8B-Instruct-GGUF

295,119 402
Calculate
ggml-org GGUF Ready

NVIDIA-Nemotron-3.5-Lightning-30B-A3B

ggml-org/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF

274,796 31
Calculate
Qwen GGUF Ready

Qwen2.5-Coder-7B-Instruct

Qwen/Qwen2.5-Coder-7B-Instruct-GGUF

274,702 452
Calculate
janhq GGUF Ready

Jan-v3.5-4B

janhq/Jan-v3.5-4B-gguf

274,492 37
Calculate
MaziyarPanahi GGUF Ready

Qwen3-0.6B

MaziyarPanahi/Qwen3-0.6B-GGUF

274,290 15
Calculate
MaziyarPanahi GGUF Ready

Qwen3-14B

MaziyarPanahi/Qwen3-14B-GGUF

260,844 11
Calculate
MaziyarPanahi GGUF Ready

Qwen3-1.7B

MaziyarPanahi/Qwen3-1.7B-GGUF

258,323 11
Calculate
bartowski GGUF Ready

Qwen_Qwen3-Next-80B-A3B-Thinking

bartowski/Qwen_Qwen3-Next-80B-A3B-Thinking-GGUF

256,235 16
Calculate
MaziyarPanahi GGUF Ready

Qwen3-30B-A3B

MaziyarPanahi/Qwen3-30B-A3B-GGUF

254,592 6
Calculate
MaziyarPanahi GGUF Ready

Qwen3-32B

MaziyarPanahi/Qwen3-32B-GGUF

254,405 2
Calculate
unsloth GGUF Ready

Llama-3.2-3B-Instruct

unsloth/Llama-3.2-3B-Instruct-GGUF

238,975 79
Calculate
MaziyarPanahi GGUF Ready

Yi-Coder-9B-Chat

MaziyarPanahi/Yi-Coder-9B-Chat-GGUF

222,384 9
Calculate
bartowski GGUF Ready

Qwen2.5-7B-Instruct

bartowski/Qwen2.5-7B-Instruct-GGUF

206,967 75
Calculate
unsloth GGUF Ready

DeepSeek-R1-Distill-Llama-70B

unsloth/DeepSeek-R1-Distill-Llama-70B-GGUF

201,056 121
Calculate
Qwen GGUF Ready

Qwen2.5-1.5B-Instruct

Qwen/Qwen2.5-1.5B-Instruct-GGUF

196,886 156
Calculate
unsloth GGUF Ready

Qwen3-Coder-Next

unsloth/Qwen3-Coder-Next-GGUF

195,568 841
Calculate
Qwen GGUF Ready

Qwen2.5-0.5B-Instruct

Qwen/Qwen2.5-0.5B-Instruct-GGUF

193,524 130
Calculate
MaziyarPanahi GGUF Ready

Qwen3-4B-Instruct-2507

MaziyarPanahi/Qwen3-4B-Instruct-2507-GGUF

188,915 3
Calculate
ggml-org GGUF Ready

gpt-oss-120b

ggml-org/gpt-oss-120b-GGUF

177,621 76
Calculate
MaziyarPanahi GGUF Ready

Yi-Coder-1.5B-Chat

MaziyarPanahi/Yi-Coder-1.5B-Chat-GGUF

168,629 20
Calculate
MaziyarPanahi GGUF Ready

Meta-Llama-3-8B-Instruct

MaziyarPanahi/Meta-Llama-3-8B-Instruct-GGUF

166,310 103
Calculate
MaziyarPanahi GGUF Ready

Phi-3.5-mini-instruct

MaziyarPanahi/Phi-3.5-mini-instruct-GGUF

160,658 34
Calculate
MaziyarPanahi GGUF Ready

Mixtral-8x22B-v0.1

MaziyarPanahi/Mixtral-8x22B-v0.1-GGUF

158,504 78
Calculate
MaziyarPanahi GGUF Ready

Mistral-7B-Instruct-v0.3

MaziyarPanahi/Mistral-7B-Instruct-v0.3-GGUF

158,477 147
Calculate
bartowski GGUF Ready

Hermes-3-Llama-3.1-70B

bartowski/Hermes-3-Llama-3.1-70B-GGUF

157,751 10
Calculate
MaziyarPanahi GGUF Ready

Llama-3.2-1B-Instruct

MaziyarPanahi/Llama-3.2-1B-Instruct-GGUF

157,150 18
Calculate
MaziyarPanahi GGUF Ready

gemma-3-4b-it

MaziyarPanahi/gemma-3-4b-it-GGUF

154,067 20
Calculate
MaziyarPanahi GGUF Ready

Llama-3-8B-Instruct-32k-v0.1

MaziyarPanahi/Llama-3-8B-Instruct-32k-v0.1-GGUF

153,817 59
Calculate
google GGUF Ready

gemma-2b

google/gemma-2b

152,870 1236
Calculate
MaziyarPanahi GGUF Ready

solar-pro-preview-instruct

MaziyarPanahi/solar-pro-preview-instruct-GGUF

152,469 29
Calculate
MaziyarPanahi GGUF Ready

DeepSeek-R1-0528-Qwen3-8B

MaziyarPanahi/DeepSeek-R1-0528-Qwen3-8B-GGUF

151,771 11
Calculate
MaziyarPanahi GGUF Ready

Mistral-Small-24B-Instruct-2501

MaziyarPanahi/Mistral-Small-24B-Instruct-2501-GGUF

151,428 12
Calculate
MaziyarPanahi GGUF Ready

Phi-4-mini-instruct

MaziyarPanahi/Phi-4-mini-instruct-GGUF

150,528 16
Calculate
MaziyarPanahi GGUF Ready

Llama-3.3-70B-Instruct

MaziyarPanahi/Llama-3.3-70B-Instruct-GGUF

150,144 22
Calculate
MaziyarPanahi GGUF Ready

phi-4

MaziyarPanahi/phi-4-GGUF

149,977 10
Calculate
MaziyarPanahi GGUF Ready

Mistral-Nemo-Instruct-2407

MaziyarPanahi/Mistral-Nemo-Instruct-2407-GGUF

149,803 55
Calculate
MaziyarPanahi GGUF Ready

gemma-3-1b-it

MaziyarPanahi/gemma-3-1b-it-GGUF

149,427 13
Calculate
MaziyarPanahi GGUF Ready

QwQ-32B

MaziyarPanahi/QwQ-32B-GGUF

149,111 4
Calculate
MaziyarPanahi GGUF Ready

Qwen2-7B-Instruct

MaziyarPanahi/Qwen2-7B-Instruct-GGUF

149,075 10
Calculate
MaziyarPanahi GGUF Ready

gemma-2-2b-it

MaziyarPanahi/gemma-2-2b-it-GGUF

148,947 15
Calculate
MaziyarPanahi GGUF Ready

Mistral-Small-Instruct-2409

MaziyarPanahi/Mistral-Small-Instruct-2409-GGUF

148,782 4
Calculate
MaziyarPanahi GGUF Ready

Ministral-3-3B-Reasoning-2512

MaziyarPanahi/Ministral-3-3B-Reasoning-2512-GGUF

148,782 6
Calculate
MaziyarPanahi GGUF Ready

gemma-3-12b-it

MaziyarPanahi/gemma-3-12b-it-GGUF

148,263 19
Calculate
MaziyarPanahi GGUF Ready

mistral-small-3.1-24b-instruct-2503-hf

MaziyarPanahi/mistral-small-3.1-24b-instruct-2503-hf-GGUF

148,225 2
Calculate
MaziyarPanahi GGUF Ready

DeepSeek-V3-0324

MaziyarPanahi/DeepSeek-V3-0324-GGUF

148,170 23
Calculate
MaziyarPanahi GGUF Ready

Meta-Llama-3.1-70B-Instruct

MaziyarPanahi/Meta-Llama-3.1-70B-Instruct-GGUF

147,892 41
Calculate
MaziyarPanahi GGUF Ready

Yi-1.5-6B-Chat

MaziyarPanahi/Yi-1.5-6B-Chat-GGUF

147,516 9
Calculate
MaziyarPanahi GGUF Ready

WizardLM-2-7B

MaziyarPanahi/WizardLM-2-7B-GGUF

147,363 83
Calculate
MaziyarPanahi GGUF Ready

gemma-3-27b-it

MaziyarPanahi/gemma-3-27b-it-GGUF

146,740 8
Calculate
MaziyarPanahi GGUF Ready

mathstral-7B-v0.1

MaziyarPanahi/mathstral-7B-v0.1-GGUF

146,644 8
Calculate
MaziyarPanahi GGUF Ready

Llama-3-8B-Instruct-64k

MaziyarPanahi/Llama-3-8B-Instruct-64k-GGUF

146,564 14
Calculate
MaziyarPanahi GGUF Ready

Mistral-Large-Instruct-2411

MaziyarPanahi/Mistral-Large-Instruct-2411-GGUF

146,455 2
Calculate
MaziyarPanahi GGUF Ready

INTELLECT-2

MaziyarPanahi/INTELLECT-2-GGUF

146,281 3
Calculate
MaziyarPanahi GGUF Ready

Meta-Llama-3.1-405B-Instruct

MaziyarPanahi/Meta-Llama-3.1-405B-Instruct-GGUF

146,060 15
Calculate
MaziyarPanahi GGUF Ready

firefunction-v2

MaziyarPanahi/firefunction-v2-GGUF

146,014 19
Calculate
unsloth GGUF Ready

GLM-4.7-Flash

unsloth/GLM-4.7-Flash-GGUF

134,735 705
Calculate
LiquidAI GGUF Ready

LFM2.5-2.6B-DSpark

LiquidAI/LFM2.5-2.6B-DSpark-GGUF

126,342 49
Calculate
Qwen GGUF Ready

Qwen2.5-Coder-1.5B-Instruct

Qwen/Qwen2.5-Coder-1.5B-Instruct-GGUF

113,790 114
Calculate
Qwen GGUF Ready

Qwen2.5-Coder-32B-Instruct

Qwen/Qwen2.5-Coder-32B-Instruct-GGUF

107,613 223
Calculate
bartowski GGUF Ready

Qwen2.5-72B-Instruct

bartowski/Qwen2.5-72B-Instruct-GGUF

106,511 47
Calculate
MaziyarPanahi GGUF Ready

Qwen3-30B-A3B-Instruct-2507

MaziyarPanahi/Qwen3-30B-A3B-Instruct-2507-GGUF

105,258 4
Calculate
Qwen GGUF Ready

Qwen2.5-Coder-14B-Instruct

Qwen/Qwen2.5-Coder-14B-Instruct-GGUF

101,439 216
Calculate
bartowski GGUF Ready

Qwen_Qwen3-14B

bartowski/Qwen_Qwen3-14B-GGUF

99,311 40
Calculate
unsloth GGUF Ready

DeepSeek-R1-Distill-Qwen-1.5B

unsloth/DeepSeek-R1-Distill-Qwen-1.5B-GGUF

89,539 160
Calculate
ggml-org GGUF Ready

stories15M_MOE

ggml-org/stories15M_MOE

87,095 8
Calculate
MaziyarPanahi GGUF Ready

GLM-4.6V-Flash

MaziyarPanahi/GLM-4.6V-Flash-GGUF

84,160 6
Calculate
MaziyarPanahi GGUF Ready

gpt-oss-20b-Derestricted

MaziyarPanahi/gpt-oss-20b-Derestricted-GGUF

83,102 3
Calculate
LiquidAI GGUF Ready

LFM2.5-350M

LiquidAI/LFM2.5-350M-GGUF

82,562 99
Calculate
MaziyarPanahi GGUF Ready

Nemotron-Orchestrator-8B

MaziyarPanahi/Nemotron-Orchestrator-8B-GGUF

81,961 6
Calculate
MaziyarPanahi GGUF Ready

Trinity-Mini

MaziyarPanahi/Trinity-Mini-GGUF

81,794 1
Calculate
bartowski GGUF Ready

Qwen_Qwen3-0.6B

bartowski/Qwen_Qwen3-0.6B-GGUF

77,303 26
Calculate
Qwen GGUF Ready

Qwen2.5-Coder-3B-Instruct

Qwen/Qwen2.5-Coder-3B-Instruct-GGUF

76,026 122
Calculate
bartowski GGUF Ready

Ling-3.0-tiny

bartowski/Ling-3.0-tiny-GGUF

75,927 40
Calculate
LiquidAI GGUF Ready

LFM2.5-8B-A1B-DSpark

LiquidAI/LFM2.5-8B-A1B-DSpark-GGUF

75,099 25
Calculate
unsloth GGUF Ready

gemma-3-270m-it

unsloth/gemma-3-270m-it-GGUF

72,395 174
Calculate
bartowski GGUF Ready

Qwen2.5-14B-Instruct

bartowski/Qwen2.5-14B-Instruct-GGUF

70,168 72
Calculate
bartowski GGUF Ready

DeepSeek-R1-Distill-Qwen-7B

bartowski/DeepSeek-R1-Distill-Qwen-7B-GGUF

70,079 145
Calculate
MaziyarPanahi GGUF Ready

Ministral-3-14B-Reasoning-2512

MaziyarPanahi/Ministral-3-14B-Reasoning-2512-GGUF

66,968 1
Calculate
MaziyarPanahi GGUF Ready

NVIDIA-Nemotron-Nano-12B-v2

MaziyarPanahi/NVIDIA-Nemotron-Nano-12B-v2-GGUF

66,828 4
Calculate
MaziyarPanahi GGUF Ready

Qwen3-4B-Thinking-2507

MaziyarPanahi/Qwen3-4B-Thinking-2507-GGUF

66,365 2
Calculate
google GGUF Ready

gemma-2b-it

google/gemma-2b-it

59,086 952
Calculate
bartowski GGUF Ready

DeepSeek-R1-Distill-Qwen-32B

bartowski/DeepSeek-R1-Distill-Qwen-32B-GGUF

53,629 325
Calculate
bartowski GGUF Ready

DeepSeek-R1-Distill-Qwen-14B

bartowski/DeepSeek-R1-Distill-Qwen-14B-GGUF

51,875 248
Calculate
unsloth GGUF Ready

DeepSeek-R1-Distill-Llama-8B

unsloth/DeepSeek-R1-Distill-Llama-8B-GGUF

46,578 312
Calculate
unsloth GGUF Ready

Kimi-K2-Instruct

unsloth/Kimi-K2-Instruct-GGUF

45,571 229
Calculate
TheBloke GGUF Ready

Llama-2-7B-Chat

TheBloke/Llama-2-7B-Chat-GGUF

45,070 518
Calculate
bartowski GGUF Ready

Qwen_Qwen3-1.7B

bartowski/Qwen_Qwen3-1.7B-GGUF

42,173 22
Calculate
bartowski GGUF Ready

microsoft_Phi-4-mini-instruct

bartowski/microsoft_Phi-4-mini-instruct-GGUF

40,965 45
Calculate
unsloth GGUF Ready

Qwen3-Next-80B-A3B-Instruct

unsloth/Qwen3-Next-80B-A3B-Instruct-GGUF

40,932 199
Calculate
bartowski GGUF Ready

Qwen_Qwen3-8B

bartowski/Qwen_Qwen3-8B-GGUF

40,604 39
Calculate
TheBloke GGUF Ready

Mistral-7B-Instruct-v0.2

TheBloke/Mistral-7B-Instruct-v0.2-GGUF

39,942 517
Calculate
bartowski GGUF Ready

Qwen2.5-32B-Instruct

bartowski/Qwen2.5-32B-Instruct-GGUF

39,601 76
Calculate
unsloth GGUF Ready

Ornith-1.0-35B

unsloth/Ornith-1.0-35B-GGUF

37,407 152
Calculate
Qwen GGUF Ready

Qwen3-235B-A22B

Qwen/Qwen3-235B-A22B-GGUF

34,980 10
Calculate
bartowski GGUF Ready

moonshotai_Kimi-Linear-48B-A3B-Instruct

bartowski/moonshotai_Kimi-Linear-48B-A3B-Instruct-GGUF

32,754 28
Calculate
TheBloke GGUF Ready

Mistral-7B-Instruct-v0.1

TheBloke/Mistral-7B-Instruct-v0.1-GGUF

31,362 614
Calculate
bartowski GGUF Ready

HuggingFaceTB_SmolLM3-3B

bartowski/HuggingFaceTB_SmolLM3-3B-GGUF

31,003 11
Calculate
bartowski GGUF Ready

gemma-2-9b-it

bartowski/gemma-2-9b-it-GGUF

29,905 234
Calculate
bartowski GGUF Ready

SmolLM2-135M-Instruct

bartowski/SmolLM2-135M-Instruct-GGUF

29,870 14
Calculate
google GGUF Ready

gemma-7b-it

google/gemma-7b-it

28,998 1250
Calculate
bartowski GGUF Ready

cognitivecomputations_Dolphin-Mistral-24B-Venice-Edition

bartowski/cognitivecomputations_Dolphin-Mistral-24B-Venice-Edition-GGUF

28,573 168
Calculate
unsloth GGUF Ready

MiniMax-M2.7

unsloth/MiniMax-M2.7-GGUF

28,563 221
Calculate
unsloth GGUF Ready

DeepSeek-R1

unsloth/DeepSeek-R1-GGUF

28,492 1118
Calculate
bartowski GGUF Ready

L3-8B-Stheno-v3.2

bartowski/L3-8B-Stheno-v3.2-GGUF

27,030 55
Calculate
google GGUF Ready

gemma-7b

google/gemma-7b

27,018 3435
Calculate
unsloth GGUF Ready

GLM-4.7-Flash-REAP-23B-A3B

unsloth/GLM-4.7-Flash-REAP-23B-A3B-GGUF

26,622 271
Calculate
bartowski GGUF Ready

WhiteRabbitNeo_WhiteRabbitNeo-V3-7B

bartowski/WhiteRabbitNeo_WhiteRabbitNeo-V3-7B-GGUF

26,010 24
Calculate
bartowski GGUF Ready

Qwen_Qwen3-4B-Instruct-2507

bartowski/Qwen_Qwen3-4B-Instruct-2507-GGUF

25,596 23
Calculate
bartowski GGUF Ready

TheDrummer_Cydonia-24B-v4.3

bartowski/TheDrummer_Cydonia-24B-v4.3-GGUF

23,165 68
Calculate
unsloth GGUF Ready

Laguna-S-2.1

unsloth/Laguna-S-2.1-GGUF

22,484 311
Calculate
microsoft GGUF Ready

Phi-3-mini-4k-instruct

microsoft/Phi-3-mini-4k-instruct-gguf

22,427 604
Calculate
ibm-granite GGUF Ready

granite-4.1-8b-fp8

ibm-granite/granite-4.1-8b-fp8

21,416 14
Calculate
bartowski GGUF Ready

openai_gpt-oss-20b

bartowski/openai_gpt-oss-20b-GGUF

21,233 39
Calculate
bartowski GGUF Ready

Hermes-3-Llama-3.2-3B

bartowski/Hermes-3-Llama-3.2-3B-GGUF

19,886 16
Calculate
bartowski GGUF Ready

dolphin-2.9-llama3-8b

bartowski/dolphin-2.9-llama3-8b-GGUF

19,351 17
Calculate
bartowski GGUF Ready

NemoMix-Unleashed-12B

bartowski/NemoMix-Unleashed-12B-GGUF

19,240 131
Calculate
bartowski GGUF Ready

NousResearch_Hermes-4-14B

bartowski/NousResearch_Hermes-4-14B-GGUF

19,148 55
Calculate
bartowski GGUF Ready

Qwen2.5-Math-1.5B-Instruct

bartowski/Qwen2.5-Math-1.5B-Instruct-GGUF

19,080 3
Calculate
TheBloke GGUF Ready

CodeLlama-7B-Instruct

TheBloke/CodeLlama-7B-Instruct-GGUF

18,281 151
Calculate
lmstudio-community GGUF Ready

Qwen2.5-Math-7B-Instruct

lmstudio-community/Qwen2.5-Math-7B-Instruct-GGUF

18,163 3
Calculate
TheBloke GGUF Ready

phi-2

TheBloke/phi-2-GGUF

18,009 233
Calculate
bartowski GGUF Ready

deepseek-ai_DeepSeek-R1-0528

bartowski/deepseek-ai_DeepSeek-R1-0528-GGUF

17,148 11
Calculate
TheBloke GGUF Ready

Llama-2-7B

TheBloke/Llama-2-7B-GGUF

16,880 207
Calculate
Qwen GGUF Ready

Qwen2.5-Coder-0.5B-Instruct

Qwen/Qwen2.5-Coder-0.5B-Instruct-GGUF

16,874 29
Calculate
bartowski GGUF Ready

Ling-3.0-flash

bartowski/Ling-3.0-flash-GGUF

16,597 17
Calculate
lmstudio-community GGUF Ready

Phi-4-mini-reasoning

lmstudio-community/Phi-4-mini-reasoning-GGUF

16,547 6
Calculate
lmstudio-community GGUF Ready

Qwen2.5-VL-32B-Instruct

lmstudio-community/Qwen2.5-VL-32B-Instruct-GGUF

16,355 1
Calculate
unsloth GGUF Ready

Qwen3.8-2.4T-A95B

unsloth/Qwen3.8-2.4T-A95B-GGUF

16,141 116
Calculate
Qwen GGUF Ready

Qwen2-1.5B-Instruct

Qwen/Qwen2-1.5B-Instruct-GGUF

15,999 31
Calculate
unsloth GGUF Ready

NVIDIA-Nemotron-3-Nano-Omni-30B-A3B-Reasoning

unsloth/NVIDIA-Nemotron-3-Nano-Omni-30B-A3B-Reasoning-GGUF

15,988 143
Calculate
unsloth GGUF Ready

Qwen3-Coder-30B-A3B-Instruct-1M

unsloth/Qwen3-Coder-30B-A3B-Instruct-1M-GGUF

15,541 162
Calculate
mradermacher GGUF Ready

Wanabi-Novelist-24B

mradermacher/Wanabi-Novelist-24B-GGUF

15,466 0
Calculate
bartowski GGUF Ready

zai-org_GLM-4.7-Flash

bartowski/zai-org_GLM-4.7-Flash-GGUF

15,458 57
Calculate
bartowski GGUF Ready

Qwen_Qwen3-Coder-Next

bartowski/Qwen_Qwen3-Coder-Next-GGUF

15,145 27
Calculate
bartowski GGUF Ready

Dolphin3.0-Llama3.1-8B

bartowski/Dolphin3.0-Llama3.1-8B-GGUF

14,892 25
Calculate
lmstudio-community GGUF Ready

Qwen2.5-7B-Instruct-1M

lmstudio-community/Qwen2.5-7B-Instruct-1M-GGUF

14,644 45
Calculate
janhq GGUF Ready

Jan-code-4b

janhq/Jan-code-4b-gguf

14,585 70
Calculate
bartowski GGUF Ready

cognitivecomputations_Dolphin3.0-R1-Mistral-24B

bartowski/cognitivecomputations_Dolphin3.0-R1-Mistral-24B-GGUF

14,571 86
Calculate
bartowski GGUF Ready

granite-4.2-30b

bartowski/granite-4.2-30b-GGUF

14,456 7
Calculate
bartowski GGUF Ready

Nanbeige_Nanbeige4.2-3B

bartowski/Nanbeige_Nanbeige4.2-3B-GGUF

14,446 32
Calculate
unsloth GGUF Ready

Qwen3-Next-80B-A3B-Thinking

unsloth/Qwen3-Next-80B-A3B-Thinking-GGUF

13,853 83
Calculate
unsloth GGUF Ready

Nemotron-3-Nano-30B-A3B

unsloth/Nemotron-3-Nano-30B-A3B-GGUF

13,761 329
Calculate
QuantFactory GGUF Ready

Cotype-Nano

QuantFactory/Cotype-Nano-GGUF

13,676 5
Calculate
unsloth GGUF Ready

GLM-4.5-Air

unsloth/GLM-4.5-Air-GGUF

13,369 185
Calculate
bartowski GGUF Ready

Qwen_Qwen3-30B-A3B

bartowski/Qwen_Qwen3-30B-A3B-GGUF

13,229 60
Calculate
Qwen GGUF Ready

Qwen2-0.5B-Instruct

Qwen/Qwen2-0.5B-Instruct-GGUF

13,061 77
Calculate
bartowski GGUF Ready

qwen2.5-7b-ins-v3

bartowski/qwen2.5-7b-ins-v3-GGUF

13,014 9
Calculate
bartowski GGUF Ready

allenai_Olmo-3.1-32B-Think

bartowski/allenai_Olmo-3.1-32B-Think-GGUF

13,014 6
Calculate
bartowski GGUF Ready

Qwen_Qwen3-30B-A3B-Instruct-2507

bartowski/Qwen_Qwen3-30B-A3B-Instruct-2507-GGUF

13,002 29
Calculate
lmstudio-community GGUF Ready

Qwen2.5-Coder-32B

lmstudio-community/Qwen2.5-Coder-32B-GGUF

12,960 6
Calculate
bartowski GGUF Ready

google_gemma-3-1b-it

bartowski/google_gemma-3-1b-it-GGUF

12,895 18
Calculate
bartowski GGUF Ready

Hermes-3-Llama-3.1-8B

bartowski/Hermes-3-Llama-3.1-8B-GGUF

12,813 20
Calculate
LiquidAI GGUF Ready

LFM2-2.6B

LiquidAI/LFM2-2.6B-GGUF

12,754 61
Calculate
bartowski GGUF Ready

Qwen_Qwen3-32B

bartowski/Qwen_Qwen3-32B-GGUF

12,505 41
Calculate
Qwen GGUF Ready

Qwen1.5-7B-Chat

Qwen/Qwen1.5-7B-Chat-GGUF

12,400 71
Calculate
lmstudio-community GGUF Ready

Hermes-4-70B

lmstudio-community/Hermes-4-70B-GGUF

11,842 12
Calculate
LiquidAI GGUF Ready

LFM2-1.2B

LiquidAI/LFM2-1.2B-GGUF

11,759 111
Calculate
unsloth GGUF Ready

GLM-4.7

unsloth/GLM-4.7-GGUF

11,571 224
Calculate
bartowski GGUF Ready

openai_gpt-oss-120b

bartowski/openai_gpt-oss-120b-GGUF

11,306 13
Calculate
bartowski GGUF Ready

OLMoE-1B-7B-0924-Instruct

bartowski/OLMoE-1B-7B-0924-Instruct-GGUF

11,299 10
Calculate
lmstudio-community GGUF Ready

gemma-3-1B-it-qat

lmstudio-community/gemma-3-1B-it-qat-GGUF

11,240 18
Calculate

Frequently asked questions about local models

What is the GGUF model format?

GGUF (GPT-Generated Unified Format) is the standard format used by llama.cpp, Ollama, LM Studio, and Jan to run large language models on consumer hardware. It stores weights and quantization metadata in a single, fast-loading binary file.

How do I know which model fits on my computer?

Select any model above to open our interactive calculator. It reads the model's exact weight sizes and calculates required memory for KV cache and system overhead across various context windows (4K, 8K, 32K+).

Can't find the model you need?

You can search for any GGUF repository directly on the homepage search bar, or browse models categorized by graphics card.