Published by meta-llama. 8 members have been measured on real machines. Every number below was signed by the machine that produced it — none of it is from a model card.
what it takes, what it returns
text → text
Declared by the network’s catalog, not measured — a signature says what this family would be asked. Whether it answered is the signed capability flags on each member.
Declared, not measured — this is what the catalog says exists and what it would cost a machine to hold. Nothing here is signed, and none of it claims the build works.
how much memory does your machine have
38 of 39 builds fit a 16 GB machine — budgeting 70% of it, or 11.2 GiB, because the cache and the runtime want the rest.
hf.co/MaziyarPanahi/Llama-3.2-1B-Instruct-GGUF:IQ1_SIQ1_Sollamameasured herequantized by MaziyarPanahi
0.4 GiB
hf.co/MaziyarPanahi/Llama-3.2-1B-Instruct-GGUF:IQ3_XSIQ3_XSollamameasured herequantized by MaziyarPanahi
0.6 GiB
mlx-community/Llama-3.2-1B-Instruct-4bit4BITmlxderived by mlx-community
0.6 GiB
hf.co/MaziyarPanahi/Llama-3.2-1B-Instruct-GGUF:Q3_K_LQ3_K_Lollamaquantized by MaziyarPanahi
0.7 GiB
hf.co/MaziyarPanahi/Llama-3.2-1B-Instruct-GGUF:Q4_K_SQ4_K_Sollamaquantized by MaziyarPanahi
0.7 GiB
hf.co/MaziyarPanahi/Llama-3.2-1B-Instruct-GGUF:Q4_K_MQ4_K_Mollamameasured herequantized by MaziyarPanahi
0.8 GiB
hf.co/hugging-quants/Llama-3.2-1B-Instruct-Q4_K_M-GGUF:Q4_K_MQ4_K_Mollamaderived by hugging-quants
0.8 GiB
hf.co/unsloth/Llama-3.2-3B-Instruct-GGUF:IQ1_SIQ1_Sollamaquantized by unsloth
0.8 GiB
hf.co/unsloth/Llama-3.2-3B-Instruct-GGUF:IQ2_XXSIQ2_XXSollamaquantized by unsloth
1.0 GiB
hf.co/MaziyarPanahi/Llama-3.2-1B-Instruct-GGUF:Q8_0Q8_0ollamaquantized by MaziyarPanahi
1.2 GiB
hf.co/unsloth/Llama-3.2-3B-Instruct-GGUF:IQ3_XXSIQ3_XXSollamaquantized by unsloth
1.3 GiB
hf.co/bartowski/Llama-3.2-3B-Instruct-uncensored-GGUF:Q2_KQ2_Kollamaderived by bartowski
1.4 GiB
hf.co/bartowski/Llama-3.2-3B-Instruct-uncensored-GGUF:IQ3_XSIQ3_XSollamameasured herederived by bartowski
1.5 GiB
mlx-community/Llama-3.2-3B-Instruct-4bit4BITmlxderived by mlx-community
1.7 GiB
hf.co/unsloth/Llama-3.2-3B-Instruct-GGUF:Q4_0Q4_0ollamaquantized by unsloth
1.8 GiB
hf.co/bartowski/Llama-3.2-3B-Instruct-uncensored-GGUF:Q3_K_LQ3_K_Lollamaderived by bartowski
1.8 GiB
hf.co/unsloth/Llama-3.2-3B-Instruct-GGUF:Q4_K_MQ4_K_Mollamaquantized by unsloth
1.9 GiB
hf.co/hugging-quants/Llama-3.2-3B-Instruct-Q4_K_M-GGUF:Q4_K_MQ4_K_Mollamaderived by hugging-quants
1.9 GiB
RedHatAI/Llama-3.2-1B-Instruct-FP8-dynamicFP8vllmquantized by RedHatAI
1.9 GiB
hf.co/bartowski/Llama-3.2-3B-Instruct-uncensored-GGUF:Q4_0_8_8Q4_0_8_8ollamaderived by bartowski
2.0 GiB
hf.co/bartowski/Llama-3.2-3B-Instruct-uncensored-GGUF:Q4_0_4_4Q4_0_4_4ollamaderived by bartowski
2.0 GiB
hf.co/bartowski/Llama-3.2-3B-Instruct-uncensored-GGUF:Q4_K_MQ4_K_Mollamaderived by bartowski
2.1 GiB
unsloth/Llama-3.2-3B-Instruct-unsloth-bnb-4bitBITSANDBYTESvllmderived by unsloth
2.2 GiB
mlx-community/Llama-3.2-1B-Instruct-bf16BF16mlxderived by mlx-community
2.3 GiB
unsloth/Llama-3.2-1B-InstructBF16vllmderived by unsloth
2.3 GiB
hf.co/bartowski/Llama-3.2-1B-Instruct-GGUF:F16F16ollamameasured herequantized by bartowski
2.3 GiB
mlx-community/Llama-3.2-3B-Instruct-uncensored-6bit6BITmlxderived by mlx-community
2.4 GiB
mlx-community/Llama-3.2-3B-Instruct-8bitBF16mlxderived by mlx-community
3.2 GiB
hf.co/unsloth/Llama-3.2-3B-Instruct-GGUF:Q8_0Q8_0ollamaquantized by unsloth
3.2 GiB
hf.co/bartowski/Llama-3.2-3B-Instruct-uncensored-GGUF:Q8_0Q8_0ollamaderived by bartowski
3.6 GiB
RedHatAI/Llama-3.2-3B-Instruct-FP8-dynamicFP8vllmquantized by RedHatAI
4.1 GiB
hf.co/leafspark/Llama-3.2-11B-Vision-Instruct-GGUF:Q4_K_MQ4_K_Mollamaquantized by leafspark
5.6 GiB
unsloth/Llama-3.2-3B-InstructBF16vllmderived by unsloth
6.0 GiB
hf.co/unsloth/Llama-3.2-3B-Instruct-GGUF:F16F16ollamaquantized by unsloth
6.0 GiB
hf.co/unsloth/Llama-3.2-3B-Instruct-GGUF:BF16BF16ollamaquantized by unsloth
6.0 GiB
hf.co/bartowski/Llama-3.2-3B-Instruct-uncensored-GGUF:F16F16ollamaderived by bartowski
6.7 GiB
hf.co/leafspark/Llama-3.2-11B-Vision-Instruct-GGUF:Q8_0Q8_0ollamaquantized by leafspark
9.7 GiB
mlx-community/Llama-3.2-11B-Vision-Instruct-8bitBF16mlxderived by mlx-community
10.6 GiB
hf.co/leafspark/Llama-3.2-11B-Vision-Instruct-GGUF:F16F16ollamaquantized by leafspark
18.2 GiB
Sizes are what the catalog declares, not what a machine measured — fitting on disk is necessary and not sufficient. Nobody here has run these on your machine, and this narrows the candidates rather than certifying one.
the same builds, by member
Llama 3.2 11B
hf.co/leafspark/Llama-3.2-11B-Vision-Instruct-GGUF:Q4_K_MQ4_K_Mollama5.6 GiB on diskquantized by leafspark
hf.co/leafspark/Llama-3.2-11B-Vision-Instruct-GGUF:Q8_0Q8_0ollama9.7 GiB on diskquantized by leafspark
hf.co/leafspark/Llama-3.2-11B-Vision-Instruct-GGUF:F16F16ollama18.2 GiB on diskquantized by leafspark
Llama 3.2 11B · mlx-community
mlx-community/Llama-3.2-11B-Vision-Instruct-8bitBF16mlx10.6 GiB on diskderived by mlx-community
Llama 3.2 1B1.2B
hf.co/MaziyarPanahi/Llama-3.2-1B-Instruct-GGUF:IQ1_SIQ1_Sollama375 MiB on diskquantized by MaziyarPanahi
hf.co/MaziyarPanahi/Llama-3.2-1B-Instruct-GGUF:IQ3_XSIQ3_XSollama592 MiB on diskquantized by MaziyarPanahi
hf.co/MaziyarPanahi/Llama-3.2-1B-Instruct-GGUF:Q3_K_LQ3_K_Lollama698 MiB on diskquantized by MaziyarPanahi
hf.co/MaziyarPanahi/Llama-3.2-1B-Instruct-GGUF:Q4_K_SQ4_K_Sollama739 MiB on diskquantized by MaziyarPanahi
hf.co/MaziyarPanahi/Llama-3.2-1B-Instruct-GGUF:Q4_K_MQ4_K_Mollama770 MiB on diskquantized by MaziyarPanahi
hf.co/MaziyarPanahi/Llama-3.2-1B-Instruct-GGUF:Q8_0Q8_0ollama1.2 GiB on diskquantized by MaziyarPanahi
RedHatAI/Llama-3.2-1B-Instruct-FP8-dynamicFP8vllm1.9 GiB on diskquantized by RedHatAI
hf.co/bartowski/Llama-3.2-1B-Instruct-GGUF:F16F16ollama2.3 GiB on diskquantized by bartowski
Llama 3.2 1B · hugging-quants
hf.co/hugging-quants/Llama-3.2-1B-Instruct-Q4_K_M-GGUF:Q4_K_MQ4_K_Mollama770 MiB on diskderived by hugging-quants
Llama 3.2 1B · mlx-community
mlx-community/Llama-3.2-1B-Instruct-4bit4BITmlx663 MiB on diskderived by mlx-community
mlx-community/Llama-3.2-1B-Instruct-bf16BF16mlx2.3 GiB on diskderived by mlx-community
Llama 3.2 1B · unsloth
unsloth/Llama-3.2-1B-InstructBF16vllm2.3 GiB on diskderived by unsloth
Llama 3.2 3B3.2B
hf.co/unsloth/Llama-3.2-3B-Instruct-GGUF:IQ1_SIQ1_Sollama870 MiB on diskquantized by unsloth
hf.co/unsloth/Llama-3.2-3B-Instruct-GGUF:IQ2_XXSIQ2_XXSollama998 MiB on diskquantized by unsloth
hf.co/unsloth/Llama-3.2-3B-Instruct-GGUF:IQ3_XXSIQ3_XXSollama1.3 GiB on diskquantized by unsloth
hf.co/unsloth/Llama-3.2-3B-Instruct-GGUF:Q4_0Q4_0ollama1.8 GiB on diskquantized by unsloth
hf.co/unsloth/Llama-3.2-3B-Instruct-GGUF:Q4_K_MQ4_K_Mollama1.9 GiB on diskquantized by unsloth
hf.co/unsloth/Llama-3.2-3B-Instruct-GGUF:Q8_0Q8_0ollama3.2 GiB on diskquantized by unsloth
RedHatAI/Llama-3.2-3B-Instruct-FP8-dynamicFP8vllm4.1 GiB on diskquantized by RedHatAI
hf.co/unsloth/Llama-3.2-3B-Instruct-GGUF:F16F16ollama6.0 GiB on diskquantized by unsloth
hf.co/unsloth/Llama-3.2-3B-Instruct-GGUF:BF16BF16ollama6.0 GiB on diskquantized by unsloth
Llama 3.2 3B · bartowski
hf.co/bartowski/Llama-3.2-3B-Instruct-uncensored-GGUF:Q2_KQ2_Kollama1.4 GiB on diskderived by bartowski
hf.co/bartowski/Llama-3.2-3B-Instruct-uncensored-GGUF:IQ3_XSIQ3_XSollama1.5 GiB on diskderived by bartowski
hf.co/bartowski/Llama-3.2-3B-Instruct-uncensored-GGUF:Q3_K_LQ3_K_Lollama1.8 GiB on diskderived by bartowski
hf.co/bartowski/Llama-3.2-3B-Instruct-uncensored-GGUF:Q4_0_8_8Q4_0_8_8ollama2.0 GiB on diskderived by bartowski
hf.co/bartowski/Llama-3.2-3B-Instruct-uncensored-GGUF:Q4_0_4_4Q4_0_4_4ollama2.0 GiB on diskderived by bartowski
hf.co/bartowski/Llama-3.2-3B-Instruct-uncensored-GGUF:Q4_K_MQ4_K_Mollama2.1 GiB on diskderived by bartowski
hf.co/bartowski/Llama-3.2-3B-Instruct-uncensored-GGUF:Q8_0Q8_0ollama3.6 GiB on diskderived by bartowski
hf.co/bartowski/Llama-3.2-3B-Instruct-uncensored-GGUF:F16F16ollama6.7 GiB on diskderived by bartowski
Llama 3.2 3B · hugging-quants
hf.co/hugging-quants/Llama-3.2-3B-Instruct-Q4_K_M-GGUF:Q4_K_MQ4_K_Mollama1.9 GiB on diskderived by hugging-quants
Llama 3.2 3B · mlx-community
mlx-community/Llama-3.2-3B-Instruct-4bit4BITmlx1.7 GiB on diskderived by mlx-community
mlx-community/Llama-3.2-3B-Instruct-uncensored-6bit6BITmlx2.4 GiB on diskderived by mlx-community
mlx-community/Llama-3.2-3B-Instruct-8bitBF16mlx3.2 GiB on diskderived by mlx-community
Llama 3.2 3B · unsloth
unsloth/Llama-3.2-3B-Instruct-unsloth-bnb-4bitBITSANDBYTESvllm2.2 GiB on diskderived by unsloth
unsloth/Llama-3.2-3B-InstructBF16vllm6.0 GiB on diskderived by unsloth
ask your own gram
Your concierge can answer this from the network rather than from this page. The ask carries the operator id, so it resolves to a capability and an engine — not to a name somebody has to recognise.
What has this network measured about the Llama 3.2 family (epn.family.llama-3-2), and can my gram run any of it? See /inferences/llama-3-2