snowflake-arctic-embed snowflake-arctic-embed:22m
Est. memory: 0.6 GB
Download: 0.0 GB
0.0 GB 0.6 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 44 MB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.
all-minilm all-minilm:22m
Est. memory: 0.6 GB
Download: 0.0 GB
0.0 GB 0.6 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 44 MB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.
granite-embedding granite-embedding:30m
Est. memory: 0.6 GB
Download: 0.1 GB
0.1 GB 0.6 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 60 MB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.
smollm smollm:135m-base-v0.2-q4_K_M
Exceeds documented context maximum (2,048). Memory estimate only.
Est. memory: 0.6 GB
Download: 0.1 GB
0.1 GB 0.6 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 101 MB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.
smollm2 smollm2:135m-instruct-q4_K_M
Est. memory: 0.6 GB
Download: 0.1 GB
0.1 GB 0.6 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 101 MB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.
embeddinggemma embeddinggemma:300m-qat-q4_0
Exceeds documented context maximum (2,048). Memory estimate only.
Est. memory: 0.8 GB
Download: 0.2 GB
0.2 GB 0.8 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 227 MB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.
nomic-embed-text nomic-embed-text:137m-v1.5-fp16
Exceeds documented context maximum (2,048). Memory estimate only.
Est. memory: 0.9 GB
Download: 0.3 GB
0.3 GB 0.9 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 262 MB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.
functiongemma functiongemma:270m
Est. memory: 0.9 GB
Download: 0.3 GB
0.3 GB 0.9 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 287 MB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.
qwen2 qwen2:0.5b-instruct-q4_K_M
Est. memory: 1.0 GB
Download: 0.4 GB
0.4 GB 1.0 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 379 MB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.
reader-lm reader-lm:0.5b-q4_K_M
Est. memory: 1.0 GB
Download: 0.4 GB
0.4 GB 1.0 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 379 MB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.
qwen2.5 qwen2.5:0.5b-base-q4_K_M
Est. memory: 1.0 GB
Download: 0.4 GB
0.4 GB 1.0 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 379 MB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.
qwen2.5-coder qwen2.5-coder:0.5b-base-q4_K_M
Est. memory: 1.0 GB
Download: 0.4 GB
0.4 GB 1.0 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 379 MB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.
qwen qwen:0.5b-text-v1.5-q4_K_M
Est. memory: 1.0 GB
Download: 0.4 GB
0.4 GB 1.0 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 388 MB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.
qwen3 qwen3:0.6b
Est. memory: 1.2 GB
Download: 0.5 GB
0.5 GB 1.2 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 498 MB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.
minicpm-v4.6 minicpm-v4.6:1b
Est. memory: 1.2 GB
Download: 1.5 GB
1.5 GB 1.2 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 1.5 GB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.
paraphrase-multilingual paraphrase-multilingual:278m
Est. memory: 1.2 GB
Download: 0.5 GB
0.5 GB 1.2 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 537 MB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.
tinydolphin tinydolphin:1.1b-v2.8-q4_K_M
Est. memory: 1.4 GB
Download: 0.6 GB
0.6 GB 1.4 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 637 MB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.
tinyllama tinyllama:1.1b-chat-v0.6-q4_K_M
Exceeds documented context maximum (2,048). Memory estimate only.
Est. memory: 1.4 GB
Download: 0.6 GB
0.6 GB 1.4 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 638 MB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.
mxbai-embed-large mxbai-embed-large:335m
Est. memory: 1.4 GB
Download: 0.6 GB
0.6 GB 1.4 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 639 MB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.
bge-large bge-large:335m
Est. memory: 1.4 GB
Download: 0.6 GB
0.6 GB 1.4 GB Excellent fit low confidence
Details How this result was calculated
GPU memory: Pass Model fits within the unified memory GPU allocation.
System RAM: Pass Unified memory shared with GPU allocation above; no separate RAM requirement.
Storage: Pass 500.0 GB free, 639 MB required.
Acceleration: Pass metal acceleration available. Assumptions and missing information
KV cache size is estimated using a fallback formula because architecture metadata is incomplete. Actual memory may differ significantly. This is a hypothetical PC. No inference benchmark was run. GB uses 1,024³ bytes. Memory includes weights, context cache and runtime overhead.