HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

965,049 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

reinforcement-learning

mradermacher/ATLAS-8B-Thinking-i1-GGUF

mradermacher/ATLAS-8B-Thinking-i1-GGUF is a reinforcement learning model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,676♡ 1
8.0B8 GB+ VRAMapache-2.0
Deployment details →
general AI

mradermacher/gpt-oss-20b-i1-GGUF

mradermacher/gpt-oss-20b-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 16 GB. It is publicly listed on Hugging Face.

↓ 1,676♡ 1
20.0B16 GB+ VRAMapache-2.0
Deployment details →
general AI

jgebbeken/gemma-4-coder-gguf

jgebbeken/gemma-4-coder-gguf is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,676♡ 22
Unknown8 GB+ VRAMapache-2.0
Deployment details →
general AI

tencent/Hy-MT2-1.8B-2Bit-GGUF

tencent/Hy-MT2-1.8B-2Bit-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 1,675♡ 14
1.8B4 GB+ VRAMapache-2.0
Deployment details →
voice-activity-detection

aufklarer/Silero-VAD-v6.2.1-CoreML

aufklarer/Silero-VAD-v6.2.1-CoreML is a voice activity detection model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,674♡ 1
Unknown8 GB+ VRAMmit
Deployment details →
text-to-speech

Thorsten-Voice/Kokoro

Thorsten-Voice/Kokoro is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,674♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

arcanicai/Con0-GGUF

arcanicai/Con0-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,674♡ 2
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-generation

Atomic-Germ/Qwen3.6-35B-A3B-NPU2

Atomic-Germ/Qwen3.6-35B-A3B-NPU2 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

↓ 1,674♡ 0
35.0B32 GB+ VRAMapache-2.0
Deployment details →
general AI

mradermacher/Qwen3.5-4B-GGUF

mradermacher/Qwen3.5-4B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

↓ 1,672♡ 1
4.0B6 GB+ VRAMapache-2.0
Deployment details →
general AI

sweepai/sweep-next-edit-1.5B

sweepai/sweep-next-edit-1.5B is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 1,671♡ 336
1.5B4 GB+ VRAMapache-2.0
Deployment details →
sentence-similarity

sdadas/mmlw-retrieval-e5-small

sdadas/mmlw-retrieval-e5-small is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 1,671♡ 1
Unknown4 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

RedHatAI/Qwen3.8-27B

RedHatAI/Qwen3.8-27B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

↓ 1,671♡ 0
27.0B64 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

rapid-mlx/Qwen3.8-27B-mixed-3.5bpw-MLX

rapid-mlx/Qwen3.8-27B-mixed-3.5bpw-MLX is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

↓ 1,671♡ 4
27.0B64 GB+ VRAMapache-2.0
Deployment details →
automatic-speech-recognition

onnx-community/cohere-transcribe-03-2026-ONNX

onnx-community/cohere-transcribe-03-2026-ONNX is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,670♡ 20
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-generation

Hatim2221/Fikr-7B-Reasoning

Hatim2221/Fikr-7B-Reasoning is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 1,670♡ 0
7.0B24 GB+ VRAMapache-2.0
Deployment details →
general AI

mradermacher/Qwen3.6-27B-i1-GGUF

mradermacher/Qwen3.6-27B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 1,669♡ 16
27.0B24 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

catlilface/Gemma-4-26B-A4B-NVFP4-GGUF

catlilface/Gemma-4-26B-A4B-NVFP4-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 1,669♡ 10
26.0B24 GB+ VRAMapache-2.0
Deployment details →