HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

845,060 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

AEON-7/AEON-DFlash-Qwen3.6-35B-A3B

AEON-7/AEON-DFlash-Qwen3.6-35B-A3B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

3,2732
35.0B80 GB+ VRAMmit
Deployment details
image-text-to-text

unsloth/Qwen3.8-Flash-Next-FP8

unsloth/Qwen3.8-Flash-Next-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,27234
Unknown8 GB+ VRAMother
Deployment details
text-generationGated

saricles/Qwen3-Coder-Next-NVFP4-GB10

saricles/Qwen3-Coder-Next-NVFP4-GB10 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.

3,27048
Unknown8 GB+ VRAMapache-2.0
Deployment details
audio-to-audio

LocalAI-io/LocalVQE

LocalAI-io/LocalVQE is a audio to audio model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,26967
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-textGated

bottlecapai/ThinkingCap-Qwen3.6-27B

bottlecapai/ThinkingCap-Qwen3.6-27B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. Access approval is required on Hugging Face.

3,266681
27.0B64 GB+ VRAMapache-2.0
Deployment details
text-to-image

cusiman/Krea2-realism-V2

cusiman/Krea2-realism-V2 is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

3,2651
Unknown12 GB+ VRAMmit
Deployment details
image-text-to-text

bartowski/ProCreations_grug-35b-v2-GGUF

bartowski/ProCreations_grug-35b-v2-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

3,2631
35.0B32 GB+ VRAMapache-2.0
Deployment details
image-to-video

rzgar/minimax_h3_ref2va_fp8_e4m3fn

rzgar/minimax_h3_ref2va_fp8_e4m3fn is a image to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,2632
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

bartowski/zai-org_GLM-4.5-Air-GGUF

bartowski/zai-org_GLM-4.5-Air-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,26121
Unknown8 GB+ VRAMLicense unknown
Deployment details
automatic-speech-recognition

FunAudioLLM/Fun-ASR-Nano-GGUF

FunAudioLLM/Fun-ASR-Nano-GGUF is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,2605
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-classification

furiosa-ai/Qwen3-Reranker-4B

furiosa-ai/Qwen3-Reranker-4B is a text classification model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

3,2600
4.0B12 GB+ VRAMapache-2.0
Deployment details
general AI

mlx-community/Qwen3-ASR-0.6B-4bit

mlx-community/Qwen3-ASR-0.6B-4bit is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

3,25716
600M4 GB+ VRAMapache-2.0
Deployment details
text-ranking

cross-encoder/ettin-reranker-1b-v1

cross-encoder/ettin-reranker-1b-v1 is a text ranking model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

3,2567
1.0B6 GB+ VRAMapache-2.0
Deployment details