HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

210,822 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

image-text-to-text

RedHatAI/gemma-4-26B-A4B-it-NVFP4

RedHatAI/gemma-4-26B-A4B-it-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

247,53641
26.0B64 GB+ VRAMapache-2.0
Deployment details
audio-text-to-text

OpenMOSS-Team/MOSS-Transcribe-Diarize

OpenMOSS-Team/MOSS-Transcribe-Diarize is a audio text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

245,269413
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

ornith-ai/Ornith-1.5-35B-A3B-FP8

ornith-ai/Ornith-1.5-35B-A3B-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

243,69424
35.0B80 GB+ VRAMmit
Deployment details
text-generation

z-lab/Qwen3.8-27B-DFlash2

z-lab/Qwen3.8-27B-DFlash2 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

241,660269
27.0B64 GB+ VRAMapache-2.0
Deployment details
sentence-similarity

lightonai/GTE-ModernColBERT-v1

lightonai/GTE-ModernColBERT-v1 is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

240,710177
Unknown4 GB+ VRAMapache-2.0
Deployment details
general AI

nvidia/Cosmos3-Nano

nvidia/Cosmos3-Nano is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

240,171348
Unknown8 GB+ VRAMother
Deployment details
visual-document-retrieval

vidore/colqwen2.5-v0.2

vidore/colqwen2.5-v0.2 is a visual document retrieval model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

238,273100
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

lightonai/LightOnOCR-2-1B

lightonai/LightOnOCR-2-1B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

237,414800
1.0B6 GB+ VRAMapache-2.0
Deployment details
text-generation

ReadyArt/gemma-4-31B-it-scotoma-2-GGUF

ReadyArt/gemma-4-31B-it-scotoma-2-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

236,42036
31.0B24 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

Qwen/Qwen3-ASR-1.7B-hf

Qwen/Qwen3-ASR-1.7B-hf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

236,06188
1.7B6 GB+ VRAMapache-2.0
Deployment details
general AI

Abiray/Nemotron-3-Embed-8B-GGUF

Abiray/Nemotron-3-Embed-8B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

235,6087
8.0B8 GB+ VRAMopenmdw-1.1
Deployment details
general AI

kernels-community/triton-layer-norm

kernels-community/triton-layer-norm is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

235,0320
Unknown8 GB+ VRAMbsd-3-clause
Deployment details
text-to-speech

drbaph/Higgs-Audio-v3-Studio

drbaph/Higgs-Audio-v3-Studio is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

234,6757
Unknown8 GB+ VRAMother
Deployment details
text-generation

cyankiwi/GLM-4.7-Flash-AWQ-4bit

cyankiwi/GLM-4.7-Flash-AWQ-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

233,10156
Unknown8 GB+ VRAMmit
Deployment details
general AI

Comfy-Org/flux1-schnell

Comfy-Org/flux1-schnell is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

232,366279
Unknown8 GB+ VRAMapache-2.0
Deployment details
sentence-similarity

MongoDB/mdbr-leaf-ir

MongoDB/mdbr-leaf-ir is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

232,10566
Unknown4 GB+ VRAMapache-2.0
Deployment details
any-to-any

openbmb/MiniCPM-o-2_6

openbmb/MiniCPM-o-2_6 is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

229,7511,299
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

poolside/Laguna-XS-2.1-GGUF

poolside/Laguna-XS-2.1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

229,32287
Unknown8 GB+ VRAMopenmdw-1.1
Deployment details
image-text-to-text

bartowski/Fara1.5-4B-GGUF

bartowski/Fara1.5-4B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

224,4632
4.0B6 GB+ VRAMmit
Deployment details