HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

208,822 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

image-text-to-text

infly/Infinity-Parser2-Pro

infly/Infinity-Parser2-Pro is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

267,74191
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

ornith-ai/Ornith-1.5-35B-A3B

ornith-ai/Ornith-1.5-35B-A3B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

267,682549
35.0B80 GB+ VRAMmit
Deployment details
image-text-to-text

ggml-org/Qwen3.6-35B-A3B-GGUF

ggml-org/Qwen3.6-35B-A3B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

266,03818
35.0B32 GB+ VRAMapache-2.0
Deployment details
text-generation

AtomicChat/Ling-3.0-flash-GGUF

AtomicChat/Ling-3.0-flash-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

265,90458
Unknown8 GB+ VRAMmit
Deployment details
text-generation

ai-sage/GigaChat3.1-Audio-10B-A1.8B

ai-sage/GigaChat3.1-Audio-10B-A1.8B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

264,84353
10.0B24 GB+ VRAMmit
Deployment details
text-generation

FastFlowLM/GPT-OSS-20B-NPU2

FastFlowLM/GPT-OSS-20B-NPU2 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.

262,1511
20.0B48 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

cyankiwi/gemma-4-31B-it-AWQ-4bit

cyankiwi/gemma-4-31B-it-AWQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

261,29754
31.0B24 GB+ VRAMapache-2.0
Deployment details
token-classification

protectai/unbiased-toxic-roberta-onnx

protectai/unbiased-toxic-roberta-onnx is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

260,0217
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

bartowski/Fara1.5-27B-GGUF

bartowski/Fara1.5-27B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

259,2079
27.0B24 GB+ VRAMmit
Deployment details
image-text-to-text

bartowski/Fara1.5-9B-GGUF

bartowski/Fara1.5-9B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

253,5857
9.0B8 GB+ VRAMmit
Deployment details
text-to-video

drbaph/MiniMax-H3-Turbo-Lora-ComfyUI

drbaph/MiniMax-H3-Turbo-Lora-ComfyUI is a text to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

250,888386
Unknown8 GB+ VRAMapache-2.0
Deployment details
visual-document-retrieval

ModernVBERT/colmodernvbert

ModernVBERT/colmodernvbert is a visual document retrieval model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

248,46936
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

RedHatAI/gemma-4-26B-A4B-it-NVFP4

RedHatAI/gemma-4-26B-A4B-it-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

247,53641
26.0B64 GB+ VRAMapache-2.0
Deployment details
audio-text-to-text

OpenMOSS-Team/MOSS-Transcribe-Diarize

OpenMOSS-Team/MOSS-Transcribe-Diarize is a audio text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

245,269413
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

ornith-ai/Ornith-1.5-35B-A3B-FP8

ornith-ai/Ornith-1.5-35B-A3B-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

243,69424
35.0B80 GB+ VRAMmit
Deployment details
text-generation

z-lab/Qwen3.8-27B-DFlash2

z-lab/Qwen3.8-27B-DFlash2 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

241,660269
27.0B64 GB+ VRAMapache-2.0
Deployment details
sentence-similarity

lightonai/GTE-ModernColBERT-v1

lightonai/GTE-ModernColBERT-v1 is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

240,710177
Unknown4 GB+ VRAMapache-2.0
Deployment details
visual-document-retrieval

vidore/colqwen2.5-v0.2

vidore/colqwen2.5-v0.2 is a visual document retrieval model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

238,273100
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

lightonai/LightOnOCR-2-1B

lightonai/LightOnOCR-2-1B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

237,414800
1.0B6 GB+ VRAMapache-2.0
Deployment details
text-generation

ReadyArt/gemma-4-31B-it-scotoma-2-GGUF

ReadyArt/gemma-4-31B-it-scotoma-2-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

236,42036
31.0B24 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

Qwen/Qwen3-ASR-1.7B-hf

Qwen/Qwen3-ASR-1.7B-hf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

236,06188
1.7B6 GB+ VRAMapache-2.0
Deployment details
general AI

kernels-community/triton-layer-norm

kernels-community/triton-layer-norm is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

235,0320
Unknown8 GB+ VRAMbsd-3-clause
Deployment details