HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

969,049 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

any-to-any

coder3101/gemma-4-E4B-it-heretic

coder3101/gemma-4-E4B-it-heretic is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,661♡ 32
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-ranking

nlpai-lab/LAMAR-600m

nlpai-lab/LAMAR-600m is a text ranking model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,661♡ 19
Unknown8 GB+ VRAMmit
Deployment details →
text-generation

osunlp/QUEST-35B-SFT

osunlp/QUEST-35B-SFT is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

↓ 1,660♡ 1
35.0B80 GB+ VRAMapache-2.0
Deployment details →
token-classification

kormilitzin/en_core_med7_lg

kormilitzin/en_core_med7_lg is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,659♡ 25
Unknown8 GB+ VRAMmit
Deployment details →
image-text-to-text

Jackrong/Qwopus3.6-27B-v2-GGUF

Jackrong/Qwopus3.6-27B-v2-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 1,659♡ 251
27.0B24 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

bairongz/QianfanOCR

bairongz/QianfanOCR is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,658♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-generation

batiai/Qwen3.6-35B-A3B-GGUF

batiai/Qwen3.6-35B-A3B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

↓ 1,658♡ 5
35.0B32 GB+ VRAMapache-2.0
Deployment details →
general AI

cstr/granite-speech-4.1-2b-GGUF

cstr/granite-speech-4.1-2b-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 1,658♡ 5
2.0B4 GB+ VRAMapache-2.0
Deployment details →
automatic-speech-recognition

cstr/voxtral-mini-4b-realtime-GGUF

cstr/voxtral-mini-4b-realtime-GGUF is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

↓ 1,658♡ 1
4.0B6 GB+ VRAMapache-2.0
Deployment details →
text-generation

inclusionAI/LLaDA2.2-flash

inclusionAI/LLaDA2.2-flash is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,658♡ 88
Unknown8 GB+ VRAMapache-2.0
Deployment details →
feature-extraction

braindecode/eegdino-small-pretrained

braindecode/eegdino-small-pretrained is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 1,657♡ 0
Unknown4 GB+ VRAMbsd-3-clause
Deployment details →
text-generation

Abhisingh-18/Sutra-1.3B-Chat

Abhisingh-18/Sutra-1.3B-Chat is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

↓ 1,657♡ 1
1.3B6 GB+ VRAMapache-2.0
Deployment details →
text-generation

InternScience/Agents-A1-FP8

InternScience/Agents-A1-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,656♡ 23
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-classification

latam-gpt/Wayra-Perplexity-Estimator-55M

latam-gpt/Wayra-Perplexity-Estimator-55M is a text classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,655♡ 21
Unknown8 GB+ VRAMapache-2.0
Deployment details →