HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

635,071 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

BayesRL/Llama3.1-IVON-SFT-8B

BayesRL/Llama3.1-IVON-SFT-8B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,4980
8.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

OrisTeam/Vyuhu-280M-Base-3012m

OrisTeam/Vyuhu-280M-Base-3012m is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,4921
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

unsloth/Ornith-1.0-397B-GGUF

unsloth/Ornith-1.0-397B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.

2,49023
397.0B256 GB+ VRAMmit
Deployment details
token-classification

stanfordnlp/stanza-hi

stanfordnlp/stanza-hi is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,4900
Unknown8 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

seanghay/Qwen3-ASR-0.6B-Khmer

seanghay/Qwen3-ASR-0.6B-Khmer is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

2,48913
600M4 GB+ VRAMapache-2.0
Deployment details
fill-mask

nvidia/esm2_t48_15B_UR50D

nvidia/esm2_t48_15B_UR50D is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.

2,4858
15.0B48 GB+ VRAMmit
Deployment details
general AI

SoraExplora/VideoMae

SoraExplora/VideoMae is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,4850
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

kenotron/Qwen3.8-27B-mlx-4Bit

kenotron/Qwen3.8-27B-mlx-4Bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,4842
27.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

MuXodious/Qwen3.8-27B-absolute-heresy

MuXodious/Qwen3.8-27B-absolute-heresy is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

2,47713
27.0B64 GB+ VRAMapache-2.0
Deployment details
general AI

meituan/EvoCUA-8B-20260105

meituan/EvoCUA-8B-20260105 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,47616
8.0B24 GB+ VRAMapache-2.0
Deployment details
token-classification

OpenMed/privacy-filter-nemotron-mlx-8bit

OpenMed/privacy-filter-nemotron-mlx-8bit is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,4737
Unknown8 GB+ VRAMapache-2.0
Deployment details