HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

1,628,986 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

usermma/TwIL-LM3-mlx-8Bit

usermma/TwIL-LM3-mlx-8Bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,867♡ 0
Unknown8 GB+ VRAMother
Deployment details →
image-text-to-text

Vontra/GLM-5.3-Flash-MLX-oQ2-MTP

Vontra/GLM-5.3-Flash-MLX-oQ2-MTP is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,867♡ 2
Unknown8 GB+ VRAMmit
Deployment details →
text-generation

cjvt/GaMS-9B-Instruct

cjvt/GaMS-9B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 1,866♡ 2
9.0B24 GB+ VRAMgemma
Deployment details →
text-generation

AETHER-LAB/akie-350m

AETHER-LAB/akie-350m is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,866♡ 0
Unknown8 GB+ VRAMunknown
Deployment details →
general AI

mistralai/Devstral-Small-2505

mistralai/Devstral-Small-2505 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,865♡ 868
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-generation

McGill-NLP/AfriqueQwen3.5-4B-50Langs

McGill-NLP/AfriqueQwen3.5-4B-50Langs is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

↓ 1,865♡ 7
4.0B12 GB+ VRAMcc-by-4.0
Deployment details →
image-text-to-text

Hcompany/Holo-3.1-9B

Hcompany/Holo-3.1-9B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 1,864♡ 27
9.0B24 GB+ VRAMapache-2.0
Deployment details →
text-generation

fraQtl/Qwen3.6-35B-A3B-Hi-Fi-GGUF

fraQtl/Qwen3.6-35B-A3B-Hi-Fi-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

↓ 1,863♡ 4
35.0B32 GB+ VRAMapache-2.0
Deployment details →
general AI

mradermacher/Piko-9b-i1-GGUF

mradermacher/Piko-9b-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,863♡ 3
9.0B8 GB+ VRAMapache-2.0
Deployment details →
text-to-image

nvidia/Cosmos3-Super-Text2Image

nvidia/Cosmos3-Super-Text2Image is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

↓ 1,863♡ 183
Unknown12 GB+ VRAMother
Deployment details →
image-text-to-text

mlx-community/Qwen3.8-Flash-Next-4bit

mlx-community/Qwen3.8-Flash-Next-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,863♡ 3
Unknown8 GB+ VRAMLicense unknown
Deployment details →
text-generation

amd/MiniMax-M2.7-MXFP4

amd/MiniMax-M2.7-MXFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,862♡ 1
Unknown8 GB+ VRAMother
Deployment details →
text-to-speech

CC-TM/Qwen3-TTS-GGUF

CC-TM/Qwen3-TTS-GGUF is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,862♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →