HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

1,159,035 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

johninthepool/Qwen3.8-27B-MTPLX-8bit

johninthepool/Qwen3.8-27B-MTPLX-8bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

↓ 1,383♡ 2
27.0B32 GB+ VRAMapache-2.0
Deployment details →
text-generation

Myric/KAT-Coder-V2.5-Dev-APEX-GGUF

Myric/KAT-Coder-V2.5-Dev-APEX-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,383♡ 1
Unknown8 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

topabaem/Qwen3.8-27B-STQ

topabaem/Qwen3.8-27B-STQ is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 1,383♡ 0
27.0B24 GB+ VRAMapache-2.0
Deployment details →
general AI

OpenVINO/Qwen3-8B-int4-cw-ov

OpenVINO/Qwen3-8B-int4-cw-ov is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,381♡ 13
8.0B8 GB+ VRAMapache-2.0
Deployment details →
text-generation

saidutta69/Qwen3-0.6B-heretic

saidutta69/Qwen3-0.6B-heretic is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 1,381♡ 1
600M4 GB+ VRAMapache-2.0
Deployment details →
text-generation

specklabs/Speck1-140M-Instruct

specklabs/Speck1-140M-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,381♡ 4
Unknown8 GB+ VRAMmit
Deployment details →
text-generation

nvidia/Qwen3-Nemotron-235B-A22B-GenRM

nvidia/Qwen3-Nemotron-235B-A22B-GenRM is a text generation model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.

↓ 1,380♡ 31
235.0B256 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

onnx-community/Qwen3.5-0.8B-ONNX-OPT

onnx-community/Qwen3.5-0.8B-ONNX-OPT is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 1,380♡ 1
800M4 GB+ VRAMapache-2.0
Deployment details →
text-generation

logic65/whittle-next-moe-test

logic65/whittle-next-moe-test is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,380♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

majentik/Qwen3.8-27B-MLX-2bit

majentik/Qwen3.8-27B-MLX-2bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

↓ 1,379♡ 2
27.0B64 GB+ VRAMapache-2.0
Deployment details →
text-generation

IvmeLabs/Ivme-Conversate-XL-v1-Base

IvmeLabs/Ivme-Conversate-XL-v1-Base is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,379♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →