antoinelouis/colbert-xm
antoinelouis/colbert-xm is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
300,305 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
antoinelouis/colbert-xm is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
Avuja/Qwen3.8-27B-int4-AutoRound is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
lynaNSFW/DaSiWa_MiniMax_H3 is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
handy-computer/Fun-ASR-Nano-2512-gguf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
agentionai/Qwen3.8-Flash-Next-ROCmFP4-FAST-imatrix-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
z-lab/Qwen3.5-9B-DFlash is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Jommarn/UNSEEN_Gemma_4_26B_NSFW-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Inferact/GLM-5.3-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
noctrex/Ling-3.0-tiny-MXFP4_MOE-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
molbal/Minimax-Music3-GGUF is a text to audio model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
LiquidAI/LFM2.5-230M-Base is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
handy-computer/parakeet-ctc-0.6b-gguf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
antirez/glm-5.2-gguf is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
PocketAiHub/Qwen3.8-27B-Abliterated-MTPLX-Optimized-Speed is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
rdtand/Qwen3.8-27B-PrismaAQUA-5.5bit-vllm is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
bigscience/bloom is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
deepseek-ai/dspark_qwen3_8b_block7 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
malaiwah/GLM-5.2-EXL3-FQ-segments is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
optimum-intel-internal-testing/tiny-random-gemma4-unified is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
dzannotti/Qwen3.8-Flash-Next-MTP-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
syvai/hviske-v5.3 is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
zerodigest/Qwen3.8-27B-Cold-Fusion-GAIN-V1.1-YMQ-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
google/tipsv2-b14-dpt is a depth estimation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
lemuralabs/Qwen3.6-27B-V2-abliterated-uncensored-Q4_K_M-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.