nvidia/parakeet-rnnt-0.6b
nvidia/parakeet-rnnt-0.6b is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
264,510 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
nvidia/parakeet-rnnt-0.6b is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
leejet/MiniMax-H3-GGUF is a image to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
nineninesix/gepard-1.0 is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
AtomicChat/Qwen3.5-9B-DFlash-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mistralai/Ministral-3-14B-Instruct-2512-BF16 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
google/gemma-4-E2B is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
CrashOverrideX/Quillan-Ronin is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
AtomicChat/Qwen3-4B-DFlash-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
Lightricks/LTX-2.5-Diffusers is a text to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.
AtomicChat/Qwen3.5-4B-DFlash-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
utter-project/EuroLLM-22B-Instruct-2512 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
MangoGoes/libero4in1_wan2.2vae_latent_cosmos_style is a robotics model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
SC117/Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-Preserved-APEX-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
QuantTrio/gemma-4-31B-it-AWQ is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
unsloth/Qwen3.8-27B-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
BAAI/seggpt-vit-large is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
allenai/Olmo-3-7B-Think is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
cyankiwi/Qwen3-VL-2B-Instruct-AWQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
protectai/MoritzLaurer-roberta-base-zeroshot-v2.0-c-onnx is a zero shot classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
handy-computer/gigaam-v3-e2e-rnnt-gguf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
AiHub4MSRH-Hash/hash-MedGemma-4B-16bit-eng-swa-it is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
LiquidAI/LFM2-1.2B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
twolven/Qwen3.8-27B-abliterated-AWQ-MTP is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
goldhub/Qwen3.8-27B-BF16-INT4-W4A16-G32-AutoRound is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.