AEON-7/AEON-DFlash-Qwen3.6-35B-A3B
AEON-7/AEON-DFlash-Qwen3.6-35B-A3B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
845,060 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
AEON-7/AEON-DFlash-Qwen3.6-35B-A3B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
unsloth/Qwen3.8-Flash-Next-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
saricles/Qwen3-Coder-Next-NVFP4-GB10 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.
LocalAI-io/LocalVQE is a audio to audio model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Lucebox/DeepSeek-V4-Flash-0731-DSpark-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
amd/gpt-oss-20b-MoE-Quant-W-MXFP4-A-FP8-KV-FP8 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
Rabbit-Hole-Ai/Qwen3.8-27B-5090-goldilocks-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
bottlecapai/ThinkingCap-Qwen3.6-27B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. Access approval is required on Hugging Face.
nappa114514/Qwen-Image-Edit-2509-Manga-Tone is a image to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
cusiman/Krea2-realism-V2 is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
bartowski/ProCreations_grug-35b-v2-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
rzgar/minimax_h3_ref2va_fp8_e4m3fn is a image to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
bartowski/zai-org_GLM-4.5-Air-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
PeppX/gemma-4-e2b-uncensored-litertlm is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
reaperdoesntknow/Dualmind-Qwen-1.7B-Thinking is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
Forturne/Qwen3-VL-Embedding-8B-NVFP4 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
FunAudioLLM/Fun-ASR-Nano-GGUF is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
furiosa-ai/Qwen3-Reranker-4B is a text classification model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mradermacher/Mistral-NeMo-12B-Abliterated-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mlx-community/Qwen3-ASR-0.6B-4bit is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
cross-encoder/ettin-reranker-1b-v1 is a text ranking model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
trl-internal-testing/tiny-Qwen2AudioForConditionalGeneration is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.8-27B-Uncensored-Cyber-i1-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
bartowski/bottlecapai_ThinkingCap-Qwen3.6-27B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.