← Model index
image-text-to-text

barozp/Qwen3.8-27B-Opus-Distill-v2-FP8

barozp/Qwen3.8-27B-Opus-Distill-v2-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

Parameters27.0B
Downloads1,424
Likes3
Licenseapache-2.0
Librarytransformers
Hardware tierDatacenter
DEPLOYMENT NOTES

What to verify before deployment

  • Confirm the exact weight format and quantization.
  • Measure memory at your intended context or image size.
  • Review model-card limitations and evaluation methodology.
  • Test latency and throughput on your target runtime.
LICENSE SIGNAL

Likely commercial-friendly

License metadata is an index signal, not legal advice. Follow the repository license and any model-specific acceptable-use terms.

Verify at the source →
RUN THIS MODEL

Choose how you want to deploy

GPU CLOUD · SIMPLE START

Rent a GPU sized for this model

Start with at least 96 GB VRAM, then measure memory and latency with your workload.

Check RunPod availability ↗Referral link. AI Pentium may earn credits or commission from qualifying new users.
GPU MARKETPLACE · PRICE SHOP

Compare marketplace GPU offers

Filter Vast.ai offers for a GPU with at least 96 GB VRAM and compare availability by region.

Compare GPUs on Vast.ai ↗Referral link. AI Pentium may earn commission from qualifying new users.
ORIGINAL SOURCE

Inspect weights and model card

Verify files, license terms, limitations and usage instructions before deploying.

View on Hugging Face ↗Source link; AI Pentium does not host model weights.
SIMILAR OPTIONS

More image-text-to-text models

image-text-to-text

Qwen/Qwen3-VL-8B-Instruct

Qwen/Qwen3-VL-8B-Instruct is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

15,231,4941,084
8.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Qwen/Qwen3.5-9B

Qwen/Qwen3.5-9B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

11,389,2111,912
9.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

google/gemma-4-26B-A4B-it

google/gemma-4-26B-A4B-it is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

8,588,4681,484
26.0B64 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Qwen/Qwen2.5-VL-7B-Instruct

Qwen/Qwen2.5-VL-7B-Instruct is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

7,747,6151,699
7.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Qwen/Qwen3.5-4B

Qwen/Qwen3.5-4B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

7,264,578897
4.0B12 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Qwen/Qwen3.6-27B-FP8

Qwen/Qwen3.6-27B-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

7,198,015354
27.0B64 GB+ VRAMapache-2.0
Deployment details

Operate a GPU cloud or inference API?

Reach developers after they have selected a model and are ready to run it.

Become a platform partner