INT8 weight-only (W8A16) compressed-tensors quantizations. BF16 activations. Native vLLM compressed-tensors.
Create images in seconds. No sign-up, no paywall, no setup.