Chat · Free · collected from HuggingFace
Qwen3-8B-AWQ is a 4-bit quantized version of the Qwen3-8B model, optimized for memory-constrained environments without significant performance loss. It supports long-context understanding and diverse tool-use capabilities, making it suitable for local inference tasks. The AWQ format ensures faster loading times and lower resource requirements compared to full-precision models.
ai huggingface text-generation
Listed as Free. Pricing changes often — confirm on the official site.
Listings are collected automatically from public sources and refreshed daily. We do not take payment for placement.