Empero AI has released several versions of its Qwen3.8-9B model, including base, distilled, and quantized formats, on Hugging Face. These models are designed for compatibility with various popular inference frameworks such as llama.cpp, vLLM, Ollama, and Transformers. The releases provide detailed instructions and code snippets for integrating these models into different development environments, from local applications to cloud-based notebooks. AI
IMPACT Facilitates broader adoption and experimentation of Qwen3.8-9B models across diverse AI development stacks.
RANK_REASON The cluster consists of Hugging Face model repository pages providing instructions for using specific model versions with various inference tools.
Read on Hugging Face Trending Models →
- Dockerdocker
- empero-ai/Qwen3.8-9B
- Google Colab
- Hugging Face
- Kaggle
- lmsysorg
- OpenAI
- Qwen3.8
- SGLang
- transformers
- vLLM
- empero-ai/Qwen3.8-9B-GGUF
- llama.cpp
- Ollama
- Unsloth Studio
- empero-ai/Qwen3.8-9B-Distill
- empero-ai/Qwen3.8-9B-Distill-GGUF
- Jan
- LM Studio
AI-generated summary · Google Gemini · from 4 sources. How we write summaries →