Instructions to use nikita-savelyev-cerebras/tiny-random-qwen3.8-flash-next-fp8 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use nikita-savelyev-cerebras/tiny-random-qwen3.8-flash-next-fp8 with Transformers:
# pip install -U transformers accelerate # Load model directly from transformers import AutoModel model = AutoModel.from_pretrained("nikita-savelyev-cerebras/tiny-random-qwen3.8-flash-next-fp8", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Tiny random Qwen3.8 Flash-Next FP8 text model
This synthetic Qwen4Exp text model uses the architecture and fine-grained FP8 expert layout of Qwen/Qwen3.8-Flash-Next-FP8, reduced to four decoder blocks and four routed experts per block. Its weights are random and intended for loader tests, not inference quality. A compact tokenizer is adapted from malaiwah/qwen4-exp-tiny-random-bf16.
- Downloads last month
- 44
Inference Providers NEW
This model isn't deployed by any Inference Provider. 馃檵 Ask for provider support