KittyLM-4B (qwen3-4b)

A kitten. In a language model. This is a LoRA finetune of Qwen/Qwen3-4B that answers everything in kitten language (mrrp, nya~, prrr, *actions*, occasional :3) while staying factually correct underneath.

Training

  • Data: 900 ShareGPT-style pairs + 100 held-out eval (see KittyLM/kittylm-data), system prompt baked in
  • Method: LoRA SFT (scripts/train.py in the KittyLM project), RTX 3060 12GB
  • Files: merged bf16 weights + the LoRA adapter (adapter_*.safetensors) live side by side. AutoModelForCausalLM loads the merged model; PeftModel picks up the adapter.
  • GGUF quants for Ollama / llama.cpp / LM Studio: KittyLM/kittylm-qwen3-4b-gguf
  • Eval: v2 (prompt-free): ablation full/none/generic style 0.75/0.78/0.81, 15/15 factual; train loss 5.74->0.48, eval loss 0.885->0.869->0.937 (best @ epoch 2). Voice e.g. 'plant drinks sunlight... solar dinner buffet. envy. owo.'

Limitations

  • Persona is stylistic, not a refusal behavior: an explicit "answer in plain English" can make it drop character — it was never trained to resist.
  • Small-model knowledge gaps persist: off-distribution facts may confabulate (tracked per model by the 20-probe suite in eval/).
  • Kitten flavor adds no capability: reasoning/coding limits are the base model's limits.

Usage (transformers)

from transformers import AutoModelForCausalLM, AutoTokenizer
tok = AutoTokenizer.from_pretrained("KittyLM/kittylm-qwen3-4b")
model = AutoModelForCausalLM.from_pretrained("KittyLM/kittylm-qwen3-4b", device_map="auto", dtype="bfloat16")
msgs = [{"role": "user", "content": "What's the capital of Japan?"}]
x = tok.apply_chat_template(msgs, tokenize=False, add_generation_prompt=True)
print(tok.decode(model.generate(**tok(x, return_tensors="pt").to(model.device), max_new_tokens=120)[0]))
# mrrp... Tokyo. big city. lots of cats. nya~

Usage (Ollama)

hf download KittyLM/kittylm-qwen3-4b-gguf --include "*.gguf" --local-dir ./gguf
ollama create kittylm-qwen3-4b -f Modelfile   # see Modelfile template in project
ollama run kittylm-qwen3-4b "Good night!"
Downloads last month
695
Safetensors
Model size
4B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for KittyLM/kittylm-qwen3-4b

Finetuned
Qwen/Qwen3-4B
Adapter
(1170)
this model
Quantizations
1 model

Dataset used to train KittyLM/kittylm-qwen3-4b

Collections including KittyLM/kittylm-qwen3-4b