Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

Spaces:
logic65
/
whittle-45b-chat
Running on Zero

App Files Files Community
Fetching metadata from the HF Docker repository...
whittle-45b-chat
20.4 kB
Ctrl+K
Ctrl+K
  • 1 contributor
History: 8 commits
logic65's picture
logic65
Request at most 140 s of GPU (x2 on xlarge fits a free account's 300 s/day); max 1024 new tokens
9b47d2f verified 5 days ago
  • .gitattributes
    1.52 kB
    initial commit 5 days ago
  • README.md
    690 Bytes
    Whittle 45B ZeroGPU chat demo: stock transformers + whittle_load.py, model mounted at /models/w45 5 days ago
  • app.py
    9.54 kB
    Request at most 140 s of GPU (x2 on xlarge fits a free account's 300 s/day); max 1024 new tokens 5 days ago
  • requirements.txt
    79 Bytes
    Small files from the Hub, shards from the mount (size-checked); drop flash-linear-attention (no GPU at import on ZeroGPU) 5 days ago
  • whittle_load.py
    8.6 kB
    Table to RAM with 64 MB sequential reads (page-faulting the mmap over the mount was far too slow) 5 days ago