Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

Spaces:

Duplicated from  Leon4gr45/gemma-inference

Leon4gr45
/
fable5-inference
Paused

App Files Files Community
Fetching metadata from the HF Docker repository...
fable5-inference
74.2 kB
Ctrl+K
Ctrl+K
  • 1 contributor
History: 38 commits
Leon4gr45's picture
Leon4gr45
Upload folder using huggingface_hub
fd44567 verified 25 days ago
  • __pycache__
    Upload folder using huggingface_hub 25 days ago
  • .gitattributes
    1.52 kB
    initial commit 4 months ago
  • .gitignore
    45 Bytes
    Upload gemma-inference space 4 months ago
  • .hfignore
    45 Bytes
    Upload gemma-inference space 4 months ago
  • Dockerfile
    2.2 kB
    build: use prebuilt llama.cpp CPU binary (b9895) instead of source compile — fixes mtmd OOM hang, ~20x faster build about 1 month ago
  • README.md
    3.88 kB
    Upload folder using huggingface_hub 25 days ago
  • app.py
    19 kB
    Upload folder using huggingface_hub 25 days ago
  • requirements.txt
    38 Bytes
    v2.0: single-instance CPU proxy, native OpenAI tools+streaming, flash-attn+q4_0 KV, all-cores about 1 month ago
  • test_chat.py
    1.1 kB
    Upload folder using huggingface_hub 25 days ago