AI & ML interests

None defined yet.

cbensimonΒ 
updated a model 4 months ago
cbensimonΒ 
published a model 4 months ago
cbensimonΒ 
updated a model 4 months ago
cbensimonΒ 
published a model 4 months ago
cbensimonΒ 
updated a model 4 months ago
cbensimonΒ 
published a model 4 months ago
cbensimonΒ 
posted an update over 1 year ago
view post
Post
5441
πŸš€ ZeroGPU now supports PyTorch native quantization via torchao

While it hasn’t been battle-tested yet, Int8WeightOnlyConfig is already working flawlessly in our tests.

Let us know if you run into any issues β€” and we’re excited to see what the community will build!

import spaces
from diffusers import FluxPipeline
from torchao.quantization.quant_api import Int8WeightOnlyConfig, quantize_

pipeline = FluxPipeline.from_pretrained(...).to('cuda')
quantize_(pipeline.transformer, Int8WeightOnlyConfig()) # Or any other component(s)

@spaces.GPU
def generate(prompt: str):
    return pipeline(prompt).images[0]
  • 5 replies
Β·
cbensimonΒ 
posted an update over 1 year ago
view post
Post
6237
πŸš€ ZeroGPU medium size is now available as a power-user feature

Nothing too fancy for nowβ€”ZeroGPU Spaces still default to large (70GB VRAM)β€”but this paves the way for:
- πŸ’° size-based quotas / pricing (medium will offer significantly more usage than large)
- 🦣 the upcoming xlarge size (141GB VRAM)

You can as of now control GPU size via a Space variable. Accepted values:
- auto (future default)
- medium
- large (current default)

The auto mode checks total CUDA tensor size during startup:
- More than 30GB β†’ large
- Otherwise β†’ medium
  • 3 replies
Β·
cbensimonΒ 
posted an update about 2 years ago
view post
Post
4810
Hello everybody,

We've rolled out a major update to ZeroGPU! All the Spaces are now running on it.

Major improvements:

1. GPU cold starts about twice as fast!
2. RAM usage reduced by two-thirds, allowing more effective resource usage, meaning more GPUs for the community!
3. ZeroGPU initializations (coldstarts) can now be tracked and displayed (use progress=gr.Progress(track_tqdm=True))
4. Improved compatibility and PyTorch integration, increasing ZeroGPU compatible spaces without requiring any modifications!

Feel free to answer in the post if you have any questions

πŸ€— Best regards,
Charles