|
Download .venv/transformers/docs/source/en/main_classes/deepspeed.md from DrDavis/PythonProject1: direct link, hf CLI and curl.
- Browser
- Download file 1.56 kB
-
https://huggingface.co/DrDavis/PythonProject1/resolve/main/.venv/transformers/docs/source/en/main_classes/deepspeed.md
- Command line
-
hf download hf://DrDavis/PythonProject1/.venv/transformers/docs/source/en/main_classes/deepspeed.md
-
curl -L -o deepspeed.md https://huggingface.co/DrDavis/PythonProject1/resolve/main/.venv/transformers/docs/source/en/main_classes/deepspeed.md
1.56 kB
DeepSpeed
DeepSpeed, powered by Zero Redundancy Optimizer (ZeRO), is an optimization library for training and fitting very large models onto a GPU. It is available in several ZeRO stages, where each stage progressively saves more GPU memory by partitioning the optimizer state, gradients, parameters, and enabling offloading to a CPU or NVMe. DeepSpeed is integrated with the [Trainer] class and most of the setup is automatically taken care of for you.
However, if you want to use DeepSpeed without the [Trainer], Transformers provides a [HfDeepSpeedConfig] class.
Learn more about using DeepSpeed with [Trainer] in the DeepSpeed guide.
HfDeepSpeedConfig
[[autodoc]] integrations.HfDeepSpeedConfig - all