You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
Tracking issue for retiring unused / unmaintained DeepSpeed features so the runtime, docs, tests, and packaging stay consistent.
Goal: remove code paths that are no longer supported, then sweep the leftover docs, extras, CODEOWNERS, and tests so they cannot be re-enabled by accident.
Compression leftovers: 1-bit comm backends, MoQ/eigenvalue, fake-quant kernels, docs/CODEOWNERS/issue template
There is no PR for the full follow-up sweep described below (MoQ + 1-bit comm + PLD + Apex AMP as one bundle). #8535 covers the 1-bit comm and MoQ/eigenvalue leftovers. Progressive Layer Dropping and Apex AMP still have no PR.
Coverage statistics
Line counts are from git diff --numstat. Code = runtime/library sources; tests = tests/; docs = docs/ plus CODEOWNERS / issue templates / CI where noted.
Three independent classifications (Opus, Codex, Grok) over the same GitHub pull agreed on the next deprecation set below.
Method
Issues/PRs: unauthenticated /search/issues on 2026-09-14. Primary column is -org:deepspeedai. Relative score vs "zero_optimization" -org:deepspeedai = 1,362.
Code: authenticated /search/code on 2026-09-14 via gh. Relative score vs "zero_optimization" -org:deepspeedai = 49,280. Code search over-counts vendored DeepSpeed copies (e.g. DeepSpeed-0.9.5/ inside other repos) and generic identifiers. Treat as ordinal rank, not user counts.
Not consensus / do not add from this scan: inference v1/v2, Windows, ZeRO offload, pipeline, DeepCompile, Hybrid Engine, autotuning, DataStates, ZenFlow.
Apex AMP removal is still local / un-PRed; PyTorch torch_autocast is unchanged. Local import apex remains only to detect apex.optimizers.FusedAdam for FP16 optimizer wrapping.
Stats methodology: Deprecate unused DeepSpeed features #8490 from git diff --numstat 71d316d60^1 71d316d60; follow-up from git diff --numstat origin/master on the working tree at the time; combined vs b726f4edb.
Summary
Tracking issue for retiring unused / unmaintained DeepSpeed features so the runtime, docs, tests, and packaging stay consistent.
Goal: remove code paths that are no longer supported, then sweep the leftover docs, extras, CODEOWNERS, and tests so they cannot be re-enabled by accident.
PRs
71d316d60, +169 / −8,298)elastic_checkpoint(use Universal Checkpointing). Stage 1/2 unchangedsparse_gradientszeropp_loco_param)There is no PR for the full follow-up sweep described below (MoQ + 1-bit comm + PLD + Apex AMP as one bundle). #8535 covers the 1-bit comm and MoQ/eigenvalue leftovers. Progressive Layer Dropping and Apex AMP still have no PR.
Coverage statistics
Line counts are from
git diff --numstat. Code = runtime/library sources; tests =tests/; docs =docs/plusCODEOWNERS/ issue templates / CI where noted.Removed in #8490
config.py/engine.py/ launchers / CODEOWNERS)recursive_getattr→module_utils)Follow-up sweep (no dedicated PR)
Local working-tree estimate vs
masterwhen this was written: +104 / −2,814. Combined vs pre-#8490: +253 / −11,092.amppath)Leftover configs should fail at parse time (
DeepSpeedConfigErroror ZeRO-3ValidationError) instead of silently training without the feature.Features to deprecate
Training / checkpointing
deepspeed/nebula,NebulaCheckpointEngine) — #8490MiCS_Init,mics_shard_size,mics_hierarchical_params_gather) — #8490zero_optimization.elastic_checkpointwith stage 3; use Universal Checkpointing). ZeRO-1/2 elastic checkpoints remain. — #8099Optimizers
OneBitAdam) — #8490OneBitLamb) — #8490ZeroOneAdam) — #8490Compression / quantization
deepspeed.compression, compression scheduler, related tests) — #8490quantize_training/ eigenvalue scheduling — #8535Mixed precision
ampconfig). Usefp16,bf16, ortorch_autocast. No PR.Training schedule
progressive_layer_drop). No PR.1-bit communication leftovers (optimizers are gone; backends remain)
deepspeed/runtime/comm/{nccl,mpi,compressed}.py,deepspeed/runtime/compression/cupy.py) — #8535tests/onebit/and remaining unit tests for those backends — #8535setup.pyextras1bit/1bit_mpiandrequirements/requirements-1bit-mpi.txt— #8535Docs / repo hygiene
.github/ISSUE_TEMPLATE/compression_bug_report.md— #8535CODEOWNERS/deepspeed/runtime/fp16/onebit/(removed in Deprecate unused DeepSpeed features #8490)CODEOWNERS/deepspeed/runtime/compression/— #8535GitHub usage consensus (2026-09-14)
Three independent classifications (Opus, Codex, Grok) over the same GitHub pull agreed on the next deprecation set below.
Method
/search/issueson 2026-09-14. Primary column is-org:deepspeedai. Relative score vs"zero_optimization" -org:deepspeedai= 1,362./search/codeon 2026-09-14 viagh. Relative score vs"zero_optimization" -org:deepspeedai= 49,280. Code search over-counts vendored DeepSpeed copies (e.g.DeepSpeed-0.9.5/inside other repos) and generic identifiers. Treat as ordinal rank, not user counts.Consensus candidates
"sparse_attention""mode")"sparse_attention"sparse_gradients"sparse_gradients""sparse_gradients""max_train_batch_size" "micro_batch_sizes"graph_harvesting"graph_harvesting""graph_harvesting"curriculum_learning"curriculum_learning""curriculum_learning""zeropp_loco_param""zeropp_loco_param"2-of-3 only (not unanimous; do not treat as consensus)
Data Efficiency / Random-LTD,
dump_state(debug flag;"dump_state"code search is 164k unrelated hits),disable_allgather, FusedLion.Other unused / already-deprecated candidates
deepspeed/ops/sparse_attention,extras_require['sparse_attn']) — #8493sparse_gradients(config-json already calls this essentially deprecated) — #8504elasticity/max_train_batch_size+micro_batch_sizes)graph_harvestingzero_optimization.zeropp_loco_param) — #8534Notes
coalesced_collectives.pystays for ZeRO-3).deepspeed/compression/helper.pyremains as aFutureWarningshim for DeepSpeed-Chat after Deprecate unused DeepSpeed features #8490.torch_autocastis unchanged. Localimport apexremains only to detectapex.optimizers.FusedAdamfor FP16 optimizer wrapping.git diff --numstat 71d316d60^1 71d316d60; follow-up fromgit diff --numstat origin/masteron the working tree at the time; combined vsb726f4edb.