-
Notifications
You must be signed in to change notification settings - Fork 4.5k
All issues
Issue creation is restricted in this repository
Issues
is:issue state:open
is:issue state:open
Search results
[ENHANCEMENT] Accept ProcessGroupCollection in inference communication_utils broadcast helpers
waiting-on-customerWaiting on the original author to respondWaiting on the original author to respondStatus: Open.#7119 In NVIDIA/Megatron-LM;[BUG] InferenceInterface.base_generate uses assert NotImplementedError and never raises
waiting-on-customerWaiting on the original author to respondWaiting on the original author to respondStatus: Open.#7118 In NVIDIA/Megatron-LM;- Status: Open.#7106 In NVIDIA/Megatron-LM;
- Status: Open.#7105 In NVIDIA/Megatron-LM;
- Status: Open.#7103 In NVIDIA/Megatron-LM;
- Status: Open.#7100 In NVIDIA/Megatron-LM;
Add trainable QSA sparse attention with a TileLang backend
enhancementNew feature or requestNew feature or requestwaiting-on-customerWaiting on the original author to respondWaiting on the original author to respondStatus: Open.#7060 In NVIDIA/Megatron-LM;- Status: Open.#7039 In NVIDIA/Megatron-LM;
RuntimeError: Trying to resize storage that is not resizable in fine-grained activation offloading when force-releasing MoE expert_fc1 input that is a CUDA-graph static output
bugSomething isn't workingSomething isn't workingwaiting-on-maintainersWaiting on maintainers to respondWaiting on maintainers to respondStatus: Open.#7009 In NVIDIA/Megatron-LM;- Status: Open.#7006 In NVIDIA/Megatron-LM;
[REGRESSION] core_v0.19.0: HF->mcore checkpoint converter broken twice — GTP_remat process groups uninitialized, and saver inherits target-derived weight-shard args
bugSomething isn't workingSomething isn't workingStatus: Open.#6989 In NVIDIA/Megatron-LM;- Status: Open.#6977 In NVIDIA/Megatron-LM;