-
Notifications
You must be signed in to change notification settings - Fork 285
All issues
Issue creation is restricted in this repository
- #3278 路 HuiyingLi opened
on Jul 28, 2026 - #2958 路 khazic opened
on Jul 7, 2026 7 - #3576 路 yuhezhang-ai opened
on Aug 18, 2026 1
Issues
is:issue state:open
is:issue state:open
Search results
- Status: Open.#3778聽In NVIDIA-NeMo/Automodel;
kd loss under tensor parallelism trains with gradients scaled by tp_size
bugSomething isn't workingSomething isn't workingwaiting-on-customerWaiting on the original author to respondWaiting on the original author to respondStatus: Open.#3776聽In NVIDIA-NeMo/Automodel;- Status: Open.#3774聽In NVIDIA-NeMo/Automodel;
- Status: Open.#3767聽In NVIDIA-NeMo/Automodel;
Agent docs describe the pre-0.4.0 CLI and a PR title format that fails CI
waiting-on-customerWaiting on the original author to respondWaiting on the original author to respondStatus: Open.#3761聽In NVIDIA-NeMo/Automodel;group_by_length silently does nothing for lazily-tokenized datasets
waiting-on-maintainersWaiting on maintainers to respondWaiting on maintainers to respondStatus: Open.#3755聽In NVIDIA-NeMo/Automodel;- Status: Open.#3754聽In NVIDIA-NeMo/Automodel;
Loading Qwen/Qwen3.8-27B failed with NeMoAutoModelForCausalLM(OOM), but successfully load with Qwen3_5ForCausalLM
bugSomething isn't workingSomething isn't workingStatus: Open.#3745聽In NVIDIA-NeMo/Automodel;mask_reasoning_content does not mask template-inserted empty
bugSomething isn't workingSomething isn't workingwaiting-on-customerWaiting on the original author to respondWaiting on the original author to respondStatus: Open.#3738聽In NVIDIA-NeMo/Automodel;Qwen3.5 4B TE packed seq
enhancementNew feature or requestNew feature or requestStatus: Open.#3735聽In NVIDIA-NeMo/Automodel;- Status: Open.#3727聽In NVIDIA-NeMo/Automodel;
- Status: Open.#3726聽In NVIDIA-NeMo/Automodel;