[https://nvbugs/6535779][fix] Remove stale Qwen3.5 and DeepSeekV32 waivers - #17649
[https://nvbugs/6535779][fix] Remove stale Qwen3.5 and DeepSeekV32 waivers#17649VALLIS-NERIA wants to merge 2 commits into
Conversation
|
/bot run --extra-stage "DGX_B200-4_GPUs-PyTorch-Post-Merge-1, DGX_B200-4_GPUs-PyTorch-Post-Merge-2, DGX_B200-4_GPUs-PyTorch-Post-Merge-3, DGX_B200-4_GPUs-PyTorch-Post-Merge-4" |
|
PR_Github #66118 [ run ] triggered by Bot. Commit: |
|
PR_Github #66118 [ run ] completed with state
|
|
No actionable comments were generated in the recent review. 🎉 ℹ️ Recent review info⚙️ Run configurationConfiguration used: Path: .coderabbit.yaml Review profile: CHILL Plan: Enterprise Run ID: 📒 Files selected for processing (1)
🚧 Files skipped from review as they are similar to previous changes (1)
WalkthroughThe integration test waive list removes three Qwen3.5 397B NVFP4 skip entries and one B300 DeepSeekV32 chunked-prefill skip entry. ChangesIntegration test waive-list cleanup
Estimated code review effort: 1 (Trivial) | ~2 minutes Merge Risk: ⚪ Minimal · up to This change restores coverage by removing four stale test waivers, and no actionable merge-blocking risk remains after normal checks and review. 🚥 Pre-merge checks | ✅ 5✅ Passed checks (5 passed)
✨ Finishing Touches🧪 Generate unit tests (beta)
Comment |
Signed-off-by: Xiwen Yu <13230610+VALLIS-NERIA@users.noreply.github.com>
Signed-off-by: Xiwen Yu <13230610+VALLIS-NERIA@users.noreply.github.com>
5972811 to
3784d5a
Compare
|
/bot run |
|
PR_Github #66465 [ run ] triggered by Bot. Commit: |
|
PR_Github #66465 [ run ] completed with state
|
What
Why
NVBugs 6535779 and 6483369 are fixed, so retaining these entries hides valid integration coverage. The NVBug 6483369 waiver removal is moved from #17438; its MoE unit-test changes are intentionally discarded. This fresh change supersedes #17611 because that branch now conflicts with the current
mainwaiver list.Validation
1 passed.python3 -m pre_commit run --files tests/integration/test_lists/waives.txtTestQwen3_5_397B_A17B::test_nvfp4_mtp3_gdn_replay_tep41 passedin 809.08 s; executor evaluation time 74.219 sMambaHybridCacheManagerV2,KVCacheV2Scheduler, cached replay, and replay state updates enabledReplay manager performance sanity check
Qwen3.5-35B-A3B-FP8 on one B200, MTP3 cached replay, fixed 128-request synthetic workload (ISL 2048 / OSL 512, concurrency 64), two reversed-order A/B pairs:
This is a directional executor-level comparison; GPU clocks were not locked, and the second pair used an otherwise idle GPU on a shared host.
Dev Engineer Review
tests/integration/test_lists/waives.txt.QA Engineer Review
tests/integration/test_lists/waives.txt.test-db/orqa/files.