Conversation
🔗 Helpful Links🧪 See artifacts and rendered test results at hud.pytorch.org/pr/pytorch/executorch/23385
Note: Links to docs will display an error until the docs builds have been completed. ❌ 9 New Failures, 1 Pending, 1 Unrelated Failure, 19 Unclassified FailuresAs of commit 8214424 with merge base 0b3d26d ( NEW FAILURES - The following jobs have failed:
UNCLASSIFIED FAILURES - DrCI could not classify the following jobs because the workflow did not run on the merge base. The failures may be pre-existing on trunk or introduced by this PR:
FLAKY - The following job failed but was likely due to flakiness present on trunk:
This comment was automatically generated by Dr. CI and updates every 15 minutes. |
|
@Gasoonjia has exported this pull request. If you are a Meta employee, you can view the originating Diff in D123135327. |
This PR needs a
|
…3385) Summary: cuda-perf.yml is retired. Its benchmark + upload pipeline now runs inside cuda.yml's test-model-cuda-e2e cells, reusing each cell's exported pte/ptd and built runner in place (new helper .ci/scripts/cuda_benchmark_from_e2e.sh, allowlisted to the 7 perf models; whisper-medium backfilled into the accuracy matrix). New jobs: - upload-benchmark-results: aggregates cuda-bench-* artifacts to the unchanged s3://gha-artifacts/executorch-cuda-perf/ prefix (history preserved) + HUD dashboard v3 upload. - check-perf-regression: fails the PR when decode or prefill tok/s drops >5% vs the median of the last 5 successful main runs (same model, quant, GPU, min 3 history points; new .ci/scripts/cuda_check_regression.py). Bypass with the bypass-perf-regression PR label. Cleanup: delete cuda-perf.yml, rename trigger_cuda_perf.sh to trigger_cuda_benchmark.sh (retargeted at cuda.yml dispatch inputs), drop ciflow/cuda-perf from pytorch-probot.yml, update the _ci-run-decision.yml caller comment. Differential Revision: D123135327
cb2bbc8 to
8214424
Compare
Summary:
cuda-perf.yml is retired. Its benchmark + upload pipeline now runs inside
cuda.yml's test-model-cuda-e2e cells, reusing each cell's exported
pte/ptd and built runner in place (new helper
.ci/scripts/cuda_benchmark_from_e2e.sh, allowlisted to the 7 perf
models; whisper-medium backfilled into the accuracy matrix).
New jobs:
unchanged s3://gha-artifacts/executorch-cuda-perf/ prefix (history
preserved) + HUD dashboard v3 upload.
drops >5% vs the median of the last 5 successful main runs (same
model, quant, GPU, min 3 history points; new
.ci/scripts/cuda_check_regression.py). Bypass with the
bypass-perf-regression PR label.
Cleanup: delete cuda-perf.yml, rename trigger_cuda_perf.sh to
trigger_cuda_benchmark.sh (retargeted at cuda.yml dispatch inputs),
drop ciflow/cuda-perf from pytorch-probot.yml, update the
_ci-run-decision.yml caller comment.
Differential Revision: D123135327