Pull Request已成功合入, 合并人@ascend-robot
(感谢 xfeng 的贡献)Thanks for your pull-request.
The full list of commands accepted by me can be found at here。
You can get sig-info at here
PR Approval Progress
✅ Congratulations! All modules have met the lgtm and approve requirements.
Module Approval Details
| module | lgtm status | approve status |
|---|---|---|
| test | ✅ li_jing_hw, 陈豪 (2/2) | ✅ li_jing_hw, 陈豪 (2/1) |
| torch_npu/profiler | ✅ 陈豪, li_jing_hw (2/2) | ✅ 陈豪, li_jing_hw (2/1) |
💡 Tip:
- Committer can comment
/approveor/lgtm- Commenting
/approveimplies both code review (lgtm) and intent to merge (approve)
CLA Signature Pass
zyb_230, thanks for your pull request. All authors of the commits have signed the CLA. 👍


当前仓库存在以下 保护分支 :
| Protected Branch | Version | Release |
|---|---|---|
| master | ||
| v2.8.0 | ||
| v2.7.1 | ||
| v2.9.0 | ||
| v2.11.0 | ||
| v2.10.0 |
评论 /sync <branch1> <branch2> ... 可将当前 PR 修改同步到其它分支(创建同步 PR):
a) 如果当前 PR 是 Open 状态,同步操作将延迟到 PR 被合并时执行
b) 如果当前 PR 已经 Merged,将立即执行同步操作
注意:
- /sync 命令可以指定同步到多个分支,仅最后一个 /sync 命令生效
- 如果创建的同步 PR 不正确,可通过向同步 PR 的源分支提交轻量级 PR 完善,或使用 /close 命令关闭


本PR中共发现代码检查告警抑制1处,请Committer检视合理性
本评论自动扫描PR中使用的开源代码检查工具(ruff、clang-tidy、CodeQL等)的屏蔽注释。
这些屏蔽注释会阻止开源代码检查工具对特定代码区域的检测,可能导致潜在问题被忽略。
| 文件路径 | 行号 | 屏蔽类型 | 代码片段 | 工具名称 |
|---|---|---|---|---|
| torch_npu/profiler/profiler.py | 1 | 行级屏蔽 | 1: # ruff: noqa: UP045, UP007 3: import os.path 4: from collections.abc import Callable, Iterable 6: from typing import Any, Optional, Union 12: _enable_profiler_in_child_thread, 14: NpuProfilerConfig, |
ruff |


Review Guide
This pull-request passes review.
Committers who wrote a comment of /approve are: 陈豪, li_jing_hw.
Reviewers who wrote a comment of /lgtm are: li_jing_hw, 陈豪.


In response to this:
/sync v2.9.0 v2.10.0 v2.11.0 master
同步操作执行结果:
| Branch | Status | Pull Request |
|---|---|---|
| v2.9.0 | 同步失败:请手动创建 PR 进行同步,我们会继续完善分支之间同步操作,尽量避免同步失败的情况 | |
| v2.10.0 | 同步失败:请手动创建 PR 进行同步,我们会继续完善分支之间同步操作,尽量避免同步失败的情况 | |
| v2.11.0 | 同步失败:请手动创建 PR 进行同步,我们会继续完善分支之间同步操作,尽量避免同步失败的情况 | |
| master | 同步失败:请手动创建 PR 进行同步,我们会继续完善分支之间同步操作,尽量避免同步失败的情况 |


【合入来源】
【修改方案】
对齐pytorch profiler的部分接口:

一个是增加一个
skip_first_wait:一个是
custom_trace_id_callback,然后增加两个对应的set和get的接口set_custom_trace_id_callback和get_trace_id【资料变更】
涉及,需要和资料沟通
【接口变更】
涉及,前面已经说明
【功能验证】
场景1


设置
schedule = torch_npu.profiler.schedule(wait=2, warmup=1, active=1, repeat=2, skip_first=0, skip_first_wait=1)之前:采集第3和第7个step
现在:采集第1和第5个step
场景2
测试get_trace_id接口:
默认是一个uuid,是直接从pytorch里面拷贝过来的,现在会在profiler_metadata.json里面落盘,db里面也有
场景3


异常的skip_first_wait参数不生效,reset为0
正常:
异常:必须设置为整数,否则有警告信息,reset为0
场景4:
设置custom_trace_id_callback,这个trace_id,我们是想和每一份ascend_pt数据或者repeat参数绑定的,
如果call_back类型不对,会有警告信息,然后使用默认的uuid

【CheckList】