mlas/arm64: add NEON conv asm kernels and tune NCHWC kernel selection #10400
windows_tensorrt.yml
on: pull_request
Windows GPU TensorRT CI Pipeline
29m 53s
Windows GPU TensorRT CI Pipeline Test Job
38m 27s
Annotations
6 warnings
|
Windows GPU TensorRT CI Pipeline:
onnxruntime/core/mlas/lib/amd64/QgemmU8X8KernelAvx2.asm#L1234
epilog offset from end of function exceeds 4095
|
|
Windows GPU TensorRT CI Pipeline:
onnxruntime/core/mlas/lib/amd64/QgemmU8X8KernelAvx2.asm#L1227
epilog offset from end of function exceeds 4095
|
|
Windows GPU TensorRT CI Pipeline:
onnxruntime/core/mlas/lib/amd64/QgemmU8X8KernelAvx2.asm#L1220
epilog offset from end of function exceeds 4095
|
|
Windows GPU TensorRT CI Pipeline:
onnxruntime/core/mlas/lib/amd64/QgemmU8X8KernelAvx2.asm#L1213
epilog offset from end of function exceeds 4095
|
|
Windows GPU TensorRT CI Pipeline:
onnxruntime/core/mlas/lib/amd64/QgemmU8X8KernelAvx2.asm#L1206
epilog offset from end of function exceeds 4095
|
|
Windows GPU TensorRT CI Pipeline:
onnxruntime/core/mlas/lib/amd64/QgemmU8X8KernelAvx2.asm#L1199
epilog offset from end of function exceeds 4095
|
Artifacts
Produced during runtime
| Name | Size | Digest | |
|---|---|---|---|
|
build-artifacts
|
1.89 GB |
sha256:2b9a5c9f539615099dfde9c73b1491d1916ffec5532fe1a2a83754848851db11
|
|