python

Enabled configurable auto Tensor Parallelism (TP) for the inference of diverse models #11491

# to view logs

Re-run triggered February 5, 2025 01:47

delock

#6553

gyou2021:configurable_autoTP

Status Success

Total duration 2m 5s

Artifacts –

python.yml

on: pull_request

Matrix: unit-tests