nv-accelerate-v100

Enabled configurable auto Tensor Parallelism (TP) for the inference of diverse models #13173

Sign in to view logs

Summary
Jobs
- unit-tests
Run details
- Usage
- Workflow file

Re-run triggered February 20, 2025 13:18

delock

#6553

gyou2021:configurable_autoTP

Status Success

Total duration 8m 18s

Artifacts –

nv-accelerate-v100.yml

on: pull_request