onnxruntime

mirror of https://github.com/saymrwulf/onnxruntime.git synced 2026-05-14 20:48:00 +00:00

Author	SHA1	Message	Date
Yifan Li	5c3c7643db	Update range of gpu arch (#23309 ) ### Description <!-- Describe your changes. --> * Remove deprecated gpu arch to control nuget/python package size (latest TRT supports sm75 Turing and newer arch) * Add 90 to support blackwell series in next release (86;89 not considered as adding them will rapidly increase package size) \| arch_range \| Python-cuda12 \| Nuget-cuda12 \| \| -------------- \| ------------------------------------------------------------ \| ---------------------------------- \| \| 60;61;70;75;80 \| Linux: 279MB Win: 267MB \| Linux: 247MB Win: 235MB \| \| 75;80 \| Linux: 174MB Win: 162MB \| Linux: 168MB Win: 156MB \| \| 75;80;90 \| Linux: 299MB Win: 277MB \| Linux: 294MB Win: 271MB \| \| 75;80;86;89 \| [Linux: MB Win: 390MB](https://aiinfra.visualstudio.com/Lotus/_build/results?buildId=647457&view=results) \| Linux: 416MB Win: 383MB \| \| 75;80;86;89;90 \| [Linux: MB Win: 505MB](https://aiinfra.visualstudio.com/Lotus/_build/results?buildId=646536&view=results) \| Linux: 541MB Win: 498MB \| ### Motivation and Context <!-- - Why is this change required? What problem does it solve? - If it fixes an open issue, please link to the issue here. --> Callout: While adding sm90 support, the build of cuda11.8+cudnn8 will be dropped in the coming ORT release, as the build has issue with blackwell (mentioned in comments) and demand on cuda 11 is minor, according to internal ort-cuda11 repo.	2025-01-14 14:27:34 -08:00
Yulong Wang	7b0fa407eb	fix requirements.txt path (#22946 ) ### Description #22380 removes the file `tools/ci_build/github/linux/docker/inference/x86_64/python/cpu/scripts/requirements.txt` but it is still used in `dockerfiles/Dockerfile.cuda`. This change updates the file path of the requirements.txt fixes #22945.	2024-12-04 13:08:29 -08:00
Tianlei Wu	72186bbb71	[CUDA] Build nhwc ops by default (#22648 ) ### Description * Build cuda nhwc ops by default. * Deprecate `--enable_cuda_nhwc_ops` in build.py and add `--disable_cuda_nhwc_ops` option Note that it requires cuDNN 9.x. If you build with cuDNN 8, NHWC ops will be disabled automatically. ### Motivation and Context In general, NHWC is faster than NCHW for convolution in Nvidia GPUs with Tensor Cores, and this could improve performance for vision models. This is the first step to prefer NHWC for CUDA in 1.21 release. Next step is to do some tests on popular vision models. If it help in most models and devices, set `prefer_nhwc=1` as default cuda provider option.	2024-11-06 09:54:55 -08:00
Tianlei Wu	071b607807	[CUDA] Add CUDA_VERSION and CUDNN_VERSION etc. arguments to Dockerfile.cuda (#22351 ) ### Description * Add a few arguments CUDA_VERSION, CUDNN_VERSION, OS, GIT_COMMIT, GIT_BRANCH and ONNXRUNTIME_VERSION to the Dockerfile.cuda to allow for more flexibility in the build process. * Update README.md to include the new arguments and their usage. * Output labels to image so that it is easy to inspect the image. Available CUDA versions for ubuntu 24.04 can be found [here](https://hub.docker.com/r/nvidia/cuda/tags), and available CUDNN versions can be found [here](https://pypi.org/project/nvidia-cudnn-cu12/#history). Example command line to build docker image: ``` docker build -t onnxruntime-cuda --build-arg CUDA_VERSION=12.6.1 \ --build-arg CUDNN_VERSION=9.5.0.50 \ --build-arg GIT_BRANCH=$(git rev-parse --abbrev-ref HEAD) \ --build-arg GIT_COMMIT=$(git rev-parse HEAD) \ --build-arg ONNXRUNTIME_VERSION=$(cat ../VERSION_NUMBER) \ -f Dockerfile.cuda .. ``` Example labels from `docker inspect onnxruntime-cuda`: ``` "Labels": { "CUDA_VERSION": "12.6.1", "CUDNN_VERSION": "9.5.0.50", "maintainer": "Changming Sun <chasun@microsoft.com>", "onnxruntime_git_branch": "main", "onnxruntime_git_commit": "bc84958dcef5c6017ae58085f55b669efd74f4a5", "onnxruntime_version": "1.20.0", "org.opencontainers.image.ref.name": "ubuntu", "org.opencontainers.image.version": "24.04" } ``` ### Motivation and Context https://github.com/microsoft/onnxruntime/pull/22339 has hard-coded the cuda and cudnn versions. User might want to choose specified cuda and cudnn version during building docker image.	2024-10-09 12:06:33 -07:00
Tianlei Wu	8595e56d8e	[CUDA] Update Docker to use Ubuntu 24.04, cuda 12.6, cudnn 9.4 and python 3.12 (#22339 ) ### Description Serve as example to build and run onnxruntime-gpu with latest software stack. To build docker image: ``` git clone https://github.com/microsoft/onnxruntime cd onnxruntime/dockerfiles docker build -t onnxruntime-cuda -f Dockerfile.cuda .. ``` To launch the docker image built from previous step (and mount the code directory to run a unit test below): ``` cd .. docker run --rm -it --gpus all -v $PWD:/code onnxruntime-cuda /bin/bash ``` Then run the following in docker image to verify that the cuda provider is good: ``` python /code/onnxruntime/test/python/onnxruntime_test_python_cudagraph.py ``` ### Motivation and Context https://github.com/microsoft/onnxruntime/issues/22335	2024-10-08 09:54:46 -07:00
Tianlei Wu	c7d0ded079	[CUDA] Update Dockerfile.cuda with cuda 12.5.1 and cudnn 9 (#21987 ) ### Description Previous image is based on cuda 12.1 and cudnn 8, which is out of date since we have moved to cudnn 9 since 1.19 release. (1) Upgrade base image to cuda 12.5.1 and cudnn 9. (2) Update CMAKE_CUDA_ARCHITECTURES from 52;60;61;70;75;86 to 61;70;75;80;86;90 to support A100 and H100 (3) Make the build faster: exclude unit test; use ninja etc. (4) upgrade some packages (like packaging etc) before building to avoid build error. ### Motivation and Context https://github.com/microsoft/onnxruntime/issues/21792 https://github.com/microsoft/onnxruntime/issues/21532	2024-09-05 15:25:40 -07:00
zkep	7313accd44	Update Dockerfile.cuda (#21042 )	2024-06-13 23:50:03 -07:00
Changming Sun	c6b0d185b4	Update cmake to 3.27 and upgrade Linux CUDA docker files from CentOS7 to UBI8 (#16856 ) ### Description 1. Update docker files and their build instructions. ARM64 and x86_64 can use the same docker file. 2. Upgrade Linux CUDA pipeline's base docker image from CentOS7 to UBI8 AB#18990	2023-09-05 18:12:10 -07:00
Changming Sun	7c58d013aa	Remove Ubuntu 18.04 usages (#15781 ) ### Description Remove Ubuntu 18.04 usages because it will be EOL this month. ### Motivation and Context	2023-05-11 11:44:00 -07:00
Changming Sun	5b826b1bc3	Update cmake version in Linux build (#15707 ) ### Description All our Windows build pipelines already uses cmake 3.26 except one pipeline: QNN ARM64. This PR does the same for Linux build pipelines. ### Motivation and Context This change is related to #15704 .	2023-04-27 20:02:33 -07:00
Edward Chen	ea40dc3ad6	Update build.py to disallow running as root user by default. (#15164 ) Try to address intermittent permissions issues that show up in non-transient CI environments.	2023-03-27 14:46:04 -07:00
Jian Zeng	928c44f71b	fix(cuda): install missing python3-packaging in Dockerfile Signed-off-by: Jian Zeng <anonymousknight96@gmail.com>	2022-12-12 16:26:29 -08:00
Changming Sun	23da468154	Upgrade cmake version to 3.24 (#13569 ) ### Description Upgrade cmake version to 3.24 because I need to use a new feature that is only provided in that version and later. Starting from cmake 3.24, the [FetchContent](https://cmake.org/cmake/help/latest/module/FetchContent.html#module:FetchContent) module and the [find_package()](https://cmake.org/cmake/help/latest/command/find_package.html#command:find_package) command now support integration capabilities, which means calls to "FetchContent" can be implicitly redirected to "find_package", and vice versa. Users can use a cmake variable to control the behavior. So, we don't need to provide such a build option. We can delete our "onnxruntime_PREFER_SYSTEM_LIB" build option and let cmake handle it. And it would be easier for who wants to use vcpkg. ### Motivation and Context Provide a unified package management method, and get aligned with the community. This change is split from #13523 for easier review.	2022-11-04 22:58:51 -07:00
vade	bacae967a2	Update Cuda to 11.4.2, update architectures, support Ubuntu 20.04 (#10169 )	2022-01-07 13:00:44 -08:00
Olivia Jain	60089f7093	Cuda11.4 (#8709 ) * initial update from 11.1 to 11.4 * change 11.4.1 to 11.4.0 * adjusting to match nvidia/cuda image tags * adjusting to match nvidia/cuda image tags centos7 * correction to 11.4.0 * correction to 11.4.0 * update to cuda 11.4 * change training back to 11.1 * change training back to 11.1 * point to correct nvcr.io/nvidia/cuda 11.4.1 image * change centos8 to centos7 * correct cudnn path * Update linux-gpu-ci-pipeline.yml for Azure Pipelines * Update c-api-noopenmp-packaging-pipelines.yml * need to resolve centos images but remove space and change to 11.4 * Update linux-gpu-ci-pipeline.yml * add cudnn to docker image * bump devtoolset to 10 * revert cuda 11.4 change to setup_env_trt * orttraining back to 11.1 * use nvcr.io * Fix previous change back to cuda 11.1 * update cudnn path * use cudnn image (revert if failure)	2021-08-17 16:36:26 -07:00
Changming Sun	0510688411	Update compliance tasks in python packaging pipeline and fix some compile warnings (#8471 ) 1. Update SDLNativeRules from v2 to v3. The new one allows us setting excluded paths. 2. Update TSAUpload from v1 to v2. And add a config file ".gdn/.gdntsa" for it. 3. Fix some parentheses warnings 4. Update cmake to the latest. 5. Remove "--x86" build option from pipeline yaml files. Now we can auto-detect cpu architecture from python. So we don't need to ask user to specify it.	2021-07-30 17:16:37 -07:00
Changming Sun	29ca08a729	Update Dockerfile.cuda: remove compute capability 30 30 is not supported in CUDA 11.1. 35,37,50 are deprecated.	2021-07-12 15:19:48 -07:00
Changming Sun	ed5fd919ef	Update dockerfiles to use the latest cmake (#7933 )	2021-06-03 18:51:00 -07:00
Changming Sun	3323fb6082	Update docker files to put 'unattended-upgrades' in a right place(#5983 )	2020-12-01 10:45:03 -08:00
Changming Sun	1dbabb2362	Update dockerfiles (#5929 ) 1. Remove conda from the images. Because conda contains a file named /opt/miniconda/lib/libcrypto.so.1.0.0 which can't pass our security scan. Also, it will be easier for us to manage the third party usage registrations. 2. Remove openssh from the images. Because the official openssh package provided by Ubuntu can't pass our security scan. 3. Reduce the image size to 1/3 by using stages. Also, because it contains less packages, it will be less often needed to update. 4. Put the LICENSE-IMAGE.txt file in right place. It is missed in current images. You can see it was added to a temp folder "/code" but it got deleted afterwards. 5. Update the CPU docker image's base image to Ubuntu 18.04. The GPU one is already 18.04. It's better to keep them the same. 6. Remove the build arg ONNXRUNTIME_REPO/ONNXRUNTIME_BRANCH. Instead, the new one always uses the local source. I feel it can reduce confusion.	2020-11-25 15:38:22 -08:00
Changming Sun	965e2b095d	Update MCR CUDA docker image to 10.2 (#5181 )	2020-09-16 09:01:31 -07:00
George Wu	8d2e22558d	unattended-upgrades (#4804 )	2020-08-14 18:12:27 -07:00
Changming Sun	0a6d9dd301	Remove Openmp from the GPU docker files	2020-05-25 14:17:48 -07:00
Changming Sun	30efe65e95	Add use_openmp back to the docker files	2020-05-25 14:17:48 -07:00
George Wu	ba0e7daf20	update dockerfiles/README (#2336 )	2019-11-06 16:54:10 -08:00
George Wu	06a6d74a67	update ngraph dockerfile. add python lib location to LD_LIBRARY_PATH for cuda/tensorrt Dockerfiles. (#2330 )	2019-11-06 11:29:55 -08:00
Pranav Sharma	69970d1f2a	Include the new Privacy.md file in all release packages. (#2200 )	2019-10-20 07:58:36 -07:00
Vinitra Swamy	2a21df2309	Updates to CUDA and TensorRT dockerfiles for v0.5.0 (#1731 ) * updates to cuda and tensorrt dockerfiles for v0.5.0 * add table of build tags	2019-09-13 14:16:47 -07:00
Hector Li	dc9c89546d	Update the docker file for OpenVINO (#1741 ) Update the docker file for OpenVINO which is used for AML	2019-08-30 22:32:24 -07:00
manashgoswami	6d783e8a07	Added license files in the base image (#1595 ) * Update Dockerfile.openvino * Update Dockerfile.cuda * Update Dockerfile.cuda * Update Dockerfile.openvino * Update Dockerfile.cuda * added ThirdParty notice file to base image. * corrected license file name	2019-08-09 13:02:06 -07:00
Vinitra Swamy	6b32c77804	Dockerfiles for TensorRT, CUDA, build from source (#922 ) * dockerfile updates for BYOC scenario * updates for 3 different build versions * updating to remove libopenblas, python3, python3-pip * Including LICENSE-IMAGE.txt for CUDA/TensorRT dockerfiles * remove unnecessary cmake files * fixing comment typo * optimizing dockerfile.source as per review suggestions (not working currently) * Optimizing dockerfiles with install_dependencies script * update dockerfile with --cmake_extra_defines version number * add &&\ for license copy lines * updates, adding miniconda to path, reincluded clearing the pycache * adding maintainer note * update readme instructions * update tensorrt versioning in dockerfile	2019-07-09 02:03:55 -07:00

31 commits