saymrwulf/onnxruntime: ONNX Runtime: cross-platform, high performance ML inferencing and training accelerator

mirror of https://github.com/saymrwulf/onnxruntime.git synced 2026-07-09 17:28:58 +00:00

ONNX Runtime: cross-platform, high performance ML inferencing and training accelerator

Find a file

Hector Li 5daeb5e0b0 enable model with external data be loaded from memory buffer (#19089 ) ### Description Background: User save large model with initializer data in external file. e.g: onnx.save_model(onnx_model, "path/to/save/the/model.onnx", save_as_external_data=True, all_tensors_to_one_file=True, location="filename", size_threshold=1024). In that case, Ort loads the model, get the external initializer information (external file name, offset, length) and use the model path to find the external file, and locate to the tensor data via the offset and length. But it won't work if user load the model from memory, since Ort lost track of the model path. This PR adds API/session option to let user provide a table with external initializer file name as the key, the pointer to the loaded external file in memory and the buffer length as value. So that 1. user can load the model from memory buffer with external initializers in memory buffer too. 2. the initializers can be shared across sessions, for different EPs. 3. user can load the file in any way they want, e.g mmap. Internally, 1. at session creation time, Ort goes through the external initializers in the graph, gets the file name, offset, data length of the external initializers from Tensorproto . 2. With the file name, Ort get the file in memory buffer and buffer length from the table user provided. 4. Ort locates the tensor buffer from file in memory buffer (user provided) using the offset and data length (from Tensorproto ). 5. Ort creates the Tensor and replace the existing Tensor in the graph. ### Motivation and Context https://github.com/onnx/onnx/blob/main/docs/ExternalData.md For a model with external data, the Tensorproto may have initializer data in a separate file. The external file location is set via the file path relative to the model path. With the API to load model from memory buffer, it lost track of the model path. So it causes error if the model has external data. By adding a session option to set the external data buffer, Ort can find the external data correctly if model loaded from memory buffer.		2024-04-17 19:01:01 -07:00
.config
.devcontainer
.gdn
.github	Bump gradle/wrapper-validation-action from 2 to 3 (#20305 )	2024-04-16 14:20:51 -07:00
.pipelines
.vscode
cgmanifests	Integration with ONNX 1.16.0 (#19745 )	2024-04-12 09:46:49 -07:00
cmake	Add patch for ONNX 1.16.0 shape inference bug (#20316 )	2024-04-17 10:23:22 -07:00
csharp	Bump Sixlabors.ImageSharp from 2.1.7 to 2.1.8 in /csharp/sample/Microsoft.ML.OnnxRuntime.FasterRcnnSample (#20314 )	2024-04-17 14:47:44 -07:00
dockerfiles
docs	Add Gemma Rotary Embedding (#20267 )	2024-04-16 15:31:56 -07:00
include/onnxruntime/core	enable model with external data be loaded from memory buffer (#19089 )	2024-04-17 19:01:01 -07:00
java
js	fix csum and enable ut (#20355 )	2024-04-17 15:01:06 -07:00
objectivec	[objc] Add check for ORTValue being a tensor in ORTValue methods that should only be used with tensors. (#19946 )	2024-03-18 08:54:24 -07:00
onnxruntime	enable model with external data be loaded from memory buffer (#19089 )	2024-04-17 19:01:01 -07:00
orttraining	enable model with external data be loaded from memory buffer (#19089 )	2024-04-17 19:01:01 -07:00
rust
samples
tools	More fixes on random connection excepiton in Mac Build. (#20328 )	2024-04-17 08:37:56 +08:00
winml	#19921 [Dup] LLC Core count calculations updated (#20171 )	2024-04-02 16:53:47 -07:00
.clang-format
.clang-tidy
.dockerignore
.gitattributes
.gitignore
.gitmodules
.lintrunner.toml
build.bat
build.sh
build_arm64x.bat
CITATION.cff
CODEOWNERS
CONTRIBUTING.md
lgtm.yml
LICENSE
NuGet.config
ort.wprp
ORT_icon_for_light_bg.png
packages.config
pyproject.toml
README.md
requirements-dev.txt
requirements-doc.txt
requirements-lintrunner.txt
requirements-training.txt
requirements.txt.in
SECURITY.md
setup.py
ThirdPartyNotices.txt	Fix HalideIR title in third party notices reference (#20190 )	2024-04-05 11:12:43 -07:00
VERSION_NUMBER

README.md

ONNX Runtime is a cross-platform inference and training machine-learning accelerator.

ONNX Runtime inference can enable faster customer experiences and lower costs, supporting models from deep learning frameworks such as PyTorch and TensorFlow/Keras as well as classical machine learning libraries such as scikit-learn, LightGBM, XGBoost, etc. ONNX Runtime is compatible with different hardware, drivers, and operating systems, and provides optimal performance by leveraging hardware accelerators where applicable alongside graph optimizations and transforms. Learn more →

ONNX Runtime training can accelerate the model training time on multi-node NVIDIA GPUs for transformer models with a one-line addition for existing PyTorch training scripts. Learn more →

Get Started & Resources

General Information: onnxruntime.ai
Usage documentation and tutorials: onnxruntime.ai/docs
YouTube video tutorials: youtube.com/@ONNXRuntime
Upcoming Release Roadmap
Companion sample repositories:
- ONNX Runtime Inferencing: microsoft/onnxruntime-inference-examples
- ONNX Runtime Training: microsoft/onnxruntime-training-examples

Builtin Pipeline Status

System	Inference	Training
Windows
Linux
Mac
Android
iOS
Web
Other

Third-party Pipeline Status

System	Inference	Training
Linux

Data/Telemetry

Windows distributions of this project may collect usage data and send it to Microsoft to help improve our products and services. See the privacy statement for more details.

Contributions and Feedback

We welcome contributions! Please see the contribution guidelines.

For feature requests or bug reports, please file a GitHub Issue.

For general discussion or questions, please use GitHub Discussions.

Code of Conduct

This project has adopted the Microsoft Open Source Code of Conduct. For more information see the Code of Conduct FAQ or contact opencode@microsoft.com with any additional questions or comments.

License

This project is licensed under the MIT License.