onnxruntime

mirror of https://github.com/saymrwulf/onnxruntime.git synced 2026-07-05 04:17:53 +00:00

Author	SHA1	Message	Date
Changming Sun	bd5451b4ed	Don't define USE_OPENMP if the compiler doesn't support OpenMP (#1836 )	2019-09-13 16:42:50 -07:00
KeDengMS	cf22ea6893	Upgrade TVM for a fix in tvm::Integer with int64_t input (#1824 ) * Upgrade TVM for a fix in tvm::Integer with int64_t input	2019-09-12 23:15:13 -07:00
Bowen Bao	8712a523a4	Bump onnx to latest (#1756 ) * Bump onnx to latest Update onnx.in.proto with changes for SparseTensor. * add temp skip tests * remove passed tests from skip list * skip more tests for new ops in opset 11 * skip crashing tests * update handling of new attribute types sparse tensor and sparse tensors * advance onnx commit and remove skip cpu_flaky_tests * temporarily skip yolo3 model test due to resize opset10 shape inference regression * update proto for onnxruntime server * advance onnx commit further	2019-09-12 11:46:49 -07:00
Dmitri Smirnov	fe8915863c	Implement C API entry points for creating and fetching non-standard types to OrtValue (#1714 ) C/C++ Opage APIs Add new virtual interfaces for NonTensorType Implement entry points. Add shared header for the data container. Add export symbols. Add serialization/deserialization. Implement model with Opaque types. Rework opqaue_api_test as a standalone executable.	2019-09-11 14:52:47 -07:00
shahasad	6a5b11756b	Conditionally export execution provider apis in chsarp (#1724 )	2019-09-09 11:17:44 -07:00
Tracy Sharpe	071a0c2522	MLAS: MlasSgemm refactoring (#1749 ) Refactor the SGEMM kernels to resynchronize the code between Windows/Linux and remove unneeded binary bloat from a different zero/add mode kernel. Another goal is to get to a cleaner state for then doing a DGEMM kernel.	2019-09-06 22:26:28 -07:00
KeDengMS	58fe5a6bf1	Enable Nuphar docker build, and reinstate Nuphar tests (#1757 ) Enable Nuphar EP docker build Revert back to LLVM 6.0.1 Reinstate disabled Softmax tests caused by LLVM 8.0.1 Reinstate Nuphar Python test due to stale sympy version Increase build timeout of Linux CI	2019-09-05 08:50:48 -07:00
Changming Sun	94d9161166	Add nuphar to Linux CI build (#1750 )	2019-09-03 11:39:27 -07:00
KeDengMS	c9240f4e93	Implementation of Nuphar execution provider (#881 ) * Implement Nuphar execution provider Nuphar execution provider is a TVM-based compilation provider. It has shown great speedups for RNN models using Scan. This PR is mainly for a preview of the shared codegen library for other TVM-based providers. * Fix submodules * Fix TVM submodule * Update Nuphar to latest and resolve confliction * Remove stale files caused by merge -X theirs * Revert heap buffer change to not introduce onnxruntime_framework into onnxruntime_perf_test * Fix bad merge * Merge from Nuphar * Fix warning treated as error, revert some unnecessary changes * Revert some more test changes * Some more test revert or comments to make review easier New tests could be added later * One more revert of unnecessary changes * More change revert. Test could be added back later.	2019-09-01 23:01:47 -07:00
Changming Sun	81ad48080b	Remove TaskThreadPool (#1713 )	2019-08-28 18:00:10 -07:00
Tracy Sharpe	73312b8195	MLAS: Android sgemm kernel build fix (#1710 ) Fix the aarch64 kernel to build properly with the Android NDK (specifically clang).	2019-08-28 16:14:12 -07:00
Ashwini Khade	961b14ac4a	use MLAS for QGEMM in matmulInteger and convInteger (#1692 ) * use mlas qgemm for u8u8_s32 gemms * update test	2019-08-26 18:13:22 -07:00
Changming Sun	4de0aa8049	Optimize kernel index (#1672 )	2019-08-22 10:26:35 -07:00
Changming Sun	7be5695fad	Remove --whole-archive (#1655 )	2019-08-20 12:04:10 -07:00
jywu-msft	bdc694314c	update MKLML to version which contains fix for thread hang. (#1636 ) * update MKLML which has bugfix for thread hang. move PATCH_COMMAND outside BUILD_FOR_NATIVE_MACHINE check. * MKLML_VERSION 2020.0.20190813 is for windows only.	2019-08-19 10:34:48 -07:00
Tracy Sharpe	bc72c2dba7	MLAS: add U8U8 MatMul operation (#1644 ) Implement the first round of changes for quantization inside MLAS. This adds a MatMul operation for U8xU8=S32 for x86/x64 processors.	2019-08-18 18:15:48 -07:00
Changming Sun	6b89c7ad04	Let mlas use session thread pool (#1609 ) 1.Let mlas use session thread pool 2.Remove onnxruntime_USE_MLAS cmake option 3. Remove the win32 thread pool code inside mlas mlas will: 1.use ort thread pool if it get passed in 2.use openmp if the threadpool parameter is nullptr 3.run single threaded if the threadpool parameter is nullptr and openmp is disabled.	2019-08-16 13:21:15 -07:00
Ashwini Khade	0044be6259	update onnx to latest commit (#1622 ) * update onnx to latest commit * Disable and/or fix failing tests * disable not yet implemented tests for opset 11 * disable tests * fix bug in mkldnn fp16 graph check	2019-08-15 17:10:32 -07:00
Dmitri Smirnov	17c8fe44e3	Integrate featurizers (#1573 ) Added Sample Featurizer and Infrastructure Make featurizers and unit tests compile and run with GTest. Create definitions for the first featurizer kernel. Add new operator domain. Create datetime_transformer kernel and build. Move OPAQUE types definitions for featurizers kerneles out to a separate cc. Register them with the type system. Provide unit tests for new AutoML DateTimeTransformer kernel. Make necessary adjustments to the test infrastructure to make it run with new types.	2019-08-15 13:59:59 -07:00
Pranav Sharma	a6a4c4c079	Fix perf test executable. (#1598 ) * Mention OrtCreateSessionFromArray in C API doc * Fix perf test executable due to removal of certain C APIs * fix linux build * Avoid duplication * Fix mem leak	2019-08-12 09:49:29 -07:00
Tomasz Dołbniak	69baf9e800	Update nGraph to v0.22.1 (#1582 ) * Update nGraph to 0.21 and adjust the EP * Share the graph initializers between custom ops * Update nGraph to 0.22 and exclude Gather entirely * Enable building on Windows with nGraph v0.21.1-rc.0 * Disable the unsigned input Shrink op tests for nGraph until the next update * Line-shortening code refactor * Fix for the master branch merge artifact * MKLDNN patches adjustment for Windows * Exclude MatMulInteger for non-const zero points * Exclude ConvInteger for non-const zero points * Enable full Cast op support * Use the v0.22.1 tag * Skip ConvTranspose_InvalidKernelShape test for ngraph provider * Create sub-graph ModelProto from fused_node	2019-08-10 17:41:08 -07:00
Ashwini Khade	7be40b2946	put all gemmlowp common code in one place (#1590 ) * put all gemmlowp common code in one place * fix gpu build failures * minor update	2019-08-10 17:01:07 -07:00
stevenlix	1c5b15c2b8	Remove memory copy between TensorRT and CUDA (#1561 ) * remove memory copy between CUDA and TRT * add info to RegisterExecutionProvider input * use new IDeviceAllocator for trt allocator * remove SetDefaultInputsMemoryType from TRT EP * remove onnx-tensorrt 5.0 * add submodule onnx-tensorrt branch 5.1 * remove redundancy * Update transformer_memcpy.cc * Update tensorrt_execution_provider.cc * switch to TensorRT 5.1.5.0 * update python binding * disable failed test case on TensorRT * Update activation_op_test.cc * upgrade to TensorRT container 19.06 * update according to feedback * add comments * remove tensorrt allocator and use cuda(gpu) allocator * update onnx-tensorrt submodule * change ci build cuda directory name	2019-08-08 19:31:39 -07:00
Pranav Sharma	a443b013dd	Remove unneeded C APIs + some refactoring. (#1555 ) * Mention OrtCreateSessionFromArray in C API doc * c api changes after review (1) * updates... * fixes * Reorder include	2019-08-07 11:05:29 -07:00
S. Manohar Karlapalem	05bbb3065c	[OpenVINO-EP] Update hardware branding of VAD-R as VAD-M (#1552 ) Replaces all occurrences of VAD-R/VAD_R with VAD-M/VAD_M. Aligns with the official hardware branding.	2019-08-05 15:28:46 -07:00
Ashwini Khade	b599360014	enable sse4.1 optimizations for gemmlowp (#1529 )	2019-07-31 18:44:02 -07:00
xkszltl	33ae28ccb1	Empty double quota `""` is passed to `find_package(Thread)`, causing a test command `gcc ... "" ...` failed while trying to compile a source file with empty name. (#1508 ) ``` [user@******** /]# gcc "" gcc: error: : No such file or directory gcc: fatal error: no input files compilation terminated. ```	2019-07-26 03:11:37 -07:00
xkszltl	be16b274fc	Upgrade mklml and set march with official option. (#1469 ) 1. There's formal way for setting march. 2. Upgrade to new MKLML. Besides, the mem patch can be drop for v1.0.0 since it's fixed in upstream.	2019-07-25 19:37:59 -07:00
daquexian	ec3c553501	NNAPI EP Update (#1483 ) * Update DNNLibrary * Allow fp16 by default * Add nnapi build in ci * Fix nnapi ep after #1268 * Remove unused variables * Support nnapi in onnx_test_runner * Update DNNLibrary to fix tests * Update build.py for android build support, solve conflict of tools/ci_build/build.py * Support non-ARM Android build, solve conflict of tools/ci_build/build.py * Enable android test by x86_64 android emulator * Add dnnlibrary/NNAPI support in build.py * suppress the verbose adb output * Remove debug logs * Install cmake by pip * Fix undefined host_protoc_path * cmake==3.13.2 in pypi is actually 3.12.2, so install 3.13.2.post1 instead * Fix Android ARM64 build * Use android ndk r20 instead of r19c, fix conflicts in install_deps_android.sh	2019-07-24 13:20:05 -07:00
shahasad	768ced703c	Expose provider factory C API, especially for CUDA users (#1461 ) Exposed provider factory C API, for cpu and cuda providers, into the published packages.	2019-07-22 19:03:06 -07:00
Yufeng Li	6be93f11e5	build mklml/ngraph without openmp (#1460 ) cleanup the option to build mklml/ngraph without openmp	2019-07-22 16:59:32 -07:00
Ke Zhang	638398e675	sync onnx to get equal op with float support (#1432 ) * sync onnx to get equal op with float support * doc update * fix test failure because of updated shape inference logic for roialign. * filter consum test cases since it's not implemented yet.	2019-07-19 13:19:09 -07:00
suryasidd	e9e777925f	[OpenVINO-EP] Added support for OpenVINO R1.1 (#1438 ) * Initial commit for OpenVINO R1 Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Fixed MO dynamic shape error Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Add debug messages for failure * Update install_openvino.sh script Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Try catch included. Return type of Isgraphsupported function changed to void * Removed error_msg variable and commented code * formatting cleanup * Added missing return statement Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Changed MO to be compatible with both R5 and R1 Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Updated docker scripts to include openvino version number Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Ignore compiler warnings from external headers * Updated dockerfiles Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Code cleanup using clang-format Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Suppress model optimizer info error Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Python code formatting using auto pep8 Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Updated documentation Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com>	2019-07-19 00:52:15 -07:00
Sreekanth Yalachigere	f3c74ec3e9	Reduce memory footprint of MKL-DNN EP (#1429 ) * MKL-DNN EP memory fix patch * Call default provider for Opset10 * opset 10 fix * removed email header from patch * UseSubgraph method refactored	2019-07-18 22:57:00 -07:00
Colin Versteeg	5ee0f185dc	Add GRPC support to ONNX Runtime Server (#1144 ) * add grpc * add-submodule * Revert "add-submodule" This reverts commit e35994b25035ce310a98909658582bff759ee358. * fix submodule * IT BUILDS * Initial commit of prediction_service_impl.cpp * Server builds and runs! * add request id, health and reflection. GRPC is done * enable channelz for monitoring * GRPC unit tests * clang format * add unit tests * Add function tests for GRPC * add grpc to model_zoo_tests * revert update protobuf to 3.7.0 * update submodules * builds but runs some gflags tests which fail * get build working * confine build changes to onnxruntime_server.cmake * update build files * code reveiw comments * Maik's code review comments * update cares version to fix compilation issue * update build to fix c-ares * code review comments * update cgmanifest.json * remove extraneous file * Klein comments. * update ci based on discussions for go dependency * fix tag issue * fix build issues * remove stray submodule * update dockerfile and build script * dynamic linking changes * update build script * code review comments * update dockerfile * update script for mount * code review comments	2019-07-18 11:10:38 -07:00
Yufeng Li	6c41809655	Build Shared Library with cuda 10.1 (#1418 ) Description: Describe your changes. Change the logic to find cublas dll Motivation and Context Why is this change required? What problem does it solve? The name pattern of cublas changed since 10.1. It doesn't include minor version in its name anymore. If it fixes an open issue, please link to the issue here.	2019-07-18 09:51:19 -07:00
Changming Sun	c2aa2056b5	Sample for imagenet and batch prediction (#1372 ) * Sample for imagenet and batch prediction (Will add a readme later)	2019-07-16 14:23:45 -07:00
Tracy Sharpe	719e58d831	Use MLAS to retrieve the CPU preferred tensor buffer alignment (#1377 ) Add MlasGetPreferredBufferAlignment() for use by CPUAllocator::Alloc to get the byte alignment for CPU tensors. Using MLAS allows the value to be based on the platform the binary is running on instead of a constant value fixed at compile time.	2019-07-12 22:22:46 -07:00
Ke Zhang	3bf0e364e2	Move CopyTensor out of IExecutionProvider interface. (#1268 ) * add ortdevice class * add data transfer manager for copying tensors. * update * add data trasnfer for gpu * fix constexpr build break. * update * remove unnecessary header files. * remove unnecessary header files. * add dependency * add dependency * add dependency * add dependency * fix linux build break. * update * fix build break * fix build break * fix build break * update * update * update c api. * update to not use OrtCreateAllocatorInfo * change to all eps . * fix linux build break * remove useless codes. * update * move datatransfermanager in session state * update * fix cuda build break. * fix comments * fix windows GPU build. * fix comments * fix build break * fix comments * fix test failure * update * fix comments * fix onnx runtime server. * update * fix test failure. * fix comments * fix comment	2019-07-11 14:49:20 -07:00
jignparm	e580b76305	Fix ARM64 build + Add NuGet pipeline including ARM binaries (#1335 ) * Add arm64 nocontribops pipeline * minor fix * Added new template for arm build -- disable all tests * fix build command * add arm64 flag for msbuild * add arm leg as upstream dependency * update platform to arm64 for msbuild * remove test task from arm build * remove ESRP signing of C# dlls in arm build * Updated to work for both --arm and --arm64 * Make the cross compiling cmake flags symmetric * Add dynamic check for /Wno-error flag, instead of extra build option * remove extra full-stop	2019-07-11 11:49:17 -07:00
R. G. Esteves	93528d9b3c	Reduce memory footprint of nGraph (#1296 ) * Fix unnecessary memory allocation in MKLDNN 1x1 convolution. * remove the patch header.	2019-07-07 20:23:19 -07:00
Colin Versteeg	a8ff209ab6	Refactor Onnx runtime Server to only use public APIs (#1271 ) * replace log sinks * limit headers to include dir * first changes to do dynamic linking * wip for using cxx api * remove weird dangling dependency * building with tests failing * finish updating converters * fix const * intital introduction of typedef * change logging to use spdlog * get tests passing * clang format * map logging levels better * clean up unused imports * trent cr comments * clang-format * code review comments * changing buffer use to reserve * Dynamically link * revert tvm * update binary uploading * catch exceptions by const-ref * Revert "revert tvm" This reverts commit 387676dd1018134d15eb71fa126f7caf94380800. * fix typo * update versioning of lib	2019-07-04 01:08:14 -07:00
KeDengMS	0d204f3f06	Implementation of TVM codegen library (#888 ) Description: This change adds the common part of TVM based codegen library. It includes following parts: * Microsoft TVM Inventory (MTI): a set of TVM ops for neural networks, similar to TOPI * Compiler pass for traversing ONNX graph and generate TVM ops * Compiler pass for traversing generated graph and specify TVM schedule * Compiler pass for handling weight layout * Utils for debugging Motivation and Context: TVM is an open deep learning compiler stack for cpu, gpu and specialized accelerators. To leverage it in ONNX, we built an execution provider named Nuphar. Currently, Nuphar gets good performance on CPUs with AVX2 on quantized LSTM models. This codegen library was part of Nuphar execution provider. It is split out for sharing with other execution providers, as we'd like to reuse TVM in more devices.	2019-07-03 10:32:59 -07:00
daquexian	c65489a47f	Initial PR for NNAPI execution provider (#1220 ) * init * Update DNNLibrary * Update DNNLibrary, set compiler flags, it compiles now * Add more missing flags, add test * Update DNNLibrary * Update Compile method, fix allocator and some other bugs * Update DNNLibrary * Implement CopyTensor * Not delete state explicitly since it is managed by unique_ptr * Add the missing files when SingleUnitTestProjct is ON * misc changes * Fix wrong name in provider factory * Add my own test * Update the code of add node into graph, and add the missing initializer into graph * Fix the bug that re-build the graph produces extra output * Update DNNLibrary * Transpose nchw (ONNX) -> nhwc (NNAPI) * Add license * Add GetSupportedNodes method (implement it later) * Rename onnxruntime_nnapi_test->onnxruntime_nnapi_squeezenet_test * Update squeezenet_test.cpp after rebase master * Remove squeezenet_test.cpp since it is almost same with the c++ sample * Update DNNLibrary for GetSupportedNodes * Update GetSupportedNodes * Revert "Remove squeezenet_test.cpp since it is almost same with the c++ sample" This reverts commit a97575fd9ff49e50ba1dc8d8154790d8cd86c48d. * Update DNNLibrary * Fix multiple outputs bug * Remove GetKernelRegistry * Revert "Revert "Remove squeezenet_test.cpp since it is almost same with the c++ sample"" This reverts commit 2a0670e9cbf10ea654111ce39e198a4be0ddd838. * Set default memory type of NNAPI EP * Add CPUOutput allocator * Update DNNLibrary for multiple outputs * Fix bug of nhwc->nchw * Remove GetExecutionHandle()	2019-07-02 06:03:29 -07:00
Matthieu Darbois	04d581995d	Use manylinux2010 image to build linux python wheels (#1282 ) * Update cuda for python wheels * Update cuda for python wheels * Update cuda for python wheels * Update azure-pipelines-py-packaging.yml * Update to cuda 10 * Only test win gpu * Update cuda for python wheels * Use manylinux2010 image to build linux python wheels Allow wheels built to truly be compliant with a manylinux policy	2019-06-27 15:45:06 -07:00
Scott McKay	0951f53c80	Update ONNX to d94f99d21a9a0820d58966410ceaf525132f85f1 to pickup change to checker that makes ssd_mobilenet model load 20x faster by avoiding unnecessary copies. (#1307 )	2019-06-27 08:39:41 -07:00
Tracy Sharpe	3ebad81abc	MLAS: NCHWc low-level changes (#1283 ) Implementation of the MLAS changes for NCHWc convolution/pooling support. These changes adopt the blocking format used by MKL-DNN and other convolution libraries for better performance.	2019-06-25 16:57:30 -07:00
Ashwini Khade	a571ea74a6	update onnx (#1287 )	2019-06-24 14:17:27 -07:00
Ashwini Khade	92dc5c506d	move all contrib ops to contrib ops namespace (#1190 ) * move all contrib ops to one place * namespace changes * bug fix - remove redundant file after merge master * plus more minor bug fixes * bug fix * fix extra space in include header + namespace fix * fix linux build failure: * fix test group names * remove redundant test	2019-06-24 10:19:01 -07:00
RandySheriffH	671c15a56a	Treat attribute warning as non-error on cross compiling ARM (#1261 ) * abandon attribute error on cross compiling * install dep lib	2019-06-23 17:59:38 -07:00

1 2 3 4 5

207 commits