onnxruntime

mirror of https://github.com/saymrwulf/onnxruntime.git synced 2026-05-19 21:32:23 +00:00

Author	SHA1	Message	Date
David Brownell	72cd61baae	Removed use of parameters in python wheel build scripts (#3524 )	2020-04-15 10:31:14 -07:00
Changming Sun	a2feb29b0d	Fix build break (#3528 ) Ignore some known test failures Install ONNX package before running Windows CI builds	2020-04-14 18:07:56 -07:00
David Brownell	006c5be1b1	Optionally produce a python wheel that includes featurizers (#3491 )	2020-04-14 09:00:13 -07:00
M. Zeeshan Siddiqui	5d99f179b9	Merge pull request #3486 from microsoft/sedymche/merge_master_ort_training Merge from master into ort_training	2020-04-13 10:55:36 -07:00
Sergii Dymchenko	bf3df41424	Put back SubmoduleCheckoutMode parameter into mac-ci.yml.	2020-04-12 21:49:38 -07:00
George Wu	7f6e407e09	fix python packaging manylinux1 build break. (#3482 )	2020-04-11 06:58:22 +08:00
edgchen1	cffdff6702	Publish unit test results from Linux and Mac builds (#3480 ) * Added publish test results step to Linux and Mac builds. * Fix test result file pattern.	2020-04-10 14:51:56 -07:00
liqunfu	e7297e6c9d	create pipeline for ci frontend tests (#3422 ) create pipeline for nightly python front-end e2e tests	2020-04-09 15:31:22 -07:00
Sergii Dymchenko	6ba7c99e50	Merge branch 'master' into ort_training	2020-04-09 12:42:04 -07:00
Changming Sun	33006f48c0	Update onnx submodule to 1.7.0 release candidate (#3405 ) Update onnx submodule to 1.7.0 release candidate. This isn't a release tag, but it will be released soon, in 1-2 weeks.	2020-04-04 16:23:42 -07:00
Changming Sun	a5fea26cb4	Disable model tests for Mac OS X builds	2020-04-02 15:14:32 -07:00
Thiago Crepaldi	759818f2c1	Merge remote-tracking branch 'origin/master' into thiagofc/ort_training_merge_from_master	2020-03-31 10:53:22 -07:00
stevenlix	2332a93db0	Update onnx-tensorrt parser (#3369 ) * sync onnx-tensorrt parser and update TensorRT doc * remove --msvc_toolset 14.16 in tensorrt ci pipeline	2020-03-30 20:31:59 -07:00
Xueyun Zhu	ccc3535e72	resolve conflict	2020-03-20 20:20:35 +00:00
Tiago Koji Castro Shibata	3bdb0b620a	Fix WCOS/Win32 linking bugs (#3126 ) * Fix WCOS/Win32 linking bugs * Remove unused NODEFAULTLIB flags * Avoid plain target_link_libraries signature * Avoid plain target_link_libraries signature * Fix library list escaping * Use library list instead of string * Remove duplicate link to windowsapp.lib * Remove Win32 build workarounds * Specify CMake policies before initializing language * Expose Win32 header definitions during build * Force set API family * Enable Win32 APIs in featurizer * Use MT dynamic CRT * Expose Win32 specific functions * Disable app container globally * Disable default wide functions in featurizers * Add featurizers to test include path * Workaround https://gitlab.kitware.com/cmake/cmake/issues/19428 * Revert pipeline debugging hacks * Skip /FI in CUDA sources * Default to Win32 builds * Enable WCOS when using WinML * Use generator expression to apply CMAKE_MSVC_RUNTIME_LIBRARY to C++ only	2020-03-19 08:52:40 -07:00
Changming Sun	0fceb33288	Fix onnxruntime server docker file build failure (#3219 ) 1. Fix onnxruntime server docker file build failure. Tested with the notebook in ONNX tutorial, it works well. 2. Delete the docker files for the other EPs, because currently they don't work and I don't have enough time to update them.	2020-03-15 14:46:46 -07:00
Tracy Sharpe	fe0b2b2abd	QLinearConv speed up (#3196 ) For x86/x64 builds, change the QLinearConv op to use MLAS for the u8u8=s32 GEMM, then requantize the intermediate buffer to u8.	2020-03-13 16:54:55 -07:00
Zeeshan Siddiqui	2cad08bd60	Merged PR 5688: Upgrade ONNX submodule to the latest from github ONNX master. We want to implement SoftmaxCrossentropy and NegativeLossLikelihoodLoss forward training ops for opset-12 but that requires ONNX submodule to point to the latest commit to have the latest and greatest ONNX spec! - Reverse integrate changes from *.in.proto files in github ONNX repo. - Regenerate csharp/test/Microsoft.ML.OnnxRuntime.Tests/OnnxMl.cs - Disable ONNX tests that don't have op implementation for the latest opset.	2020-03-12 16:51:45 -07:00
edgchen1	fa4dd51e3b	Add back orttraining-linux-gpu-inference-only-ci-pipeline.yml. (#3182 )	2020-03-11 18:03:58 -07:00
Edward Chen	80dd62a240	Enable CI for training.	2020-03-11 14:41:32 -07:00
Edward Chen	e542cfd0e0	Introduce training changes.	2020-03-11 14:39:03 -07:00
Hariharan Seshadri	3464801c3e	Explicitly specify NugetPackage parameter while validating nuget in some release pipelines (#3139 )	2020-03-10 15:14:09 -07:00
Dmitri Smirnov	f87b6913cd	Add package download step before pushing to feeds (#3162 ) Add package download step before publishing.	2020-03-09 14:32:18 -07:00
Changming Sun	6ed5d7c332	Update post_binary_sizes_to_dashboard.py (#3161 ) Discussed with Faith, because the data size is very small and changes are gradual, there is no need to delete the old data. We want to keep all the history.	2020-03-09 13:21:58 -07:00
Tiago Koji Castro Shibata	a59243090a	Publish release symbols (#3152 ) * Publish release symbols * Publish symbols if IsReleaseBuild	2020-03-05 22:32:18 -08:00
Dmitri Smirnov	e2894c5ffb	Fix package name overrides (#3150 ) Add env var with the package name.	2020-03-05 17:10:55 -08:00
Dmitri Smirnov	2c446a7f2f	Add push to ORT-NIGHTLY. (#3146 )	2020-03-05 11:38:22 -08:00
Dmitri Smirnov	ef8768a53f	Override native package name. Preserve managed package name the same. (#3133 ) Override native package name. Preserve managed package name the same. Specify pckage name for validation purposes. Fix up validation package name parameter.	2020-03-04 10:12:55 -08:00
Changming Sun	12605f05d1	Fix CUDA PATH (#3131 ) Previously, we put the "bin" folder of all the CUDA verions in the system PATH. And 10.2 is in the front. It's a mess. So I've removed all of them from the system PATH env. But I need to add one of them back through build scripts. (The problem only affect the C# test, not the C/C++ tests that forked from build.py).	2020-03-03 14:34:19 -08:00
smk2007	6cdd2b4934	Enable DML Nuget Package for x64 or x86 architectures (#3120 ) * add dml gpu pipelines * add x86 to the gpu dml dev build pipeline * Enable DML x86 builds * Fix uint64_t -> size_t warning * fix warnings * enable dml on x86 ci builds * operatorHelper 773 error uint32_t vs uint64_t * operatorHelper 773 error uint32_t vs uint64_t * make x86 pipeline use the gpu pool * more warnings * fix x86 directml path * make dml nuget package * disable tf_pnasnet_large * disable zfnet512 * make validation use wildcards * disable x86 dml gpu tests * add args. * update gpu.yml * change nupkg wildcard * add debug statements * package x86 dml nupkg * dont drop managed nuget again from dml pipeline build * Add DML EULA * directml license should be renamed to not clobber the existing license * casing on dml package.... * {} to () * fix license name * disable dml from x86 ci * typo and cr feedback * remove featurizers * ship the dml pdb as well	2020-03-02 20:18:46 -08:00
Dmitri Smirnov	e45326b5df	Create NuGet packaging pipeline for ORT Featurizers (#3125 ) Create a new pipeline to publish ORT with Featurizers Update pipeline for two separate packages. Change package names.	2020-03-02 17:00:56 -08:00
Hariharan Seshadri	86b755774f	Create a separate Nuget hosting just managed assemblies (#3020 ) * Initial commit * More changes * More changes * More changes 3 * More changes 4 * More changes 5 * More changes 5 * More changes 6 * More changes 7 * More changes 8 * Remove C# ifdefs * More changes 10 * More changes 11 * YAML changes for other release pipelines * Add release notes metadata * Props and Targets change * Add CSHarp proj * More changes 12 * More changes * Minor fix * Minor fix * Fix yaml * Some missing logic for winml * Minor update * Fix casing for winmd file * Fix casing * Add targets and props for managed section into native nuget * revert file * a	2020-02-27 18:00:17 -08:00
daquexian	37a905f557	Make Java API available on Android (#3030 )	2020-02-27 08:23:50 -08:00
Changming Sun	d7500b26bd	Remove Publish Build Symbols from pre-checkin CI build (#3088 )	2020-02-25 08:02:36 -08:00
stevenlix	f4a5d17294	Upgrade to CUDA10.2 for TensorRT (#3084 ) * Switch to CUDA10.2 * Update win-gpu-tensorrt-ci-pipeline.yml * Update win-gpu-tensorrt-ci-pipeline.yml * remove dynamic_shape * update onnx-tensorrt submodule * check if input shape is specified for TensorRT subgraph input and enable some TensorRT unit tests * fix format issue * add shape inference instruction for TensorRT * update according to the reviews * Update win-gpu-tensorrt-ci-pipeline.yml	2020-02-25 05:36:01 -08:00
Hariharan Seshadri	d7f2cdcc7e	Fix target platform of managed OnnxRuntime dll and enable x86 .NET testing (#3056 ) * WIP: Re-enable x86 .NET testing in Release pipelines Enabling x86 testing will make sure that ORT packages doesn’t break x86 projects of customers * Remove setting some env variables * Comment out a test failing on x86 builds * More changes * Minor fix * More changes * More changes * s * s * s * Revert minor change * More changes * More changes * More changes 2 * explicitly set platform target * Delete bin and obj folders * Clean output dirs * Add back TargetFramwork * Disable x86 .net framework tests * Skip x86 tests in MKLML pipeline	2020-02-24 23:02:59 -08:00
Dmitri Smirnov	dae9a31719	Introduce new Featurizers packaging pipeline. (#3068 ) Introduce new Featruizers packaging pipeline.	2020-02-24 13:57:38 -08:00
Changming Sun	61ae134469	Fix binary size report (#3080 )	2020-02-22 21:01:06 -08:00
Changming Sun	fb871978b5	Adjust build flags for the release pipelines (#3066 ) 1. Add LTCG back. It was set to default OFF in my previous PR to speed up Windows build. It is only needed in release pipelines. 2. Remove --use_featurizers from all the packaging pipelines 3. Make sure all the packages have openmp	2020-02-21 16:45:42 -08:00
Changming Sun	179603775f	Use CUDA 10.1 for Linux build (#3057 ) Use CUDA 10.1 for Linux build (Windows change is already in) Please note, cublas 10.2.1.243 is for CUDA SDK 10.1.243, not CUDA 10.2.x. CUDA 10.2.89 need cublas 10.2.2.89. They match on the last part of the digits. libcublas10-10.1.0.105 won't work!!! The cuda docker image by viswamy is already using 10.1, no need to change.	2020-02-21 11:55:32 -08:00
Ori Levari	be12fb3143	include winml x86 binaries in the drop-signed-nuget artifact (#3058 )	2020-02-21 11:17:23 -08:00
Changming Sun	ef2bba316b	CUDA 10.1 for Windows(#3049 )	2020-02-19 23:26:47 -08:00
Dmitri Smirnov	daf8c4bee4	Remove faturizers from CPU MLDNN and NoContribOps builds. (#3039 ) The first one is temp. The second one is permanent removal.	2020-02-19 06:23:36 -08:00
James Yuzawa	411b3aa801	Java build system enhancements (#2866 )	2020-02-18 15:41:49 -08:00
stevenlix	da653ccdac	Upgrade TensorRT to version 7.0.0.11 (#2973 ) * update onnx-tensorrt submodule to trt7 branch * add fp16 option for TRT7 * switch to master branch of onnx tensorrt * update submodule * update to TensorRT7.0.0.11 * update to onnx-tensorrt for TensorRT7.0 * switch to private branch due to issues in master branch * remove trt_onnxify * disable warnings c4804 for TensorRT parser * disable warnings c4702 for TensorRT parser * add back sanity check of shape tensort input in the parser * disable some warnings for TensorRT7 * change fp16 threshold for TensorRT * update onn-tensorrt parser * fix cycle issue in faster-rcnn and add cycle detection in GetCapability * Update TensorRT container to v20.01 * Update TensorRT image name * Update linux-multi-gpu-tensorrt-ci-pipeline.yml * Update linux-gpu-tensorrt-ci-pipeline.yml * disable rnn tests for TensorRT * disable rnn tests for TensorRT * disabled some unit test for TensorRT * update onnx-tensorrt submodule * update build scripts for TensorRT * formating the code * Update TensorRT-ExecutionProvider.md * Update BUILD.md * Update tensorrt_execution_provider.h * Update tensorrt_execution_provider.cc * Update win-gpu-tensorrt-ci-pipeline.yml * use GetEnvironmentVar function to get env virables and switch to Win-GPU-2019 agent pool for win CI build * change tensorrt path * change tensorrt path * fix win ci build issue * update code based on the reviews * fix build issue * roll back to cuda10.0 * add RemoveCycleTest for TensorRT * fix windows ci build issues * fix ci build issues * fix file permission * fix out of range issue for max_workspace_size_env	2020-02-12 07:03:58 -08:00
Dmitri Smirnov	273868eaa5	Disable NuGetPackaging on Linux GPU and remove DML from the pipelines (#3006 )	2020-02-11 20:08:18 -08:00
Dmitri Smirnov	c1997db85e	Exclude faturizers from Linux NuGet packaging.	2020-02-10 22:21:52 -08:00
Dmitri Smirnov	36915b3674	Temporarily remove Featirizers from packaging-pipelines	2020-02-10 22:21:52 -08:00
smk2007	ce713823cc	enable winml in the gpu ci pipeline (#2993 )	2020-02-10 22:21:13 -08:00
smk2007	5c5ac34b5c	Disable use_dml in nuget pipeline (#3001 )	2020-02-10 22:09:58 -08:00
Tiago Koji Castro Shibata	fb2182f3fc	Release ARM/ARM64 Nuget packages (#2987 ) * Enable ARM64 release builds * Add ARM release * Skip C# dll signing in ARM * Copy ARM binaries to Nuget * Restore nuget packages before ARM packaging * wip * Use host protoc at C# build * Set ProtocDirectory on cross-compiled builds * wip * Fix typo	2020-02-10 16:29:27 -08:00
Dmitri Smirnov	c8ea154e55	Package data_frame_tool, include featurizers into Manilinux2010 (#2989 ) * Package data_frame_tool, exclude featurizers from Manilinux2010 as their fail to build.	2020-02-07 11:38:42 -08:00
Dmitri Smirnov	0e5582bcc3	Publish nightly NuGet packages to a new Vsts feed (#2976 ) * Add a push to the internal ORT-NIGHTLY * Add a descriptive name * Specify packageToPush attribute	2020-02-06 10:08:21 -08:00
smk2007	c32cedc6c9	Merge windowsai (winml layering) into master (#2956 ) * Initial Commit * Merged PR 3985217: add onecoreuap_apiset.lib in order to avoid linking against kernel32.lib etc (#2346) add onecoreuap_apiset.lib in order to avoid linking against kernel32.lib etc and violating our OS layering requirements. We linked against onecoreuap_apiset.lib in VB so we will continue doing this, but I am still unsure why not to link against onecore instead since that is where we ship. However, since Sheil is the owner of this code we will wait to discuss with him before changing anything. * Initial changes for layering * more snipping to get core into ort * update build instructions to include --build_shared_lib (#2358) * update build instructions to include --build_shared_lib * fix line breaks * Task 23998197: add winml_lib_core into onnnxruntime.dll (#2368) * Task 23998197: add winml_lib_core into onnnxruntime.dll * PR feedback build break on perf_test * return proper error when the model path isn't found (#2391) * LearningModelSession is cleaned up to use the adapter, and parts of b… (#2382) this is a big PR. we are going to move it up to layer_dev , which is still a L3 so we are still safe to do work there agile. we are going to move this into the L3 so that ryan can start doing intergration testing. we will pause for a full code review and integration test result prior to going into the L2. >>>> raw comments from previous commits >>> * LearningModelSession is cleaned up to use the adapter, and parts of binding are. * moved everything in the winmladapter made it all nano-com using, WRL to construct objects in the ORT side. base interfaces for everythign for winml to call cleaned up a bunch of winml to use the base interfaces. * more pieces * GetData across the abi. * renamed some namepsace cleaned up OrtValue cleaned up Tensor cleaned up custom ops. everything but learnignmodel should be clean * make sure it's building. winml.dll is still a monolith. * model moved over. everything builds clean. step ! * weak ref comment * Layer dev paulm (#2408) * model moved over. everything builds clean. step ! * weak ref comment * added a wrapper for RoGetActivationFactory to hook back into winml for creating winml objects. fixes model load. * Layer dev paulm (#2414) * model moved over. everything builds clean. step ! * weak ref comment * added a wrapper for RoGetActivationFactory to hook back into winml for creating winml objects. fixes model load. * User/xianz/win ml telemetry (#2410) * add option to enable winml telemetry * add option to enable winml telemetry * clean logs while developping * clean the log of GUID * compile onnxruntime_common with winml telemetry * use option for use_telemetry * rename option winml_use_telemetry to onnxruntime_use_telemetry * little change * fixed some lifetime management. fixed the debug build. squeezenet passes using winmlrunner for CPU and GPU * Layer dev paulm (#2423) * model moved over. everything builds clean. step ! * weak ref comment * added a wrapper for RoGetActivationFactory to hook back into winml for creating winml objects. fixes model load. * fixed some lifetime management. fixed the debug build. squeezenet passes using winmlrunner for CPU and GPU * PR feedback. * Layer dev paulm (#2424) * model moved over. everything builds clean. step ! * weak ref comment * added a wrapper for RoGetActivationFactory to hook back into winml for creating winml objects. fixes model load. * fixed some lifetime management. fixed the debug build. squeezenet passes using winmlrunner for CPU and GPU * PR feedback. * couple of fixes and coded getmutabledata() * Layer dev paulm (#2425) * model moved over. everything builds clean. step ! * weak ref comment * added a wrapper for RoGetActivationFactory to hook back into winml for creating winml objects. fixes model load. * fixed some lifetime management. fixed the debug build. squeezenet passes using winmlrunner for CPU and GPU * PR feedback. * couple of fixes and coded getmutabledata() * fixed 2 more heap corruptions * Layer dev paulm (#2426) * model moved over. everything builds clean. step ! * weak ref comment * added a wrapper for RoGetActivationFactory to hook back into winml for creating winml objects. fixes model load. * fixed some lifetime management. fixed the debug build. squeezenet passes using winmlrunner for CPU and GPU * PR feedback. * couple of fixes and coded getmutabledata() * fixed 2 more heap corruptions * Add opset and IR check when loading model (#2413) * Add opset and IR check. * Add test case for future opsets. https://github.com/microsoft/onnxruntime/issues/2371 * fixed map and sequence when passing stl types across the ABI . found a leak in nvidia driver, but skipped it. all winmlapitests pass now * Moved SessionOptions over to the abi * WinML CI (#2412) * Pass flags to build/test WinML in CI * Add initial CMake config for unit tests in WinML * Set winml_unittests standard to C++17 * Add WinML API tests and port them to googletest * Install WinML test collateral * Add LearningModelSessionAPITests ported to googletest * Fix WinML test files encoding * Add GPU tests * Add parameterized test, skip GPU tests * Enable precompiled header * Remove unused code and collateral * Remove brand images * Add dllload.cpp * Remove images not used in API tests * Add LICENSE.md to image collaterals * Add models with licenses * Remove FNS Candy tests * Add API test models * Add ModelInSubdirectory * Install collaterals post-build with copy_if_different, split common lib * fix warnings * Link to gtest_main * Register WinML TraceLogging provider on Onnxruntime.dll (#2455) * Register WinML TraceLogging provider on Onnxruntime.dll * Add ifdef to make sure trace logging provider has telemetry option when LAYERING_DONE * No need for ifdef for TraceLoggingOptionMicrosoftTelemetry * PR feedback * Move etw registration into lotus environment constructor and deresgister in lotus environment destructor * Brianma/cpuwinml (#2466) * allow building winml cpu without dml. * Brianma/breaks (#2469) * fix some more breaks * learning model doesn't need lotusEnvironment and CPU shouldn't include dmlEP headers * move dml checks out of winml and into the adapter * better error handling * Brianma/fi (#2470) * learning model doesn't need lotusEnvironment and CPU shouldn't include dmlEP headers * User/xianz/win ml telemetry (#2410) * add option to enable winml telemetry * add option to enable winml telemetry * clean logs while developping * clean the log of GUID * compile onnxruntime_common with winml telemetry * use option for use_telemetry * rename option winml_use_telemetry to onnxruntime_use_telemetry * little change * Add opset and IR check when loading model (#2413) * Add opset and IR check. * Add test case for future opsets. https://github.com/microsoft/onnxruntime/issues/2371 * WinML CI (#2412) * Pass flags to build/test WinML in CI * Add initial CMake config for unit tests in WinML * Set winml_unittests standard to C++17 * Add WinML API tests and port them to googletest * Install WinML test collateral * Add LearningModelSessionAPITests ported to googletest * Fix WinML test files encoding * Add GPU tests * Add parameterized test, skip GPU tests * Enable precompiled header * Remove unused code and collateral * Remove brand images * Add dllload.cpp * Remove images not used in API tests * Add LICENSE.md to image collaterals * Add models with licenses * Remove FNS Candy tests * Add API test models * Add ModelInSubdirectory * Install collaterals post-build with copy_if_different, split common lib * fix warnings * Link to gtest_main * fix bad merge * Checking in a staging checkpoint point so that Ryan can work with me in parrallel * build break. * Brianma/testfails (#2473) * add missing ir version to dictvectorizer-string.onnx * add missing ir version to relu.onnx * add missing ir version to zipmaponnx add IR version to manually generated models * remove an unnecessary ifdef dml * Brianma/windowsai fi (#2475) * update dockerfiles/README (#2336) * Make elementwise op run 4 items per thread (#2335) Description: Describe your changes. Make elementwise op run 4 items per thread unroll for loop to leverage ILP remove unnessary N==0 check inside elementwise GPU kernel Motivation and Context Why is this change required? What problem does it solve? It can improve the performance of GPU elementwise ops. ~2% performance gain on popular NLP bert model. If it fixes an open issue, please link to the issue here. * Add CUDA GatherElements kernel (#2310) * Updates * Update test * Update * Updates * nits * PR feedback * Update * Update * PR feedback * PR comments * Update * Fix build * Fix build * Nits * Fix * Layer Normalization Fusion (#2319) basic layer normalization transform * Add FastGelu Cuda Op for Gelu and Add bias fusion (#2293) * Add FastGelu cuda op * Add AddBiasGelu for experiment * Revert "Add AddBiasGelu for experiment" This reverts commit 5c1ee019858c657e6bb75887265cb85675626e5b. * Add bias * Add unit tests * update comment * update script * fix build error * update coding style * update for CR feedback Enable half2 optimization only when cuda arch >= 7.0 * move _Tanh to common.cuh * implement CPU contrib OP Attention (#2333) * Remove unused initializer from GraphProto as well as name_to_initial_tensor_ in CleanUnusedInitializers. (#2320) * Remove unused initializer from GraphProto as well as name_to_initial_tensor_ in CleanupUnusedInitializers. This means initializers that have been replaced during graph optimizations are not left in the GraphProto when we save an optimized model. * Handle edge case where a model has an unused initializer with matching graph input by also removing the graph input. * Use non-const iterators in std::find_if calls to make centos build happy. * Nuget pipeline changes (#2305) 1. refactor the pipeline, remove some duplicated code 2. Move Windows_py_GPU_Wheels job to Win-GPU-CUDA10. We'll deprecated the "Win-GPU" pool 3. Delete cpu-nocontribops-esrp-pipeline.yml and cpu-nocontribops-pipeline.yml 4. In Linux nuget jobs, run "make install" before creating the package. So that extra RPAH info will be removed * Cuda Reverse Sequence Op, maping types of same size using same template function. (#2281) * Set ElementType to String type of node metadata, instead of byte[] (#2348) * Set ElementType to String type of node metadata, instead of byte[] * Fix spacing * Introduce PrimitiveType into a Type System along with an integer constant (#2307) Improve perf by avoiding GetType<T>() calls. Introduce MLTypeCallDispatcher to switch on Input Type. Add Tensor IsType<T>() fast method. * Fix/test dim value of 0 handling in a couple of places (#2337) * Update the CUDA Where implementation broadcasting logic to handle a dim with value of 0. Add unit test Also add unit test for unary op with dim value of 0 * Exclude ngraph from Where test with 0 dim. * Openvino EP R3.1 onnxrt server (#2357) * onnxrt server with OVEP * onnxrt server with OVEP * Update Dockerfile.server.openvino * onnxrt server OVEP fix reviews * onnxrt server OVEP fix reviews * Implement cuda nonzero op. (#2056) Implement cuda nonzero op. * Direct use python numpy array's memory if already contiguous. (#2355) * Direct use python numpy array's memory if already contiguous. This could greatly improve performance for session with large input, like big image 1920x1080 fastrcnn, 30~40% speed up could be achieved. * Add test case enforce contiguous/non-contiguos numpy array as inputs. * Add helper to create output to minimize binary size. (#2365) Add ConstEigenTensorMap typedef so we don't unnecessarily const_cast the const input Tensor. * fix builds enabling onnxruntime_DEBUG_NODE_INPUTS_OUTPUTS (#2369) * fix builds enabling onnxruntime_DEBUG_NODE_INPUTS_OUTPUTS * update * Add Tracelogging for profiling (#1639) Enabled only if onnxruntime_ENABLE_INSTRUMENT is ON * test bidaf with nuphar for avx target (#2370) increase nuphar test coverage a bit * Fix a bug in TLS refcount that may destabilized CUDA CI (#2374) * update output size calculation for resize (#2366) * change how output size is calculated for resize op * add tests for ver 10 resize * Extend OneHot CPU kernel to support more types (#2311) * Extend OneHot CPU kernel to support input int64_t, depth int32_t, output float * Skip BERT before the test data fix is picked up * Fix bug with Slice. Need to pass in flattened input dimensions so the initial offset into the input is calculated correctly. (#2372) * Add opset 11 version of Split to CUDA ops (#2376) Organize the CUDA ops definitions so all the opset 10 and 11 parts are together (same setup used for CPU ops) * Layer Norm Fusion Fix (#2379) * layer norm fusion fix * Add input shape check in code and unit tests * Fuse Add + Gelu (#2360) Implement the transformer to fuse add + gelu Implement the accurate kernel * Skip layer norm transform (#2350) * skip layer normalization transformer * Another try to stabilize CUDA CI (#2383) The root cause seems to be failure in CUDA dealloc when tear down. cudaFree return code was ignored before, so should the debug check. * fix BUILD.md typo (#2375) build.py: error: argument --config: invalid choice: 'RelWithDebugInfo' (choose from 'Debug', 'MinSizeRel', 'Release', 'RelWithDebInfo') * Fixed compilation with ngraph (#2388) * Fix reuse logic in allocation planner. (#2393) * Fix reuse logic in allocation planner. * PR comments * Add helpful comments * Don't allow reuse across string tensors. * [NupharEP] Multiple optimizations (#2380) Fuse transpose into MatMul Implement Pow and constant scalar simplification Vectorize ReduceMean Improve symbolic shape inference Minor updates for better debugging in fused function name * Avoid using the default logger in the graph lib and optimizers (#2361) 1. Use the session logger if it is available. 2. Don't disable warning 4100 globally. We should fix the warnings instead of disabling it. * Change CUDA implementation of Transpose to support all fixed size tensor types (#2387) * Change CUDA implementation of Transpose to not use a typed kernel so we can support more types with minimum binary size. Add support for 8, 16, 32 and 64 bit types. Add unit tests. Add method so the implementation can be called directly (will be used by CUDA Scan very soon). * Disable TensorRT for MLFloat16 and int8 unit tests. * Address PR comment and add support for calling cublas implementation if type is mlfloat16. * Add opset 11 versions of the existing CUDA operators that had negative axis support explicitly added. (#2398) * Add opset 11 versions of the existing CUDA operators that had negative axis support explicitly added. * [NupharEP] force some low/zero cost ops to be inlined (#2409) * fix cross compile bug (#2415) * Minor optimization: if a node has already been placed, there's no need to find a kernel for it. (#2417) * Add Reshape Fusion (#2395) * Add reshape fusion * Add some comments * update comments * update comment format * update according to feedback * update for recent logger change * fix build error * (1) Support both input and output edges in find path in graphutils (2) Add a test case of only one constant initializer of Concat input. (3) Refactor ReshapeFusion class to allow add more subgraph fusion in the future. * fix error * (1) loose constraint on initializer: non constant is allowed for reshape fusion. (2) Change versions type to vector. (3) Add logging. (4) Return false when multiple output edges matched in FindPath. Add comments. * only allow one direction (input or output) in FindPath * [NupharEP] Update notebook and docker image (#2416) Add BERT squad in Nuphar tutorial Enhance speed comparsion readability * Fix the issue in matmul_add_fusion (#2407) Fix the issue in matmul_add_fusion If Muatmul + Add has shape [K] * [K, N], reset it to [1, K] * [K, N] will make the output shape to [1, N] will also requires a reshape on the output. Fix: just remove the shape reset to not fuse it. Add a negative test case for matmul+add fusion * feat(treeregressor): Update TreeEnsembleRegressor for type support (#2389) Updates the `TreeEnsembleRegressor` to allow for `double`, `float`, `int64`, and `int32` inputs to match the upstream specification. Signed-off-by: Nick Groszewski <nicholas.groszewski@capitalone.com> * onnxrt server documentation update (#2396) * Added support for Pad-2 operator in OpenVINO-EP (#2405) * Add CUDA If operator. (#2377) * Add CUDA If operator. Uses CPU operator for implementation. By adding a CUDA version the inputs/outputs (with the exception of the 'cond' input) stay on GPU, and no other logic is required to avoid a copy to CPU across the control flow node. * Improved documentation for onnxruntime::utils::SwapByteOrderCopy(), added precondition check. * Fix the type constraints on CUDA If operator to exclude strings. (#2431) * add Im2col<uint8_t> (#2438) * Adjust codegen vectorization width from target (#2439) * Adjust codegen vectorization width from target * Add CUDA Scan operator. (#2403) * Add Scan CUDA op. Uses CPU implementation for logic. Added some device specific functors for handling when data needs to be manipulated on a different device. Added ability to override the materialization logic in the OrtValue slicer so DML can plugin their handling. * Fix Windows GPU C API packaging pipeline failure (#2440) Fix Windows GPU C API packaging pipeline failure (#2440) * Correctly handle implicit inputs for fused nodes (#2390) * Correctly handle implicit inputs for fused nodes Previously, nuphar's partitioning function didn't include node's implicit inputs into the inputs list of MetaDef, and hence a crash was triggered in the onnx graph checker. This commit fixed the issue. Furthermore, it also fixed a related issue where we didn't add implicit inputs into graph_inputs_excluding_initializers_ in Graph::SetGraphInputsOutputs. the issue was that graph_inputs_including_initializers_ populated by SetInputs (e.g. called by FunctionImpl::FunctionImpl) may contain implicit inputs which were not of any node's initializers in the graph. Because they were not part of any initializers, these implicit inputs couldn't be visited by going through all nodes' inputs. Consequently, they would not be added into graph_inputs_excluding_initializers_. We fixed the issue by first copying the populated graph_inputs_including_initializers_ into graph_inputs_excluding_initalizers_, which then had both initializers and non-initializers as its initial content. Later, we erase initializers from the list. In this way, we can ensure all implicit inputs to remain in graph_inputs_excluding_initializers_. * refined comments and fixed duplicates Address CR by revisiting comments in terms of implicit inputs Also fixed an issue by skipping duplicates while copying inputs from graph_inputs_including_initializers_. * address CR explain why we need to collect nodes' implicit inputs * don't rely on pointer values for iterating std::set Previously, openvino relied on iterating a set of NodeArg pointers to construct inputs and outputs for a fused graph. It could cause non-determinism. The reason was that although iterating std::set by itself is stable, pointer values of NodeArgs may vary. Consequently, we could end up visiting the set's elements in different orders for different runs for the same test, which resulted in constructing inputs (and outputs) with different orders to the fused graph. For example, for the same test, we may have inputs [A, B] in some runs but inputs[B, A] in others. Let's use std::string as the key type to avoid such nondeterminism. This commit also added implicit inputs into meta->inputs while returning the capability from the openvino provider. * Fixed another latent issue in openvino's GetCapability function The issue was that we couldn't simply erase fused_inputs and fused_outputs while iterating the nodes. For example, an output NodeArg may have multiple uses, and it's wrong if we erase it from fused_outputs when we encounter only one of its uses as input. * Remove DeviceAllocatorRegistry class (#2451) Remove DeviceAllocatorRegistry class * CSharp api and test for loading custom op shared library (#2420) - Added C-API test for loading custom op shared lib. - Made some changes in C++ api header and C-api implementation to get it working. - Added C# API and corresponding test for loading custom op shared library. * Parallel Gelu with ParallelFor (#2399) Parallel Gelu to get better performance for Gelu * Clean up build.py (#2446) * Pull the latest image before running docker build * Fuse SkipLayerNorm with Bias (#2453) Fuse SkipLayerNorm with Bias * Allow more than one invocation of CreateEnv in the same process. (#2467) * Allow more than one invocation of CreateEnv in the same process. * Fix centos build * Symbolic shape inference improvements: (#2460) * Symbolic shape inference improvements: - add a mode to guess unknown ops' output rank - add support for GatherND - add support for If - fix a bug in get_int_values when then tensor rank > 1D, by treating it as no sympy data - add symbol to literal merge when ONNX silently merges dims - fix a bug in Concat when input dim is 0 - fix a bug in ConstantOfShape that computed dim is not updated - add support for dynamic shape in ConstantOfShape - fix a bug in Loop output shape that loop iterator dim is not inserted at dim 0 - add support for dynamic padding in Pad - add support for dynamic shape in Reshape - add support for Resize with opset > 10, by treating output dims as dynamic - fix a bug in Slice when starts/ends are dynamic - restrict input model to opset 7 and above - make output model optional to avoid disk write when testing Run model tests for symbolic shape inference Reduce 2GB docker image size of nuphar * add additional test data set for nuget pipeline (#2448) * add SAS token to download internal test data for nuget pipeline * update azure endpoint * fix keyvault download step * fix variable declaration for secret group * fix indentation * fix yaml syntax for variables * fix setting secrets for script * fix env synctax * Fix macos pipeline * attempt to add secrets to windows download data * fix mac and win data download * fix windows data download * update test data set url and location * Revert "Brianma/windowsai fi (#2475)" This reverts commit `5780b864a1`. * Add scenario tests (#2457) * Add scenario tests * Remove TODO from model license * Add winml_api test dependency * fix model load test. fi from master changed the constructor (#2483) * make api tests all pass (#2486) * fix bad merge * fix bad model merge * Layer dev paulm (#2492) * commetns for dml graph transformer fixed ort value passing using the allocatir info * fixed and coded maps and sequences across the abi * Rename ambiguous header (#2489) * fix one more missing IR version model (#2500) * add missing IR version to 4 more models used by scenario tests (#2501) * Add CLI parameters to test runner, build WinML in ARM and x86 CI (#2479) * Support test parameters through CLI arguments * Add WinML do Windows x86/ARM CI builds * Code style fixes * Update googletest Remove GPUTEST macros everywhere now that GTEST_SKIP is supported * Refactor main.cpp * Build scenario tests without DML * Link scenario tests to DML when it's enabled (#2502) * Layer dev release pipeline (#2488) Adds winml binaries to existing cpu nuget package, and creates new gpu dml nuget package with winml binaries and DML EP. * Layer dev paulm (#2506) * commetns for dml graph transformer fixed ort value passing using the allocatir info * fixed and coded maps and sequences across the abi * cleaned up w4's cleaned up the model info ABI delayload directml.dll from winml * Remove usage of IOBinding in WinML and use C_API Run method (#2504) * remove usage of iobinding * Change data structure to use vector of Ort::Values * Polish bind input / output * Use C APIrun method * Update providers on evaluate getresults * Remove run and IObinding interface from WinMLAdapter * Remove use of IObinding * bind unbound outputs code moved to learningmodelbinding * clean up unneeded istensor adapter function * Fix comment * Check if session is closed before binding and clearing * PR feedback * Layer dev paulm (#2507) * commetns for dml graph transformer fixed ort value passing using the allocatir info * fixed and coded maps and sequences across the abi * cleaned up w4's cleaned up the model info ABI delayload directml.dll from winml * cleaned up namepsace aliases. renamed _winmla to winmla this was good PR feedback from tiago a while back. * Make tests dependend on winml_dll (#2509) * add dml binaries to DirectML package and be more explicit about condition variables (#2520) * re-enable warnings for winml builds and fix the warnings that were hiding (#2526) * turn devmode back on for winml builds * fix some warnings. include protobuf in a way that disables some warnings * undo protobufhelpers changes and just ignore 4100 errors in pb code * attempt to isolate protobufhelpers errors * add template specialization for getting tensor proto data * Layer dev paulm (#2533) * commetns for dml graph transformer fixed ort value passing using the allocatir info * fixed and coded maps and sequences across the abi * cleaned up w4's cleaned up the model info ABI delayload directml.dll from winml * cleaned up namepsace aliases. renamed _winmla to winmla this was good PR feedback from tiago a while back. * moved files from inc to lib\api.core cleaned up some of the cmake * staged changes * Spawn child process to run DeviceLostRecovery scenario test (#2530) * Spawn child process to run DeviceLostRecovery scenario test * Layer dev paulm (#2536) ori said yes * add missing namespace to winml_trace_logging_provider in lotusenvironment.h (#2542) * Handle exception thrown from all apis in WinMLAdapter (#2539) * various changes to unblock windowsai ADO build * Fix custom ops scenario tests (#2562) * Do not shutdown protobuf after ort environment gets destroyed. Lazy load lotus environment first time it is needed * comment typo * pr comment about calling phoenix singleton * Make lotus_environment static in winmladapter * Layer dev paulm (#2567) * commetns for dml graph transformer fixed ort value passing using the allocatir info * fixed and coded maps and sequences across the abi * cleaned up w4's cleaned up the model info ABI delayload directml.dll from winml * cleaned up namepsace aliases. renamed _winmla to winmla this was good PR feedback from tiago a while back. * moved files from inc to lib\api.core cleaned up some of the cmake * staged changes * making windowsAI azure dev ops work. * code review comments. * revert changes * Cmake and preprocessor fixes that where uncovered by building on agents without DML available via SDK * Layer dev dml delayload (#2580) * Brianma/cpu (#2583) * don't include dml stuff in cpu builds * tests that link the image lib also need the telemetry lib now * Throw Winml_err_invalid_binding if binding gpu resource on cpu device (#2589) * Throw Winml_err_invalid_binding if binding gpu resource on cpu device * PR comments. No need to query executionprovider for is gpu device * User/xianz/ortthrow (#2596) * thrown and handle onnxruntime exceptions * handle exception thrown from ort in winmladapter * undo changes in error.h * add message to HRESULT * User/xianz/ortthrow (#2599) * thrown and handle onnxruntime exceptions * handle exception thrown from ort in winmladapter * undo changes in error.h * add message to HRESULT * add status error message * Remove uwp onsuspending winrt call because logruntimeperf is getting removed (#2630) * User/xianz/dedup telemetry (#2631) * investigate duplication of telemetry in winml and ort * remove winml telemetry events * telemetry executionProviderEvent * remove unneccessary file and refactor code little bit * Revert back TelemetryEvent, which send up ETW event. * merge changes from layer_dev to windowsai (#2638) * Remove underscore from googletest names (#2616) * Fix leaking memory allocator Fix https://microsoft.visualstudio.com/OS/_workitems/edit/24278761 and https://microsoft.visualstudio.com/OS/_workitems/edit/24330198 * Explicitly initialize Ort::Value with nullptr * Cache WinML adapter * bad merge * define private version of dxcore enum that is added in 19H1 SDK. (#2654) * add comment for explaning private definition of dxcore d3d feature level ennum value. (#2672) * do not package directml.pdb for redist packages. (#2676) * Fix leaking operator registry (#2645) Fix https://microsoft.visualstudio.com/OS/_workitems/edit/24354916 * User/orilevari/windowsai master merge (#2674) merge resolutions included pulling in telemetry logic that was merged to master and not windowsai and dereferencing InferenceSession::sessionstate now that it is a unique pointer * Delete Ort Allocator in LearningModelBinding (#2653) * Delete OrtAllocator in LearningModelBinding * PR comments to make Ort::Allocator a smart pointer * Small comment change * PR feedback to clean up code * PR feedback on move semantics * Clean up std::move * Fix memory leaks (#2679) Fix https://microsoft.visualstudio.com/OS/_workitems/edit/24356109, https://microsoft.visualstudio.com/OS/_workitems/edit/24388361 and https://microsoft.visualstudio.com/OS/_workitems/edit/24388596 * various changes to properly organize and skip GPU tests. For now for No DML builds we will not run GPU tests at all. In the future we should adapt the tests to expect the appropiate errors. (#2695) * Windowsai without fi (#2701) * Disable Attention fusion tests when DISABLE_CONTRIB_OPS is defined (#2529) * Setup java ci (#2528) * Add provision in ORT for session options to be parsed when available via model file (#2449) * Initial commit * Fix gitmodules * Nits * Nits * Updates * Update * More changes * Updates * Update * Some updates * More changes * Update * Update * Merge * Update * Updates * More changes * Update * Fix nits * Updates * Fix warning * Fix build * Add comment * PR feedback * PR feedback * Updates * Updates * Update * More changes * Fix build break * Comment test for now * Updates * Updates * PR feedback * Updates * Nits * Add tests * Fix build * Fix build * Fix build * Fix build break * Fix build * Nits * PR feedback * More change * Expose GetSessionOptions in pybind logic and add unit test for python * Fix build * PR feedback * PR feedback * Revert "Disable thread pool creation when enabled OpenMP (#2485)" (#2535) This reverts commit `7c7d5a149c`. * Add dynamic shape support in TensorRT execution provider (#2450) * remove onnx-tensorrt submodule * add new onnx-tensorrt submodule (experiment) for trt6 * update engine build for trt6 * update compile and compute for tensorrt6.0 * Update tensorrt_execution_provider.cc * Update tensorrt_execution_provider.cc * Update tensorrt_execution_provider.cc * Update tensorrt_execution_provider.cc * switch to onnx-tensorrt master for TensorRT6' * Update tensorrt_execution_provider.cc * Handle dynamic batch size and add memcpy in TensorRT EP * update test cases * Update tensorrt_execution_provider.cc * update onnx-tensorrt submodule * Update Dockerfile.ubuntu_tensorrt * Update Dockerfile.ubuntu_tensorrt * Update run_dockerbuild.sh * Update run_dockerbuild.sh * Update install_ubuntu.sh * Update concat_op_test.cc * Update tensorrt_execution_provider.cc * Upgrade TensorRT to version 6.0.1.5 * Update onnxruntime_providers.cmake * Update CMakeLists.txt * Update reduction_ops_test.cc * Update install_ubuntu.sh * Update Dockerfile.ubuntu_tensorrt * Update Dockerfile.tensorrt * Update BUILD.md * Update run_dockerbuild.sh * Update install_ubuntu.sh * Update onnxruntime_providers.cmake * Update install_ubuntu.sh * Update install_ubuntu.sh * Update gemm_test.cc * Update gather_op_test.cc * Update CMakeLists.txt * Removed submodule * update onnx-tensorrt submodule * update header file * Removed submodule * add submodule onnx-tensorrt kevin's branch shape-test' * add debugging code * Update tensorrt_execution_provider.cc * Update tensorrt_execution_provider.cc * merge master * Removed submodule * update onnx-tensorrt submodule * add more changes for dynamic shapes * Update tensorrt_execution_provider.cc * update for dynamic shape * update dynamic shape processing * fix logger issue * remove submodule onnx-tensorrt * add submodule onnx-tensorrt * add env variable min_subgraph_size * remove redundency * update document * use onnxruntime::make_unique * fix multi-run issue * remove some tests to save CI build time * Add dynamic shape test * Update TensorRT-ExecutionProvider.md * Add example of running Faster R-CNN model on TensorRT EP * Add more details on env variables * update environment variables * Update tensorrt_basic_test.cc * Update model tests * Update tensor_op_test.cc * remove --use_full_protobuf * Update build.py * User/xianz/telemetry (#2458) * enabme telemetry * enable telemetry * set enable telemetry as default * for debugging * remove log and set disable telemetry as default back * delete private file while testing * resolve comment: mainly add license header, rename macro and update docs * rewording in privacy.md * Fix integer overflow in cuda NonMaxSuppression implementation (#2540) * add test case that should pass but fail * fix nms * extract int_max_output_boxes_per_class * Introduce container type runtime checks and other improvements (#2522) Rework TensorSeq in a manner consistent with Tensor and SparseTensor in terms of type system setup. Reduce templating. Introduce helpers to ensure the same data type. Make OrtValue __dtor not virtual. Introduce ContainerChecker * Fix C API tests for centos and mac (#2544) * change c++14 to c++11 * add ld lib path for centos * enable csharp tests on macos * fix C API test on MacOS + fix manylinux dotnet install * fix manylinux dotnet install * fix lib link * Add back executable bit to build.py * Fix a bug handling negative begin pad values in Pad op (#2550) * Fix bug in Pad op * Update * DNNL CMAKE update (#2548) * Fix android build (#2558) * Update win-x86-ci.yml (#2557) Fix build pipeline break * Re-enable Windows C# tests (#2564) * disable onnx_test_runner -x invocations for dnnl (#2568) * Allow sequence length to be symbolic (#2559) * setup java ci mac (#2570) * make layernorm fusion to support opset 11 (#2545) * Fix a warning found in the latest VS release * Add more check on SkipLayerNorm and BiasGelu fusion (#2574) * Fix file not found error during docker build. (#2569) * Add ConvTranspose1D (#2578) * Ryanunderhill/packagename test (#2582) * [Nuphar EP] fixes for some object detection models (#2581) Update notebook tutorial with multi-threaded int8 GEMM from #2517 * EmbedLayerNormalization Fusion Improvement (#2553) Embedding layer norm fusion improvements - add more checks * Update version (#2584) * Temporarily exclude vgg19 test from Python backend test 1. temporarily exclude vgg19 test which comsumes too much memory, run out of memory on Upsquared device. Single test pass for vgg19, need furture investigation (#2588) 2. Update docker file to decrease the docker image size * Update docs for Android NNAPI EP (#2586) * Fix lto bug for protobuf and ubuntu * add path to build dir before test run (#2590) * Add missig env variables for mac pipeline test (#2595) * Fixed an issue in updating realized dims (#2597) when we update realized dims for scan's output, the sliced axis also needs to be inclusive, i.e. we should check with "dim >= insert_inclusive_axis", because the offset in the symbols are based on Scan sugraph. Otherwise, we would end up with shape mismatch later. * Java API for onnxruntime (#2215) * Add support for opset 11 in reshape fusion (#2592) Support opset verion 11 in reshape fusion * Rename automl python tools folder to featurizer_ops. (#2593) * Support opset 11 subgraph of Squad model in Embed Layer Normalization (#2605) Support opset 11 Squad model that is exported from PyTorch nightly. The embed layer uses Range op which is missed in the transformer. * symbolic shape inference: fix warnings in GPT-2 model (#2608) And revise nuphar perf test on BERT squad * Dump subgraph ID and fused graph ID (#2607) * Dump subgraph ID and fused graph ID Dump subgraph ID and fused graph ID for better debugging * Remove local static fused_count added a field global_fused_count_ to NupharExecutionProvider class * EmbedLayerNormalization Fusion For Dynamic Squad Model Opset 10 (#2613) Support subgraph of SQuAD model exported from pytorch with dynamic input axes * Allow providers to be set for InferenceSession at construction (#2606) * Remove unnecessary parameter in some places in GatherElements implementation (#2612) * Remove unnecessary parameter in some places * Update * Update * Make sure fenced tensor could not reuse other tensor. (#2561) Fix random error caused by this. * Improve Embed Layer Norm Fusion for SQuAD with static input shape (#2621) * fix float16 comparison in initializer (#2629) * epsilon attribute for layernormalization fusion (#2639) * removed unnecessary batch file and fix path (#2640) * Add shape inference to ConvTransposeWithDynamicPads schema (#2632) * Improve cuda expand() opeator's performance. (#2624) * Cuda pad optimize when no padding is needed. (#2625) * Shortcut cuda Pad() when no padding is needed. * Optimize cuda scatter() on 2D compatible. (#2628) * Optimize cuda scatter() on 2D compatible. * Add some comments. * fix build error for ARM (#2648) * Improve performance of resize() in Nearest mode (#2626) Special treatment for 2D, check same size as input image. And in 2d kernel, template use_expolation. * Fix memory exception in Layer Norm Fusion (#2644) * Windows CI changes(#2650) * Revert "User/orilevari/windowsai master merge (#2674)" This reverts commit `fe26146311`. * Revert "Windowsai without fi (#2701)" This reverts commit `285d4c85ff`. * Revert "User/orilevari/windowsai master merge (#2674)" This reverts commit `fe26146311`. * Deref unique pointer for session_state * send shutdown event when dll is unloaded and EvaluationStop, SessionC… (#2704) * send shutdown event when dll is unloaded and EvaluationStop, SessionCreationStart Events. * Add EvalutationStart Event * add comment * use correct type for for loop (#2755) * ARM CI (#2759) * Set ARM agent pool * Set CMake generator to VS 2019 in ARM * Use system-wide CMake instead of custom version Our custom version is too old for VS 2019 * Use DML and build shared lib in ARM CI * Restore nuget packages in ARM CI * Disable DML * Refactor ARM debug/release builds * Use system packaged Python version * Remove hardcoded Python path * Downgrade Python to 3.7 for build * Remove explicit CMake path * Fix invalid JSON in cgmanifest.json (#2760) * Fix cgmanifest.json generating script (#2770) * Fix protobuf submodule name * Workaround pygit2 bug * Remove usage of WHOLEARCHIVE in WinML CMake and add WinMLAdapterFactory (#2726) * Remove usage of WHOLEARCHIVE in WinMLAdapter CMake and add WinMLAdapterFactory * PR feedback, no need for dll(export) since using def file * PR comments * Small comment in gen_def.py * User/orilevari/32bit comparison warning (#2800) * use correct type for for loop * explicitly specify void for parameters of OrtGetApiBase because the function is defined in c, so when the function is just (), it is interpreted as having an unknown number of parameters. This was causing compiler warning C4276. * Move winml_provider_factory.h to proper location (#2801) * Scneario Test : Build Google Test and Taef Test based on preprocessor definition (#2809) * Add winml macro wrappers on top of google test macros * change test methods to disabled * Add custom winml macros for both taef and google tests * PR comments * Filter CPU case for IsFloat16Supported (#2802) * Merge fixes * CMake cross-generator fixes (#2790) * Fix compilation w/ non-VS CMake generators * Fix custom WINMD target in Ninja * Remove usage of msbuild .targets file * Fix linking using DML in Ninja * Automate SDK kit version choice * Cleanup DML package install * Fix SDK version detection * Fix comment * Revert unittest linkage changes * Fix latest SDK detection * Don't link to non-uapcore libraries * Remove MessageBoxA reference and unused link libs * Refactor WinMLAPI Tests to build both google and taef test based on preprocessor definition (#2829) * Add winml macro wrappers on top of google test macros * change test methods to disabled * Add custom winml macros for both taef and google tests * PR comments * Refactor winml api tests * Move additional gtest specific macro definition into googleTestMacros.h * Fix test build break since winml_lib_api needs to be statically linked to tests since winmlp::learningmodeldevice::iscpu() is being used in devicehelpers.cpp (#2837) * Enforce WINML_TEST_CLASS_BEGIN_* matches w/ a WINML_TEST_CLASS_END (#2841) * Fix warnings that cause build to fail * Fix test warnings and delayload linking (#2843) * Ortmemoryinfo struct changed * mark the camera scenario test as edgecore because it uses d3d11 (#2852) * User/orilevari/pipeline fi breaks (#2853) * remove conflicting artifact names. Decided to stop using drop-nuget-cuda since this may have implications on other dependent pipelines. * change job name in gpu.yml back to Windows_CI_GPU_CUDA_Dev * Remove internal libs from tests (#2864) * Support custom DML in onnxruntime_providers.cmake (#2867) * Make DML include path global (#2882) * Make DML include path global * Add generated cppwinrt headers to winml_lib_common * Integrate changes to WindowsAI to make ADO Build (#2886) * Revert "CMake cross-generator fixes (#2790)" This reverts commit `dbe7d97fa1`. * add additional suppress warning in onnx_proto * ignore /wd4996 warning * DML execution provider fixes * Revert "Revert "CMake cross-generator fixes (#2790)"" This reverts commit `1ae7b4bcbc`. * Update func signature of custom op function overloads * common devicehelpers fixes * Add pch.h for winml_lib_common * re-add winml_lib_common_dir/inc to include path for winml_adapter * User/orilevari/dml redist shared folder (#2890) * move dml nuget package directory up one level to make it shared between build flavors * Merge conflict fix * Revert "Merge conflict fix" This reverts commit 142fa72cf9ce4344ad717b50b7ea2b8582aadc7c. * Revert "Merge remote-tracking branch 'origin/master' into windowsai" This reverts commit 6e2126d46e5e5f564d65da37dd4f70c93dd81165, reversing changes made to b3f5583dc9249834b947c8ea905f6a98060d5bd6. * Make winml_test_common free of test macros (#2902) * Add option to build winml_test_common without googletest specifics * remove test macros from squeezenet * comment change * Make cmake functions to get scenario and api source * PRcomments about hresult * Build errors fixed * Fix cmake variable * Make winml_google_test_lib to build main.cpp once * PRcomments * Don't generate files outside the build root (#2914) * Don't generate files outside the build root * Add onnxruntime_EXTERNAL_DEPENDENCIES to WinML * Add DML depedency on RESTORE_PACKAGES * User/orilevari/fix yaml merge bugs (#2918) * Add winml test source parameter into cmake function (#2919) * Add option to build winml_test_common without googletest specifics * remove test macros from squeezenet * comment change * Make cmake functions to get scenario and api source * PRcomments about hresult * Build errors fixed * Fix cmake variable * Make winml_google_test_lib to build main.cpp once * PRcomments * Add arguments to unittest cmake functions * remove comment * Revert "Revert "Merge remote-tracking branch 'origin/master' into windowsai"" This reverts commit ade5abe72a4234fdbc3623093c61c02c6b0bdc26. * Fix breaks from merge with ORT master * Brianma/linux (#2917) * don't include windows.h in cross-plat header * add default case for switch statement * signed/unsigned mismatch fix Co-authored-by: Brian Martin <42186431+martinb35@users.noreply.github.com> * User/sheilk/winml adapter c api (#2891) * Create winml adapter c api * fix build * make it build * move adapter into onnxruntime core/session * entry point not exported * minor changes * make model metadata work * make tests pass * implement all the model reflection apis on the adapter c abi * update the new ort interface to create a lotus ennvironment with a logging sink * start adding ort env * move all winml code into adapter folder/lib to isolate it * ensure a single logging manager at a time * start refactoring session * refactor session creation interface * add cpu and dml session option methods to adapter * finish session init * stub out interfaces in ort lib to perform similar mechanics of iinference session * enable profiling, and enable schema override * update session register graph transformers * turn back on custom registry for custom ops * Add sync api * add last c api stubs * should build... but all feature values are broken since this is in flight to moving all implementation details into ivalue * remove ep adapter header * Implement DML execution provider functions from adapter (#2846) * Implement DML execution provider functions from adapter * Use functions in OnnxruntimeEngine.cpp * make map/sequence type_infos freeable, and start implementing ivalue * make it build again * implement value methods * implement remaining methods * remove com adapter abi * check dml session * cache the allocator on ivalue * check if resource is cpu/gpu when access its mutable data * update tensor * mismatched parentheses * fix tensor base and binding obj * it evaluates tensors! sometimes... * minor fixes * enable gpu evals * wrapper all existing winml adapter apis with API_IMPL to try catch (#2854) * update winml... tensor strings are broken, need to template tensorbase to do different things for strings * make tensor strings work with 2 copies in/2 copies out * Fix tensor string and allocator bug * make maps work again... needs some fixes still * Make it build! * enable map inputs * map outputs * unbound outputs for sequences and maps * User/xianz/merge windowsai (#2883) * Packaging pipeline changes for VS 2019 (#2711) * Tiny fix to codegen * Simplify cache implementation and avoid static variables that may carry over between models * Extend DML kernels (#2641) * Additional DML operators * Check unsupported attributes and inputs * Address PR comments * Add kernel capability function used for partitioning, and re-enable stride-based int64 support based on value range * Fix test failures * Build fix * PR comments * Update Nuphar tutorial notebook (#2721) 1. Reflect int8 GEMV improvements for multi-threading from #2696 2. Add notes on multi-threading control using OpenMP 3. Add samples of running multi-isa AOT, and show int8 GEMM differences between AVX and AVX2 4. Add rnn_benchmark example to resolve #1993 * Add schema for new Qops (#2611) * Add schema for new Qops * adding shape inference + qlinearaveragepool * plus review comments * plus review comments * updates per review comments * plus review comments * [server] Add supposed for model_name and model_version as cli parameter (#2708) * remove 64bit warning message from python validation. (#2727) * MLAS: ARM64 build fix (#2734) fix bad usage of vreinterpret to cast vector element types * Fix broken python docs links (#2740) * Fix build on Mac OS (#2731) mac os ld doesn't support --while-archive, correct option is -all_load * fix ngraph wheel (#2737) * fix ngraph wheel 1.1.0 onnxruntime_ngraph wheel doesn't work * remove libdnnl.so in nGraph Libs * make it easy to compare * Split onnxruntime server to a separated folder (#2744) * Fix build for Python 3.8 (#2747) * Fix build for Python 3.8 * Update protobuf to 3.11.2 (#1928) Update protobuf to 3.11.2 (#1928) * Change default optimization level to All (from Basic) (#2745) * change default optimization level to All (from Basic) * fix test * fix c# test * Update numpy to 1.18 (#2758) * Update numpy to 1.18 * Pipeline changes for python 3.8 (#2753) 1. Pipeline changes for python 3.8 2. Fix a regression in setup.py which was just introduced in the previous commit. Please notice, we still haven't made python 3.8 + Windows + CUDA work. * Add basic stacktrace output for posix debug builds. (#2749) * [NupharEP] fix a race condition when multiple sessions running different models concurrently (#2772) * Revert "Change default optimization level to All (from Basic) (#2745)" This reverts commit `56bb503c2f`. * Fix typo in error message (#2736) * Rename MKL-DNN to DNNL to fix broken link (#2730) * Fix nightly build version number issue * Pass BUILD_BUILDNUMBER to linux docker * Disable featurizers in python packages * Import more featurizers (#2781) Make kernels non-template. Add input constraint for learnt data. Add min_max_scalar_transformer, robust_scalar_transformer, inputation_marker_transfomer, label_encoder_transformer, missing_dummies_transformer along with tests. Advance Featurizers library commit. * Implement a more stable softmax (#2715) * Implement a more stable SoftMax e^x is represented as infinity if x is large enough, like 100.f. Infinity divided by Infinity is a NAN. Thus, softmax gets a NAN if one or more item are large enough. A math transform as below is leveraged to get a stable softmax: e^xi/(e^x1 + ...e^xn) = e^(xi - max) / (e^(x1 - max) + ... + e^(xn - max)) And for convenience, force max to 0.f if all xi are negative * Contributing: Fix a typo (#2784) * ACL EP GEMM improvements (#2780) When it is posible we use a fully connected layer instead of the gemm implementation. This will let the library use the best implementation based on the input data. * ACL EP convolution improvements (#2774) Added the optimized implementation for depthwise convolution for both ACL v19.02 and ACL 19.05. Also the pointwise convolution seems to be more optimal in the CPU implementation so we opted for that instead. * Add script for release Nuget validation (#2719) * Initial commit * Nits * Disable a test temporarily * Change working directory * Test * Add download python step * Test update * More changes * Fix space issue * Fix * Verify nuget signing * Fix * Spaces * PR feedback * Nit * Fix * Fix * Remove temporary changes * add uint8 support to where op (#2792) * Improve bert optimization script: (#2712) (1) Move input int64=>int32 conversion to embed layer fusion. (2) Output epsilon attribute for LayerNormalization fusion. * add session creation time cost. (#2798) * ML.NET team needs featurizers within a package (#2789) Add auto ml featurizers to Windows, MacOS as well as to GPU packaging-pipelines. * Initialize max of softmax with lowest of float (#2786) * MLAS: update SGEMM threading parameters (#2808) * add interface to copy batch tensors. (#2807) * add interface to copy batch tensors. * onnxruntime * speed up Windows TRT CI (#2811) * don't run cuda tests if building with tensorrt * remove unnecessary build options for win trt ci * refactor win gpu tensorrt ci yml * --numpy_version=1.17 * update * update * azcopy and cuda path * Update test data (#2356) * Add timeseries imputer transformer featurizer kernel (#2813) Make kernels non-template. Add input constraint for learnt data. Fixup tests. Add two more featurizers along with tests. Tests fail. min_max_scalar_transformer robust_scalar_transformer Fix tests serialized stream by prepending version bytes. Add inputation_marker_transfomer and the test. Fix up float/double type designations. Added label_encoder_transformer along with a test. string_throw case is broken at the momement. Fix labelencodertransfomer_test.cc string_throw case Rename maxabsscalertransformer_test.cc Add MissingDummiesTransformer along with the test. Update manifest. Add TimeSeriesImputerTransformer definition, implementation and tests * Fix memory leak in TRT (#2815) * fix memory leak issue * revert EP_FAIL on enueueV2 * Add manifest missing comma * Run static code analyzer on most of our code (#2817) * Scneario Test : Build Google Test and Taef Test based on preprocessor definition (#2809) * Add winml macro wrappers on top of google test macros * change test methods to disabled * Add custom winml macros for both taef and google tests * PR comments * update quantization doc (#2783) * update documentation for quantization script * plus some spell corrections * Filter CPU case for IsFloat16Supported (#2802) * update default optimization level + fix gemm_activation fusion (#2791) * update defualt optimization level + fix gemm_activation fusion * fix typo * add unit test and incorporate review comments * fix test comment * Fix dnnl wheel package name (#2823) * Append '-dnnl' to whl package name when --use_dnnl * Update build.py * Update Ubuntu & TensorRT version in README (#2820) Dockerfile.tensorrt is using nvcr.io/nvidia/tensorrt:19.09-py3 as base Image, update Ubuntu and TensorRT version according to https://docs.nvidia.com/deeplearning/sdk/tensorrt-container-release-notes/rel_19-09.html#rel_19-09 * Merge fixes * Add OneHotEncoder and HashOneHotEncoder kernels. (#2830) Add defs and imlementation for OneHotEncoders, adjuist date_time_transformer kernel and test. Add OneHotEncoder kernel test. Add HashOneHotVectorizerTransformer unit test. This does not link due to multiple definitions of functions that are included into header from a CPP file. * Upgrade gtest to the latest version (#2827) WinML would like to update the googletest submodule. They want some newer features (namely GTEST_SKIP to skip tests programmatically and be able to skip entire fixtures easily) and would need to update the submodule version. However, because the new version of code hit a bug in gcc, even though the bug is already fixed in the latest gcc but we're using gcc 4.8.x and it won't get patched for the bug, so we have to do a compromise, change our code a little bit to make it work. The gcc bug: https://gcc.gnu.org/bugzilla/show_bug.cgi?id=51213 * Add support for int64_t for topk CPU. Fixes github issue #2806. (#2833) * Ignore allocator type in ExecutionProviders allocator map. Make default initialization of OrtMemoryInfo more clearly invalid. (#2768) * Remove allocator type from the key comparison in ExecutionProviders. Remove usage of DummyArena as it's no longer necessary. * Fix x86 tests where arena allocator is disabled. Make initialization of OrtMemoryInfo clearer by adding Invalid enum value. * Make OrtValueNameIdxMap::MaxIdx more intuitive. * Convert ExternalProject Featurizers into git submodule (#2834) Add git submodule for Featurizer library. Update cmake to build for git submodule. * add domain check for nodes + update documentation (#2831) * Fix cgmanifest.json generating script (#2770) * Fix protobuf submodule name * Workaround pygit2 bug * User/orilevari/32bit comparison warning (#2800) * use correct type for for loop * explicitly specify void for parameters of OrtGetApiBase because the function is defined in c, so when the function is just (), it is interpreted as having an unknown number of parameters. This was causing compiler warning C4276. * CMake cross-generator fixes (#2790) * Fix compilation w/ non-VS CMake generators * Fix custom WINMD target in Ninja * Remove usage of msbuild .targets file * Fix linking using DML in Ninja * Automate SDK kit version choice * Cleanup DML package install * Fix SDK version detection * Fix comment * Revert unittest linkage changes * Fix latest SDK detection * Don't link to non-uapcore libraries * Remove MessageBoxA reference and unused link libs * Fix Linux CUDA nuget packaging pipeline break * Refactor WinMLAPI Tests to build both google and taef test based on preprocessor definition (#2829) * Add winml macro wrappers on top of google test macros * change test methods to disabled * Add custom winml macros for both taef and google tests * PR comments * Refactor winml api tests * Move additional gtest specific macro definition into googleTestMacros.h * Fix test build break since winml_lib_api needs to be statically linked to tests since winmlp::learningmodeldevice::iscpu() is being used in devicehelpers.cpp (#2837) * Enforce WINML_TEST_CLASS_BEGIN_* matches w/ a WINML_TEST_CLASS_END (#2841) * update optimization doc for BERT related fusions (#2819) * Add bert related transformers to doc * Add execution provider and comment for bert optimizations * Add comment about accuracy impact of approximation * Fix warnings that cause build to fail * MLAS: enable threading for quantized GEMMs (#2844) * Fix test warnings and delayload linking (#2843) * Ortmemoryinfo struct changed * mark the camera scenario test as edgecore because it uses d3d11 (#2852) * User/orilevari/pipeline fi breaks (#2853) * remove conflicting artifact names. Decided to stop using drop-nuget-cuda since this may have implications on other dependent pipelines. * change job name in gpu.yml back to Windows_CI_GPU_CUDA_Dev * Remove internal libs from tests (#2864) * Support custom DML in onnxruntime_providers.cmake (#2867) * remove old winmladapter cpp Co-authored-by: Changming Sun <chasun@microsoft.com> Co-authored-by: KeDengMS <kedeng@microsoft.com> Co-authored-by: Jeff <38966965+jeffbloo@users.noreply.github.com> Co-authored-by: Ashwini Khade <askhade@microsoft.com> Co-authored-by: Andrey <andrey.lompart@gmail.com> Co-authored-by: George Wu <jywu@microsoft.com> Co-authored-by: Tracy Sharpe <42477615+tracysh@users.noreply.github.com> Co-authored-by: Faith Xu <txsafx@gmail.com> Co-authored-by: zhanyi-ms <zhanyi@microsoft.com> Co-authored-by: Changyoung Koh <gkcy1019@gmail.com> Co-authored-by: Scott McKay <Scott.McKay@microsoft.com> Co-authored-by: Takeshi Watanabe <take-cheeze@users.noreply.github.com> Co-authored-by: Dmitri Smirnov <yuslepukhin@users.noreply.github.com> Co-authored-by: Yufeng Li <liyufeng1987@gmail.com> Co-authored-by: Maher Jendoubi <maher.jendoubi@gmail.com> Co-authored-by: Andrews548 <32704142+Andrews548@users.noreply.github.com> Co-authored-by: Hariharan Seshadri <shariharan91@gmail.com> Co-authored-by: Nathan <7902510+ybrnathan@users.noreply.github.com> Co-authored-by: Tianlei Wu <tlwu@microsoft.com> Co-authored-by: Ke Zhang <kezhan@microsoft.com> Co-authored-by: stevenlix <38092805+stevenlix@users.noreply.github.com> Co-authored-by: Ryan Lai <ryalai96@gmail.com> Co-authored-by: Ori Levari <ori.levari@microsoft.com> Co-authored-by: Yingge WAN <y-wan@users.noreply.github.com> Co-authored-by: Qing <cwq1913@gmail.com> Co-authored-by: Pranav Sharma <emailpranav@gmail.com> Co-authored-by: Tiago Koji Castro Shibata <tiago.shibata@gmail.com> * move sequence implementation into ort lib... still commented out... need to turn back on... * begin sequence implementation * make maps and sequences work * fix broken tests * remove dead code * misc cleanup * CR feedback * User/xianz/winml adapter c api (#2869) * wrapper all existing winml adapter apis with API_IMPL to try catch * Return HR or Throw for WinML adapter APIs if failed * undo macro wrapper for two places * Wrap error macros around ort apis, too. * address CR feedback #2 * add more api throw/return macros * Revert changes no longer needed * revert changes to cxx api * format winml lib.ort and winml adapter * remove static pheonix singleton Co-authored-by: Ryan Lai <ryalai96@gmail.com> Co-authored-by: Xiang Zhang <xianz@microsoft.com> Co-authored-by: Changming Sun <chasun@microsoft.com> Co-authored-by: KeDengMS <kedeng@microsoft.com> Co-authored-by: Jeff <38966965+jeffbloo@users.noreply.github.com> Co-authored-by: Ashwini Khade <askhade@microsoft.com> Co-authored-by: Andrey <andrey.lompart@gmail.com> Co-authored-by: George Wu <jywu@microsoft.com> Co-authored-by: Tracy Sharpe <42477615+tracysh@users.noreply.github.com> Co-authored-by: Faith Xu <txsafx@gmail.com> Co-authored-by: zhanyi-ms <zhanyi@microsoft.com> Co-authored-by: Changyoung Koh <gkcy1019@gmail.com> Co-authored-by: Scott McKay <Scott.McKay@microsoft.com> Co-authored-by: Takeshi Watanabe <take-cheeze@users.noreply.github.com> Co-authored-by: Dmitri Smirnov <yuslepukhin@users.noreply.github.com> Co-authored-by: Yufeng Li <liyufeng1987@gmail.com> Co-authored-by: Maher Jendoubi <maher.jendoubi@gmail.com> Co-authored-by: Andrews548 <32704142+Andrews548@users.noreply.github.com> Co-authored-by: Hariharan Seshadri <shariharan91@gmail.com> Co-authored-by: Nathan <7902510+ybrnathan@users.noreply.github.com> Co-authored-by: Tianlei Wu <tlwu@microsoft.com> Co-authored-by: Ke Zhang <kezhan@microsoft.com> Co-authored-by: stevenlix <38092805+stevenlix@users.noreply.github.com> Co-authored-by: Ori Levari <ori.levari@microsoft.com> Co-authored-by: Yingge WAN <y-wan@users.noreply.github.com> Co-authored-by: Qing <cwq1913@gmail.com> Co-authored-by: Pranav Sharma <emailpranav@gmail.com> Co-authored-by: Tiago Koji Castro Shibata <tiago.shibata@gmail.com> * missing use_dml check in winml_adapter_session (#2930) * --use_dnnl flag was mangled in merge (#2931) * use dml macro not wrapping custom registry code (#2934) * Disable LNK4199 winml_dll to enable cuda builds (#2936) * Disable LNK4199 in winml_dll * linkler->linker * LearningModelSessionAPITestGpu.CreateSessionWithCastToFloat16InModel should return DXGI_ERROR_UNSUPPORTED when FP16 not supported (#2937) * Disable LNK4199 in winml_dll * linkler->linker * Need to return DXGI_ERROR_UNSUPPORTED when Model does not support fp16 * Publish build symbols (#2939) * Publish build symbols * Don't upload PDBs for .exe files * Make x86 build (#2943) * fix last remaining size_t/int64_t warnings->errors (#2948) * TensorString, Sequences and Maps use the first allocator, but should use the cpu default allocator. (#2952) * fix tensor string allcoator * clean up default allocator usage for strings in winml lib/api.ort Co-authored-by: Ryan Lai <ryalai96@gmail.com> * Handle tensor shape of zero (#2954) Co-authored-by: Ryan Lai <ryalai96@gmail.com> * CR feedback (#2970) * CR feedback * fix weird formatting on privacy readme * Add 'All rights reserved.' everywhere * readd all rights reserved to winml_provider_factory.h * remove extra space in comment * remove extra whitespace * fixes post master merge * remove winml from nuget gpu pipeline * set IR VERSION on generated_model in rnn_benchmark (#2972) * Fix slice conformance failures (#2908) Co-authored-by: Adrian Tsai <adtsai@microsoft.com> Co-authored-by: Brian Martin <42186431+martinb35@users.noreply.github.com> Co-authored-by: Ryan Lai <ryalai96@gmail.com> Co-authored-by: Paul McDaniel <paul_mcdaniel@hotmail.com> Co-authored-by: Xiang Zhang <xianz@microsoft.com> Co-authored-by: Dwayne Robinson <fdwr@hotmail.com> Co-authored-by: Tiago Koji Castro Shibata <tiago.shibata@gmail.com> Co-authored-by: Ori Levari <ori.levari@microsoft.com> Co-authored-by: Jeff <38966965+jeffbloo@users.noreply.github.com> Co-authored-by: Changming Sun <chasun@microsoft.com> Co-authored-by: KeDengMS <kedeng@microsoft.com> Co-authored-by: Ashwini Khade <askhade@microsoft.com> Co-authored-by: Andrey <andrey.lompart@gmail.com> Co-authored-by: George Wu <jywu@microsoft.com> Co-authored-by: Tracy Sharpe <42477615+tracysh@users.noreply.github.com> Co-authored-by: Faith Xu <txsafx@gmail.com> Co-authored-by: zhanyi-ms <zhanyi@microsoft.com> Co-authored-by: Changyoung Koh <gkcy1019@gmail.com> Co-authored-by: Scott McKay <Scott.McKay@microsoft.com> Co-authored-by: Takeshi Watanabe <take-cheeze@users.noreply.github.com> Co-authored-by: Dmitri Smirnov <yuslepukhin@users.noreply.github.com> Co-authored-by: Yufeng Li <liyufeng1987@gmail.com> Co-authored-by: Maher Jendoubi <maher.jendoubi@gmail.com> Co-authored-by: Andrews548 <32704142+Andrews548@users.noreply.github.com> Co-authored-by: Hariharan Seshadri <shariharan91@gmail.com> Co-authored-by: Nathan <7902510+ybrnathan@users.noreply.github.com> Co-authored-by: Tianlei Wu <tlwu@microsoft.com> Co-authored-by: Ke Zhang <kezhan@microsoft.com> Co-authored-by: stevenlix <38092805+stevenlix@users.noreply.github.com> Co-authored-by: Yingge WAN <y-wan@users.noreply.github.com> Co-authored-by: Qing <cwq1913@gmail.com> Co-authored-by: Pranav Sharma <emailpranav@gmail.com>	2020-02-04 17:12:19 -08:00
Changming Sun	7ff5c0e5a3	CMake changes (#2961 ) 1. Add support for vstest. 2. Add support for vcpkg. To use it: ```bat vcpkg install zlib:x64-windows benchmark:x64-windows gtest:x64-windows protobuf:x64-windows pybind11:x64-windows re2:x64-windows mkdir build cmake ..\cmake -DCMAKE_BUILD_TYPE=Debug -A x64 -T host=x64 -DCMAKE_TOOLCHAIN_FILE=C:\vcpkg\scripts\buildsystems\vcpkg.cmake -DVCPKG_TARGET_TRIPLET=x64-windows -Donnxruntime_PREFER_SYSTEM_LIB=ON ``` 3. New cmake option: onnxruntime_PREFER_SYSTEM_LIB, which allows user using the preinstall libs instead of the things in onnxruntime submodule. 4. New cmake option: onnxruntime_ENABLE_MEMLEAK_CHECKER, which allows user turn on/off the memory leak checker by @RyanUnderhill in Windows Debug Build. The checker doesn't work with vstest. 4. Fix the post merge pipeline(Mainly for test coverage report). 5. Ignore the compile warning from the Featurizer library code 6. Apply "/utf-8" VC compile flag to our code. Without this, you can't build onnxruntime on Chinese Windows. 7. Remove the SingleUnitTestProject cmake option because it's deprecated more than one year and nobody is using it. 8. Move opaque api tests to onnxruntime_test_all 9. Enable "/W4" on CUDA ep's C++ code(Not the *.cu files), and fix some warnings, add some extra checks. 10. Delete the onnxruntime::test::TestEnvironment class. 11. Add a DLLmain for onnxruntime.dll. 12. Allow dynamic link to libprotobuf	2020-02-03 19:33:14 -08:00
jignparm	645d8fb213	Jignparm/upgrade macos vm image (#2928 ) * update MacOS image to 10.14 * Update to macos 10.14	2020-01-28 20:24:45 -08:00
Changming Sun	e0c9cdaa73	Fix the nuget pipelines (#2901 )	2020-01-23 20:02:18 -08:00
Jeff	ba336b5583	Disable DML EP on software adapter, fix float16 fallback bug, re-enable DML in CI (#2896 ) * Re-enable DML in CI pipeline * Fix bug with float16 fallback + fusion, and disallow DML EP with software adapter * Address PR comments	2020-01-23 15:18:28 -08:00
Changming Sun	201b089a36	Fix some warnings on Windows (#2560 ) 1. Enable warning "4503" # Decorated name length exceeded. 2. Enable warning "4146" # unary minus operator applied to unsigned type. 3. Enable float64 support for the Softmax operator 4. Enable compliance checks for Windows x86 32bits build 5. Use TryBatchParallelFor to replace some fallback code in mlas pooling.cc 6. Fix Android CI pipeline.	2020-01-22 15:59:11 -08:00
Pranav Sharma	49725f896c	Disable openmp for the nocontribops pipeline. (#2888 )	2020-01-22 12:07:44 -08:00
KeDengMS	f9f25ec047	Fix spurious component detection warning (#2857 ) Fix spurious component detection warning Use component detection template for all pipelines	2020-01-18 20:10:35 -08:00
Changming Sun	e6f7658ade	Update Windows GPU build to use cudnn 7.6	2020-01-17 12:23:13 -08:00
Changming Sun	47e27ec9a1	Disable DML in Windows GPU CI build (#2856 ) Disable DML in Windows GPU CI build for now, because there are some wired model test failure and I don't know how to fix it. Will seek help from WinML team.	2020-01-16 18:47:30 -08:00
Changming Sun	c4e4abce73	Run static code analyzer on most of our code (#2817 )	2020-01-10 22:17:17 -08:00
Changming Sun	48e042868f	Update test data (#2356 )	2020-01-10 10:52:23 -08:00
George Wu	31200ed92c	speed up Windows TRT CI (#2811 ) * don't run cuda tests if building with tensorrt * remove unnecessary build options for win trt ci * refactor win gpu tensorrt ci yml * --numpy_version=1.17 * update * update * azcopy and cuda path	2020-01-10 08:40:40 -08:00
Dmitri Smirnov	2c8179bee4	ML.NET team needs featurizers within a package (#2789 ) Add auto ml featurizers to Windows, MacOS as well as to GPU packaging-pipelines.	2020-01-09 10:54:12 -08:00
Hariharan Seshadri	ebfcad1c90	Add script for release Nuget validation (#2719 ) * Initial commit * Nits * Disable a test temporarily * Change working directory * Test * Add download python step * Test update * More changes * Fix space issue * Fix * Verify nuget signing * Fix * Spaces * PR feedback * Nit * Fix * Fix * Remove temporary changes	2020-01-08 18:42:22 +05:30
Changming Sun	e3f674b563	Disable featurizers in python packages	2020-01-06 11:16:44 -08:00
Changming Sun	7ace7a5bcd	Pass BUILD_BUILDNUMBER to linux docker	2020-01-06 11:16:44 -08:00
Changming Sun	382fa86af8	Pipeline changes for python 3.8 (#2753 ) 1. Pipeline changes for python 3.8 2. Fix a regression in setup.py which was just introduced in the previous commit. Please notice, we still haven't made python 3.8 + Windows + CUDA work.	2020-01-02 15:25:25 -08:00
Changming Sun	fd334aff44	Update numpy to 1.18 (#2758 ) * Update numpy to 1.18	2019-12-30 14:51:01 -08:00
Changming Sun	c7a9c6b488	Split onnxruntime server to a separated folder (#2744 )	2019-12-27 11:21:23 -08:00
Changming Sun	b42cb61904	Packaging pipeline changes for VS 2019 (#2711 )	2019-12-20 19:53:51 -08:00
Changming Sun	89d6bfaa94	VS 2019 build pipeline changes (#2693 ) 1. Move Win GPU pipeline to VS2019 2. Move C API pipeline to VS 2019 3. Move nuget mklml pipeline to VS 2019 4. Move windows no contrib ops pipeline to VS 2019	2019-12-18 15:34:58 -08:00
Dmitri Smirnov	ce7a180f21	Import more featurizers with tests (#2685 ) Advance commit to 4df80d5865a9d4e97f6d0b9304d4316115a04d9e Add generated code for the commit before editing. Import more featurizers. Rename Automl ops domain to mlfeaturizers. Rename conditional compilation macro. Move and rename files getting rid of automl Rename --use_automl build switch to --use_featurizers Rename CMake option accordingly. Rename automl CMake targets. Adjust CI and packaging pipeline switches. Rename namespace automl to featurizers.	2019-12-17 22:17:40 -08:00
Hector Li	47503ec7a6	Initiate the build scripts for ARM ACL (#2652 ) 1. Add scripts to build Yocto image & toolchain 2. Update docker build scripts to support Onnxruntime build with ARM ACL 19.02/19.05	2019-12-16 09:44:19 -08:00
Dmitri Smirnov	7c87070b24	Import Featurizers (#2643 ) Import FeaturizerLibrary as ExternalPorject which is optional and is not registered as git submodule.	2019-12-13 16:07:12 -08:00
Changming Sun	a46a28b7d8	Windows CI changes(#2650 )	2019-12-13 12:23:49 -08:00
shahasad	4dbf9442cc	removed unnecessary batch file and fix path (#2640 )	2019-12-12 14:21:02 -08:00
Adam Pocock	35ceb1a6a6	Java API for onnxruntime (#2215 )	2019-12-10 08:28:46 -08:00
Ashwini Khade	78099701b4	Add missig env variables for mac pipeline test (#2595 )	2019-12-09 21:10:43 -08:00
shahasad	41fc820f76	add path to build dir before test run (#2590 )	2019-12-09 18:58:55 -08:00
Hector Li	0ab54521f4	Temporarily exclude vgg19 test from Python backend test 1. temporarily exclude vgg19 test which comsumes too much memory, run out of memory on Upsquared device. Single test pass for vgg19, need furture investigation (#2588) 2. Update docker file to decrease the docker image size	2019-12-09 12:25:46 -08:00
shahasad	eeb28a80c0	setup java ci mac (#2570 )	2019-12-06 11:43:40 -08:00
Changming Sun	7eddac16c2	Re-enable Windows C# tests (#2564 )	2019-12-05 21:22:31 -08:00
Ryan Hill	854362cf05	Update win-x86-ci.yml (#2557 ) Fix build pipeline break	2019-12-05 18:44:12 -08:00
Ashwini Khade	281933fa1c	Fix C API tests for centos and mac (#2544 ) * change c++14 to c++11 * add ld lib path for centos * enable csharp tests on macos * fix C API test on MacOS + fix manylinux dotnet install * fix manylinux dotnet install * fix lib link	2019-12-04 18:01:35 -08:00
Xiang Zhang	3e7aaf8fa1	User/xianz/telemetry (#2458 ) * enabme telemetry * enable telemetry * set enable telemetry as default * for debugging * remove log and set disable telemetry as default back * delete private file while testing * resolve comment: mainly add license header, rename macro and update docs * rewording in privacy.md	2019-12-03 23:34:53 -08:00
stevenlix	293b15480b	Add dynamic shape support in TensorRT execution provider (#2450 ) * remove onnx-tensorrt submodule * add new onnx-tensorrt submodule (experiment) for trt6 * update engine build for trt6 * update compile and compute for tensorrt6.0 * Update tensorrt_execution_provider.cc * Update tensorrt_execution_provider.cc * Update tensorrt_execution_provider.cc * Update tensorrt_execution_provider.cc * switch to onnx-tensorrt master for TensorRT6' * Update tensorrt_execution_provider.cc * Handle dynamic batch size and add memcpy in TensorRT EP * update test cases * Update tensorrt_execution_provider.cc * update onnx-tensorrt submodule * Update Dockerfile.ubuntu_tensorrt * Update Dockerfile.ubuntu_tensorrt * Update run_dockerbuild.sh * Update run_dockerbuild.sh * Update install_ubuntu.sh * Update concat_op_test.cc * Update tensorrt_execution_provider.cc * Upgrade TensorRT to version 6.0.1.5 * Update onnxruntime_providers.cmake * Update CMakeLists.txt * Update reduction_ops_test.cc * Update install_ubuntu.sh * Update Dockerfile.ubuntu_tensorrt * Update Dockerfile.tensorrt * Update BUILD.md * Update run_dockerbuild.sh * Update install_ubuntu.sh * Update onnxruntime_providers.cmake * Update install_ubuntu.sh * Update install_ubuntu.sh * Update gemm_test.cc * Update gather_op_test.cc * Update CMakeLists.txt * Removed submodule * update onnx-tensorrt submodule * update header file * Removed submodule * add submodule onnx-tensorrt kevin's branch shape-test' * add debugging code * Update tensorrt_execution_provider.cc * Update tensorrt_execution_provider.cc * merge master * Removed submodule * update onnx-tensorrt submodule * add more changes for dynamic shapes * Update tensorrt_execution_provider.cc * update for dynamic shape * update dynamic shape processing * fix logger issue * remove submodule onnx-tensorrt * add submodule onnx-tensorrt * add env variable min_subgraph_size * remove redundency * update document * use onnxruntime::make_unique * fix multi-run issue * remove some tests to save CI build time * Add dynamic shape test * Update TensorRT-ExecutionProvider.md * Add example of running Faster R-CNN model on TensorRT EP * Add more details on env variables * update environment variables * Update tensorrt_basic_test.cc * Update model tests * Update tensor_op_test.cc * remove --use_full_protobuf * Update build.py	2019-12-03 23:18:33 -08:00
shahasad	178d059111	Setup java ci (#2528 )	2019-12-03 14:21:51 -08:00
Ashwini Khade	e32eff826c	enable nuget package testing on centos7 (#2527 ) * add centos tests to linux cpu ci pipeline * Disable failing test * use centos6 instead of centos7 * change back to centos7 * add dotnet runtime dependency * fix dotnet runtime dependencies * install dotnet sdk instead of runtimes * add more dotnet dependencies * temporary skip failing test * ix lib path * reenable failing test	2019-12-03 10:16:45 -08:00
Sreekanth Yalachigere	31ea11a696	Renaming MKL-DNN as DNNL (#2515 ) * DNNL: Moving Files to rename file names * DNNL name change * azure pipeline updated * disable ceil/dialation and enable Opset10 * disable ceil/dialation tests in Python * mlperf_ssd_resnet34_1200 disabled	2019-12-03 07:34:23 -08:00
Changming Sun	3d627362a0	Upgrade Windows CPU CI pipeline to use VS 2019 (#2519 )	2019-12-02 23:05:35 -08:00
shahasad	882f28a74b	Fix NuGet end to end tests for custom op dll (#2472 )	2019-11-25 15:26:09 -08:00
Ashwini Khade	4caf5c9c13	add additional test data set for nuget pipeline (#2448 ) * add SAS token to download internal test data for nuget pipeline * update azure endpoint * fix keyvault download step * fix variable declaration for secret group * fix indentation * fix yaml syntax for variables * fix setting secrets for script * fix env synctax * Fix macos pipeline * attempt to add secrets to windows download data * fix mac and win data download * fix windows data download * update test data set url and location	2019-11-25 13:08:03 -08:00
Changming Sun	e7131f6d12	Pull the latest image before running docker build	2019-11-22 13:48:37 -08:00
Changming Sun	6bcf95477e	Fix Windows GPU C API packaging pipeline failure (#2440 ) Fix Windows GPU C API packaging pipeline failure (#2440)	2019-11-20 14:00:37 -08:00
Changming Sun	080a0a3186	Nuget pipeline changes (#2305 ) 1. refactor the pipeline, remove some duplicated code 2. Move Windows_py_GPU_Wheels job to Win-GPU-CUDA10. We'll deprecated the "Win-GPU" pool 3. Delete cpu-nocontribops-esrp-pipeline.yml and cpu-nocontribops-pipeline.yml 4. In Linux nuget jobs, run "make install" before creating the package. So that extra RPAH info will be removed	2019-11-08 09:45:52 -08:00
Patrick Foley	151075790d	[OpenVINO-EP] Update to latest version: OpenVINO 2019 R3.1 (#2308 ) * Updates OpenVINO EP to latest version: 2019 R3.1 * Reviews fixed * Update Dockerfile.openvino * Addressed PR comments and disabled model tests temporarily * Update Dockerfile.ubuntu_openvino	2019-11-05 19:55:46 -08:00
Changming Sun	138a7f194e	Add cleanup step	2019-10-30 08:13:09 -07:00
Changming Sun	ce14b07b1c	Fix the GPU nuget pipeline failure (#2255 )	2019-10-25 13:55:38 -07:00
Tomasz Dołbniak	63acd4e89b	Adjust the nGraph EP to the newest CI test data (#2180 ) * Adjust the nGraph EP to the newest CI test data * Increase the linux pipeline timeout for nGraph	2019-10-24 16:44:03 -07:00
Changming Sun	4b62241c77	Update ONNX to 1.6.1 (#2235 )	2019-10-23 13:47:45 -07:00
Jeff	ab39f7ec99	Jeffbloo/fix dml rnn failures (#2234 ) * Address a possible cause of incorrect DML kernel registrations and re-enable tests * Re-enable DML build	2019-10-23 13:46:16 -07:00
Changming Sun	aef055ebe8	Update nuget pipeline to use CentOS6 (#2211 )	2019-10-21 17:55:36 -07:00
shahasad	fcf50ca081	Fix nuget mklml pipeline (#2204 ) * some fixes on nuget CPU pipeline * revert `d738c89536` * fix for MKLML package * fix if else	2019-10-21 08:46:28 -07:00
Pranav Sharma	69970d1f2a	Include the new Privacy.md file in all release packages. (#2200 )	2019-10-20 07:58:36 -07:00
Changming Sun	cff7879d89	Update C API pipeline to use CentOS 6 (#2198 )	2019-10-19 22:25:42 -07:00
shahasad	7efc9bdcc7	Some condition fixes on nuget pipeline, to get it green (#2195 )	2019-10-19 18:28:12 -07:00
Dmitri Smirnov	acec4b446f	Make CentOS 6 CUDA build and run (#2159 ) * Add manylinux1 source code changes * Disable a python test	2019-10-19 15:33:31 -07:00
Pranav Sharma	f8c30b8aa9	Disable DML builds for now until further investigation since the tests are very flaky. (#2194 )	2019-10-19 12:13:25 -07:00
Changming Sun	021073b5e5	Update python packaging pipelines (#2167 )	2019-10-19 07:42:54 -07:00
shahasad	35dae992f1	Fix nuget gpu ci test error (#2164 ) * fix nuget version extraction script for Gpu packages * fix cuda version in gpu end-to-end test	2019-10-18 23:01:26 -07:00
Changming Sun	00e2d1c604	update (#2140 )	2019-10-17 19:28:10 -07:00
Pranav Sharma	70e7eaf1e8	Update DML transformers with the new Graph API and re-enable DML in the GPU CI build. (#2147 )	2019-10-17 11:46:14 -07:00
Pranav Sharma	6445e7182c	Disable DML build temporarily until we fix the removal of the IsNodeOutputsInGraphOutputs Graph API. (#2144 )	2019-10-16 15:04:40 -07:00
Adrian Tsai	4090d0d0de	Add DirectML Execution Provider (#2057 ) This change adds a new execution provider powered by [DirectML](https://aka.ms/DirectML). DirectML is a high-performance, hardware-accelerated DirectX 12 library for machine learning on Windows. DirectML provides GPU acceleration for common machine learning tasks across a broad range of supported hardware and drivers. The DirectML execution provider is capable of greatly improving evaluation time of models using commodity GPU hardware, without sacrificing broad hardware support or requiring vendor-specific extensions to be installed. Note that the DML EP code was moved verbatim from the existing WindowsAI project, which is why it doesn't yet conform to the onnxruntime coding style. This is something that can be fixed later; we would like to keep formatting/whitespace changes to a minimum for the time being to make it easier to port fixes from WindowsAI to ORT during this transition. Summary of changes: * Initial commit of DML EP files under onnxruntime/core/providers/dml * Add cmake entries for building the DML EP and for pulling down the DirectML redist using nuget * Add a submodule dependency on the Windows Implementation Library (WIL) * Add docs under docs/execution_providers/DirectML-ExecutionProvider.md * Add support for DML EP to provider tests and perf tests * Add support for DML EP to fns_candy_style_transfer sample * Add entries to the C ABI for instantiating the DML EP	2019-10-15 06:13:07 -07:00
Hector Li	640f71c91b	Enable Gpu multi-device test for CUDA EP and Trt EP Enable multi-device test for GPU * Add build pipeline for TensorRT multi-GPU test * Add code to disable fp16 test if hardware architecture not supported * Add option to set the device id in onnx_test_runner for model tests	2019-10-14 11:16:34 -07:00
Changming Sun	5558b80774	clean up ubuntu docker scripts (#2103 )	2019-10-14 07:20:20 -07:00
shahasad	8803f6fff4	C# end to end test fix, and make end to end tests mandatory (#2079 )	2019-10-10 19:23:43 -07:00
Changming Sun	a314402097	Downgrade python gpu package to CUDA 10.0 (#2086 )	2019-10-10 18:31:24 -07:00
Changming Sun	ccaf692ff2	Run auditwheel for manylinux1 (#2063 )	2019-10-09 09:23:00 -07:00
Changming Sun	a00ca56ae1	Remove gcc from manylinux1 docker image (#2048 )	2019-10-08 13:49:15 -07:00
Changming Sun	e9bed8b23b	Change python packaging pipeline to use manylinux1 (#2035 ) 1. Change the python packaing pipeline to use manylinux1 2. Temporarily disable model test in the python pipeline.	2019-10-08 10:03:54 -07:00
stevenlix	544e53e24e	Update TensorRT to version 6.0.1.5 (#1966 ) * remove onnx-tensorrt submodule * add new onnx-tensorrt submodule (experiment) for trt6 * update engine build for trt6 * update compile and compute for tensorrt6.0 * Update tensorrt_execution_provider.cc * Update tensorrt_execution_provider.cc * Update tensorrt_execution_provider.cc * Update tensorrt_execution_provider.cc * switch to onnx-tensorrt master for TensorRT6' * Update tensorrt_execution_provider.cc * Handle dynamic batch size and add memcpy in TensorRT EP * update test cases * Update tensorrt_execution_provider.cc * update onnx-tensorrt submodule * Update Dockerfile.ubuntu_tensorrt * Update Dockerfile.ubuntu_tensorrt * Update run_dockerbuild.sh * Update run_dockerbuild.sh * Update install_ubuntu.sh * Update concat_op_test.cc * Update tensorrt_execution_provider.cc * Upgrade TensorRT to version 6.0.1.5 * Update onnxruntime_providers.cmake * Update CMakeLists.txt * Update reduction_ops_test.cc * Update install_ubuntu.sh * Update Dockerfile.ubuntu_tensorrt * Update Dockerfile.tensorrt * Update BUILD.md * Update run_dockerbuild.sh * Update install_ubuntu.sh * Update onnxruntime_providers.cmake * Update install_ubuntu.sh * Update install_ubuntu.sh * Update gemm_test.cc * Update gather_op_test.cc * Update CMakeLists.txt * Removed submodule * update onnx-tensorrt submodule * Add Ubuntu18.04 build option * Add Ubuntu18.04 build option * Add Ubuntu18.04 build option * Add Ubuntu18.04 build option * Remove redundency * Fix issue that it does not add memcopy node correctly if some nodes fall back to CUDA EP. e.g. after partition, there's TRT_Node -> Cuda_node (with CPU memory expected), we still need to add memcpy node between them. * update for Trt Windows build * Update onnxruntime_providers.cmake * Disable opset11 tests on TensorRT * Update pad_test.cc * Update build.py * update scripts for ubuntu18.04 * Disable warning for Windows build	2019-10-06 10:40:53 -07:00
Hariharan Seshadri	f528da35f2	Update ONNX to a newer commit (#2015 ) * Update ONNX to a newer version * PR comments	2019-10-04 19:41:00 -07:00
daquexian	e071a1249b	Android CI (#1600 )	2019-10-04 17:39:51 -07:00
Changming Sun	ace0b2ca1c	CentOS CI (#1998 )	2019-10-04 10:48:43 -07:00
Changming Sun	c86d17754a	Dockerfile for CentOS CI build (#1986 )	2019-10-03 11:46:27 -07:00
Dmitri Smirnov	d1b1cdc5c4	Replace GSL with GSL-LITE submodule and fix up refs (#1920 ) Remove gsl subodule and replace with a local copy of gsl-lite Refactor for onnxruntime::make_unique gsl::span size and index are now size_t Remove lambda auto argument type detection. Remove constexpr from fail_fast in gsl due to Linux not being happy. Comment out std::stream support due to MacOS std lib broken. Move make_unique into include/core/common so it is accessible for server builds. Relax requirements for onnxruntime/test/providers/cpu/ml/write_scores_test.cc due to x86 build. Add ONNXRUNTIME_ROOT to Server Lib includes so gsl is recognized	2019-10-01 12:43:29 -07:00
shahasad	b355193841	Add Date-time stamp in NuGet package versioning for appropriate ordering of the packages (#1951 )	2019-09-30 16:24:16 -07:00
baowenlei	9fc5598b7e	update nuphar ci llvm version and uncomment unit tests (#1954 )	2019-09-29 23:35:14 -07:00
shahasad	c3ffd1f47d	added continue on error for the linux cleanup step, to mitigate the build failure. root cause unknown (#1936 )	2019-09-26 15:57:24 -07:00
Changming Sun	09cdbe9d76	Update test data (#1932 )	2019-09-26 10:13:53 -07:00
Changming Sun	d46e023ee4	Remove toolset=14.11 for CUDA build (#1921 )	2019-09-25 15:19:37 -07:00
Pranav Sharma	052339d9dc	Fix python packaging pipeline (#1922 ) * Mention OrtCreateSessionFromArray in C API doc * Fix python packaging pipeline broken by this commit id dc03ce.	2019-09-25 14:55:17 -07:00
Hariharan Seshadri	dbff8272e7	Update ONNX to newer commit (#1907 )	2019-09-24 19:25:34 -07:00
Pranav Sharma	1f4190de3c	Disable Linux x86 builds as they're not required any more. (#1887 ) * Mention OrtCreateSessionFromArray in C API doc * Disable Linux x86 builds as they're not required any more. * more ...	2019-09-23 00:26:23 -07:00
Hariharan Seshadri	aacfa2af65	Bump up ONNX to the latest commit (#1868 ) * Initial commit * Delete unnecessary files * Update generated proto files * Update server proto file * Update submodule onnx * Update OnnxMl.cs * update OnnxMl.cs * Update OnnxMl.cs * Comment one test * Update disabled test list * Update backend tests * Formatting fix * Formatting * Disable a test * More tests updated * commit id update * Update to a newer commit * More updates * More test updates * Update * Update * Updates * Update	2019-09-20 18:15:16 -07:00
Hector Li	582a27f546	remove sudo from the cleanup step for Linux so that we don't need the sudo access for vstsagent build user 1. remove sudo from the cleanup step for Linux so that we don't need the sudo access for vstsagent build user 2. a minor fix in the install_ubuntu.sh to make the image smaller for openvino	2019-09-18 11:22:37 -07:00
Changming Sun	dc03ce0278	New OP: CDist (#1808 ) Add a new op for scikit-learn converter. It's for scikit's cdist function: https://docs.scipy.org/doc/scipy/reference/generated/scipy.spatial.distance.cdist.html Will add docs and shape-inference function later. Will convert it to an ONNX function before pushing into ONNX.	2019-09-17 10:55:31 -07:00
shahasad	aac6021549	Add NuGet feed publish to nuget pipeline (#1833 )	2019-09-13 15:27:35 -07:00
Bowen Bao	8712a523a4	Bump onnx to latest (#1756 ) * Bump onnx to latest Update onnx.in.proto with changes for SparseTensor. * add temp skip tests * remove passed tests from skip list * skip more tests for new ops in opset 11 * skip crashing tests * update handling of new attribute types sparse tensor and sparse tensors * advance onnx commit and remove skip cpu_flaky_tests * temporarily skip yolo3 model test due to resize opset10 shape inference regression * update proto for onnxruntime server * advance onnx commit further	2019-09-12 11:46:49 -07:00
Hector Li	2b8677b210	Enable Openvino nightly build on edge device (#1684 ) 1. Add openvino GPU nightly build pipeline, this test is running on Intel Up square Edge device. The device are host locally not from Azure VM. We persist a smaller model test data on Edge device. 2. Update the build condition for openvino GPU so it works for GPU_FP32, GPU_FP16 3. add option to install_ubuntu.sh to exclude the package used for nuphar, so that we can save some disk space as the Edge device usually have limited disk space.	2019-09-11 16:36:12 -07:00
shahasad	6a5b11756b	Conditionally export execution provider apis in chsarp (#1724 )	2019-09-09 11:17:44 -07:00
KeDengMS	58fe5a6bf1	Enable Nuphar docker build, and reinstate Nuphar tests (#1757 ) Enable Nuphar EP docker build Revert back to LLVM 6.0.1 Reinstate disabled Softmax tests caused by LLVM 8.0.1 Reinstate Nuphar Python test due to stale sympy version Increase build timeout of Linux CI	2019-09-05 08:50:48 -07:00
Changming Sun	94d9161166	Add nuphar to Linux CI build (#1750 )	2019-09-03 11:39:27 -07:00
shahasad	833e18345d	Publish perf tool with nightly build (#1728 )	2019-08-30 11:25:55 -07:00
shahasad	f25847bccd	More fixes on the NuGet CPU CI pipeline (#1688 ) - Fix the Windows end-to-end test in NuGet CI - Skip the TestModelSerialization, because it is failing on Linux. Must be fixed before API is released for use. Owner is notified.	2019-08-23 18:13:13 -07:00
shahasad	6f70a78e1f	Fix a few errors in the NuGet pipeline (still broken) (#1656 )	2019-08-21 15:42:23 -07:00
Dmitri Smirnov	fbd790f703	Add AutoML to 3 main builds. (#1631 ) Add AutoML to 3 main builds. Fix unit tests. Enable copy elision, do not move movable object on return by value.	2019-08-16 18:06:16 -07:00
jywu-msft	372b657900	update TRT EP CI's to use latest model.zip (#1637 )	2019-08-16 17:44:22 -07:00
shahasad	f9834105aa	removed --gen_doc (#1633 )	2019-08-16 09:52:36 -07:00
shahasad	c9eb13a638	Copy System.Numerics.Tensors sources from dotnet/corefx into onnxruntime (#1605 ) Copy System.Numerics.Tensors sources from dotnet/corefx into onnxruntime	2019-08-15 17:28:47 -07:00
Ashwini Khade	0044be6259	update onnx to latest commit (#1622 ) * update onnx to latest commit * Disable and/or fix failing tests * disable not yet implemented tests for opset 11 * disable tests * fix bug in mkldnn fp16 graph check	2019-08-15 17:10:32 -07:00
shahasad	0c5d2c998b	Generate documentation from the registered operator kernels (#1395 ) - Added python script for generating markdown doc from the registered opkernels. - Made some conditional changes in the pybind to expose necessary python API - Added some missing type-constraints in the op kernel registrations	2019-08-14 18:12:24 -07:00
Hariharan Seshadri	09db1e06b5	Make changes to pipeline template to include missing headers in tars/zips (#1617 )	2019-08-14 13:51:29 -07:00
stevenlix	1c5b15c2b8	Remove memory copy between TensorRT and CUDA (#1561 ) * remove memory copy between CUDA and TRT * add info to RegisterExecutionProvider input * use new IDeviceAllocator for trt allocator * remove SetDefaultInputsMemoryType from TRT EP * remove onnx-tensorrt 5.0 * add submodule onnx-tensorrt branch 5.1 * remove redundancy * Update transformer_memcpy.cc * Update tensorrt_execution_provider.cc * switch to TensorRT 5.1.5.0 * update python binding * disable failed test case on TensorRT * Update activation_op_test.cc * upgrade to TensorRT container 19.06 * update according to feedback * add comments * remove tensorrt allocator and use cuda(gpu) allocator * update onnx-tensorrt submodule * change ci build cuda directory name	2019-08-08 19:31:39 -07:00
Changming Sun	aeb0bcb4a3	parallel build	2019-08-07 08:38:26 -07:00
Changming Sun	65ff02fdb0	Set job timeout for code coverage pipeline to 120min(#1563 )	2019-08-06 07:49:31 -07:00
Changming Sun	7ee8aca1bf	Avoid downloading test data into C:\ (#1562 )	2019-08-05 19:53:15 -07:00
jywu-msft	8a6bfe00af	roll back model test update for ngraph provider. (#1551 )	2019-08-02 15:53:32 -07:00
daquexian	93cb29f958	[WIP] NNAPI EP Update (#1540 )	2019-08-01 22:25:56 -07:00
Hariharan Seshadri	624411bb69	Upload correct ESRP signed package (#1531 ) (#1534 )	2019-08-01 10:56:18 -07:00
Changming Sun	3045a5f88b	Update test data (#1512 ) * Update test data	2019-08-01 10:42:08 -07:00
Hariharan Seshadri	28a6f6b11b	Add back MacOS leg of the Python packaging job (#1523 ) (#1526 ) * Add MacOS leg of Python packaging job * Update copy files source directory for Mac OS leg * Add a task to display the binaries directories contents after build wheel creation * Revert some changes * Add task to log * Update * Remove unnecessary logs	2019-07-31 15:57:26 -07:00
Hariharan Seshadri	4d768b3a0f	Fix inclusion of ARM binary in the release pkg (#1513 ) (#1521 ) * Fix inclusion of ARM binary in the release pkg * Add lib and pdb as well	2019-07-31 15:57:03 -07:00
shahasad	fb5d0fc538	Publish nuget package to azure blob store (#1525 ) Publish daily build NuGet package to Azure blob store for sharing among internal partners	2019-07-31 14:17:54 -07:00
shahasad	a86486ab7f	Post binary sizes to dashboard database (#1517 ) Python script and necessary changes in the azure-pipelines yaml file to post the binary size data from NuGet package build. Currently only posted from CPU pipeline. GPU and other pipelines may be added as necessary.	2019-07-30 08:59:43 -07:00
Hariharan Seshadri	6df4bc2ebe	Update scripts to access pipeline variables correctly (#1499 ) * Update scripts to access IsReleaseBuild pipeline variable correctly * Correct access of PACKAGENAME pipeline variable * Fix Linux CUDA 10 package tests * Enable C# GPU test * Update	2019-07-25 15:30:32 -07:00
Changming Sun	a7223ed801	Fix android build (#1489 )	2019-07-25 13:00:00 -07:00
shahasad	258ff06e42	Revert "publish nuget package to azure blob (#1309 )" (#1485 ) This reverts commit `1601650161`.	2019-07-24 18:07:33 -07:00
daquexian	ec3c553501	NNAPI EP Update (#1483 ) * Update DNNLibrary * Allow fp16 by default * Add nnapi build in ci * Fix nnapi ep after #1268 * Remove unused variables * Support nnapi in onnx_test_runner * Update DNNLibrary to fix tests * Update build.py for android build support, solve conflict of tools/ci_build/build.py * Support non-ARM Android build, solve conflict of tools/ci_build/build.py * Enable android test by x86_64 android emulator * Add dnnlibrary/NNAPI support in build.py * suppress the verbose adb output * Remove debug logs * Install cmake by pip * Fix undefined host_protoc_path * cmake==3.13.2 in pypi is actually 3.12.2, so install 3.13.2.post1 instead * Fix Android ARM64 build * Use android ndk r20 instead of r19c, fix conflicts in install_deps_android.sh	2019-07-24 13:20:05 -07:00
jignparm	a8e9e1878e	Reduce artifacts size (#1477 ) * Update wildcard pattern to match only relevant archives * Update TensorRT build to add CUDA VS extensions	2019-07-23 22:23:51 -07:00
shahasad	1601650161	publish nuget package to azure blob (#1309 )	2019-07-23 11:07:35 -07:00
jignparm	b41f6eef52	Jignparm/copy cuda extensions (#1462 ) * Add CUDA extensions for v 10.0 * Add CUDA extensions for v 10.0 * update path * change 'vsts' to 'github'	2019-07-22 23:27:48 -07:00
shahasad	768ced703c	Expose provider factory C API, especially for CUDA users (#1461 ) Exposed provider factory C API, for cpu and cuda providers, into the published packages.	2019-07-22 19:03:06 -07:00
Yufeng Li	6be93f11e5	build mklml/ngraph without openmp (#1460 ) cleanup the option to build mklml/ngraph without openmp	2019-07-22 16:59:32 -07:00
Klein Hu	227734139a	Fix ORTSRV nightly build (#1440 ) * Update the build_dir * Fix indent in the model_zoo_tests.py * Remove unnecessary tests in the server build.	2019-07-21 19:12:05 -07:00
jignparm	2c05291908	Jignparm/patch 0001 (#1419 ) * remove extra $ from 8592 * fix	2019-07-20 17:07:56 -07:00
jignparm	1a957e0642	Update C-API packaging pipeline to use CUDa 10 (#1445 )	2019-07-20 14:27:43 -07:00
jignparm	9e4ac8c66a	remove mkldnn from gpu nuget package (#1443 )	2019-07-20 11:51:00 -07:00
Ke Zhang	638398e675	sync onnx to get equal op with float support (#1432 ) * sync onnx to get equal op with float support * doc update * fix test failure because of updated shape inference logic for roialign. * filter consum test cases since it's not implemented yet.	2019-07-19 13:19:09 -07:00
suryasidd	e9e777925f	[OpenVINO-EP] Added support for OpenVINO R1.1 (#1438 ) * Initial commit for OpenVINO R1 Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Fixed MO dynamic shape error Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Add debug messages for failure * Update install_openvino.sh script Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Try catch included. Return type of Isgraphsupported function changed to void * Removed error_msg variable and commented code * formatting cleanup * Added missing return statement Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Changed MO to be compatible with both R5 and R1 Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Updated docker scripts to include openvino version number Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Ignore compiler warnings from external headers * Updated dockerfiles Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Code cleanup using clang-format Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Suppress model optimizer info error Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Python code formatting using auto pep8 Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Updated documentation Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com>	2019-07-19 00:52:15 -07:00
Colin Versteeg	5ee0f185dc	Add GRPC support to ONNX Runtime Server (#1144 ) * add grpc * add-submodule * Revert "add-submodule" This reverts commit e35994b25035ce310a98909658582bff759ee358. * fix submodule * IT BUILDS * Initial commit of prediction_service_impl.cpp * Server builds and runs! * add request id, health and reflection. GRPC is done * enable channelz for monitoring * GRPC unit tests * clang format * add unit tests * Add function tests for GRPC * add grpc to model_zoo_tests * revert update protobuf to 3.7.0 * update submodules * builds but runs some gflags tests which fail * get build working * confine build changes to onnxruntime_server.cmake * update build files * code reveiw comments * Maik's code review comments * update cares version to fix compilation issue * update build to fix c-ares * code review comments * update cgmanifest.json * remove extraneous file * Klein comments. * update ci based on discussions for go dependency * fix tag issue * fix build issues * remove stray submodule * update dockerfile and build script * dynamic linking changes * update build script * code review comments * update dockerfile * update script for mount * code review comments	2019-07-18 11:10:38 -07:00
Changming Sun	fbdd905440	Switch some of the linux pipelines to use the new data download script (#1379 )	2019-07-17 16:06:02 -07:00
Yufeng Li	a7b1a8969c	simply nocontribops-ci and fix build break (#1422 ) simply nocontribops-ci and fix build break	2019-07-17 13:43:40 -07:00
Changming Sun	d38badffdb	Disable mklml in Windows Build	2019-07-16 11:09:17 -07:00
Raymond Yang	a203077dcd	Relax timeout in CI system (#1394 ) * Relax timeout in CI system (temporary) * Relax timeout on TensorRT pipeline	2019-07-15 15:10:08 -07:00
jignparm	e580b76305	Fix ARM64 build + Add NuGet pipeline including ARM binaries (#1335 ) * Add arm64 nocontribops pipeline * minor fix * Added new template for arm build -- disable all tests * fix build command * add arm64 flag for msbuild * add arm leg as upstream dependency * update platform to arm64 for msbuild * remove test task from arm build * remove ESRP signing of C# dlls in arm build * Updated to work for both --arm and --arm64 * Make the cross compiling cmake flags symmetric * Add dynamic check for /Wno-error flag, instead of extra build option * remove extra full-stop	2019-07-11 11:49:17 -07:00
Changming Sun	20f6c84fd2	Switch to use nvidia-docker2 command format	2019-07-10 13:11:07 -07:00
jignparm	57225cd4ee	Add C++ API test for NuGet package (#1364 )	2019-07-09 13:51:51 -07:00
Changming Sun	58d6ff3f13	Remove AgentPool setting in CI yaml	2019-07-08 15:40:54 -07:00
Colin Versteeg	a8ff209ab6	Refactor Onnx runtime Server to only use public APIs (#1271 ) * replace log sinks * limit headers to include dir * first changes to do dynamic linking * wip for using cxx api * remove weird dangling dependency * building with tests failing * finish updating converters * fix const * intital introduction of typedef * change logging to use spdlog * get tests passing * clang format * map logging levels better * clean up unused imports * trent cr comments * clang-format * code review comments * changing buffer use to reserve * Dynamically link * revert tvm * update binary uploading * catch exceptions by const-ref * Revert "revert tvm" This reverts commit 387676dd1018134d15eb71fa126f7caf94380800. * fix typo * update versioning of lib	2019-07-04 01:08:14 -07:00
Changming Sun	28759e2f6f	Uninstall the preinstalled cmake in tensorrt image because it's too old (#1316 )	2019-07-01 15:08:01 -07:00
Matthieu Darbois	04d581995d	Use manylinux2010 image to build linux python wheels (#1282 ) * Update cuda for python wheels * Update cuda for python wheels * Update cuda for python wheels * Update azure-pipelines-py-packaging.yml * Update to cuda 10 * Only test win gpu * Update cuda for python wheels * Use manylinux2010 image to build linux python wheels Allow wheels built to truly be compliant with a manylinux policy	2019-06-27 15:45:06 -07:00
Scott McKay	0951f53c80	Update ONNX to d94f99d21a9a0820d58966410ceaf525132f85f1 to pickup change to checker that makes ssd_mobilenet model load 20x faster by avoiding unnecessary copies. (#1307 )	2019-06-27 08:39:41 -07:00
jignparm	a56b294428	Activate compliance tasks for private builds, and also set a daily scheduler (#1280 ) * add compliance and build schedules * cpu-esrp-pipeline.yml * update schedule time for testing * add schedules to all pipelines	2019-06-25 11:13:55 -07:00
Ashwini Khade	a571ea74a6	update onnx (#1287 )	2019-06-24 14:17:27 -07:00
jignparm	d3e5474c1d	Refactor CI pipelines - add GPU NuGet pipelines and ESRP code signing steps (#1247 ) * Simplify linux gpu pipeline * Refactor win-gpu-ci-pipeline.yml * Set cuda environment variables for testing and version * Remove variables from starter script * minor fix * Add GPU Nuget pipeline * Set DisableContribOps environment variable for Linux package tests * Add ESRP tasks * Add ESRP signing templates * Test out hardcode value of ERSP * Test out hardcode value of ERSP * Test out hardcode value of ERSP * Test out hardcode value of ERSP * test variable expansion * test variable expansion * test variable expansion * test variable expansion * test variable expansion * test variable expansion * test out variable expansion * test variable expansion * test variable expansion * test variable expansion * test variable expansion * test variable expansion * test variable expansion * test variable expansion * test variable expansion * update cpu pipeline to conditionally esrp sign * Set C# GPU tests to run only if env var is set * Refactor for easy parameter passing * refactored esrp templates * remove variables from template * Add packaging variables back to pipelines * update C# for cuda 10 * Merge vars ana parameters for gpu pipeline * remove vars from mklml pipeline * display envvars on terminal * Clean up C# cuda tests, and upgrade to Cuda10 * Introduce CUDNN_PATH pipeline varaible * YAML variable are always uppercased (not true with classic) * Update C# GPU test to be more meaningful * remove macos from gpu tests * remove debugging info for DisableContribOps option * Remove DisableContrib ops parameters -- use variables only * Fix typo from = to - * remove debug steps * fix typo * remove unused variable TESTONGPU from some templates * clean up CUDA env setup scripts * Remove CUDNN_PATH from setup_env_cuda.bat	2019-06-20 19:41:30 -07:00
Raymond Yang	c96049fe4a	Update ONNX version to include new fixes/changes (#1250 )	2019-06-18 14:39:36 -07:00
RandySheriffH	a4148c85a5	call install_onnx.sh with relative path (#1225 ) * call install_onnx with relative path * use caller path * get abs path of current script * process unicode char * replace script name * add suffix as part of match * match only the end of path	2019-06-18 13:28:09 -07:00
S. Manohar Karlapalem	8d15ffd8f5	Initial commit for OpenVINO Execution Provider (#935 ) * Initial commit for OpenVINO Execution Provider OpenVINO Execution Provider provides the interface for ONNX Runtime applications to access Intel's hardware accelerators using Intel's OpenVINO Toolkit. * Fixed bug in GetCapability to disable custom ops Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Added OPENVINO ci pipeline Added new pipeline for openvino provider, made changes to support the docker build and onnxruntime build with openvino. Signed-off-by: Luis Daniel Castellanos <luis.daniel.castellanos@intel.com> * Enabled all unit tests for OpenVINO EP Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Fixed syntax issue in run_docker_build.sh file * Added missing default OPENVINO_VERSION Default value for OPENVINO_VERSION env was missing causing the build to fail * Added install Model Optimizer deps step * Fixed python unit tests and some tests from onnx_backend_test_series Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Fixed indentation bug Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Disabled some of the python backend tests for OpenVINO Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Disabled some model tests Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Remove Duplicate checks for openvino in build.py Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Modified GetCapability for FP16 Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Disabled GPU FP32 tests that are not supported Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Convert modelProto to string and use it in compile Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Pass byte-array input args to MO * Serialized ModelProto passed in-memory to MO ModelOptimizer python module receives the serialized ModelProto in-memory. Uses appropriate ONNX function to load the serialized bytes. * Make Py_Finalize compatible with older python versions Also, remove pFunc unassigned variable possibility. * Fallback if input dims of Matmul is greater than 2 Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * fixup: Device #define syntax * Updated the documentation Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Enable dynamic dim value * removed commented out code * Added Dockerfile for openvino EP Updated instructions on dockerfiles/README.md file Signed-off-by: Luis Daniel Castellanos <luis.daniel.castellanos@intel.com> * Disabled fp16_inception_v1 test Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Code formatting with clang-format Uses style from the .clang-format file in root directory. * fixup: docker tag and build error fixes * Heuristics to automatically detect batching Distributes slices from batch into parallel infer-request objects. * Handle disabled tests in GetCapability Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Disabled average pool and max pool if ceil_mode is 1 Also dilations are not supported if they are greater than 1 Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Disabled Unsqueeze int32 test Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * changes to fix output results bug * Disabled a few C++ unit tests for MYRIAD FP16 Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Manually revert '9fe162bb Enable dynamic dim value' Reverts compile time setting of dynamic shape Reverting manually due to significantly huge auto-revert conflicts. * Fixed unused variable warning Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Disabled Mul test for GPU_FP16 due to accuracy issue Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * VPU documentation update * Disabled inception_v1 for MYRIAD and HDDL Also disabled few C++ accuracy tests for HDDL Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> updates from upstream * use the new CustomOpApis for I/O interfacing * Pass initializers as subgraph meta-def inputs in GetCapability() Requirement due to API changes introduced with PR# 1019. * Remove obsolete functions * Save indexes of graph inputs from fused_node info Both inputs and initializers are passed as data inputs to the infer function. To identify only inputs among them, save thier index info from fused_node in Compile function. * Documentation changes to enable VPU * Fix VPU related changes in documentation * Fix minor changes in documentation * Fix VPU related changes in documentation * Use Node.In/OutputDefs() to track graph inputs and outputs. Don't use graph_viewer's GetInputs() or GetInputsIncludingInitializers(). * Permit "SAME_UPPER" auto_pad attribute from MaxPool * Disabled fp16_tiny_yolov2 in onnx model tests * Updated documentation to include configuration guides for myriad and hddl Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Use 8 Infer requests only for VAD-R * disable debug prints * Clang-format source files * Updated BUILD.md with OpenVINO R5 links Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Disabled same upper python tests Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Update test exclusion syntax * Change path of install_onnx.sh Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Disable tiny_yolov2 in broken tests Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Revert "Change path of install_onnx.sh" This reverts commit ba9db165f3be430f2aff1ef413299ed04637196a. This change is only required for Intel internal CI pipeline until the settings are matched with the upstream's CI pipeline. * Added debug statements for debugging CI error Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Add --build_wheel to linux openvino pipeline Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Added -v option to onnx_test_runner for debugging Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Removed path change patch Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Added -c 1 to onnx_test_runner Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Refactor MO python invocation in separate function Cleans up Model Optimizer python invocation check and conversion logic. Invokes MO only once in GetCapability() and passes the IR strings (xml and bin) to the Compiler as meta-def attributes. * Add comments * code cleanup and comments * Code cleanup for GetCapability Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Removed unnecessary files Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Revert "Added -v option to onnx_test_runner for debugging" This reverts commit d1dd70938a94d648df1a1dbbc2e48d0b97e49ec8. * Revert "Added debug statements for debugging CI error" This reverts commit b86d41afed2aa29c3508155d6f9c8d3a7263cc60. * incorporate Status Code changes * ComputeFunc returns Status::OK() on success * Use test names to disable tests for MYRIAD and VAD-R Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Rename local identifiers from CNNNetwork to OpenVINO network CNNNetwork is an OpenVINO's API class that represents more than just convolutional neural networks (CNNs). Renaming helps to avoid confusion that the API's only support CNN type models. * Added error message if building on windows * Removed duplicate option in Cmake * Removed unnecessary parameters in activation_opt_test Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Refactor Map search and access logic for efficiently and cleanliness. * use C++ style casts * Use os.path.join for python directory path operations * use C++ style casts * EP classes should use onnxruntime namespace * Clean up fixes from PR comments * Don't explicitly shutdown Py interpreter * Remove debug print statements Prints will be re-enabled later with a logging mechanism with debug/verbose printing options. * Decrement ref counts for used pyObjects * Restore build instructions for other compilers Content under the "Using other compilers" section has been accidentally deleted by a previous commit. Restoring back that content from the latest upstream repo. * CMake code cleanup Code clean up, commenting and formatting of CMake code. * Don't pass the unused device_info parameter to OpenVINOGraph ctor. * Add support for multiple I/O data types Adds support for the following tensor data types for graph inputs and outputs: 1) float 2) float16 3) int32 4) int16 5) int8 6) uint16 7) uint8 * cleanup setup.py module list definition * Deduce index of input using tracked input index map Ignores initializers in case they are ordered before inputs. * Removed debug statement in MO code Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * PR feedback * Removed per_sample_tolerance for openvino * Removed unnecessary disabled tests Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Removed debug function Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Disabled tiny_yolo_v2 due to accuracy issues Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Changed the disabled reason for broken tests Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Disabled Reshape with no input Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Python formatting with Autopep8 * Minor fix for MYRIAD devices * Added zero dimension check Removed setting batch size for the network Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> Set the threshold to larger value for MNIST Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Removed setting higher threshold in provider_test_utils Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com> * Check for --use_openvino in python wheel setup.py Add openvino modules to the setup script for building the wheel package only for --use_openvino a build option. * Removed nullptr checks for GetNode() Signed-off-by: suryasidd <surya.siddharth.pemmaraju@intel.com>	2019-06-18 08:58:53 -07:00
RandySheriffH	6d7081f33f	Add PyOp doc and UT (#1200 ) * Enable Interop Op test in ci pipelines * add doc * update doc * update doc * add supported types * format doc * update doc * update doc * update doc * update doc * update doc * update doc * update doc * remove comment * catch up with latest C API * fix cmake issue * update doc * add comments * update doc * update doc * fix compile err on mac * update doc * update doc * update doc * update doc * add sequence chart * format sequence chart * format sequence chart * update doc * update doc * update doc * update doc * update doc * update doc * fix cmake issue * fix cmake issue * fix cmake issue * fix cmake issue * fix cmake issue * fix cmake issue * fix cmake issue * add graph dependency * add graph dependency * add graph dependency * add links * enable test from yml * revert change in yml * renable test in yml * update doc	2019-06-17 11:15:44 -07:00
shahasad	5f21eedcbd	Script for uploading code coverage data to dashboard (#1209 ) - Added Python script to post the code coverage data to the MySQL table used for dashboard - Added a build job to run a windows cpu debug build on every merge on master, and run the script - Removed the code coverage step from the CI build	2019-06-15 18:33:22 -07:00
jignparm	08731589c9	Refactor CI pipelines, and add YAML NuGet package generation pipelines ( for CPU, MKLML, NoContribOps) (#1223 ) * Initial check in * Add win x86 * minor update to x86 * update win-ci * update win-ci * update win-x86ci * add linux and mac templates * add nuget pipelines and test templates * remove buildConfig * add compliance template * fix minor typos * update pool for macos * update mac agent pool * update macos pool * update agent pools for tests * turn off debug build for testing * some modifications to packaging scripts * change ordering of compliance tasks * Add mklml pipeline * Add packagename variable to mklml pipeline * remove unrequired dependent jobs from mklml pipeline * Update build command for macOS legs in mklml and cpu pipeline * Set vcvars to true * Add no contrib ops pipeline * Add no-contrib-ops pipeline * set vcvars to true for package tests * remove repetition in nuget templates * get buildarch correct * get name of test template correct * remove steps from test_all_os.yml * add parameters to test_all_os.yml * Need jobs, not steps * set envars for disablecontrib ops * add cleanup tasks and CG to package tests * fix path to cleanup script for macos * remove buildDirectory -- not needed * remove fp16tiny_yolov2 model from nocontribops tests * remove debugging info * fix individual linux pipelines to use correct template * remove unneeded bak_latest2 * increase timeout to 120 to allow for variance * turn off code coverage report	2019-06-14 14:51:03 -07:00
RandySheriffH	6850e55966	Fix install_onnx.sh issue (#1201 ) * switch from map to array to keep visiting sequence * add comment * improve comment	2019-06-13 13:37:59 -07:00
Ryan Hill	6c17567d7b	Add C++ headers to nuget package (#1218 )	2019-06-13 11:38:19 -07:00
Dmitri Smirnov	a92998c235	Uncomment ConstantOfShape tests. (#1059 ) Advance ONNX submodule to 5c51f0dbbe88ee1536f17ee7bd462b2ab3772c52 This commit in ONNX contains a fix to ConstantOfShape test data. Uncomment ConstantOfShape. Update test script, make sure exclusions are uniform.	2019-06-10 14:36:36 -07:00
shahasad	97dfd5ee21	Add code coverage (#1192 ) * added the runcoverage powershell script * updated the run coverage script. added installation to the windows CI for trying * exclude other parts of win ci * fix in the download script * fix in the download script * fix in the download script * fix in the download script * fix in the download script * fix in the download script * fix in the download script * fix in the download script * fix in the download script * added the runtestcoverage script to the pipeline * some typo fix * formatting * re-commenting previously commented block * cleaned up the powershell script * fix path in pipeline * fix path in pipeline * fixed model path * some fixes * excluded long running tests * add the publish job * uncomment other tasks * fixed excluded tests * some format correction * stopped running the test debug * try placing the tes-all at the beginning * try running the failing test only * edit run_coverage * some fix * skip onnx_model_test * Added memory size log in powershell script * try running the onnxruntime_test_all.exe separately from codecov * enable error reporting, and double memory size in powershell * corrected the set-item * remove memory resize, since we are already at max 2 GB * fixed the tvm.dll issue * added back the onnx tests in codecov. added back the regular test run * cleanup * remove * from the the module path * add junction target resolution for modules dir * remove junction-resolution * reduced tests * added target extraction for the junction paths in build machine * added the appropriate change in win ci pipeline to call the updated ps script * fix typo * added back all the tests that were disabled * try fixing the source root * cleanup and enable all tests * increase timeout for windows CPU CI due to codecoverage * templatized the code coverage steps. Conitnue on error with any codecoverage step * change quote marks	2019-06-09 22:30:41 -07:00
Changming Sun	be36385a8c	Delete docker/scripts/install_deps_x86.sh and enable onnx tests for x86 (#1191 )	2019-06-08 16:17:18 -07:00
Raymond Yang	6b586bc041	Avoid warning status in python release pipeline (#1195 )	2019-06-08 00:22:32 -07:00
Changming Sun	ccab8165eb	Delete scripts/install_ubuntu_x86.sh (#1189 ) * Delete scripts/install_ubuntu_x86.sh to reduce duplicated code	2019-06-07 15:48:52 -07:00
Klein Hu	1a86421aff	Create a syslog sink for logging in !Win32 env (#1163 ) * Create a syslog sink for logging in !Win32 env * Move syslog level logic to syslog_sink.c	2019-06-06 16:35:06 +10:00
Changming Sun	6c9d815de5	Revert "Remove openmp flag (#1140 )" (#1146 ) This reverts commit `a7137a0f9d`.	2019-05-31 18:48:14 -07:00
Changming Sun	a7137a0f9d	Remove openmp flag (#1140 ) * Remove openmp flag	2019-05-31 12:41:08 -07:00
Klein Hu	6c408c3a75	Simplify ONNX Runtime Server CI build (#1136 )	2019-05-30 17:31:41 -07:00
RandySheriffH	4757933afe	Exclude test by onnx version tag (#1073 ) * add version filter to failed tests * exclude test from backend * exclude shrink from opset 9 * fix compile err * exclude certain version of constant shape * enable flatten test * fix compile err * comment mvn test * disable constantofshape test in x86 * disable x86 test * get model version from imported opset * test linux x86 case * disable nonzero opset 10 * make mutex const * test filter by commit id * adjust substr offset * Limit test platform * remove change impacting TFModleInfo.h * refactoring * refactoring * test x86 pipeline with filter * add comment * restrict version extraction on non-win * restrict version extraction on non-win * add tag * exclude case from backend test * remove dup * remove dup * make script runnable * hard code adsolute path * refactor log * fix x86 compile err * fix x86 compile err * fix x86 compile err * sync with latest tensorrt * switch to regex * fix cpu pipeline err * test filter * disable nonzero from all versions	2019-05-30 16:19:06 -07:00
Klein Hu	4e231ad907	Split binary/symbol and then upload ortsrv nightly build to blob storage (#1120 ) * Upload ortsrv nightly build to blob storage * Fix the binary directory * Temporarily disable binary split * Split binary in build container * Update azcopy command * Update getopts * Pass blob sas url as string * Avoid binary split on Windows * Update build_server logic in build.py	2019-05-29 17:46:07 -07:00
Klein Hu	b54a292ba2	Add version and latest commit id to ONNX Runtime Server (#1078 ) * Add version and latest commit id to ORT Server * Update cmake * Change build id to build number * Use target_compile_definitions instead of add_definitions	2019-05-24 17:30:40 -07:00
daquexian	ea29a664cd	Fix android build when API<23, fix android test, update build doc and pipeline (#884 )	2019-05-24 16:26:32 -07:00
R. G. Esteves	f4a9ccae99	Enable nGraph Debug ci test. (#1000 ) * Enable nGraph Debug ci test. * nGraph doesn't work with stack trace. * Fix corrupt patch.	2019-05-23 19:58:35 -07:00
jignparm	376a8240ac	Add x86 legs to CI builds (#1005 ) * add x86 legs to ci * minor update * update platform from x86 to Win32 * remove --use_mklml from x86 build * add win x64 mklml pipeline * remove pybind and use_tvm from win x86 * Add build pipelines for --disable_contrib_ops (mac, lnx, win) * remove --gen-doc generation for x86 * set environment variables during build to disablecontribops=on * update to match aiinfra pipelines * update test data url * update mac pipeline test data * remove gen_doc from win x64 leg * update model files for nocontribops * reset win-ci-pipeline.yml * remove confidential models	2019-05-22 14:08:28 -07:00
Changming Sun	9f79ff52ba	update nuget version (#1075 )	2019-05-21 22:47:06 -07:00
Bowen Bao	a42222f9de	bump onnx version & fix conv/pool tests (#1067 )	2019-05-21 09:52:41 -07:00
jignparm	32da12491d	x86 support for C# API (#962 ) * Refactor C# to handle x86 * update run script * Add Native win x86 tests * Add native x86 tests for Linux * Update linux tests scripts to control which tests are run * update linux image name for x86 to prevent using cached image * update to not run unit python unit tests unless pybind is specified * remove --build_wheel as a core common arg. Python cannot run on x86 build * update OrtGetNumOfDimensions to OrtGetDimensionsCount in rest of C#	2019-05-20 15:48:14 -07:00
jignparm	d14e65a224	Finer control over when Python tests are run (#1023 ) * Finer control over when Python tests are run * add --build_wheel to linux pipeline, instead of run_build.sh * add --build_wheel to all ci configurations * update per review comments	2019-05-15 18:30:18 -07:00
Klein Hu	406770c484	Enable ONNX Runtime Server Model Tests (#1002 ) * Enable model tests for ORT Server in build script * Nightly build pipeline definition * Force prune docker images * Clean up containers before prune images	2019-05-13 23:08:51 -07:00
Klein Hu	f7e57a3d16	Prune containers and images (#1003 ) * `docker image prune -f`to clean up unused images * Prune exited containers	2019-05-13 16:57:17 -07:00
Changming Sun	9a4128efac	Enable model test in Mac pipeline (#990 )	2019-05-08 14:43:13 -07:00
Klein Hu	c2b412f7be	Update the ONNX Runtime Server CI pipeline setup (#986 ) * Update the ORT-SRV ci pipeline setup * Update pip package installation for server tests * Install requests package in build setup * Check if python dependencies exists before install	2019-05-08 11:37:39 -07:00
R. G. Esteves	17690355ed	Support for building ngraph EP on Windows (#978 ) * Enable windows support for ngraph ep * Slight rewording of building for windows * Streamline build of ngraph ep, disable C# bindings build	2019-05-07 14:14:50 -07:00
Ashwini Khade	f4fd36ee91	merge rel-0.4.0 into master (#959 ) * Accomodate missing optional 'axes' when 'steps' is present in Slice op (#946) * Accomodate missing optional axes when steps is present in Slice implementation * PR feedback * Update package links (#937) * Update package links * Minor fix * Update README.md * Minor edit * Update onnx commit (#949) * Update onnx commit * disable failing tests which don't have to be fixed for this release * dummy change to fix file permission * fix file permission	2019-05-03 09:07:19 -07:00
Vinitra Swamy	0b5f06b0fd	removing LLVM dependency for ubuntu tensor_rt build dockerfiles (#954 )	2019-05-02 10:41:04 -07:00
shahasad	2c46fff69a	Enable gen-doc on windows CI (#716 ) * add --gen_doc to ci_build * make gen-doc conditional to build/test step * some fix in the git diff check * some more trick on doc diff * updated for input/output * updated the contrib operator doc * fix on missing input output descriptions * fixed the problem of missing doc string, due to protobuf optimization * fix * revert last change * moved gen_doc.py to /tools/python * fixed typo	2019-05-01 14:58:21 -07:00
tmccrmck	1978b3c953	Add an HTTP server for hosting of ONNX models (#806 ) * Simple integration into CMake build system * Adds vcpkg as a submodule and updates build.py to install hosting dependencies * Don't create vcpkg executable if already created * Fixes how CMake finds toolchain file and quick changes to build.py * Removes setting the CMAKE_TOOLCHAIN_FILE in build.py * Adds Boost Beast echo server and Boost program_options * Fixes spacing problem with program_options * Adds Microsoft headers to all the beast server headers * Removes CXX 14 from CMake file * Adds TODO to create configuration class * Run clang-format on main * Better exception handling of program_options * Remove vckpg submodule via ssh * Add vcpkg as https * Adds onnxruntime namespace to call classes * Fixed places where namespaces were anonymous * Adds a TODO to use the logger * Moves all setting namespace shortnames outside of onnxruntime namespace * Add onnxruntime session options to force app to link with it * Set CMAKE_TOOLCHAIN_FILE in build.py * Remove whitespace * Adds initial ONNX Hosting tests (#5) * Add initial test which is failing linking with no main * Adds test_main to get hosting tests working * Deletes useless add_executable line * Merge changes from upstream * Enable CI build in Vienna environment * make hosting_run.sh executable Add boost path in unittest * Add boost to TEST_INC_DIR * Add component detection task in ci yaml * Get tests and hosting to compile with re2 (#7) * Add finding boost packages before using it in unit tests * Add predict.proto and build * Ignore unused parameters in generated code * Removes std::regex in favor of re2 (#8) * Removes std::regex in favor of re2 * Adds back find_package in unit tests and fixes regexes * Adds more negative test cases * Adding more protos * Fix google protobuf file path in the cmake file * Ignore unused parameters for pb generated code * Updates onnx submodule (#10) * Remove duplicated lib in link * Follow Google style guide (#11) * Google style names * Adds more * Adds an additional namespace * Fixes header guards to match filepaths * Consume protobuf * Unit Test setup * Json deserialization simple test cases * Split hosting app to lib and exe for testability * Add more cases * Clean up * Add more comments * Update namespace and format the cmake files * Update cmake/external/onnx to checkout 1ec81bc6d49ccae23cd7801515feaadd13082903 * Separate h and cc in http folder * Clean up hosting application cmake file * Enable logging and proper initialize the session * Update const position for GetSession() * Take latest onnx and onnx-tensorrt * Creates configuration header file for program_options (#15) * Sets up PredictRequest callback (#16) * Init version, porting from prototype, e2e works * More executor implementation * Adds function on application startup (#17) * Attempts to pass HostingEnvironment as a shared_ptr * Removes logging and environment from all http classes * Passes http details to OnStart function * Using full protobuf for hosting app build * MLValue2TensorProto * Revert back changes in inference_session.cc * Refactor logger access and predict handler * Create an error handling callback (#19) * Creates error callback * Logs error and returns back as JSON * Catches exceptions in user functions * Refactor executor and add some test cases * Fix build warning * Add onnx as a dependency and in includes to hosting app (#20) * Converter for specific types and more UTs * More unit tests * Update onnx submodule * Fix string data test * Clean up code * Cleanup code * Refactor logging to use unique id per request and take logging level from user (#21) * Removes capturing env by reference in main * Uses uuid for logging ids * Take logging_level as a program argument * Pass logging_level to default_logging_manager * Change name of logger to HostingApp * Log if request id is null * Update GetHttpStatusCode signature * Fix random result issue and camel-case names * Rollback accidentally changed pybin_state.cc * Rollback pybind_state.cc * Generate protobuf status from onnxruntime status * Fix function name in error message * Clean up comments * Support protobuf byte array as input * Refactor predict handler and add unit tests * Add one more test * update cmake/external/onnx * Accept more protobuf MIME types * Update onnx-tensorrt * Add build instruction and usage doc * Address PR comments * Install g++-7 in the Ubuntu 16.04 build image for vcpkg * Fix onnx-tensorrt version * Check return value during initialization * Fix infinite loop when http port is in use (#29) * Simplify Executor.cc by breaking up Run method (#27) * Move request id to Executor constructor * Refactor the logger to respect user verbosity level * Use Arena allocator instead of device * Creates initial executor tests * Merge upstream master (#31) * Remove all possible shared_ptrs (#30) * Changes GetLogger to unique_ptr * Reserve BFloat raw data vector size * Change HostingEnvironment to being passed by lvalue and rvalue references * Change routes to getting passed by const references * Enable full protobuf if building hosting (#32) * Building hosting application no longer needs use_full_protobuf flag * Improve hosting application docs * Move server core into separate folder (#34) * Turn hosting project off by default (#38) * Remove vcpkg as a submodule and download/install Boost from source (#39) * Remove vcpkg * Use CMake script to download and build Boost as part of the project * Remove std::move for const references * Remove error_code.proto * Change wording of executable help description * Better GenerateProtobufStatus description * Remove error_code protobuf from CMake files * Use all outputs if no filter is given * Pass MLValue by const reference in MLValueToTensorProto * Rename variables to argc and argv * Revert "Use all outputs if no filter is given" This reverts commit 7554190ab8e50ba6947648c2f3e2a3d4d9606ce0. * Remove all header guards in favor of #pragma once * Reserve size for output vector and optimize for-loop * Use static libs by default for Boost * Improves documentation for GenerateResponseInJson function * Start Result enum at 0 instead of 1 * Remove g++ from Ubuntu's install.sh * Update cmake files * Give explanation for Result enum type * Remove all program options shortcuts except for -h * Add comments for predict.proto * Fix JSON for error codes * Add notice on hosting application docs that it's in beta * Change HostingEnvironment back to a shared_ptr * Handle empty output_filter field * Fix build break * Refactor unit tests location and groups * First end-to-end test * Add missing log * Missing req id and client req id in error response * Add one test case to validate failed resp header * Add build flag for hosting app end to end tests * Update pipeline setup to run e2e test for CI build * Model Zoo data preparation and tests * Add protobuf tests * Remove mention of needing g++-7 in BUILD.md * Make GetAppLogger const * Make using_raw_data_ match the styling of other fields * Avoid copy of strings when initializing model * Escape JSON strings correctly for error messages (#44) * Escape JSON strings correctly * Add test examples with lots of carriage returns * Add result validation * Remove temporary path * Optimize model zoo test execution * Improve reliability of test cases * Generate _pb2.py during the build time * README for integration tests * Pass environment by pointer instead of shared_ptr to executor (#49) * More Integration tests * Remove generated files * Make session private and use a getter instead (#53) * logging_level to log_level for CLI * Single model prediction shortcut * Health endpoint * Integration tests * Rename to onnxruntime server * Build ONNX Server application on Windows (#57) * Gets Boost compiling on Windows * Fix integer conversion and comparison problems * Use size_t in converter_tests instead of int * Fix hosting integration tests on Windows * Removes checks for port because it's an unsigned short * Fixes comparison between signed and unsigned data types * Pip install protobuf and numpy * Missing test data from the rename change * Fix server app path (#58) * Pass shared_ptr by const reference to avoid ref count increase (#59) * Download test model during test setup * Make download into test_util * Rename ci pipeline for onnx runtime server * Support up to 10MiB http request (#61) * Changes minimum request size to 10MB to support all models in ONNX Model Zoo	2019-04-30 18:21:23 -07:00
Raymond Yang	01cd7eaca8	Bump up onnx version (#936 ) * bump up onnx version	2019-04-30 08:44:32 -07:00
jignparm	bb58806872	Adding versioned dlls to tar/zip packages (#928 ) * Adding versioned dlls to tar/zip packages * fix syntax error * fix version name of dylib * minor fix in the target * update pattern for versioned dylib files	2019-04-28 22:44:49 -07:00
Raymond Yang	38f1f69432	Add a temporary bypass of artifacts permission issue (#921 ) * Try using blob * Try using blob * Update working directory * Update windows-build-tools-setup-steps.yml	2019-04-26 13:34:41 -07:00
nivas-x86	ba3b82648e	ng ep update1 (#895 )	2019-04-24 10:35:26 -07:00
Changming Sun	1f066d4dc4	Update onnx (#893 )	2019-04-24 21:31:49 +10:00
Changming Sun	11806529d0	Update test data (#864 ) Add: 1. mxnet_arcface 2. tf_mobilenet_v1_1.0_224 3. tf_mobilenet_v2_1.0_224 4. tf_mobilenet_v2_1.4_224 5. tf_inception_v2	2019-04-23 13:24:24 -07:00
Hector Li	e8d722003a	Move NMS to Onnx domain (#865 ) * move files * move files * Remove NonMaxSuppression from Contrib op, move it to Onnx domain, opset 10 * move NMS out of namespace contrib * update data type in UT * update to latest onnx * white list the node test for Mod which is not implemented yet	2019-04-22 13:24:27 -07:00
nivas-x86	a4d7052aeb	Add nGraph Execution Provider (#832 ) * Add nGraph Execution Provider * feedback changes 1 * feedback2 * Feedback and upgrade nGraph * Feedback 4 * Fix CI * Disable new ops	2019-04-20 17:02:35 -07:00
Changming Sun	d78c340eac	update onnx (#861 ) * update onnx * ignore some tests	2019-04-19 10:52:47 -07:00
Changming Sun	687bac455d	Convert eigen to a submodule and update it to the latest version	2019-04-18 21:24:56 -07:00
Bowen Bao	ed0c86cd90	update onnx to fix matmul shape inference (#847 ) * update onnx to fix matmul shape inference * update onnx submodule hash in cgmanifest.json and ci scripts	2019-04-18 14:52:48 -07:00
daquexian	ac82c1f483	enable android build (#715 ) * enable android build * Add 'log' to onnxruntime_EXTERNAL_LIBRARIES * Remove cmake about header_files_test.cc * Add Android CI pipeline * Remove some ms-specific(?) ci * Fix bash error * Add execute flag for install_deps_android.sh * Add install_ubuntu_for_android.sh * Remove python in deps for android * Add comment for BUILD_ARCH * Set BUILD_SERVICE to cpu * Set BUILD_OS in run_build.sh * Fix -o bug in run_build.sh * Android -> android * Correct the android ndk location * Checkout submodules in my own azure pipelines * Revert "Remove some ms-specific(?) ci" This reverts commit 302463213480487d8944c3127a3b311c591d55c0. * Revert "Checkout submodules in my own azure pipelines" This reverts commit 1acfb6755f933e532b8312ca35bb4900a833903f.	2019-04-18 09:59:04 +08:00
Raymond Yang	2a2de42bb2	Add docker image clean script (#844 ) * Add docker image clean script * Change the command not to generate warning if no such image presents * Update linux-gpu-ci-pipeline.yml * Update linux-ci-pipeline.yml * Update azure-pipelines-py-packaging.yml	2019-04-17 11:20:41 -07:00
Ashwini Khade	07e6dfa7ab	update onnx and enable tests for qlinearconv (#840 )	2019-04-16 09:43:17 -07:00
Pranav Sharma	4b4a359943	Exclude unreferenced global data and op doc strings in the opschema object. The first causes a decrease in the binary size by at least 85k. The latter reduces resident memory size. (#823 ) * Exclude unreferenced global data and op doc strings in the opschema object. The first causes a decrease in the binary size by at least 85k. The latter reduces resident memory size. * Update onnx to incorporate my PR that fixes SetDoc compiler warnings	2019-04-15 15:57:19 -07:00
Raymond Yang	fabdbdc130	Update test retrieval following #828 (#836 ) * Enable nightly build * Update fetch file names * Fix * Update setup.py * Update run_dockerbuild.sh * Resolve comments * Update test data	2019-04-15 14:51:20 -07:00
Changming Sun	2c0b8e965e	Disable test data local cache in Linux CI pipelines	2019-04-12 22:23:16 -07:00
Raymond Yang	1936d141a7	Create nightly build for python packages (#817 ) * Enable nightly build * Update fetch file names * Fix * Update setup.py * Update run_dockerbuild.sh * Resolve comments	2019-04-11 22:06:18 -07:00
Pranav Sharma	6577c3dddf	Extract debug symbols in a separate file and strip the binary. (#811 ) * Ensure Linux binaries are built with debug info. Extract debug info out of the main binaries. Strip the main binaries. * add binutils * add uname * add binutils * remove linux portion	2019-04-11 12:02:50 -07:00
Ashwini Khade	10b113f144	update onnx to bring in quantized ops (#808 ) * update onnx + move quantized ops kernels and test to onnx + remove exp ops * update onnx * Revert "update onnx" This reverts commit 533abfc297e75473a74505fb89921ffc05c46a1c. * add generated csharp test file	2019-04-10 17:20:35 -07:00
Yufeng Li	39951f35f4	Use template windows-build-tools-setup-steps.yml in win pipelines (#794 ) 1. Update nuget restore to 4.3 for capi pipeline 2. Use template windows-build-tools-setup-steps.yml in win piplines.	2019-04-08 21:35:33 -07:00
jywu-msft	571291c323	build.sh: don't require user to set --use_full_protobuf with --use_tensorrt option. we can set it implicitly. (#780 ) * use_full_protobuf if tensorrt build option is enabled. * update BUILD.md sections on MKLDNN and TensorRT/full_protobuf option	2019-04-06 10:11:57 -07:00
Yufeng Li	ef9a4d98cb	Expose parallel execution option in C# API (#767 ) * Expose parallel execution option * delete unnesary file * add doc * update nuget retore to 4.3.0 * resolve comments * remove unnessary file * make git ignore csharp/Directory.Build.props * fix yaml config for nuget 4.3	2019-04-05 12:05:56 -07:00
Changming Sun	290112d614	Update onnx (#761 ) * update onnx	2019-04-04 10:58:45 -07:00
Ashwini Khade	8bc532bfb9	update onnx and add removed experimental ops to contrib ops (#723 )	2019-04-02 22:30:00 -07:00
Raymond Yang	6cbf5bcb04	[Minor] Enable pybind in mac build (#732 ) * Enable pybind in mac build * Add wheel build option * Add numpy installation * Add numpy installation * Update mac-ci-pipeline.yml	2019-03-28 21:46:41 -07:00
Changming Sun	f6a77617c1	update test data	2019-03-27 21:56:20 -07:00
KeDengMS	deaea702ff	Bump up cmake_minimum_required to 3.13 (#722 ) This is consistent with CI version. cmake 3.11 has issues with CUDA build in Linux.	2019-03-27 14:45:24 -07:00
Raymond Yang	c35b605b8d	Support updated opschema with functionbody (#640 ) * Update onnx * Support updated function schema in ORT * Update onnx related commit hash * Check out an older commit in ONNX * Add support for subgraph attribute * Add comments	2019-03-27 11:38:10 -07:00
Changming Sun	a26696fb0e	Enable LTO on Linux	2019-03-22 15:30:37 -07:00
stevenlix	af389593be	Add Windows CI pipeline for TensorRT (#687 ) * Update win-gpu-tensorrt-ci-pipeline.yml * Update win-gpu-tensorrt-ci-pipeline.yml * Update symbols.txt * Update CMakeLists.txt * Update build.py * Update win-gpu-tensorrt-ci-pipeline.yml * Update win-gpu-tensorrt-ci-pipeline.yml * Update win-gpu-tensorrt-ci-pipeline.yml * Update tensorrt_execution_provider.cc * Update CMakeLists.txt * Update win-gpu-tensorrt-ci-pipeline.yml	2019-03-22 14:46:57 -07:00
jywu-msft	8d782582f4	fix build_wheel option (#684 )	2019-03-22 11:18:23 -07:00
Pranav Sharma	5d452b3029	Use protobuf-lite to reduce onnxruntime.dll size. (#639 ) * Test protobuf-lite * Test protobuf-lite * Test protobuf-lite * Optimize protobuf usage for LITE_RUNTIME to reduce the binary size of onnxruntime.dll. More details can be found here https://developers.google.com/protocol-buffers/docs/proto. The reduction is significant. For commit id: 4873b452151bafe49da332aaeab639ef0318fc1ca28d728, the size reduced by ~700K; from 4873728 to 4172800. * Add LITE_RUNTIME flag in in.proto files * Fix merge conflict. * Address PR comments * Forgot to add 2 files + fix linux and gpu build errors. * Fix build errors + test failures * Fix cuda tests * Fix tensor rt build * Use full protobuf for trt * Address PR comments * Print tensor shape proto as text string for easier debugging	2019-03-21 14:06:38 -07:00
stevenlix	639aa6d97a	Update run_dockerbuild.sh (#631 )	2019-03-14 16:06:21 -07:00
stevenlix	e8b0ae8923	Trt execution provider (#382 ) * updated cmake files for trt * added trt execution provider * added trt basic test * removed trt_path action attribute * Add files via upload * Update build.py * Update trt_allocator.h * fixed issues found by reviewers * changed cast operator * added comment for custom kernel implementation * changed auto to auto& * changed to function compile APIs for TRT execution provider * changed to function compile APIs for TRT execution provider * added new DType DInt64 * adapted to the changes of onnxruntime_c_api * removed trt kernel (use function compile instead) * updated onnx-tensorrt submodule * set default memory type to TRT fused kernel * resolve merge conflict * fixed the issue that USE_CUDA conflicts with USE_TRT * construct graph by adding nodes in topological order * made changes for Windows * change buffers type * bypass HasImplementationOf check for TRT XP because TRT kernel is not registered * added domain to version info in rebuilt model proto * added trt to test option list * added DomainToVersionMap() to GraphViewer * removed Copy() * fixed broken code * format the code to clang format * used local reference to the frequently used values * fixed a couple of issues according to reviewers feedback * fixed a couple of issues according to reviewers feedback * added python binding for TRT and enable use_cuda when use_trt is on * fixed a redefinition issue * changed shared_ptr to unique_ptr on trt engines, and made a few changes required by reviewers * enabled trtexecution provider for unit tests * renamed trt to tensorrt * added tesorrt to python binding * update submodule onnx and onnx-tensorrt * made a couple of minor changes based on reviewer's feedback * added CUDA_CHECK * removed test code * fixed broken code after merge * updated onnx-tensorrt submodule * added post processing to align trt inputs/outputs with graph inputs/outputs * updated onnx submodule * added CUDA fallback for TensorRT and fixed TensorRT cmake issue * added ci pipeline for tensorrt and removed some redundent code from trt xp * fixed syntax issue * updated onnx-tensorrt submodule * fix trt build problem by: (#602) 1. Add additional /wd for debug build 2. Add io.h for additional targets 3. Bring back mb version of getopt * Update install_ubuntu.sh * Update linux-gpu-tensorrt-ci-pipeline.yml * Update linux-gpu-tensorrt-ci-pipeline.yml * Update run_build.sh * Update run_build.sh * Update run_build.sh * Update run_build.sh * fixed the issue that GetKernelRegistry returns nullptr * merged master to this branch * moved some data types to private * fixed tensorrt CI pipeline issue * customized test data for TensorRT pipeline * added onnx-tensorrt in json file and fixed an issue in ci script * added comments	2019-03-14 12:00:39 -07:00
Hariharan Seshadri	cfb08c4848	TopK op: Promote onnx to a newer commit and handle changed TopK spec for opset 10 (#611 ) * Initial commit * Nit fix	2019-03-13 10:21:58 -07:00
Ke Zhang	5bb842538d	sync onnx and maintain old version history for removed exp ops (#588 ) * sync onnx and maintain old version history for removed exp ops in onnx runtime. * update * updating to specific onnx commit - remove exp ops. * update * disable the 3 failures to push the change as it's blocking folks. * update test	2019-03-12 18:48:27 -07:00
shahasad	bf43ac41aa	fix version number for tarball packages (#600 ) * add variables for version number and git commit hash * fix typo * fix typo * some logging * some logging * some logging * some logging * some logging * some logging * some logging * some logging * some more edits to see generic scripts can print * working * fixing windows git hash * try quoted echo * fix git rev-parse * echo without quotes * removed commit hash from artifact filename, added long commit hash as a file inside * added the missing commit id parameter * fix windows pipeline * keep only win 64, others disabled * remove disabling conditions	2019-03-12 17:55:08 -07:00
Randy	b452151baf	Rashuai/restore yml2 (#604 ) * restore capi yml * format file	2019-03-12 12:27:46 -07:00
Randy	f048fc5fb0	cross compile x86 linux (#562 ) * cross compile x86 linux * fix comments * install multilib for ubuntu cross compile * remove tailing slash * fix -fPIC relocations for x86 target too * add asm make flag * fix x86 compile err * test x86 with zlib and png * Disable zlib from x86 * install x86 python header * remove cross-compiling changes * test 32bit ubuntu * add x86 ubuntu docker file * add x86 as arch parametr for docker build * config pipeline * avoid dotnet install * install cmake * skip dep install * use latest ubuntu * install latest cmake * install x86 deps * configure cmake * install ninja * correct ninja dir * apt get re2c * install onnx * set processor x86 * disable warning * skip test * disable test * disable test * find lib * fix typo * restore test * disable backend model test * disable test * fix test err * stop installing onnx * disable onnx test on x86 * restore yml * mergef with master yml * cancel needless config setting * enable x86 flag * restore all onnx tests * fix yml typo * install onnx * add back x86 flag * disable cases * disable case * disable cases * add macro to disable cases * fix typo * print platform * remove condition	2019-03-12 09:47:45 -07:00
Raymond Yang	02ad7daa8b	Add component detection (#592 )	2019-03-11 15:30:49 -07:00
Hariharan Seshadri	867eda5262	Support Windows cross-compiling for ARM(64) in ORT build scripts (#549 ) * Initial commit * More changes * More changes * More changes * More changes * PR feedback * Commiting Azure build config file * Fix build pipeline * Cleanup build dir template addition * Remove conda modules download step * PR feedback * Revert x86 arguments to as they are currently * More changes	2019-03-08 17:42:20 -08:00
Pranav Sharma	9fa7b570da	Fix publishing of Linux and MacOSX artifacts. (#579 )	2019-03-08 05:57:41 -08:00
shahasad	51273b48f6	Use cuda9 1 in c api packaging (#571 ) * use CUDA 9.1 for both linux and windows * added powershell scripts for cuda props setup/cleanup * fix yml syntax * set path to cuda9.1 bin * correct label * ad --cuda_version * added some log to browse the directory * disabled jobs other than win gpu to save some resource while testing * add msvc_toolset=14.11 * added more logs * log the props file * remove setting vcvarsall * try some modificationi n build.py * fix typo * let the config Step modify envoronment * set some more env vars manually * try reordering vcvars after cuda props copying * use single script for build and test * single line script * remove extra quote * cleanup trial changes	2019-03-08 00:48:18 -08:00
Changming Sun	d40a9f894f	Enable Component Detection (#559 ) * Enable Component Detection	2019-03-07 11:07:35 -08:00
shahasad	b247fced3b	Linux and MacOS C api packaging (#555 ) * added linux packaging template and pipeline * Update linux-packaging-pipeline.yml for Azure Pipelines * fix path seperator * update copy command for linux * fixed linux gpu artifact name, added mac build * fixed linux gpu artifact name, added mac build * fixed vmImage syntax * use 1 model at a time for macos * added onnx test on Mac CI * some refactor of the pipeline scripts * try fixing the tensorproto for x86 build * try __cdecl * try C-style cast * use ORTAPICALL * put the deleter under the namespace	2019-03-06 14:56:53 -08:00
shahasad	a4a459477a	Windows packaging build pipeline for C-api packages (CPU and GPU) (#535 ) * added packaging pipeline * Update win-ci-pipeline.yml for Azure Pipelines * Update win-ci-pipeline.yml for Azure Pipelines * Update win-ci-pipeline.yml for Azure Pipelines * Update win-ci-pipeline.yml for Azure Pipelines * Update win-ci-pipeline.yml for Azure Pipelines * Update win-ci-pipeline.yml for Azure Pipelines * Update win-ci-pipeline.yml for Azure Pipelines * Update win-ci-pipeline.yml for Azure Pipelines * put the c-api header file at root instead of under core/session * Update win-ci-pipeline.yml for Azure Pipelines * Update win-ci-pipeline.yml for Azure Pipelines * Update win-ci-pipeline.yml for Azure Pipelines * parameterize the windows build script * Update win-package-pipeline.yml for Azure Pipelines * fixed indenting * fixed indenting * fix parameter reference syntax * try using arch = amd64 for the vcvarsall * remove duplicate tasks * use vcvarsall * some more refactor * fix typo * fix typo * factored out the packaging step into a template * add x86 build to package pipeline * use amd64 for vcvars arg * added gpu pipeline. added msbuild platform param * fix the msbuild platform * use amd64 host for x86 build * use buildarch=x86 for vcvarsall * remove vcvars from setup steps * add some logging for PNG lib, and disable fns_candy demo for win32 * set allocator alignment to 32 bit for win32 compiler * disable parallel execution test for x86 * use 64 bit toolchain for x86 build * add missing -T flag for toolset * fix string delimietr in workingdirectory name for package build test step * fix gpu pipeline * make io_types test conditional * use cuda 10 instead of cuda 9.1, similar to the ci build * try some workaround on the io test * undo inadvertent local change in build.py, also reenable the io test * make all test run single threaded * blacklist few failing tests for x86 * added some log in build.py * edit build.py to disable parallel test * add the failed tests into the blacklist for win32 * add tf_pasnet_large to blacklist * change control flow for build.py onnx tests * add README, license and TPN to the package * updated build.py test sequence for parallel executor * updated onnx test flow as per review comment * add type checking log in the compare_mlvalue * fix type cast * blacklist some failed test as of now * one more blacklisted test	2019-03-05 18:12:02 -08:00
Hariharan Seshadri	1d3fcc525a	deps: update onnx to a newer commit and update test exclusions (#542 ) * Update onnx dep to a newer commit and update test exclusions * Keeping shrink excuded in c++ tests * More changes	2019-03-05 12:03:16 -08:00
Raymond Yang	f5dfbba655	Clarify numpy version requirement (#537 ) * update packaging numpy version to 1.15.0 * update version in numpy version in linux * Install numpy 1.15.0 * Finish up numpy requirement after test * Try fix * Fix ci script	2019-03-05 11:07:28 -08:00
Randy	290c472839	remove mkldnn from linux packages (#533 )	2019-03-03 20:52:27 -08:00
Pranav Sharma	2f1e883c71	Don't use mkldnn and tvm for release pkgs. (#511 ) * support non-tensor types * support non-tensor types. * support non-tensor types. * fix compilation issues * fix compilation issues * Build without mkldnn for release packages. We'll default to MLAS. * Remove tvm as well * Add openmp	2019-02-22 19:29:06 -08:00
Raymond Yang	ec8ac04f30	Update cast op to support string <-> numeric (#379 ) * Update cast kernel to support to/from string * Update namespace * Add support for literal numeric case * Update to support -INF test * Update kernel registration for cast * Update ONNX to 1.4.1 * Update registy api * Resolve some comments * Update cast kernel implementation * Resolve comments * Fixed test data in onnx * Update cast kernel implementation * Resolve PR comments * Update cast_op.cc * Update onnx commits info * Update comments	2019-02-12 10:10:56 -08:00
shahasad	8a8d1b0cea	Fix MacOS shared library build (#447 ) * try removing the --version-script * remove --no-undefined flag * remove the -rpath linker flag * remove the -rpath linker flag, including the -Wl * remove the --whole-archive flags * added -all_load -noall_load flags in place of --whole-archive and --no-whole-archive * spell correct all-load * set the MacOS specific cmake configs with if(APPLE) condition * added --build_shared_lib to mac CI	2019-02-06 15:27:37 -08:00
Raymond Yang	7cd393d697	Fix 3.7 build; Add cuda version in README (#427 )	2019-02-06 13:38:04 -08:00
Changming Sun	6ae6853519	Update test data md5	2019-01-31 17:07:25 -08:00
Changming Sun	fb7be27096	Update test dataset	2019-01-31 13:11:45 -08:00
Faith Xu	91ffb980a2	Addl TPN updates (#403 ) * Updated TPN * Update batch_norm_op_test.cc * Update ThirdPartyNotices.txt * Update ThirdPartyNotices.txt * Update readme with package links * Update README.md * Update README.md * Update README.md * Merged Ryan and TPN changes into single PR * minor fix * added mkldnn to GPU pipeline. Required by C# library as it is the default execution provider	2019-01-30 17:28:17 -08:00
Changming Sun	6fc48c60de	Add win-ci-pipeline-cg.yml	2019-01-29 20:49:55 -08:00
Changming Sun	7c0a6f3d9c	CI: Enable C# tests	2019-01-28 11:40:00 -08:00
shahasad	f94fdad861	Fixes on the dotnet end-to-end test scripts to get it running on linux (#376 ) * fixed typo in runtest.sh * some fixes * some fixes * some fixes in the runtest.sh * added test data url * fixes on the dotnet test scripts * fix on prior mistake regarding installation of apt-transport-https * added verbosity in the test run for easy debugging * updated comment in the runtest.sh	2019-01-24 13:14:29 -08:00
Scott McKay	bca8daf762	Update ONNX. Implement Scan 9 changes (#366 ) * Update ONNX version to pickup Scan spec change that adds scan_output_axes. Add logic to transpose an output - write to temporary buffer when executing subgraph - transpose temporary buffer into Scan output when execution completes Add unit tests * Update to ONNX dbf3581835e3a05716e10587511d7ab3b2cdc386 to pickup inferencing bugfix. Update test to match. * Disable some tests for opset 9 operators that haven't been implemented yet.	2019-01-24 08:10:39 +10:00
jignparm	ea816615eb	remove use_tvm from base script. Put it in yaml configuration (#363 )	2019-01-22 16:46:38 -08:00
Changming Sun	948cc03490	upgrade onnx	2019-01-17 13:10:30 -08:00
Changming Sun	6225d5fe1e	Update test data (#334 ) * update test data	2019-01-15 17:01:46 -08:00
Changming Sun	7977871740	Split build pipeline	2019-01-15 12:30:59 -08:00
Raymond Yang	0efc48a11a	Install dotnet sdk on linux ci (#320 ) * Try install dotnet sdk on linux ci * Fix install script * Add configurable os version in docker build script * Avoid use ARG in docker	2019-01-14 17:51:45 -08:00
edgchen1	34bcc92554	Added test data URL and checksum arguments to build.py. (#302 ) * Added test data arguments to build.py, modified win-ci-pipeline build. * Updated CI builds to use template tasks, added test data args, removed AZURE_BLOB_KEY uses. * Fixed up set test data step template.	2019-01-09 22:33:14 -08:00
Changming Sun	e318c7317b	update	2019-01-09 15:49:27 -08:00
Pranav Sharma	31bbb4598e	Enable tvm in CI builds. (#285 ) * Enable tvm in CI builds * Fix tvm dll path issue	2019-01-07 19:37:06 -08:00
Changming Sun	5e113661a9	Build system upgrades (#281 ) * update * runas normal user	2019-01-07 13:15:24 -08:00
Raymond Yang	ec2cf59baa	Enable building python37 packages (#283 )	2019-01-05 18:41:40 -08:00
Dmitri Smirnov	7af1887b33	Introduce basic BFloat16 runtime support (#235 ) * Add basic support for BFloat16 type. * Advance onnx submodule for bfloat16 support. * Update install_deps for linux. * Address review comments.	2018-12-21 12:40:59 -08:00
Changming Sun	ac3a081ec5	Enable release build in Windows CI pipelines (#220 )	2018-12-19 11:12:43 -08:00
Changming Sun	dc8b37f4c4	update onnx (#209 ) * update onnx	2018-12-18 14:50:28 -08:00
jignparm	5fd9024139	[WIP] Initial checking for CSharp GPU support (#176 ) * Initial checking for CSharp GPU support * Enabled C# for GPU build * Update Onnxruntime to Ort * Add runtime check for cuda dlls for windows * Update pretrained model test, for models where name!=model.onnx * lowered tolerance for float checks to pass new models * ignore extra ._resnet34v2.onnx file in pretrained test	2018-12-17 21:18:48 +00:00
Changming Sun	fb021b9002	delete azure-pipelines.yml	2018-12-13 18:38:59 -08:00
Changming Sun	0aaaf4663d	Update data download script (#171 ) * Add cache dir * update * disable csharp * update * Revert "disable csharp" This reverts commit e1a80f3272f7e7881f3081a91467756d2769fdf4. * add output * update	2018-12-13 14:46:59 -08:00
Changming Sun	618cc51754	Update onnx_backend_test_series.py (#146 ) * Update onnx_backend_test_series.py * Update BUILD.md	2018-12-12 16:25:16 -08:00
Changming Sun	1beeccb89d	Enable data cache (#144 )	2018-12-10 19:15:03 -08:00
jignparm	6bfb195184	Jignparm/csharp test difftypes (#126 ) * Added tests for C# types * added multithread c# test * Added pretrained tests * Added function to load data from PB file * added exec task for protoc	2018-12-08 01:34:08 +00:00
Pranav Sharma	c5cbdd5a55	Miscellaneous fixes (#123 )	2018-12-06 22:21:04 -08:00
Hector Li	a68f5ccfd9	Upgrade gpu build to CUDA 10 + cudnn 7.3 (#112 ) * Upgrade gpu build to CUDA 10 + cudnn 7.3 * update the yaml file for python package building * switch to the cuda9.1 docker file if the CUDA_VER is cuda9.1-cudnn7.1	2018-12-05 17:49:16 -08:00
Hector Li	4801e67104	Hecli/cuda10 upgrade (#111 ) * Add build step to remove the cuda msbuildcutomization file after build, otherwise, the cuda high version could impact the lower version build * update vs path * update the path	2018-12-05 14:58:40 -08:00
Hector Li	3dcf344f09	switch build agents to the CUDA 10 pool (#106 )	2018-12-05 12:48:02 -08:00
Raymond Yang	3b08c6665a	Split the CI pipelines (#94 )	2018-12-04 13:51:35 -08:00
Dmitri Smirnov	fbb23a9ed0	Implement StringNormalizer (#69 ) * Imlpement StringNormalizer Add mixed language tests, test case insentive path. * Create a locale on the fly. Default locale does not seem to create well. * Add CI language-pack-en to make default locale available. Catch and translate locale creation exception to make the message meaningful. * Make sure locales are configured on Ubuntu.	2018-12-04 13:47:08 -08:00
Raymond Yang	1e59b6f1c2	Minor fixes to CI definitions (#80 ) * Align win gpu ci parameters with vsts definition * Fix linux CI script to avoid docker name conflict	2018-12-03 17:35:23 -08:00
Raymond Yang	fd0d7c5fc0	Add the mac python packaging script (#72 ) * Add pipeline for building python wheels for Windows/Linux CPU and GPU * try enable mkldnn * remove mklml * Update python packaging configuration * Try macos pywheel packaging * Try removing mkldnn from mac build * Use conda in mac agents * Change to build release only * Add the mac wheel packaging list to the packaging yaml * Add mkldnn into mac wheels	2018-11-30 17:31:15 -08:00
Raymond Yang	cb1781927f	Remove mklml in Linux python wheel packagaing (#53 )	2018-11-28 17:37:49 -08:00
Shah Asaduzzaman (ASAD)	6e2a1ceb41	merged from master	2018-11-28 01:29:08 -08:00
Raymond Yang	1b3efc36c1	Add pipeline for building python wheels (#41 ) * Add pipeline for building python wheels for Windows/Linux CPU and GPU * try enable mkldnn * remove mklml * Update python packaging configuration * Add python3.7 support * Revert to disable the py37 packaging on windows	2018-11-27 20:02:41 -08:00
Pranav Sharma	3f4589ced5	Add remaining build options and make minor changes in documentation (#39 ) * Minor changes in documentation * Synchronous, not sync * Add remaining build options after mkldnn fix	2018-11-27 19:59:40 -08:00
Shah Asaduzzaman (ASAD)	d7d43bc13f	merged from master	2018-11-27 18:54:06 -08:00
Raymond Yang	047ca7e57e	Merge branch 'master' into ryanunderhill/hide_leak	2018-11-27 16:21:46 -08:00
Raymond Yang	df1d01f853	Update CI configs to test mkldnn	2018-11-27 15:47:08 -08:00
Changming Sun	408fd21a7f	Windows CI: enable pybind (#34 )	2018-11-27 08:13:31 -08:00
Shah Asaduzzaman	930145f2d2	add mkldnn and openmp on windows CI build. C# package binary includes mkldnn	2018-11-26 11:27:47 -08:00
Shah Asaduzzaman (ASAD)	d8cf3f0e33	enabled shared-lib and csharp build in windows CI	2018-11-25 23:55:09 -08:00
Hector Li	d2a8b0e65b	update the cuda cudnn version information	2018-11-20 11:54:32 -08:00
Hector Li	c2d5b7e368	update linux gpu to use cuda 9.1 + cudnn 7.1.2	2018-11-20 11:39:06 -08:00
Hector Li	8f716cb4fc	Switch GPU build to use CUDA 9.1 + cudnn 7.0.5	2018-11-20 10:31:13 -08:00
Raymond Yang	53ce3e7b55	Enable Mac pipeline	2018-11-19 22:50:26 -08:00
Pranav Sharma	89618e8f1e	Initial bootstrap commit.	2018-11-19 16:48:22 -08:00

... 5 6 7 8 9 ...

640 commits