onnxruntime

mirror of https://github.com/saymrwulf/onnxruntime.git synced 2026-06-01 23:30:35 +00:00

Author	SHA1	Message	Date
Brian Martin	5780b864a1	Brianma/windowsai fi (#2475 ) * update dockerfiles/README (#2336) * Make elementwise op run 4 items per thread (#2335) Description: Describe your changes. Make elementwise op run 4 items per thread unroll for loop to leverage ILP remove unnessary N==0 check inside elementwise GPU kernel Motivation and Context Why is this change required? What problem does it solve? It can improve the performance of GPU elementwise ops. ~2% performance gain on popular NLP bert model. If it fixes an open issue, please link to the issue here. * Add CUDA GatherElements kernel (#2310) * Updates * Update test * Update * Updates * nits * PR feedback * Update * Update * PR feedback * PR comments * Update * Fix build * Fix build * Nits * Fix * Layer Normalization Fusion (#2319) basic layer normalization transform * Add FastGelu Cuda Op for Gelu and Add bias fusion (#2293) * Add FastGelu cuda op * Add AddBiasGelu for experiment * Revert "Add AddBiasGelu for experiment" This reverts commit 5c1ee019858c657e6bb75887265cb85675626e5b. * Add bias * Add unit tests * update comment * update script * fix build error * update coding style * update for CR feedback Enable half2 optimization only when cuda arch >= 7.0 * move _Tanh to common.cuh * implement CPU contrib OP Attention (#2333) * Remove unused initializer from GraphProto as well as name_to_initial_tensor_ in CleanUnusedInitializers. (#2320) * Remove unused initializer from GraphProto as well as name_to_initial_tensor_ in CleanupUnusedInitializers. This means initializers that have been replaced during graph optimizations are not left in the GraphProto when we save an optimized model. * Handle edge case where a model has an unused initializer with matching graph input by also removing the graph input. * Use non-const iterators in std::find_if calls to make centos build happy. * Nuget pipeline changes (#2305) 1. refactor the pipeline, remove some duplicated code 2. Move Windows_py_GPU_Wheels job to Win-GPU-CUDA10. We'll deprecated the "Win-GPU" pool 3. Delete cpu-nocontribops-esrp-pipeline.yml and cpu-nocontribops-pipeline.yml 4. In Linux nuget jobs, run "make install" before creating the package. So that extra RPAH info will be removed * Cuda Reverse Sequence Op, maping types of same size using same template function. (#2281) * Set ElementType to String type of node metadata, instead of byte[] (#2348) * Set ElementType to String type of node metadata, instead of byte[] * Fix spacing * Introduce PrimitiveType into a Type System along with an integer constant (#2307) Improve perf by avoiding GetType<T>() calls. Introduce MLTypeCallDispatcher to switch on Input Type. Add Tensor IsType<T>() fast method. * Fix/test dim value of 0 handling in a couple of places (#2337) * Update the CUDA Where implementation broadcasting logic to handle a dim with value of 0. Add unit test Also add unit test for unary op with dim value of 0 * Exclude ngraph from Where test with 0 dim. * Openvino EP R3.1 onnxrt server (#2357) * onnxrt server with OVEP * onnxrt server with OVEP * Update Dockerfile.server.openvino * onnxrt server OVEP fix reviews * onnxrt server OVEP fix reviews * Implement cuda nonzero op. (#2056) Implement cuda nonzero op. * Direct use python numpy array's memory if already contiguous. (#2355) * Direct use python numpy array's memory if already contiguous. This could greatly improve performance for session with large input, like big image 1920x1080 fastrcnn, 30~40% speed up could be achieved. * Add test case enforce contiguous/non-contiguos numpy array as inputs. * Add helper to create output to minimize binary size. (#2365) Add ConstEigenTensorMap typedef so we don't unnecessarily const_cast the const input Tensor. * fix builds enabling onnxruntime_DEBUG_NODE_INPUTS_OUTPUTS (#2369) * fix builds enabling onnxruntime_DEBUG_NODE_INPUTS_OUTPUTS * update * Add Tracelogging for profiling (#1639) Enabled only if onnxruntime_ENABLE_INSTRUMENT is ON * test bidaf with nuphar for avx target (#2370) increase nuphar test coverage a bit * Fix a bug in TLS refcount that may destabilized CUDA CI (#2374) * update output size calculation for resize (#2366) * change how output size is calculated for resize op * add tests for ver 10 resize * Extend OneHot CPU kernel to support more types (#2311) * Extend OneHot CPU kernel to support input int64_t, depth int32_t, output float * Skip BERT before the test data fix is picked up * Fix bug with Slice. Need to pass in flattened input dimensions so the initial offset into the input is calculated correctly. (#2372) * Add opset 11 version of Split to CUDA ops (#2376) Organize the CUDA ops definitions so all the opset 10 and 11 parts are together (same setup used for CPU ops) * Layer Norm Fusion Fix (#2379) * layer norm fusion fix * Add input shape check in code and unit tests * Fuse Add + Gelu (#2360) Implement the transformer to fuse add + gelu Implement the accurate kernel * Skip layer norm transform (#2350) * skip layer normalization transformer * Another try to stabilize CUDA CI (#2383) The root cause seems to be failure in CUDA dealloc when tear down. cudaFree return code was ignored before, so should the debug check. * fix BUILD.md typo (#2375) build.py: error: argument --config: invalid choice: 'RelWithDebugInfo' (choose from 'Debug', 'MinSizeRel', 'Release', 'RelWithDebInfo') * Fixed compilation with ngraph (#2388) * Fix reuse logic in allocation planner. (#2393) * Fix reuse logic in allocation planner. * PR comments * Add helpful comments * Don't allow reuse across string tensors. * [NupharEP] Multiple optimizations (#2380) Fuse transpose into MatMul Implement Pow and constant scalar simplification Vectorize ReduceMean Improve symbolic shape inference Minor updates for better debugging in fused function name * Avoid using the default logger in the graph lib and optimizers (#2361) 1. Use the session logger if it is available. 2. Don't disable warning 4100 globally. We should fix the warnings instead of disabling it. * Change CUDA implementation of Transpose to support all fixed size tensor types (#2387) * Change CUDA implementation of Transpose to not use a typed kernel so we can support more types with minimum binary size. Add support for 8, 16, 32 and 64 bit types. Add unit tests. Add method so the implementation can be called directly (will be used by CUDA Scan very soon). * Disable TensorRT for MLFloat16 and int8 unit tests. * Address PR comment and add support for calling cublas implementation if type is mlfloat16. * Add opset 11 versions of the existing CUDA operators that had negative axis support explicitly added. (#2398) * Add opset 11 versions of the existing CUDA operators that had negative axis support explicitly added. * [NupharEP] force some low/zero cost ops to be inlined (#2409) * fix cross compile bug (#2415) * Minor optimization: if a node has already been placed, there's no need to find a kernel for it. (#2417) * Add Reshape Fusion (#2395) * Add reshape fusion * Add some comments * update comments * update comment format * update according to feedback * update for recent logger change * fix build error * (1) Support both input and output edges in find path in graphutils (2) Add a test case of only one constant initializer of Concat input. (3) Refactor ReshapeFusion class to allow add more subgraph fusion in the future. * fix error * (1) loose constraint on initializer: non constant is allowed for reshape fusion. (2) Change versions type to vector. (3) Add logging. (4) Return false when multiple output edges matched in FindPath. Add comments. * only allow one direction (input or output) in FindPath * [NupharEP] Update notebook and docker image (#2416) Add BERT squad in Nuphar tutorial Enhance speed comparsion readability * Fix the issue in matmul_add_fusion (#2407) Fix the issue in matmul_add_fusion If Muatmul + Add has shape [K] * [K, N], reset it to [1, K] * [K, N] will make the output shape to [1, N] will also requires a reshape on the output. Fix: just remove the shape reset to not fuse it. Add a negative test case for matmul+add fusion * feat(treeregressor): Update TreeEnsembleRegressor for type support (#2389) Updates the `TreeEnsembleRegressor` to allow for `double`, `float`, `int64`, and `int32` inputs to match the upstream specification. Signed-off-by: Nick Groszewski <nicholas.groszewski@capitalone.com> * onnxrt server documentation update (#2396) * Added support for Pad-2 operator in OpenVINO-EP (#2405) * Add CUDA If operator. (#2377) * Add CUDA If operator. Uses CPU operator for implementation. By adding a CUDA version the inputs/outputs (with the exception of the 'cond' input) stay on GPU, and no other logic is required to avoid a copy to CPU across the control flow node. * Improved documentation for onnxruntime::utils::SwapByteOrderCopy(), added precondition check. * Fix the type constraints on CUDA If operator to exclude strings. (#2431) * add Im2col<uint8_t> (#2438) * Adjust codegen vectorization width from target (#2439) * Adjust codegen vectorization width from target * Add CUDA Scan operator. (#2403) * Add Scan CUDA op. Uses CPU implementation for logic. Added some device specific functors for handling when data needs to be manipulated on a different device. Added ability to override the materialization logic in the OrtValue slicer so DML can plugin their handling. * Fix Windows GPU C API packaging pipeline failure (#2440) Fix Windows GPU C API packaging pipeline failure (#2440) * Correctly handle implicit inputs for fused nodes (#2390) * Correctly handle implicit inputs for fused nodes Previously, nuphar's partitioning function didn't include node's implicit inputs into the inputs list of MetaDef, and hence a crash was triggered in the onnx graph checker. This commit fixed the issue. Furthermore, it also fixed a related issue where we didn't add implicit inputs into graph_inputs_excluding_initializers_ in Graph::SetGraphInputsOutputs. the issue was that graph_inputs_including_initializers_ populated by SetInputs (e.g. called by FunctionImpl::FunctionImpl) may contain implicit inputs which were not of any node's initializers in the graph. Because they were not part of any initializers, these implicit inputs couldn't be visited by going through all nodes' inputs. Consequently, they would not be added into graph_inputs_excluding_initializers_. We fixed the issue by first copying the populated graph_inputs_including_initializers_ into graph_inputs_excluding_initalizers_, which then had both initializers and non-initializers as its initial content. Later, we erase initializers from the list. In this way, we can ensure all implicit inputs to remain in graph_inputs_excluding_initializers_. * refined comments and fixed duplicates Address CR by revisiting comments in terms of implicit inputs Also fixed an issue by skipping duplicates while copying inputs from graph_inputs_including_initializers_. * address CR explain why we need to collect nodes' implicit inputs * don't rely on pointer values for iterating std::set Previously, openvino relied on iterating a set of NodeArg pointers to construct inputs and outputs for a fused graph. It could cause non-determinism. The reason was that although iterating std::set by itself is stable, pointer values of NodeArgs may vary. Consequently, we could end up visiting the set's elements in different orders for different runs for the same test, which resulted in constructing inputs (and outputs) with different orders to the fused graph. For example, for the same test, we may have inputs [A, B] in some runs but inputs[B, A] in others. Let's use std::string as the key type to avoid such nondeterminism. This commit also added implicit inputs into meta->inputs while returning the capability from the openvino provider. * Fixed another latent issue in openvino's GetCapability function The issue was that we couldn't simply erase fused_inputs and fused_outputs while iterating the nodes. For example, an output NodeArg may have multiple uses, and it's wrong if we erase it from fused_outputs when we encounter only one of its uses as input. * Remove DeviceAllocatorRegistry class (#2451) Remove DeviceAllocatorRegistry class * CSharp api and test for loading custom op shared library (#2420) - Added C-API test for loading custom op shared lib. - Made some changes in C++ api header and C-api implementation to get it working. - Added C# API and corresponding test for loading custom op shared library. * Parallel Gelu with ParallelFor (#2399) Parallel Gelu to get better performance for Gelu * Clean up build.py (#2446) * Pull the latest image before running docker build * Fuse SkipLayerNorm with Bias (#2453) Fuse SkipLayerNorm with Bias * Allow more than one invocation of CreateEnv in the same process. (#2467) * Allow more than one invocation of CreateEnv in the same process. * Fix centos build * Symbolic shape inference improvements: (#2460) * Symbolic shape inference improvements: - add a mode to guess unknown ops' output rank - add support for GatherND - add support for If - fix a bug in get_int_values when then tensor rank > 1D, by treating it as no sympy data - add symbol to literal merge when ONNX silently merges dims - fix a bug in Concat when input dim is 0 - fix a bug in ConstantOfShape that computed dim is not updated - add support for dynamic shape in ConstantOfShape - fix a bug in Loop output shape that loop iterator dim is not inserted at dim 0 - add support for dynamic padding in Pad - add support for dynamic shape in Reshape - add support for Resize with opset > 10, by treating output dims as dynamic - fix a bug in Slice when starts/ends are dynamic - restrict input model to opset 7 and above - make output model optional to avoid disk write when testing Run model tests for symbolic shape inference Reduce 2GB docker image size of nuphar * add additional test data set for nuget pipeline (#2448) * add SAS token to download internal test data for nuget pipeline * update azure endpoint * fix keyvault download step * fix variable declaration for secret group * fix indentation * fix yaml syntax for variables * fix setting secrets for script * fix env synctax * Fix macos pipeline * attempt to add secrets to windows download data * fix mac and win data download * fix windows data download * update test data set url and location	2019-11-25 15:20:53 -08:00
Changming Sun	2172a9e5ed	Fix an issue in the nuget run tests scripts	2019-10-30 08:13:09 -07:00
pulkittomar	1fa956fb3f	Undo integration test skip (#1917 )	2019-10-27 09:47:31 -07:00
shahasad	6a0ee7eff6	Fix model path marshalling in csharp, and re-enable the pretrained model tests (#2236 )	2019-10-24 20:39:16 -07:00
Ryan Hill	7494500221	Fix csharp CXX sample (#2251 )	2019-10-24 15:47:51 -07:00
shahasad	35dae992f1	Fix nuget gpu ci test error (#2164 ) * fix nuget version extraction script for Gpu packages * fix cuda version in gpu end-to-end test	2019-10-18 23:01:26 -07:00
Ashwini Khade	ecf5ae8b76	Askhade/disable csharptests (#2172 ) Disable flaky c# test For agility	2019-10-18 11:00:50 -07:00
Ashwini Khade	5eb4e81f80	move some optimizers to level1 (#1566 ) * move some optimizers to level1 * move matmul add fusion to level 1 * bug fix in the test code * fix make_uniques + add test exceptions * add exception for tests in c# too	2019-10-18 09:29:31 -07:00
shahasad	7ef02f14d2	Add missing test model file for symbolic dimensions (#2123 )	2019-10-15 06:55:51 -07:00
Hariharan Seshadri	95ab5ad39f	Support non-spatial mode in BatchNormalization (#2092 ) * Initial commit * Update * Update * Fix build break * Update * More changes * Update type * Exclude Nuphar for non-spatial tests * Update * Resolve PR comments	2019-10-14 18:14:14 -07:00
Hariharan Seshadri	80d09f0c59	Allow creation of empty tensors in c# (#1976 ) * Allow creation of empty tensors in c# * Keep test with updated behavior * Add more empty tensor tests * Nits	2019-10-14 14:47:02 -07:00
Pranav Sharma	91db840b6b	Introduce execution mode enum for clarity and extensibility; Change Python, C and C# APIs accordingly; Removed EnableSequentialExecution, DisableSequentialExecution in favor of the more general SetExecutionModeAPI. (#2098 ) * Introduce execution mode for clarity and extensibility; Change Python APIs accordingly; Replace DisableSequentialExecution API with EnableParallelExecution for clarity. * Fix cuda build * Modify the test slightly * Make C and C# APIs consistent with Python.	2019-10-14 09:48:19 -07:00
Scott McKay	eb24617d2e	Add ability to get symbolic dimension info for graph inputs and outputs. (#2051 ) * Add ability to get symbolic dimension info for graph inputs and outputs. WIP to get initial feedback. * Fix linxu build error. Update C# API and add unit test * Clarify the two different ways Tensor shape and type info is created. One is from concrete values and one is from a type proto where symbolic dimensions may exist. Doing so allows a change to default to empty strings for the symbolic dimensions if not provided.	2019-10-12 15:46:28 -07:00
Ryan Hill	e8e33977da	Ryanunderhill/customop dll (#2002 ) * Add OrtApiBase * Add RegisterCustomOpsLibrary API	2019-10-11 11:12:51 -07:00
shahasad	8803f6fff4	C# end to end test fix, and make end to end tests mandatory (#2079 )	2019-10-10 19:23:43 -07:00
shahasad	b70fc34fae	Fix C# end to end tests in NuGet pipeline, failing for missing test data file	2019-10-07 20:14:20 -07:00
shahasad	b0feaef9de	Update the C# pretrained model test to include opset9 and 10 models (#2003 )	2019-10-07 19:14:34 -07:00
shahasad	b322e072b9	added the overridableinitializers api (#1977 )	2019-10-04 16:38:00 -07:00
shahasad	b355193841	Add Date-time stamp in NuGet package versioning for appropriate ordering of the packages (#1951 )	2019-09-30 16:24:16 -07:00
Ryan Hill	7e22ed41b9	Fix sample tests (#1926 )	2019-09-26 10:31:48 -07:00
yeohan	034aa80167	InferenceSession ctor with byte array in C# (#1883 ) * add ctor overloads that accept model byte array * doxygen. mark Init method as private. * doxygen * rename test method for clarity * PR feedback - add two overloads that accept either model path or model byte array * update native signature to align with latest codebase * fix native call	2019-09-24 11:59:04 -07:00
Hariharan Seshadri	aacfa2af65	Bump up ONNX to the latest commit (#1868 ) * Initial commit * Delete unnecessary files * Update generated proto files * Update server proto file * Update submodule onnx * Update OnnxMl.cs * update OnnxMl.cs * Update OnnxMl.cs * Comment one test * Update disabled test list * Update backend tests * Formatting fix * Formatting * Disable a test * More tests updated * commit id update * Update to a newer commit * More updates * More test updates * Update * Update * Updates * Update	2019-09-20 18:15:16 -07:00
Ryan Hill	5781222456	Ryanunderhill/api interface (#1855 ) * Convert ABI to a versioned interface. * Convert ORT_THROW_ON_ERROR to inline function to fix link errors.	2019-09-20 13:39:11 -07:00
Pranav Sharma	a9ce941579	Refine threading control options and move inter op thread pool to session state. (#1841 ) Description: Refine threading control options and move inter op thread pool to session state. Added thread_utils.h/cc to centralize the decision around the thread pool size under various conditions. Motivation and Context Currently the thread pool size of the parallel executor is hardcoded to 32 for some reason. This PR makes the options to configure the thread pool sizes clearer.	2019-09-18 22:36:23 -07:00
shahasad	6e4e764146	upgraded CSharp test and sample projects to netcoreapp2.1 (#1869 )	2019-09-17 21:35:04 -07:00
Pranav Sharma	f8c3442880	Part 2 of renaming AllocatorInfo to MemoryInfo. (#1804 ) * Mention OrtCreateSessionFromArray in C API doc * Part 2 of renaming AllocatorInfo to MemoryInfo. * pr comments * fix comment	2019-09-12 08:19:29 -07:00
shahasad	6a5b11756b	Conditionally export execution provider apis in chsarp (#1724 )	2019-09-09 11:17:44 -07:00
Pranav Sharma	52fe574fed	Rename OrtAllocatorInfo to OrtMemoryInfo to make it more obvious. (#1758 ) * Mention OrtCreateSessionFromArray in C API doc * Rename OrtAllocatorInfo to OrtMemoryInfo to avoid confusion	2019-09-05 14:20:37 -07:00
shahasad	f25847bccd	More fixes on the NuGet CPU CI pipeline (#1688 ) - Fix the Windows end-to-end test in NuGet CI - Skip the TestModelSerialization, because it is failing on Linux. Must be fixed before API is released for use. Owner is notified.	2019-08-23 18:13:13 -07:00
Pranav Sharma	4035fe842e	Don't create the default allocator every single time. Rename API accordingly. Expose Session/Run log severity levels. (#1615 ) * Mention OrtCreateSessionFromArray in C API doc * Don't create the default allocator every single time. Rename API accordingly. * Don't create the default allocator every single time. Rename API accordingly. * updates... * updates... * PR comments * fix typo in license header * fix build	2019-08-23 10:33:20 -07:00
shahasad	a818740d91	Support Tensor<bool> and Tensor<Int8> in C# API. Support Tensor<string> as input. Fix a bug in the InferenceSession Run() with RunOptions (#1671 ) - Support bool-Tensor and int8-Tensor in input-output of C# api - Support string-tensor as input in C# api - Fix a bug in InferenceSession.Run() -- RunOptions was not passed into the native call	2019-08-22 10:14:50 -07:00
Changming Sun	224dde7ef1	Allow user disable multiple threading (#1647 )	2019-08-19 18:12:39 -07:00
shahasad	c9eb13a638	Copy System.Numerics.Tensors sources from dotnet/corefx into onnxruntime (#1605 ) Copy System.Numerics.Tensors sources from dotnet/corefx into onnxruntime	2019-08-15 17:28:47 -07:00
Pranav Sharma	8d12ce45cf	Use a friendly enum for graph optimization level. (#1586 ) * Mention OrtCreateSessionFromArray in C API doc * review changes * use enum for graph optimization level * Use explicit values for enums * updates... * Add friendly enum for graph optimization levels in C, C# and Python APIs. * Fix linux build * Fix build breakage due to master merge * PR comments	2019-08-14 17:12:08 -07:00
shahasad	a6a5acedda	Cleanup csharp API SessionOptions and RunOptions to be consistent with other APIs (#1570 ) - Updated SessionOptions API to use properties instead of setter/getter methods. - Added missing APIs. - Added RunOptions.	2019-08-14 12:02:02 -07:00
pulkittomar	a50a63aa9e	Serialize optimized onnx model (#1470 ) * Model serialization * Removed duplicate symbol * Minor update * Review comments * add tests * Model serialization * Removed duplicate symbol * Minor update * Merged PR 1106437: Model Serialization in onnxruntime * Review comments * Merged PR 1107226: Review comments Review comments * add tests * Fixed merge conflict * Correct python tests * InferenceSesssion Refeed Test * Replace use of widechar const literal-L * Fixed failing tests * Updated comment * Removed unnecessary session options * Spell check on comments * Do not serialize when level 3 optimization specified * Updated error logs * Changed log severity to WARN	2019-08-12 18:43:40 -07:00
Hariharan Seshadri	6df4bc2ebe	Update scripts to access pipeline variables correctly (#1499 ) * Update scripts to access IsReleaseBuild pipeline variable correctly * Correct access of PACKAGENAME pipeline variable * Fix Linux CUDA 10 package tests * Enable C# GPU test * Update	2019-07-25 15:30:32 -07:00
Pranav Sharma	4cbc6e1cf5	Validate input shapes. (#1352 ) * Validate input shapes. * Cache some input def metadata * Make some methods const and check for negative values of dims instead of just -1. * Fix shape inferencing test. * Fix testLabelEncoder test * Fix more tests * Fix more tests * Use size_t for loop variable	2019-07-19 13:42:34 -07:00
jignparm	57225cd4ee	Add C++ API test for NuGet package (#1364 )	2019-07-09 13:51:51 -07:00
xkszltl	98ea675e40	Fix typo: op[s]iops -> op[t]ions. (#1329 ) Resolve https://github.com/microsoft/onnxruntime/issues/1322	2019-07-01 21:25:38 -07:00
jignparm	d3e5474c1d	Refactor CI pipelines - add GPU NuGet pipelines and ESRP code signing steps (#1247 ) * Simplify linux gpu pipeline * Refactor win-gpu-ci-pipeline.yml * Set cuda environment variables for testing and version * Remove variables from starter script * minor fix * Add GPU Nuget pipeline * Set DisableContribOps environment variable for Linux package tests * Add ESRP tasks * Add ESRP signing templates * Test out hardcode value of ERSP * Test out hardcode value of ERSP * Test out hardcode value of ERSP * Test out hardcode value of ERSP * test variable expansion * test variable expansion * test variable expansion * test variable expansion * test variable expansion * test variable expansion * test out variable expansion * test variable expansion * test variable expansion * test variable expansion * test variable expansion * test variable expansion * test variable expansion * test variable expansion * test variable expansion * update cpu pipeline to conditionally esrp sign * Set C# GPU tests to run only if env var is set * Refactor for easy parameter passing * refactored esrp templates * remove variables from template * Add packaging variables back to pipelines * update C# for cuda 10 * Merge vars ana parameters for gpu pipeline * remove vars from mklml pipeline * display envvars on terminal * Clean up C# cuda tests, and upgrade to Cuda10 * Introduce CUDNN_PATH pipeline varaible * YAML variable are always uppercased (not true with classic) * Update C# GPU test to be more meaningful * remove macos from gpu tests * remove debugging info for DisableContribOps option * Remove DisableContrib ops parameters -- use variables only * Fix typo from = to - * remove debug steps * fix typo * remove unused variable TESTONGPU from some templates * clean up CUDA env setup scripts * Remove CUDNN_PATH from setup_env_cuda.bat	2019-06-20 19:41:30 -07:00
jignparm	08731589c9	Refactor CI pipelines, and add YAML NuGet package generation pipelines ( for CPU, MKLML, NoContribOps) (#1223 ) * Initial check in * Add win x86 * minor update to x86 * update win-ci * update win-ci * update win-x86ci * add linux and mac templates * add nuget pipelines and test templates * remove buildConfig * add compliance template * fix minor typos * update pool for macos * update mac agent pool * update macos pool * update agent pools for tests * turn off debug build for testing * some modifications to packaging scripts * change ordering of compliance tasks * Add mklml pipeline * Add packagename variable to mklml pipeline * remove unrequired dependent jobs from mklml pipeline * Update build command for macOS legs in mklml and cpu pipeline * Set vcvars to true * Add no contrib ops pipeline * Add no-contrib-ops pipeline * set vcvars to true for package tests * remove repetition in nuget templates * get buildarch correct * get name of test template correct * remove steps from test_all_os.yml * add parameters to test_all_os.yml * Need jobs, not steps * set envars for disablecontrib ops * add cleanup tasks and CG to package tests * fix path to cleanup script for macos * remove buildDirectory -- not needed * remove fp16tiny_yolov2 model from nocontribops tests * remove debugging info * fix individual linux pipelines to use correct template * remove unneeded bak_latest2 * increase timeout to 120 to allow for variance * turn off code coverage report	2019-06-14 14:51:03 -07:00
Ryan Hill	3c3186c761	Convert more C APIs to return OrtStatus (#1194 ) * Change SessionOptions APIs to always return a status, for consistency and ease of use (a couple returned 0 or -1 for success/failure)	2019-06-10 18:36:04 -07:00
Ryan Hill	b68bb51dd0	Change SessionOptions APIs to always return a status (#1171 ) * Change SessionOptions APIs to always return a status, for consistency and ease of use (a couple returned 0 or -1 for success/failure)	2019-06-06 13:24:24 -07:00
Torkel	10ea77a3d1	add details aboud adding execution providers in the C api to comments and docs (i.e. need OrtSessionOptionsAppendExecutionProvider_CUDA to get CUDA)	2019-06-02 17:38:36 -07:00
jignparm	2cf56639ed	Minor update to NuGet package tests -- allow model download in separate step (#1115 ) * Update docker scripts to not fetch model data * Update related files	2019-05-28 03:01:10 -07:00
Ryan Hill	9129a652c5	Ryanunderhill/cxx api2 (#1091 ) More C++ API improvements and cleanup Add templates to tensor creation Add run method that allows preallocated outputs Simplify CreateTensor<T> to multiply by sizeof(T) Convert io_types code Optimize away vector copies in Session::Run	2019-05-24 11:15:51 -07:00
jignparm	9673f3d494	Jignparm/minor update linux test (#1074 ) * Minor update for pipeline tests * uncomment data download	2019-05-21 19:32:29 -07:00
jignparm	32da12491d	x86 support for C# API (#962 ) * Refactor C# to handle x86 * update run script * Add Native win x86 tests * Add native x86 tests for Linux * Update linux tests scripts to control which tests are run * update linux image name for x86 to prevent using cached image * update to not run unit python unit tests unless pybind is specified * remove --build_wheel as a core common arg. Python cannot run on x86 build * update OrtGetNumOfDimensions to OrtGetDimensionsCount in rest of C#	2019-05-20 15:48:14 -07:00
Ryan Hill	3a32b0eb99	Change function kernels to use CustomOp APIs (#1020 ) * Change function signature * Convert compute to use custom op style APIs * Remove dead CustomOp function * Use CustomOp API in TensorRT EP * Switch to new API in ngraph	2019-05-20 14:57:43 -07:00

1 2 3

112 commits