pytorch

mirror of https://github.com/zebrajr/pytorch.git synced 2025-12-06 12:20:52 +01:00

Author	SHA1	Message	Date
Animesh Jain	1d094587ea	[NNC Testing] Randomized loop nest infrastructure (#70174 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/70174 Test Plan: Imported from OSS Reviewed By: ZolotukhinM Differential Revision: D33234529 fbshipit-source-id: 9019f1f1d4ca945c92bee401f7ec674b7d987de4	2021-12-22 22:07:39 -08:00
Raghavan Raman	4dec15e6d8	[nnc] Add a run method to TensorExprKernel that takes in output tensors (#69477 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/69477 This diff adds a new run method to `TensorExprKernel` which takes in output tensors as inputs and stores the output in those given tensors. ghstack-source-id: 146107009 Test Plan: buck test mode/dev-nosan //caffe2/test/cpp/tensorexpr:tensorexpr -- --exact 'caffe2/test/cpp/tensorexpr:tensorexpr - Kernel.RunWithAllocatedOutputs' Reviewed By: ZolotukhinM Differential Revision: D32823890 fbshipit-source-id: edc1f4839785124048b034060feb71cb8c1be34f	2021-12-22 00:30:15 -08:00
Hui Guo	ac92f7cc75	[tensorexpr] Remove the optional argument in LoopNest::prepareForCodeGen (#67144 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/67144 Test Plan: Imported from OSS Reviewed By: ZolotukhinM Differential Revision: D31881150 Pulled By: huiguoo fbshipit-source-id: af99087722ec71d6deb9049b63b573ae7720c9ec	2021-12-17 01:37:59 -08:00
Hui Guo	531b045446	[tensorexpr] Fix the buf size of discontiguous tensors (#69657 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/69657 Test Plan: Imported from OSS Reviewed By: ZolotukhinM Differential Revision: D32974473 Pulled By: huiguoo fbshipit-source-id: 52dcd13d0ad7f7e4f1beb69dcaabc8ceb386ffca	2021-12-10 01:26:37 -08:00
Mikhail Zolotukhin	1e9dcdd2a0	[TensorExpr] TensorExprKernel: support custom-class constants. (#68856 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/68856 Test Plan: Imported from OSS Reviewed By: navahgar Differential Revision: D32632907 Pulled By: ZolotukhinM fbshipit-source-id: e4180f8d791ba0cdf82bcb3bd11b61405c2faadd	2021-12-02 14:34:15 -08:00
Mikhail Zolotukhin	ec94bb787a	[TensorExpr] Add a way to define target triple/cpu/attrs for llvm codegen and turn on the AOT workflow. (#66527 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/66527 Differential Revision: D31593869 D31593869 Test Plan: Imported from OSS Reviewed By: navahgar Pulled By: ZolotukhinM fbshipit-source-id: e7534c11fbcf0dab5f49d01d6053caf77b833ef0	2021-11-13 00:52:20 -08:00
Mikhail Zolotukhin	e511a7a5b4	[TensorExpr] Remove non-determinism in iterating over unordered_set of intermediate buffers. (#68277 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/68277 Differential Revision: D32400553 D32400553 Test Plan: Imported from OSS Reviewed By: saketh-are, priyaramani Pulled By: ZolotukhinM fbshipit-source-id: a8fe820bbddaa19f95db432efaa6d3e36095a05e	2021-11-13 00:50:57 -08:00
Ivan Kobzarev	362c6069b9	[nnc] Lazy lowerings registration; custom classes network params (#67623 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/67623 Test Plan: Imported from OSS Reviewed By: ZolotukhinM Differential Revision: D32065076 Pulled By: IvanKobzarev fbshipit-source-id: 4945ac6483938d428c539ed1ce4fcd6988b34250	2021-11-11 09:00:23 -08:00
Raghavan Raman	e7a3bbce89	[nnc] Add support for dynamic shapes in TensorExprKernel (#67861 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/67861 Previously submitted as https://github.com/pytorch/pytorch/pull/67197. This got reverted because its failures were hidden by the failures of another PR. Test Plan: Imported from OSS Reviewed By: ZolotukhinM Differential Revision: D32178196 Pulled By: navahgar fbshipit-source-id: cc8a5c68aed360d06289e69645461cfa773e1300	2021-11-05 11:18:19 -07:00
Natalia Gimelshein	ca445645f9	Revert D31902471: [nnc] Add support for dynamic shapes in TensorExprKernel Test Plan: revert-hammer Differential Revision: D31902471 (`15a3c374e2`) Original commit changeset: d2729a38ba1a fbshipit-source-id: 4c05de82e626bbf744df84fd2b914b66fd165a19	2021-11-03 14:48:12 -07:00
Raghavan Raman	15a3c374e2	[nnc] Add support for dynamic shapes in TensorExprKernel (#67197 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/67197 Test Plan: Imported from OSS Reviewed By: eellison, ZolotukhinM Differential Revision: D31902471 Pulled By: navahgar fbshipit-source-id: d2729a38ba1ac607ff07f516ed56fbd9085715dc	2021-11-03 11:24:17 -07:00
Ivan Kobzarev	7fbcf79684	[tensorexpr][nnc] Support quantization (#66676 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/66676 Test Plan: Imported from OSS Reviewed By: navahgar Differential Revision: D31676329 Pulled By: IvanKobzarev fbshipit-source-id: 288b41ff4ed603dfaacb465f296997f14bb23c22	2021-10-31 22:49:30 -07:00
Priya Ramani	fa70d72e95	Set kernel func name from AOT Compiler (#67229 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/67229 Right now, assembly code generated for the a given method from the model is named wrapper or func by default. The function name is then replaced with a proper kernel_func_name after target specific assembly is generated. This PR propagates a desired kernel_func_name right from aotCompiler API so that the generated function has the needed name that doesn't need to be replaced later. Note: Most of this change was landed in https://github.com/pytorch/pytorch/pull/66337 which had to be reverted as it was breaking `test_profiler` in `test_jit_fuser_te` as it replaced the name generated for graph with the default kernel_func_name value. This PR fixes that as well. ``` (pytorch) ~/local/pytorch kname └─ $ python3 test/test_jit_fuser_te.py CUDA not available, skipping tests monkeytype is not installed. Skipping tests for Profile-Directed Typing ........................................<string>:3: UserWarning: torch.cholesky is deprecated in favor of torch.linalg.cholesky and will be removed in a future PyTorch release. L = torch.cholesky(A) should be replaced with L = torch.linalg.cholesky(A) and . . . ......................<string>:3: UserWarning: torch.symeig is deprecated in favor of torch.linalg.eigh and will be removed in a future PyTorch release. The default behavior has changed from using the upper triangular portion of the matrix by default to using the lower triangular portion. L, _ = torch.symeig(A, upper=upper) should be replaced with L = torch.linalg.eigvalsh(A, UPLO='U' if upper else 'L') and L, V = torch.symeig(A, eigenvectors=True) should be replaced with L, V = torch.linalg.eigh(A, UPLO='U' if upper else 'L') (Triggered internally at ../aten/src/ATen/native/BatchLinearAlgebra.cpp:2492.) ......[W pybind_utils.cpp:35] Warning: Using sparse tensors in TorchScript is experimental. Many optimization pathways have not been thoroughly tested with sparse tensors. Please include the fact that the network is running sparse tensors in any bug reports submitted. (function operator()) /data/users/priyaramani/pytorch/torch/testing/_internal/common_utils.py:403: UserWarning: Using sparse tensors in TorchScript is experimental. Many optimization pathways have not been thoroughly tested with sparse tensors. Please include the fact that the network is running sparse tensors in any bug reports submitted. (Triggered internally at ../torch/csrc/jit/python/pybind_utils.h:691.) return callable(args, *kwargs) .....................................................................[W Resize.cpp:23] Warning: An output with one or more elements was resized since it had shape [1], which does not match the required output shape [].This behavior is deprecated, and in a future PyTorch release outputs will not be resized unless they have zero elements. You can explicitly reuse an out tensor t by resizing it, inplace, to zero elements with t.resize_(0). (function resize_output_check) [W Resize.cpp:23] Warning: An output with one or more elements was resized since it had shape [1, 5], which does not match the required output shape [5].This behavior is deprecated, and in a future PyTorch release outputs will not be resized unless they have zero elements. You can explicitly reuse an out tensor t by resizing it, inplace, to zero elements with t.resize_(0). (function resize_output_check) ........................................................................s.......s...s.s....s......s..sss............................ ---------------------------------------------------------------------- Ran 503 tests in 37.536s OK (skipped=10) ``` Test Plan: Imported from OSS Reviewed By: navahgar, pbelevich Differential Revision: D31945713 Pulled By: priyaramani fbshipit-source-id: f2246946f0fd51afba5cb6186d9743051e3b096b	2021-10-27 13:10:49 -07:00
Natalia Gimelshein	b6fa998892	Revert D31514095: Use kernel_func_name from aotCompiler Test Plan: revert-hammer Differential Revision: D31514095 (`7b55dc8340`) Original commit changeset: b70c8e2c7336 fbshipit-source-id: ad4d828f33506e612b51c276149fa0e12b0565d5	2021-10-23 17:17:53 -07:00
Priya Ramani	7b55dc8340	Use kernel_func_name from aotCompiler (#66337 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/66337 Right now, assembly code generated for the a given method from the model is named wrapper or func by default. The function name is then replaced with a proper kernel_func_name after target specific assembly is generated. This PR propagates a desired kernel_func_name right from aotCompiler API so that the generated function has the needed name that doesn't need to be replaced later. Test Plan: Imported from OSS Reviewed By: navahgar Differential Revision: D31514095 Pulled By: priyaramani fbshipit-source-id: b70c8e2c733600a435cd4e8b32092d37b7bf7de5	2021-10-23 02:20:45 -07:00
Mikhail Zolotukhin	60a2a295ce	[TensorExpr] Use schema instead of op name in NNC lowerings. (#65843 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/65843 Fixes #64963. Test Plan: Imported from OSS Reviewed By: navahgar Differential Revision: D31282334 Pulled By: ZolotukhinM fbshipit-source-id: ffd0e1b6433d9360fedd9081c01ef41b21684439	2021-10-12 01:26:32 -07:00
Mikhail Zolotukhin	24b9b304d9	[TensorExpr] Nuke TE shape inference. (#65554 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/65554 We're relying on JIT based shape inference and not using the TE implementation. Question to the audience: we set `hasBroadcasts_` in that function, but this function was almost never invoked. Do we behave correctly in the presence of rand-calls and broadcasts? Test Plan: Imported from OSS Reviewed By: bertmaher Differential Revision: D31148925 Pulled By: ZolotukhinM fbshipit-source-id: 2898a57e389ea0950163122089d0fec3d92701c4	2021-10-12 01:25:14 -07:00
Scott Wolchok	2d885ab73d	[jit] Reduce refcounting of Types (#65345 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/65345 FooType::get() can return a const reference. Inconveniently, converting shared_ptr<FooType> to shared_ptr<Type> requires a copy & refcount bump, so to properly take advantage of this in unshapedType() we need to take a const Type& in isSubtypeOf(), which is good practice anyway -- don't require a shared_ptr if you don't need to take ownership. ghstack-source-id: 140044165 Test Plan: CI perf says c10::unshapedType time decreased from 2.8% to 2.2% during static runtime startup, though I expect this to be generally beneficial. Reviewed By: hlu1 Differential Revision: D31027361 fbshipit-source-id: 676feb81db9f74ad7b8651d8774f4ecb4cfa6ab8	2021-10-08 09:03:04 -07:00
Mikhail Zolotukhin	765b6a90f3	[TensorExpr] Move lowerings registration from kernel.cpp to lowerings.cpp. (#65553 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/65553 Test Plan: Imported from OSS Reviewed By: navahgar Differential Revision: D31148921 Pulled By: ZolotukhinM fbshipit-source-id: 772062155043d4be9e9a25f6259b8e4a6cb762f4	2021-09-30 22:56:22 -07:00
Mikhail Zolotukhin	015e0079e3	[TensorExpr] Move 'compute' functions to operators/... (#65552 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/65552 This PR is mostly a verbatim move of several functions to different files. The goal is to have more consistency in what resides where. With this PR: All `compute` functions defining how a given operator needs to be lowered to TE IR will reside in `operators/.{cpp,h}`. * Auxiliary functions for these functions will reside in `operators/misc.cpp`. `compute` functions for ops not belonging anywhere else can also go to that file. `operators/unary.` is renamed to `operators/pointwise.` and now includes functions like `computeTwoOperands`. * `kernel.` now contains only JIT-related* logic and implementations of `TensorExprKernel` methods. Test Plan: Imported from OSS Reviewed By: navahgar Differential Revision: D31148923 Pulled By: ZolotukhinM fbshipit-source-id: e36ad8e779b8d30a33b49ea4ebf6d6a7438989f4	2021-09-30 22:56:20 -07:00
Mikhail Zolotukhin	3a0165da49	[TensorExpr] Port NNC lowerings to the new registry mechanism. (#65551 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/65551 Previously we had a big switch on Op kind to decide how to lower a given JIT operator to NNC. This PR changes this switch to a hash table lookup. Why? This helps us with at least two things: 1) With this approach we can easily check if we know how to handle a given node in advance - i.e. we can inspect the entire graph and tell whether it's possible to compile it or not without actually trying to do that and dying in the middle. This would allow us to, say, provide user-friendly error messages in AOT workflow. 2) We can switch to use schema instead of op kind to determine correct lowering. Unlike op schema, op kind might be ambigous (see e.g. #64963) and using it instead of schema can lead to bugs. Test Plan: Imported from OSS Reviewed By: navahgar Differential Revision: D31148926 Pulled By: ZolotukhinM fbshipit-source-id: ac12684e2126c899426ef5e4cc1e3f70fa01f704	2021-09-30 22:56:18 -07:00
Mikhail Zolotukhin	eee9ad0fdd	[TensorExpr] Add a skeleton for a registry of NNC lowerings. (#65550 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/65550 This PR adds the source files and the class for the registry, subsequent PRs actually port existing lowerings to this mechanism. Test Plan: Imported from OSS Reviewed By: albanD Differential Revision: D31148922 Pulled By: ZolotukhinM fbshipit-source-id: 4c087b22ee898d5a5a18a5d2a4bb795aa2ffd655	2021-09-30 22:56:16 -07:00
Mikhail Zolotukhin	d84191fcc6	[TensorExpr] Kernel: make prim::ConstantChunk handled like other ops. (#65549 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/65549 Previously it had a special handling, with this change it follows the same mechanism as other ops. Test Plan: Imported from OSS Reviewed By: navahgar Differential Revision: D31148924 Pulled By: ZolotukhinM fbshipit-source-id: 572d8ae5e123e7a0e2a656154d7bd0f73c785a06	2021-09-30 22:55:00 -07:00
Ivan Kobzarev	43d47bdcca	[tensorexpr] conv2d handle optional bias (#64750 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/64750 conv2d bias is optional. It will be ArgNone in processing of the graph. This bias is prim::constant NoneType, so we do not know shape at the moment of constant binding. This adding it as a constant zeros Tensor at the moment of graph processing => for that adding `std::vector<TensorExprKernel::ConstantDescr>& constants and std::vector<at::Tensor>& constant_tensors` to `computeOperandValue` as it is not in `TensorExprKernel` Test Plan: Imported from OSS Reviewed By: navahgar Differential Revision: D30842101 Pulled By: IvanKobzarev fbshipit-source-id: 88020f6934e43fe606f8eae928b7e21b7c3f15f6	2021-09-27 20:00:53 -07:00
Ivan Kobzarev	31ea4358d8	[tensorexpr] Add Op handling for mobilenetv3 large (#64741 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/64741 Test Plan: Imported from OSS Reviewed By: navahgar Differential Revision: D30839110 Pulled By: IvanKobzarev fbshipit-source-id: d8e89c086c713fbe816dd8c8096cd64c05dc7431	2021-09-27 20:00:51 -07:00
Mikhail Zolotukhin	7e9c599784	[TensorExpr] Add a method for sanitizing Var and Buf names in Stmt. (#65010 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/65010 This pass ensures all names are legal and not-duplicated. Fixes #52727. Test Plan: Imported from OSS Reviewed By: bertmaher, navahgar Differential Revision: D30939717 Pulled By: ZolotukhinM fbshipit-source-id: 7dbe7f937de41f22ad49137a5e067d698443ed63	2021-09-15 17:15:06 -07:00
Mikhail Zolotukhin	f23f21dafe	[TensorExpr] Remove 'Placeholder' class. (#64887 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/64887 BufHandle has exactly the same functionality and should be used instead. Differential Revision: D30889483 D30889483 Test Plan: Imported from OSS Reviewed By: navahgar Pulled By: ZolotukhinM fbshipit-source-id: 365fe8e396731b88920535a3de96bd3301aaa3f3	2021-09-14 00:22:44 -07:00
Mikhail Zolotukhin	82ac3f108d	[TensorExpr] Move 2 graph passes from kernel.cpp to graph_opt.cpp (#64828 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/64828 Also, make `removeUnusedSelfArgument` more consistent with other passes by mutating the graph in-place rather than returning a copy. Test Plan: Imported from OSS Reviewed By: ngimel Differential Revision: D30870776 Pulled By: ZolotukhinM fbshipit-source-id: 4873f01b013921143a5aa43746d655a2d8d620c9	2021-09-11 10:23:15 -07:00
Raghavan Raman	cad7a4b0ea	[nnc] Added an implementation of sign op (#64033 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/64033 Test Plan: Imported from OSS Reviewed By: bertmaher Differential Revision: D30579197 Pulled By: navahgar fbshipit-source-id: f9f7fa7f2ffa109cf4e441eb1af821b8b891d4d3	2021-09-10 16:49:04 -07:00
Mikhail Zolotukhin	a17d6c7f80	[TensorExpr] Simplify TE IR before applying any transformations. (#64717 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/64717 This also exposed several bugs, which are fixed in this PR. Differential Revision: D30826408 D30826408 Test Plan: Imported from OSS Reviewed By: navahgar Pulled By: ZolotukhinM fbshipit-source-id: a67ec5739aceed9ffdf0d24f77eb3787cefe4560	2021-09-09 18:50:51 -07:00
Raghavan Raman	652a8bf7d0	[nnc] Updated indices during broadcast to use int64_t (#64627 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/64627 This fixes the root cause of S242719 Test Plan: Imported from OSS Reviewed By: ZolotukhinM Differential Revision: D30801686 Pulled By: navahgar fbshipit-source-id: b6d3ebdc7eb57116eaced53c2f35c7798bb17e80	2021-09-09 08:29:37 -07:00
Hui Guo	5c27a580ec	[tensorexpr] Allocate intermediate buffers at compile time (#64227 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/64227 Test Plan: Imported from OSS Reviewed By: ZolotukhinM Differential Revision: D30652220 Pulled By: huiguoo fbshipit-source-id: cd75005cdfa42751318de7174b44e14a3a01634e	2021-09-08 15:34:44 -07:00
Animesh Jain	18d24bb537	[NNC] Add Softplus operator (#64589 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/64589 Adding softplus operator lowering for NNC. Enabling element wise fusion as well. Test Plan: Added a test in test_jit_fuser.py Reviewed By: bertmaher Differential Revision: D30736449 fbshipit-source-id: 6c5fc3bceb5cef2322ecd4449f827e4af018ea93	2021-09-08 10:49:58 -07:00
Bert Maher	7f0feafa55	[nnc] Provide helpful error messages about turning off the fuser (#64516 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/64516 If fuser compilation fails due to a bug (which should be highly unlikely at this point) we want to direct the user how to unblock themselves by disabling fusion, in addition to requesting that they report a bug. ghstack-source-id: 137398537 Test Plan: existing tests Reviewed By: ZolotukhinM Differential Revision: D30758051 fbshipit-source-id: 98be89f1b1d4fb3bc816f5b2634c618b9297930e	2021-09-08 08:10:22 -07:00
Hui Guo	9214450b7f	[tensorexpr] Wrap error msgs with buildErrorMessages for internal asserts (#64409 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/64409 Test Plan: Imported from OSS Reviewed By: navahgar Differential Revision: D30717786 Pulled By: huiguoo fbshipit-source-id: a3b147d339ff4927f14efa24407cd3b63d80001d	2021-09-02 11:30:34 -07:00
Raghavan Raman	87d8ab6e50	[nnc] Updated generic error message with info about turning off the fuser (#64316 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/64316 Test Plan: Imported from OSS Reviewed By: bertmaher Differential Revision: D30683942 Pulled By: navahgar fbshipit-source-id: d86607563672213f99a1436dcf4f5dc28053b713	2021-09-01 10:31:50 -07:00
Bert Maher	e7fb35021a	[nnc] Enable fusion of bfloat16 ops (#64196 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/64196 Test Plan: Imported from OSS Reviewed By: ZolotukhinM Differential Revision: D30643864 Pulled By: bertmaher fbshipit-source-id: e95edeaf7089464d713ea1d1f951743d3e5f61c5	2021-08-30 20:09:36 -07:00
Raghavan Raman	093a12aaa9	[nnc] Updated internal asserts to include more detailed error messages (#64118 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/64118 Test Plan: Imported from OSS Reviewed By: ZolotukhinM Differential Revision: D30616944 Pulled By: navahgar fbshipit-source-id: 35289696cc0e7faa01599304243b86f0febc6daf	2021-08-30 04:40:51 -07:00
Bert Maher	2e6221a232	[nnc] Make 64-bit dimensions work (#64077 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/64077 We were assuming kernel dimensions fit in 32 bits (the old fuser made this assumption too), but we should be able to support 64. ghstack-source-id: 136933272 Test Plan: unit tests; new IR level test with huge sizes Reviewed By: ZolotukhinM Differential Revision: D30596689 fbshipit-source-id: 23b7e393a2ebaecb0c391a6b1f0c4b05a98bcc94	2021-08-28 19:59:47 -07:00
Raghavan Raman	6d31ba6ddc	[nnc] Sanitized the names of constants in the input graph. (#63990 ) Summary: Fixes https://github.com/pytorch/pytorch/issues/63923 The input graph can contain constants whose names contain special characters. So, all names of constants in the input graph need to be sanitized. Pull Request resolved: https://github.com/pytorch/pytorch/pull/63990 Reviewed By: ZolotukhinM Differential Revision: D30558432 Pulled By: navahgar fbshipit-source-id: de5b0c23d50ee8997f40f2c0fc605dda3719186f	2021-08-26 09:52:02 -07:00
Bert Maher	ba5f1b1076	[nnc] Fix dtype promotion involving scalars (#64002 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/64002 Fixes https://github.com/pytorch/vision/issues/4315 Test Plan: Imported from OSS Reviewed By: navahgar Differential Revision: D30566979 Pulled By: bertmaher fbshipit-source-id: eaa98b9534a926be7fcd337d46c5a0acb3243179	2021-08-26 09:43:15 -07:00
Bert Maher	8dda299d96	Re-apply: [nnc] Support thread level parallelism in fused kernels (#63776 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/63776 I reverted this out of an abundance of caution because some test failures occurred, but they were all due to precision issues fixed lower in this stack. Let's try again. I've rolled the elimination of the allow-parallelism-in-fusions toggle into this diff since they're pretty tightly coupled. ghstack-source-id: 136529847 Test Plan: CI Reviewed By: huiguoo Differential Revision: D30484555 fbshipit-source-id: 38fd33520f710585d1130c365a8c60c9ce794a59	2021-08-24 18:56:55 -07:00
Mikhail Zolotukhin	f0d274294d	[TensorExpr] Nuke KernelArena and KernelScope. (#63587 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/63587 Now that there is no classes using KernelArena for memory management we can remove it. Differential Revision: D30429115 D30429115 Test Plan: Imported from OSS Reviewed By: navahgar Pulled By: ZolotukhinM fbshipit-source-id: 375f6f9294d27790645eeb7cb5a8e87047a57544	2021-08-24 00:32:16 -07:00
Mikhail Zolotukhin	62d02f2b57	[TensorExpr] Make 'Tensor' a value type. (#63586 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/63586 This is another commit in transition from KernelArena memory management. Tensor is essentially just a pair of <BufPtr, StmtPtr> and we don't need to dynamically allocate it at all - it's cheap to pass it by value, and that's what we're switching to in this commit. After this change nothing uses KernelScope/KernelArena and they can be safely removed. Differential Revision: D30429114 D30429114 Test Plan: Imported from OSS Reviewed By: navahgar Pulled By: ZolotukhinM fbshipit-source-id: f90b859cfe863692b7beffbe9bd0e4143df1e819	2021-08-24 00:32:13 -07:00
Bert Maher	37d60c08e5	Revert D30360382: [nnc] Support thread level parallelism in fused kernels Test Plan: revert-hammer Differential Revision: D30360382 (`d6d86efb1c`) Original commit changeset: 29acf4e932c6 fbshipit-source-id: e0531113135d30eabb172dc1537d5dd6d65dc438	2021-08-21 03:46:43 -07:00
Bert Maher	d6d86efb1c	[nnc] Support thread level parallelism in fused kernels (#63386 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/63386 Test Plan: Imported from OSS Reviewed By: navahgar Differential Revision: D30360382 Pulled By: bertmaher fbshipit-source-id: 29acf4e932c669ce0f35823faea9099bcd8119b6	2021-08-20 11:18:17 -07:00
Mikhail Zolotukhin	1dc2b52764	[TensorExpr] Add a wrapper for all expr and stmt pointers. (#63195 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/63195 This helps us to later switch from using KernelArena with raw pointers to shared pointers without having to change all our source files at once. The changes are mechanical and should not affect any functionality. With this PR, we're changing the following: * `Add` --> `AddPtr` `new Add(...)` --> `alloc<Add>(...)` * `dynamic_cast<Add>` --> `to<Add>` `static_cast<Add>` --> `static_to<Add>` Due to some complications with args forwarding, some places became more verbose, e.g.: `new Block({})` --> `new Block(std::vector<ExprPtr>())` Test Plan: Imported from OSS Reviewed By: navahgar Differential Revision: D30292779 Pulled By: ZolotukhinM fbshipit-source-id: 150301c7d2df56b608b035827b6a9a87f5e2d9e9	2021-08-17 13:44:45 -07:00
Richard Barnes	d1f9c03cef	Use `const auto` with irange (#62990 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/62990 Test Plan: Sandcastle Reviewed By: zhouzhuojie Differential Revision: D30199748 fbshipit-source-id: 284b208ffa3c6c4749e5ac9b1fccb28914590f2c	2021-08-10 17:59:01 -07:00
Raghavan Raman	59dd12042e	[nnc] Removed const from all fields in IR. (#62336 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/62336 This PR was generated by removing `const` for all types of nodes in NNC IR, and fixing compilation errors that were the result of this change. This is the first step in making all NNC mutations in-place. Test Plan: Imported from OSS Reviewed By: iramazanli Differential Revision: D30049829 Pulled By: navahgar fbshipit-source-id: ed14e2d2ca0559ffc0b92ac371f405579c85dd63	2021-08-03 11:44:36 -07:00
Nikita Shulga	a9b0a921d5	Disable `avoid-non-const-global-variables` lint check (#62008 ) Summary: As GoogleTest `TEST` macro is non-compliant with it as well as `DEFINE_DISPATCH` All changes but the ones to `.clang-tidy` are generated using following script: ``` for i in `find . -type f -iname ".c" -or -iname "*.h"\|xargs grep cppcoreguidelines-avoid-non-const-global-variables\|cut -f1 -d:\|sort\|uniq`; do sed -i "/\/\/ NOLINTNEXTLINE(cppcoreguidelines-avoid-non-const-global-variables)/d" $i; done ``` Pull Request resolved: https://github.com/pytorch/pytorch/pull/62008 Reviewed By: driazati, r-barnes Differential Revision: D29838584 Pulled By: malfet fbshipit-source-id: 1b2f8602c945bd4ce50a9bfdd204755556e31d13	2021-07-22 18:04:40 -07:00

1 2 3 4 5

245 Commits