pytorch

mirror of https://github.com/zebrajr/pytorch.git synced 2025-12-07 12:21:27 +01:00

Author	SHA1	Message	Date
Zhengxu Chen	8aeb247a3d	[export] Remove WrapperModule. (#121042 ) Summary: WrapperModule seems a good idea but may introduce some surprising behavior to users, for example, it never registers enclosed modules as submodules and therefore it's unclear that's the state dict for the exported program should look like, because some people may argue to include every state in state dict but others want to keep them as constants. Test Plan: CI Reviewed By: tugsbayasgalan Differential Revision: D54326331 Pull Request resolved: https://github.com/pytorch/pytorch/pull/121042 Approved by: https://github.com/angelayi	2024-03-05 18:10:22 +00:00
ydwu4	306642b66d	[export] fix test_passes on ci (#120322 ) We put the test cases generation in unitest.setUp to avoid running export on machines that runs with Python 3.12, where dynamo is not supported. Pull Request resolved: https://github.com/pytorch/pytorch/pull/120322 Approved by: https://github.com/angelayi, https://github.com/huydhn, https://github.com/malfet	2024-02-21 21:23:40 +00:00
ydwu4	ac2ba7889d	[export] turn on replace_set_grad_with_hop_pass in pre_dispatch (#119915 ) This PR turns on replace_set_grad_with_hop_pass for pre_dispatch export. To do that, we need to propagate the meta-data from original submodule to the new higher order op and fix the names of nodes as is required by the _sig_to_specs pass. Pull Request resolved: https://github.com/pytorch/pytorch/pull/119915 Approved by: https://github.com/tugsbayasgalan ghstack dependencies: #119732, #119736, #119810, #119913, #119914	2024-02-17 02:18:35 +00:00
ydwu4	737630268c	[export] manuually create test cases for split and inline (#119914 ) This PR makes the tests for inline and sequential_split stop relying on set_grad_enabled to be in the graph. Because they'll be gone if we turn on the replace_set_grad_with_hop_pass in the following diff. Instead, we'll manually insert them into the graph. Pull Request resolved: https://github.com/pytorch/pytorch/pull/119914 Approved by: https://github.com/tugsbayasgalan ghstack dependencies: #119732, #119736, #119810, #119913	2024-02-17 02:18:35 +00:00
ydwu4	8d81e61fb6	[export] make node_inline_ also inline the get_item calls (#119913 ) As titled. Before the PR, after we split then inline_, there will be getitem calls in the graph while the original graph module doesn't have them. This PR removes the additional get_item calls by inlining. Test Plan: Added new test cases for graphs that return multiple outputs and takes multiple inputs Pull Request resolved: https://github.com/pytorch/pytorch/pull/119913 Approved by: https://github.com/tugsbayasgalan ghstack dependencies: #119732, #119736, #119810	2024-02-17 02:18:27 +00:00
ydwu4	812f05d731	[export] add replace_set_grad_with_hop_pass (#119810 ) Pull Request resolved: https://github.com/pytorch/pytorch/pull/119810 Approved by: https://github.com/tugsbayasgalan ghstack dependencies: #119732, #119736	2024-02-17 02:18:19 +00:00
ydwu4	4769e6916a	[export] add node_inline_ to prepare replacing set_grad_enabled with hop (#119736 ) Pull Request resolved: https://github.com/pytorch/pytorch/pull/119736 Approved by: https://github.com/tugsbayasgalan ghstack dependencies: #119732	2024-02-17 02:18:11 +00:00
ydwu4	068659ddc2	[export] add sequential_split to prepare replacing set_grad_enabled with hop (#119732 ) This pr is the 1/N pr of transforming the global state mutating ops such as torch._C.set_grad_enabled calls in pre-dispatch graph into a higher order op so that the graph becomes more functional. We make use of split_module to help us do the transformation. This pr preserves the node.name in original module by adding a new kwarg `keep_original_node_name` to split_module. For a graph looks like this: ```python def forward(self, arg_0): arg0_1, = fx_pytree.tree_flatten_spec(([arg_0], {}), self._in_spec) add = torch.ops.aten.add.Tensor(arg0_1, 1); arg0_1 = None sin = torch.ops.aten.sin.default(add); add = None sum_1 = torch.ops.aten.sum.default(sin); sin = None _set_grad_enabled = torch._C._set_grad_enabled(False) add_1 = torch.ops.aten.add.Tensor(sum_1, 1); sum_1 = None _set_grad_enabled_1 = torch._C._set_grad_enabled(True) sub = torch.ops.aten.sub.Tensor(add_1, 1) return pytree.tree_unflatten((add_1, sub), self._out_spec) ``` Before the change, split graph returns the following graphs and subgraphs (notice the change from `add` -> `add_tensor`, `sin` -> `sin_default`: ```python def forward(self, arg_0): arg0_1, = fx_pytree.tree_flatten_spec(([arg_0], {}), self._in_spec) submod_0 = self.submod_0(arg0_1); arg0_1 = None submod_1 = self.submod_1(submod_0); submod_0 = None submod_2 = self.submod_2(submod_1) return pytree.tree_unflatten((submod_1, submod_2), self._out_spec) # submod_0 def forward(self, arg0_1): add_tensor = torch.ops.aten.add.Tensor(arg0_1, 1); arg0_1 = None sin_default = torch.ops.aten.sin.default(add_tensor); add_tensor = None sum_default = torch.ops.aten.sum.default(sin_default); sin_default = None return sum_default # submod_1 def forward(self, sum_1): _set_grad_enabled = torch._C._set_grad_enabled(False) add_tensor = torch.ops.aten.add.Tensor(sum_1, 1); sum_1 = None return add_tensor # submod_2 def forward(self, add_1): _set_grad_enabled = torch._C._set_grad_enabled(True) sub_tensor = torch.ops.aten.sub.Tensor(add_1, 1); add_1 = None return sub_tensor """) ``` After the change, the test produce the following graph, all the node names in original graph module are preserved in sub_modules. ```python def forward(self, arg_0): sub, = fx_pytree.tree_flatten_spec(([arg_0], {}), self._in_spec) submod_0 = self.submod_0(sub); sub = None submod_1 = self.submod_1(submod_0); submod_0 = None submod_2 = self.submod_2(submod_1) return pytree.tree_unflatten((submod_1, submod_2), self._out_spec) # submod_0 def forward(self, arg0_1): add = torch.ops.aten.add.Tensor(arg0_1, 1); arg0_1 = None sin = torch.ops.aten.sin.default(add); add = None sum_1 = torch.ops.aten.sum.default(sin); sin = None return sum_1 # submod_1 def forward(self, sum_1): _set_grad_enabled = torch._C._set_grad_enabled(False) add_1 = torch.ops.aten.add.Tensor(sum_1, 1); sum_1 = None return add_1 # submod_2 def forward(self, add_1): _set_grad_enabled_1 = torch._C._set_grad_enabled(True) sub = torch.ops.aten.sub.Tensor(add_1, 1); add_1 = None return sub ``` Note that currently, we call split_module on the graph after pre-dispatch aot. The difference is even larger if we `split_module` the graph module produced by dynamo, where all the original variables names in user program are preserved after dynamo but lost after `split_module` without this change. Pull Request resolved: https://github.com/pytorch/pytorch/pull/119732 Approved by: https://github.com/tugsbayasgalan	2024-02-17 02:18:04 +00:00
gs-olive	e0f6fa6a7c	Windows Dynamo Error Removal CI Check (#115969 ) Rebase of #111313 onto `main`, for CI validation Co-authored-by: Stella Laurenzo <stellaraccident@gmail.com> Pull Request resolved: https://github.com/pytorch/pytorch/pull/115969 Approved by: https://github.com/PaliC, https://github.com/thiagocrepaldi	2024-02-14 21:14:36 +00:00
PyTorch MergeBot	4a5b2cd6cb	Revert "Windows Dynamo Error Removal CI Check (#115969 )" This reverts commit `45e7af5818`. Reverted https://github.com/pytorch/pytorch/pull/115969 on behalf of https://github.com/PaliC due to this pr ended up breaking some of our periodic tests ([comment](https://github.com/pytorch/pytorch/pull/115969#issuecomment-1942934386))	2024-02-14 01:11:46 +00:00
gs-olive	45e7af5818	Windows Dynamo Error Removal CI Check (#115969 ) Rebase of #111313 onto `main`, for CI validation Co-authored-by: Stella Laurenzo <stellaraccident@gmail.com> Pull Request resolved: https://github.com/pytorch/pytorch/pull/115969 Approved by: https://github.com/ezyang	2024-02-08 21:23:45 +00:00
Michael Suo	bf4e171539	[export] support non-persistent buffers (#118969 ) Summary: X-link: https://github.com/pytorch/executorch/pull/1817 Basic support for non-persistent buffers, which are buffers that do not show up in the state dict. One weird twist is that most of our other systems (FX, aot_export, dynamo) have completely buggy handling of non-persistent buffers. I tried to go on a wild goose chase to fix them all, but it got to be too much. So I introduced some sad rewrite passes in `_export` make the final state dict correctly align with the original module's state dict. This exposed some bugs/ambiguous handling of parameters/buffers in existing test code. For example, `TestSaveLoad.test_save_buffer` traced over a module that was not in the root module hierarchy and caused some weird behavior. I think we should error explicitly on use cases like this: https://github.com/pytorch/pytorch/issues/118410. For now I just rewrote the tests or skipped them. As a side effect, this diff tightened up quite a few sloppy behaviors around state dict handling: - Tensor attributes were getting promoted to be buffers—bad! - Tracing through a module not in the children of the root module would add its parameters/buffers to the state dict—bad! This behavior is unlikely to show up in user code since the model would be totally broken, but did show up in a bunch of tests. #buildmore Test Plan: unit tests sandcastle Differential Revision: D53340041 Pull Request resolved: https://github.com/pytorch/pytorch/pull/118969 Approved by: https://github.com/guangy10, https://github.com/huydhn, https://github.com/titaiwangms	2024-02-02 19:16:08 +00:00
PyTorch MergeBot	221747507d	Revert "[export] support non-persistent buffers (#118612 ) (#118722 )" This reverts commit `a43c28368c`. Reverted https://github.com/pytorch/pytorch/pull/118722 on behalf of https://github.com/atalman due to broke linux-jammy-py3-clang12-executorch ([comment](https://github.com/pytorch/pytorch/pull/118722#issuecomment-1921484565))	2024-02-01 14:39:29 +00:00
Michael Suo	a43c28368c	[export] support non-persistent buffers (#118612 ) (#118722 ) Summary: X-link: https://github.com/pytorch/executorch/pull/1769 Basic support for non-persistent buffers, which are buffers that do not show up in the state dict. One weird twist is that most of our other systems (FX, aot_export, dynamo) have completely buggy handling of non-persistent buffers. I tried to go on a wild goose chase to fix them all, but it got to be too much. So I introduced some sad rewrite passes in `_export` make the final state dict correctly align with the original module's state dict. This exposed some bugs/ambiguous handling of parameters/buffers in existing test code. For example, `TestSaveLoad.test_save_buffer` traced over a module that was not in the root module hierarchy and caused some weird behavior. I think we should error explicitly on use cases like this: https://github.com/pytorch/pytorch/issues/118410. For now I just rewrote the tests or skipped them. Test Plan: added a unit test Differential Revision: D53253905 Pull Request resolved: https://github.com/pytorch/pytorch/pull/118722 Approved by: https://github.com/SherlockNoMad, https://github.com/angelayi	2024-02-01 00:36:09 +00:00
suo	4ee8aa6028	[export] adopt KeyPath API in nonstrict mode (#118609 ) This PR rewrites two paths to use the newly-added keypaths API in pytree: First: we were hand-rolling a tree_map during fakification because we wanted to track sources. This PR uses keypaths instead, which can do the same thing without needing custom code. Second: our constraint error formatting was referencing placeholder names in error messages. These placeholder names are not otherwise user-visible, so they are super confusing to users (e.g. "which input does arg1_3 correspond to?"). This diff uses the `keystr` API to format the error message. This necessitated some small refactors—generating the keystr is expensive so doing it in an f-string was very bad. It can also be further improved—we can inspect the signature so that instead of `*args[0]` we can give people the actual argument name, which would be the ideal UX. But leaving that for later. Differential Revision: [D53139358](https://our.internmc.facebook.com/intern/diff/D53139358/) Pull Request resolved: https://github.com/pytorch/pytorch/pull/118609 Approved by: https://github.com/zhxchen17 ghstack dependencies: #118607, #118608	2024-01-30 19:14:11 +00:00
Angela Yi	413a434846	[export] Convert all export tests to .module() (#118425 ) Test Plan: CI Differential Revision: D53075379 Pull Request resolved: https://github.com/pytorch/pytorch/pull/118425 Approved by: https://github.com/suo	2024-01-29 23:06:54 +00:00
Edward Z. Yang	903e1913ff	Rename unbacked SymInt prefix to u (#117859 ) Currently, it conflicts with Inductor's naming convention for index variables Signed-off-by: Edward Z. Yang <ezyang@meta.com> Pull Request resolved: https://github.com/pytorch/pytorch/pull/117859 Approved by: https://github.com/lezcano, https://github.com/jansel, https://github.com/avikchaudhuri	2024-01-22 20:53:47 +00:00
suo	2ae66ddba0	[export] fix test ownership (#117886 ) as title Differential Revision: [D52924188](https://our.internmc.facebook.com/intern/diff/D52924188/) Pull Request resolved: https://github.com/pytorch/pytorch/pull/117886 Approved by: https://github.com/ydwu4	2024-01-21 01:18:16 +00:00
suo	02c96f6949	[export] modify torch.export tests to pass a Module in (#117572 ) We have a lot of tests that pass a function to torch.export. We are planning to disallow this, so fix up the tests to pass a module in. Differential Revision: [D52791309](https://our.internmc.facebook.com/intern/diff/D52791309/) Pull Request resolved: https://github.com/pytorch/pytorch/pull/117572 Approved by: https://github.com/tugsbayasgalan ghstack dependencies: #117570, #117571	2024-01-18 03:40:40 +00:00
Zhengxu Chen	5ac57a06eb	[export] Refactor ExportPassBase. (#116778 ) Summary: X-link: https://github.com/pytorch/executorch/pull/1532 as title. This diff decouple the pass base library from torch export and exir, so that different layers can evolve in their own fashion, and we have more head room to divide and conquer in the future. Test Plan: CI Reviewed By: angelayi Differential Revision: D52514517 Pull Request resolved: https://github.com/pytorch/pytorch/pull/116778 Approved by: https://github.com/angelayi	2024-01-04 21:32:14 +00:00
Zhengxu Chen	43fb1b671c	[export] Improve verifier to not specialize on dialect. (#116705 ) Summary: Currently we have a very ugly specialization on edge dialect in verifier like the following: ``` # TODO Remove this branch. if ep.dialect == "EDGE": # !!! Don't change this allowlist. !!! pass else: raise e ``` In this diff we do some additional work to make signature checking also work in exir. We decouple the transformation stack in torch export and exir so that different layers of the stack can evolve in their own fashion and the team can divide and conquer them seperately. Test Plan: CI Differential Revision: D52499225 Pull Request resolved: https://github.com/pytorch/pytorch/pull/116705 Approved by: https://github.com/tugsbayasgalan	2024-01-04 17:17:23 +00:00
Angela Yi	8e2d63cbc3	[export][reland] Remove runtime assertion pass (#115597 ) Summary: Reland of https://github.com/pytorch/pytorch/pull/115196 D52054112 to fix internal failures. Test Plan: CI Differential Revision: D52054110 Pull Request resolved: https://github.com/pytorch/pytorch/pull/115597 Approved by: https://github.com/ydwu4, https://github.com/zhxchen17	2023-12-15 03:22:03 +00:00
angelayi	92fd3927b0	[export][reland] Add math.* ops to pass base (#115559 ) Reland of https://github.com/pytorch/pytorch/pull/115271/ Fixes https://github.com/pytorch/pytorch/issues/115209 Pull Request resolved: https://github.com/pytorch/pytorch/pull/115559 Approved by: https://github.com/zhxchen17, https://github.com/atalman ghstack dependencies: #115556, #115557, #115558	2023-12-12 10:46:41 +00:00
angelayi	36199747f3	[export][reland][refactor][2/n] Move tracing logic (#115557 ) Reland of https://github.com/pytorch/pytorch/pull/114768 Pull Request resolved: https://github.com/pytorch/pytorch/pull/115557 Approved by: https://github.com/zhxchen17 ghstack dependencies: #115556	2023-12-12 05:37:07 +00:00
atalman	24a463c46c	Revert "[export][refactor][2/n] Move tracing logic (#114768 )" (#115503 ) Github first oncall. This reverts commit `0ab57ee7ea`. Fixes #ISSUE_NUMBER Pull Request resolved: https://github.com/pytorch/pytorch/pull/115503 Approved by: https://github.com/angelayi, https://github.com/kit1980	2023-12-10 19:30:15 +00:00
PyTorch MergeBot	af925a56a1	Revert "[export] Add math.* ops to pass base (#115271 )" This reverts commit `6c0a4ced53`. Reverted https://github.com/pytorch/pytorch/pull/115271 on behalf of https://github.com/atalman due to ghfirst issue when importing, will reland this PR ([comment](https://github.com/pytorch/pytorch/pull/115271#issuecomment-1847852211))	2023-12-08 21:17:56 +00:00
PyTorch MergeBot	4186932bac	Revert "[export] Remove runtime assertion pass (#115196 )" This reverts commit `c163b3c035`. Reverted https://github.com/pytorch/pytorch/pull/115196 on behalf of https://github.com/atalman due to Broke internal test ([comment](https://github.com/pytorch/pytorch/pull/115196#issuecomment-1847778344))	2023-12-08 20:07:04 +00:00
angelayi	6c0a4ced53	[export] Add math.* ops to pass base (#115271 ) Fixes https://github.com/pytorch/pytorch/issues/115209 Pull Request resolved: https://github.com/pytorch/pytorch/pull/115271 Approved by: https://github.com/ydwu4	2023-12-07 02:47:04 +00:00
angelayi	c163b3c035	[export] Remove runtime assertion pass (#115196 ) Reland of https://github.com/pytorch/pytorch/pull/111949/ Pull Request resolved: https://github.com/pytorch/pytorch/pull/115196 Approved by: https://github.com/avikchaudhuri	2023-12-07 01:44:11 +00:00
angelayi	0ab57ee7ea	[export][refactor][2/n] Move tracing logic (#114768 ) 2/n of refactoring export code: * Moved tracing logic in torch/_export/init.py to torch/export/_tracer.py Differential Revision: [D51823961](https://our.internmc.facebook.com/intern/diff/D51823961) Pull Request resolved: https://github.com/pytorch/pytorch/pull/114768 Approved by: https://github.com/ydwu4 ghstack dependencies: #114764	2023-12-06 16:46:47 +00:00
Zhengxu Chen	e0d2a24967	Reland "[export] Support user input mutation. [1/2]" (#114496 ) (#114596 ) Summary: Serialization not implemented yet. Will do in the next diff. Resolving Github issues: https://github.com/pytorch/pytorch/issues/112429 https://github.com/pytorch/pytorch/issues/114142 Test Plan: onnx doc test ``` python -m xdoctest /opt/conda/envs/py_3.8/lib/python3.8/site-packages/torch/onnx/_internal/exporter.py ONNXProgram.model_signature:0 ``` Differential Revision: D51588558 Pull Request resolved: https://github.com/pytorch/pytorch/pull/114596 Approved by: https://github.com/angelayi	2023-11-27 20:19:04 +00:00
PyTorch MergeBot	fa1ccc34c4	Revert "[export] Support user input mutation. [1/2] (#114496 )" This reverts commit `b62c0d96bc`. Reverted https://github.com/pytorch/pytorch/pull/114496 on behalf of https://github.com/facebook-github-bot due to Diff reverted internally ([comment](https://github.com/pytorch/pytorch/pull/114496#issuecomment-1827289635))	2023-11-27 07:52:21 +00:00
Zhengxu Chen	b62c0d96bc	[export] Support user input mutation. [1/2] (#114496 ) Summary: Serialization not implemented yet. Will do in the next diff. Resolving Github issues: https://github.com/pytorch/pytorch/issues/112429 https://github.com/pytorch/pytorch/issues/114142 Test Plan: buck2 run mode/opt caffe2/test:test_export -- -r test_export_ input_mutation Differential Revision: D51556962 Pull Request resolved: https://github.com/pytorch/pytorch/pull/114496 Approved by: https://github.com/tugsbayasgalan	2023-11-27 04:53:38 +00:00
Peter Bell	bbd5b935e4	Use `pytree.tree_leaves` everywhere (#112324 ) This changes all the instances I could find of `tree_flatten(...)[0]` or `x, _ = tree_flatten` to use `tree_leaves`. Pull Request resolved: https://github.com/pytorch/pytorch/pull/112324 Approved by: https://github.com/lezcano ghstack dependencies: #112327, #112323	2023-10-30 03:39:04 +00:00
Tugsbayasgalan Manlaibaatar	547a116fcf	Fix redundant asserts (#111445 ) Fixes: https://github.com/pytorch/pytorch/issues/109852 Pull Request resolved: https://github.com/pytorch/pytorch/pull/111445 Approved by: https://github.com/zhxchen17	2023-10-18 23:57:31 +00:00
PyTorch MergeBot	7a740e2b85	Revert "direct runtime assertions (#111262 )" This reverts commit `e6d9350d7f`. Reverted https://github.com/pytorch/pytorch/pull/111262 on behalf of https://github.com/jeanschmidt due to Breaking internal builds ([comment](https://github.com/pytorch/pytorch/pull/111262#issuecomment-1765881675))	2023-10-17 08:04:36 +00:00
Avik Chaudhuri	e6d9350d7f	direct runtime assertions (#111262 ) Previously we were generating a graph to add runtime assertions on inputs and then running that graph to check input constraints. This PR checks input constraints directly. Differential Revision: D50289970 Pull Request resolved: https://github.com/pytorch/pytorch/pull/111262 Approved by: https://github.com/zhxchen17	2023-10-15 05:15:09 +00:00
Tugsbayasgalan Manlaibaatar	5614023f5e	Move export.constrain_as_* to torch._constrain_as_* (#110757 ) Pull Request resolved: https://github.com/pytorch/pytorch/pull/110757 Approved by: https://github.com/avikchaudhuri ghstack dependencies: #109859	2023-10-12 05:37:44 +00:00
PyTorch MergeBot	6ce3a38050	Revert "Move export.constrain_as_* to torch._constrain_as_* (#110757 )" This reverts commit `5aee22e0e0`. Reverted https://github.com/pytorch/pytorch/pull/110757 on behalf of https://github.com/kit1980 due to Depends on https://github.com/pytorch/pytorch/pull/109859 that needs to be reverted ([comment](https://github.com/pytorch/pytorch/pull/110757#issuecomment-1758908371))	2023-10-12 04:53:29 +00:00
Tugsbayasgalan Manlaibaatar	5aee22e0e0	Move export.constrain_as_* to torch._constrain_as_* (#110757 ) Pull Request resolved: https://github.com/pytorch/pytorch/pull/110757 Approved by: https://github.com/avikchaudhuri ghstack dependencies: #109859	2023-10-11 02:37:55 +00:00
Avik Chaudhuri	5da5e068f3	deprecate constraints in favor of dynamic_shapes (#110143 ) Recently we updated the `export` API to take an experimental `dynamic_shapes` argument that was meant to subsume the existing `constraints` argument. This PR deprecates `constraints` (with a warning on its use, but without actually removing it). Simultaneously it replaces all uses of `constraints` in docs, examples, and tests with corresponding uses of `dynamic_shapes` (preserving behavior). This exercise fortunately revealed some minor bugs in the implementation which have also been fixed in this PR. Some uses of `constraints` still remain, e.g., when `torch._dynamo.export` is called directly. (Meta-internal uses will be updated in a separate diff.) Differential Revision: D49676049 Pull Request resolved: https://github.com/pytorch/pytorch/pull/110143 Approved by: https://github.com/tugsbayasgalan	2023-09-28 10:26:21 +00:00
Zhengxu Chen	138fafe72d	[export] Fix torch.export() issues for server use cases. (#108275 ) Test Plan: In D48788843 Differential Revision: D48811793 Pull Request resolved: https://github.com/pytorch/pytorch/pull/108275 Approved by: https://github.com/tugsbayasgalan	2023-08-31 07:19:18 +00:00
gmagogsfm	9af0e47653	Hide `transform` method by renaming it (#107940 ) Pull Request resolved: https://github.com/pytorch/pytorch/pull/107940 Approved by: https://github.com/tugsbayasgalan	2023-08-25 16:31:44 +00:00
Angela Yi	92f6454ff8	[export][reland] ExportedProgram.transform updates graph_signature automatically (#107792 ) Summary: Reland of https://github.com/pytorch/pytorch/pull/107080 Test Plan: CI Differential Revision: D48533622 Pull Request resolved: https://github.com/pytorch/pytorch/pull/107792 Approved by: https://github.com/gmagogsfm	2023-08-23 22:16:56 +00:00
eellison	c88775b937	Make Nd tensors hit fused addmm pass (#106911 ) Replace https://github.com/pytorch/pytorch/pull/106433 since I had a bad cla commit. Speeds up eager convnext bfloat16 inference by 35%., and eager timm bfloat16 inference average by `.5%` Pull Request resolved: https://github.com/pytorch/pytorch/pull/106911 Approved by: https://github.com/ezyang	2023-08-16 17:12:11 +00:00
PyTorch MergeBot	b860c8c5b8	Revert "ExportedProgram.transform now updates graph_signature automatically (#107080 )" This reverts commit `8c9b2fe8f0`. Reverted https://github.com/pytorch/pytorch/pull/107080 on behalf of https://github.com/izaitsevfb due to Breaks executorch tests, see D48333170 ([comment](https://github.com/pytorch/pytorch/pull/107080#issuecomment-1679588292))	2023-08-15 20:47:35 +00:00
Tugsbayasgalan Manlaibaatar	20c5add133	[export] Refactor `constrain_as_value` and `constrain_as_size` (#106591 ) Some notable changes: 1. `constrain_as_size` allows min value to be less than 2 as it will unconditionally assume min >= 2 for compiler purposes. Instead, we add additional check to make sure max value is always greater than 2. 2. Previously, we used to runtime assert on the unbacked symint's val range which would be always between [2, max]. I modified this logic to assert on [0, max] unless user explicitly specifies the min range. Pull Request resolved: https://github.com/pytorch/pytorch/pull/106591 Approved by: https://github.com/gmagogsfm, https://github.com/ezyang	2023-08-15 05:41:43 +00:00
gmagogsfm	8c9b2fe8f0	ExportedProgram.transform now updates graph_signature automatically (#107080 ) Update graph_signature according to graph after transformation. Transformations can lead to node name changes, which are used in graph_signature to identify inputs and outputs. Therefore, after each transformation, we need to update the graph_signature according to new node names. WARNING: This implementation makes a few assumptions - The transformation doesn't change number of inputs/outputs - Each input/output still has the same meaning. - For inputs, that means that the inputs in transformed graph map to the same lifted parameter/buffer or user input as the input of the same position in the graph before transformation. - Similarly for outputs, each output should correspond to the same mutated buffer or user output as the output value of the same position in the graph before transformation. It is difficult to programatically validate these assumptions, but they should hold true most of the time as inputs/outputs of the graph rarely need to be changed. Pull Request resolved: https://github.com/pytorch/pytorch/pull/107080 Approved by: https://github.com/tugsbayasgalan	2023-08-14 19:52:41 +00:00
gmagogsfm	f26aa2dcd9	Keep fx node name consistent with aot_export (#107068 ) torch.export() starts initially with node names in aot_export, if we don't make this change, any no-op transformation would break name consistency, thus breaking GraphSignature correctness. Pull Request resolved: https://github.com/pytorch/pytorch/pull/107068 Approved by: https://github.com/tugsbayasgalan	2023-08-12 23:12:03 +00:00
PyTorch MergeBot	745d29b0cc	Revert "[export] Refactor `constrain_as_value` and `constrain_as_size` (#106591 )" This reverts commit `18989890bf`. Reverted https://github.com/pytorch/pytorch/pull/106591 on behalf of https://github.com/izaitsevfb due to Breaks inductor test on trunk ([comment](https://github.com/pytorch/pytorch/pull/106591#issuecomment-1675069091))	2023-08-11 16:37:47 +00:00
Tugsbayasgalan Manlaibaatar	18989890bf	[export] Refactor `constrain_as_value` and `constrain_as_size` (#106591 ) Some notable changes: 1. `constrain_as_size` allows min value to be less than 2 as it will unconditionally assume min >= 2 for compiler purposes. Instead, we add additional check to make sure max value is always greater than 2. 2. Previously, we used to runtime assert on the unbacked symint's val range which would be always between [2, max]. I modified this logic to assert on [0, max] unless user explicitly specifies the min range. Pull Request resolved: https://github.com/pytorch/pytorch/pull/106591 Approved by: https://github.com/gmagogsfm, https://github.com/ezyang	2023-08-11 05:29:22 +00:00
Zhengxu Chen	2dbadd1eae	[export] Remove experimental runtime assertion configs from export API. (#105043 ) Test Plan: CI Differential Revision: D47390794 Pull Request resolved: https://github.com/pytorch/pytorch/pull/105043 Approved by: https://github.com/larryliu0820	2023-07-26 16:21:29 +00:00
xuanqi	3707fbf63b	[RFC]: Add test for graph partition after assertion ops functionalization. (#104287 ) This PR: * Address comment at https://github.com/pytorch/pytorch/pull/103887/files#r1244128266. * Add test for graph partition to make sure assertion ops functionalization won't break graph partition in unexpected way. NOTE: In the context of export, it's totally up to the user to any type of graph partition based on specific use case. It's hard to anticipate the concrete downstream use case nor provide any specific functionality to facilitate handling assertion ops (functional / non-functional). So this PR limit to itself to [`CapabilityBasedPartitioner`](`2da6cae43c/torch/fx/passes/infra/partitioner.py (L34)`) and make sure it doesn't break graph partition unexpectedly (by adding some test). For the test case used in PR, a few things to highlight: * Without assertion, the fused graph is roughly like: ``` class fused(torch.nn.Module): def forward(self, a, b): fused_1 = self.fused_1(a, b); relu = fused_1.relu() fused_0 = self.fused_0(fused_1, relu) return (fused_0, fused_1) class fused_0(torch.nn.Module): def forward(self, add_2, relu): ... # Logic after relu return add_4 class fused_1(torch.nn.Module): def forward(self, a, b): ... # Logic before relu, `add_1` is only exposed within this submodule. return add_2 ``` * With the assertion, the fused graph is roughly like: ``` class fused(torch.nn.Module): def forward(self, arg0_1: i64[s0], arg1_1: i64[s0]): dep_token0 = ... ... fused_1 = self.fused_1(arg0_1, arg1_1); arg0_1 = arg1_1 = None ... getitem: i64[s0] = fused_1[0] # `getitem` is actually `add_1` ... relu_default: i64[s0] = torch.ops.aten.relu.default(getitem_1) ... # For inline assertion. Note that `getitem` which is an output of `fused_1`, is consumed by it. select_int: i64[] = torch.ops.aten.select.int(getitem, 0, 0) eq_scalar: b8[] = torch.ops.aten.eq.Scalar(select_int, 5) dep_token2: f32[] = torch.ops.aten._functional_assert_async.msg( eq_scalar, 'assertion error', dep_token = dep_token1 ) ... getitem_1: i64[s0] = fused_1[1] # `getitem_1` is actually `add_2` fused_0: i64[s0] = self.fused_0(getitem_1, relu_default) ... return (fused_0, getitem_1, dep_token2) class fused_0(torch.nn.Module): def forward(self, add_tensor_2: i64[s0], relu_default: i64[s0]): ... # Logic after relu return add_tensor_4 class fused_1(torch.nn.Module): def forward(self, arg0_1: i64[s0], arg1_1: i64[s0]): ... # Logic before relu # `add_tensor_1` (basically `add_1`) is returned to allow downstream assertion op consumes it. return (add_tensor_1, add_tensor_2) ``` As shown above, the extra assertion added (actually regardless whether it's funtionalized or not), it won't case extra submodule breakage if the asserted node is an intermediate node within the submodule - here the intermediate node will be returned as extra output of submodule so downstream assertion node can consume it. Pull Request resolved: https://github.com/pytorch/pytorch/pull/104287 Approved by: https://github.com/tugsbayasgalan	2023-06-28 22:13:27 +00:00
xuanqi	bf34ecd0c8	[RFC]: Integrate assertions functionalization to export (after AOT export) (#103887 ) This PR integrated the assertion functionalization logic into current export logic. NOTE: I finally decided to do the assertion functionalization after AOT export instead of before for the following reasons: * The benefit of AOT export is that the graph is already functionalized so things like method call is already transformed to function call. However, if we do it before AOT export, the graph is still in torch level and extra logic like `bab21d20eb/torch/_export/pass_base.py (L201-L204C17)` will need to be implemented. * The graph signature is kind of already incorrect after adding runtime assertions currently (this doesn't seem break logic since we already depend on positions instead of FQNs of outputs). This PR also fixed this. Pull Request resolved: https://github.com/pytorch/pytorch/pull/103887 Approved by: https://github.com/avikchaudhuri, https://github.com/tugsbayasgalan	2023-06-27 18:14:29 +00:00
xuanqi	344bab2669	[RFC]: Functionalize assertions (#103757 ) The idea here is to create do a graph mutation to: * Create an initial dependency token at the beginning of the program. * Replace non-functional version of assertion statements to functional version. * The functional version of assertion statement will: * Accept a dependency token from output of previous functional assertion statement (or the initial dependency token if there isn't any). * Generate a dependency token as the output of assertion statement. * Augment the output to include the dependency token generated by last assertion statement. The goal here is to: * Form an explicit dependency chain and avoid potential reordering during other passes of compiling. * Make the assertions a part of overall execution graph will affect the final output (or it could potentially be DCEed). NOTE: * Currently only cover `contrain_range` and WIP to support other assertions. Send out this PR to collect feedback first. * Here it only focus on implementation itself. Will integrate it with current export in future PR. Pull Request resolved: https://github.com/pytorch/pytorch/pull/103757 Approved by: https://github.com/avikchaudhuri	2023-06-24 00:23:35 +00:00
xuanqi	b27c3558a4	[RFC]: Create aten native op for constrain_range (#103346 ) At high current implementation of constrains functions (constrain_as_) will raise exception for the following code snippets: ``` def f(x): a = x.item() constrain_as_size(a, 4, 7) return torch.empty((a, 4)) inp = torch.tensor([5]) ep = torch._export.export(f, (inp,)) ``` The reason is because current constrain logic is: 1) Purely python so it won't survive AOT export (the full node is gone after AOT export since AOT export only maintains aten level op). 2) Utilize side effect to add range constraints for traced symbol's shape env ([code](`9591e52880/torch/fx/experimental/symbolic_shapes.py (L370-L372)`)). 3) If runtime assertion is turned on (by default). [`_AddRuntimeAssertionsForConstraintsPass`](`9591e52880/torch/_export/passes/add_runtime_assertions_for_constraints_pass.py (L98-L100)`) will try to append assertion node based on range constrains extracted from shape env of symbol during another interpretation round. 4). However, since 1), in the round of AOT export, range constraints logic won't run for symbols generated during this round. And later there is no range constrains information available for assertion round and caused issue. 5) As a result of above, it will failure at `torch.empty((a, 4))` (there is no constrains for `a` that it must be positive). The fix here is just to implement range constrain logic as a native aten op (CPU implementation as no-op) to make it be able to survive AOT export. NOTE:** [Logic](`2d745b95d7/torch/fx/experimental/symbolic_shapes.py (L350-L365C15)`) within [`constrain_range`](`2d745b95d7/torch/fx/experimental/symbolic_shapes.py (LL313C74-L313C74)`) is split out as `constrain_range_int` to capture case when non `SymInt` is passed in and reused in the new `_constrain_range`. The reason is when non `SymInt` is provided: * If it directly calls `sym_constrain_range`, the C++ version will be called which will be no-op. * So in this case it calls `constrain_range_int` instead to be able to capture issue like user provides a input whose tensor's shape could be out of range during exporting, like the following for above code example: ``` ... inp = torch.tensor([10]) ep = torch._export.export(f, (inp,)) # immediately raise error ``` Differential Revision: [D46734204](https://our.internmc.facebook.com/intern/diff/D46734204) Pull Request resolved: https://github.com/pytorch/pytorch/pull/103346 Approved by: https://github.com/tugsbayasgalan	2023-06-16 14:55:40 +00:00
Tugsbayasgalan Manlaibaatar	4bb2b65ea4	Turn on add_runtime_assertion by default (#102671 ) Pull Request resolved: https://github.com/pytorch/pytorch/pull/102671 Approved by: https://github.com/angelayi, https://github.com/avikchaudhuri	2023-06-05 16:27:44 +00:00
Angela Yi	7a569f86a0	[export] Cleanup constraints (#102666 ) Redo of https://github.com/pytorch/pytorch/pull/102432 because idk how to push to that other branch... Pull Request resolved: https://github.com/pytorch/pytorch/pull/102666 Approved by: https://github.com/zhxchen17	2023-06-01 04:22:31 +00:00
Tugsbayasgalan Manlaibaatar	d9f75dded1	[export] Add aot_export 1/N (#101490 ) This PR adds aot_export_module as the lowering path from torch.level graph to aten graph. Some known limitations that need to be addressed in the follow up PRs: 1. Store param/buffer data in ExportedProgram 2. Fully support torch.cond with params/buffers 3. Making sure no duplicated ExportMetaData entry 4. This API will break Executorch if used on PyE, we will figure out a plan internally. Pull Request resolved: https://github.com/pytorch/pytorch/pull/101490 Approved by: https://github.com/avikchaudhuri	2023-05-31 20:56:21 +00:00
Angela Yi	c4028de462	[export] ExportedProgram (#102259 ) Pull Request resolved: https://github.com/pytorch/pytorch/pull/102259 Approved by: https://github.com/ydwu4, https://github.com/avikchaudhuri, https://github.com/tugsbayasgalan, https://github.com/zhxchen17	2023-05-26 23:36:38 +00:00
Avik Chaudhuri	8751002215	equality assertions (#102256 ) Previously we had runtime asserts for range constraints. This diff adds runtime asserts for equality constraints. This requires a bit of refactoring that is worth calling out. 1. [Minor] Some of the data structures produced by export and consumed by the runtime assertion pass need to be broadened. This is a WIP. There are some associated code improvements that are included in this diff, but by and large the structures are similar to what exists now. Meanwhile @angelayi and I are chatting about how to make it qualitatively better: briefly, we want to index everything by symbols, which are 1-1 with (name, dim) pairs. 2. [Major] The order in which runtime asserts are emitted is changed. Previously we used to do the work in `placeholder`, now this diff adds a hook for "post-processing" after processing of all placeholders is done. This is needed because equality constraints can mention different placeholders. This change also opens the way to optimizing codegen: e.g., each (name, dim) pair should correspond to a single intermediate variable that is reused across runtime asserts. This is future work. Differential Revision: [D46177642](https://our.internmc.facebook.com/intern/diff/D46177642/) Pull Request resolved: https://github.com/pytorch/pytorch/pull/102256 Approved by: https://github.com/tugsbayasgalan, https://github.com/angelayi	2023-05-26 14:57:31 +00:00
Tugsbayasgalan Manlaibaatar	47f43ed84a	Actually functionalize torch.export (#101433 ) I thought i enabled this, but apparently not. This PR makes the export fully functional for real this time :) Pull Request resolved: https://github.com/pytorch/pytorch/pull/101433 Approved by: https://github.com/angelayi	2023-05-17 05:09:24 +00:00
PyTorch MergeBot	eac5f2a8e4	Revert "Actually functionalize torch.export (#101433 )" This reverts commit `eec752ed05`. Reverted https://github.com/pytorch/pytorch/pull/101433 on behalf of https://github.com/PaliC due to causing failures on functorch macOS tests ([comment](https://github.com/pytorch/pytorch/pull/101433#issuecomment-1550111671))	2023-05-16 17:51:45 +00:00
Tugsbayasgalan Manlaibaatar	eec752ed05	Actually functionalize torch.export (#101433 ) I thought i enabled this, but apparently not. This PR makes the export fully functional for real this time :) Pull Request resolved: https://github.com/pytorch/pytorch/pull/101433 Approved by: https://github.com/angelayi	2023-05-16 16:22:13 +00:00
Tugsbayasgalan Manlaibaatar	194d360329	Add more canonical way of adding runtime pass (#100956 ) * #100955 Pull Request resolved: https://github.com/pytorch/pytorch/pull/100956 Approved by: https://github.com/ydwu4, https://github.com/guangy10	2023-05-16 03:23:04 +00:00
Tugsbayasgalan Manlaibaatar	9ffad5b62b	Remove input tracker from runtime assertion pass (#100955 ) Pull Request resolved: https://github.com/pytorch/pytorch/pull/100955 Approved by: https://github.com/ydwu4	2023-05-15 21:26:47 +00:00
Tugsbayasgalan Manlaibaatar	f542b31c9d	[export] More robust view->view_copy pass (#100908 ) Pull Request resolved: https://github.com/pytorch/pytorch/pull/100908 Approved by: https://github.com/ydwu4	2023-05-10 14:25:17 +00:00
ydwu4	26cd958718	Support runtime assertion for inline constraints (#100763 ) This pr does the following: 1. previously, inline constraints is not properly set for tensor output data-dependent ops such as a.nonzero because of its return value is not symint. This pr just uses all the unbacked symbols i.e.those start with "i"/"f" in create_unbacked_sym* functions. Note that these symbols are guaranteed to be a super set of inline user constraints. 2. add inline assertions support by checking. Currently, it only deal with tensor, SymInt, SymFloat, SymBool output data-dependent ops and ignore the rest. It's good enough for now as we only have a limited number of data-dependent ops (.item and .nonzero are explicitly tested). The examples for graph that is added assertions is shown below: ``` class ExportGraphModule(torch.nn.Module): def forward(self, x): arg0: i64[s0], = fx_pytree.tree_flatten_spec(([x], {}), self._in_spec) nonzero_default: i64[i0, 1] = torch.ops.aten.nonzero.default(arg0); arg0 = None return pytree.tree_unflatten([nonzero_default], self._out_spec) class GraphModule(torch.nn.Module): def forward(self, x): arg0: i64[s0], = fx_pytree.tree_flatten_spec(([x], {}), self._in_spec) sym_size: Sym(s0) = torch.ops.aten.sym_size(arg0, 0) nonzero_default: i64[i1, 1] = torch.ops.aten.nonzero.default(arg0); arg0 = None sym_size_1: Sym(i1) = torch.ops.aten.sym_size(nonzero_default, 0) ge: Sym(i1 >= 3) = sym_size_1 >= 3 scalar_tensor_default: f32[] = torch.ops.aten.scalar_tensor.default(ge); ge = None _assert_async_msg = torch.ops.aten._assert_async.msg(scalar_tensor_default, 'nonzero_default.shape[0] is outside of inline constraint [3, 5].'); scalar_tensor_default = None le: Sym(i1 <= 5) = sym_size_1 <= 5; sym_size_1 = None scalar_tensor_default_1: f32[] = torch.ops.aten.scalar_tensor.default(le); le = None _assert_async_msg_1 = torch.ops.aten._assert_async.msg(scalar_tensor_default_1, 'nonzero_default.shape[0] is outside of inline constraint [3, 5].'); scalar_tensor_default_1 = None return pytree.tree_unflatten([nonzero_default], self._out_spec) ``` Pull Request resolved: https://github.com/pytorch/pytorch/pull/100763 Approved by: https://github.com/tugsbayasgalan	2023-05-09 04:19:57 +00:00
Angela Yi	2d2f716ddc	[export] Fix cond for pass_base (#100836 ) I ported over the code for the inline interpreter incorrectly in the pass base 😅 Originally the function `make_inline_interpreter` is supposed to take in a fx.Interpreter type but I accidentally passed in an fx.Interpreter object. Also realized while modifying this diff (and comments from Tugsuu) that we don't really need this InlineInterpreter. Pull Request resolved: https://github.com/pytorch/pytorch/pull/100836 Approved by: https://github.com/zhxchen17, https://github.com/tugsbayasgalan	2023-05-08 21:51:03 +00:00
Tugsbayasgalan Manlaibaatar	9b3552eb2c	Add runtime assertions for input shape constraints (#100247 ) This PR adds runtime assertions as an extra pass in the exported graph. Several high level information: 1. We specialize all dimensions that were not added to the user input constraints 2. We haven't added relational constraints as runtime assertions (e.g x[1] == x[0]), will do in a follow up diff Differential Revision: [D45408971](https://our.internmc.facebook.com/intern/diff/D45408971) Pull Request resolved: https://github.com/pytorch/pytorch/pull/100247 Approved by: https://github.com/guangy10, https://github.com/avikchaudhuri	2023-05-04 13:26:58 +00:00
Angela Yi	7bece142a9	[export] Port over const prop pass (#100102 ) Stacked on top of https://github.com/pytorch/pytorch/pull/100000 Pull Request resolved: https://github.com/pytorch/pytorch/pull/100102 Approved by: https://github.com/gmagogsfm	2023-04-27 17:06:47 +00:00
Angela Yi	9bbd3d6489	[export] ExportPassBase + view_copy pass (#100000 ) * Added ExportPassBase, an interpreter based helper pass writing class * It can also help maintain the dialect based on the operator namespace through having users override the `get_valid_dialects` function (returning an empty lists implies the pass works for any dialect). * Added a `ReplaceBrokenOpsWithFunctionalOpsPass` to replace all ops that have not been converted with functionalization with their functional ones. Pull Request resolved: https://github.com/pytorch/pytorch/pull/100000 Approved by: https://github.com/gmagogsfm	2023-04-26 21:01:25 +00:00

1 2 3

122 Commits