pytorch

mirror of https://github.com/zebrajr/pytorch.git synced 2025-12-07 12:21:27 +01:00

Author	SHA1	Message	Date
PyTorch MergeBot	bce21ee06a	Revert "Fix bug in check required output size in _as_strided_scatter_meta (#98483 )" This reverts commit `5b692fd819`. Reverted https://github.com/pytorch/pytorch/pull/98483 on behalf of https://github.com/malfet due to Broke inductor, see https://hud.pytorch.org/hud/pytorch/pytorch/main/1?per_page=50&name_filter=inductor%2C%201%2C%201	2023-04-18 18:59:47 +00:00
Richard Zou	57e1a50da3	Fix FakeTensor printing (#99205 ) I got too confused by the FakeTensor printing, so this PR fixes it to print normally. Before: ``` with FakeTensorMode(): x = torch.empty(2, 2, device="cpu") print(x) # FakeTensor(FakeTensor(..., device='meta', shape=(2, 2)), cpu) ``` After (Tensor printing doesn't print the default device): ``` FakeTensor(..., shape=(2, 2)) ``` Test Plan: - new test Pull Request resolved: https://github.com/pytorch/pytorch/pull/99205 Approved by: https://github.com/eellison	2023-04-18 13:26:27 +00:00
lantiankaikai	5b692fd819	Fix bug in check required output size in _as_strided_scatter_meta (#98483 ) Original Issue from #92670 pytest ./generated/test_XuyangBai_PointDSC.py -k test_004 ==> RuntimeError: as_strided_scatter: sizes [4], strides [85], storage offset 256 and itemsize 4 requiring a storage size of 2048 are out of bounds for storage of size 1024 Repro: ``` class Model(nn.Module): def __init__(self): super(Model, self).__init__() def forward(self, x): x[1].fill_diagonal_(0) # this check size failed device = torch.device("cpu") model = Model() model.to(device) torch._dynamo.reset() compiled_model = torch._dynamo.optimize("inductor")(model) arg = [torch.rand([4, 1, 1])] compiled_model(*arg) ``` The error was raised at the checking required size in as_strided_scatter. https://github.com/pytorch/pytorch/blob/master/torch/_prims/__init__.py#L1818 In the case of input is a tensor with storage offset(a view), when compute input's storage length, should also take input's base tensor's size/stride/offset into account instead of compare it with number of element of input. This diff fix the bug and add test. Fixes #ISSUE_NUMBER Pull Request resolved: https://github.com/pytorch/pytorch/pull/98483 Approved by: https://github.com/ngimel	2023-04-18 05:07:57 +00:00
Natalia Gimelshein	888c65b6a4	fix fake tensor propagation for cross_entropy with smoothing (#99255 ) Fixes #99250, unfortunately I haven't figured out how to handle cross-entropy with smooth loss and weights. Pull Request resolved: https://github.com/pytorch/pytorch/pull/99255 Approved by: https://github.com/jansel, https://github.com/malfet	2023-04-17 00:31:26 +00:00
andrewor14	651c1be885	Recompute flat_arg_fake_tensors after fakification (#98769 ) Summary: This fixes the case when some of the input tensors were real tensors and fakified in `validate_and_convert_non_fake_tensors`, but `flat_arg_fake_tensors` would not contain all the inputs because it was computed before the fakification. We fix this by recomputing `flat_arg_fake_tensors` after fakification as well. Test Plan: python test/dynamo/test_export.py ExportTests.test_mixed_real_and_fake_inputs Reviewers: Chillee, voznesenskym Pull Request resolved: https://github.com/pytorch/pytorch/pull/98769 Approved by: https://github.com/voznesenskym	2023-04-14 19:14:29 +00:00
Elias Ellison	fc53472ce4	Move/Fix FakeTensor logic for detecting multiple fake modes (#97186 ) This was leftover for when we had more logic in the FakeTensor and not FakeTensorMode, and wasn't firing correctly. It also makes more sense for it to be in the other validation function. Pull Request resolved: https://github.com/pytorch/pytorch/pull/97186 Approved by: https://github.com/bdhirsh	2023-04-13 19:20:01 +00:00
PyTorch MergeBot	4828585019	Revert "Move/Fix FakeTensor logic for detecting multiple fake modes (#97186 )" This reverts commit `8a057c445d`. Reverted https://github.com/pytorch/pytorch/pull/97186 on behalf of https://github.com/huydhn due to This breaks ONNX test in trunk and it looks like a landrace as the CI signal is green	2023-04-12 19:24:54 +00:00
Elias Ellison	8a057c445d	Move/Fix FakeTensor logic for detecting multiple fake modes (#97186 ) This was leftover for when we had more logic in the FakeTensor and not FakeTensorMode, and wasn't firing correctly. It also makes more sense for it to be in the other validation function. Pull Request resolved: https://github.com/pytorch/pytorch/pull/97186 Approved by: https://github.com/bdhirsh	2023-04-12 17:40:41 +00:00
Elias Ellison	445863128b	Use .to instead of contiguous to generate channels last tensor (#96791 ) Fix for https://github.com/pytorch/pytorch/issues/95693. From https://pytorch.org/tutorials/intermediate/memory_format_tutorial.html: > There are minor difference between the two APIs to and contiguous. We suggest to stick with to when explicitly converting memory format of tensor. For general cases the two APIs behave the same. However in special cases for a 4D tensor with size NCHW when either: C==1 or H==1 && W==1, only to would generate a proper stride to represent channels last memory format. We hit this case in convolution_backward in calling `contiguous()`. Even though we were determining that we should run the backward in channels_last forward, as FakeTensor had gathered from the output of [determine_backend_memory_format](https://github.com/pytorch/pytorch/blob/master/torch/_subclasses/fake_tensor.py#L559), we were still outputting a contiguous tensor. That led to the mismatch in strides in the issue. Should we be calling `to` instead of `contiguous` more liberally throughout the codebase, especially in convolution related code ? Not sure if there are reasons not to do this. Another fix would be to update `cudnn_conv_suggest_memory_format` so that it would output a contiguous_format in this case. Pull Request resolved: https://github.com/pytorch/pytorch/pull/96791 Approved by: https://github.com/ngimel	2023-03-15 19:12:04 +00:00
Xuehai Pan	046e88a291	[BE] [3/3] Rewrite `super()` calls in test (#94592 ) Rewrite Python built-in class `super()` calls. Only non-semantic changes should be applied. - #94587 - #94588 - #94592 Also, methods with only a `super()` call are removed: ```diff class MyModule(nn.Module): - def __init__(self): - super().__init__() - def forward(self, ...): ... ``` Some cases that change the semantics should be kept unchanged. E.g.: `f152a79be9/caffe2/python/net_printer.py (L184-L190)` `f152a79be9/test/test_jit_fuser_te.py (L2628-L2635)` Pull Request resolved: https://github.com/pytorch/pytorch/pull/94592 Approved by: https://github.com/ezyang, https://github.com/seemethere	2023-02-12 22:20:53 +00:00
Wei-Sheng Chin	c5c7687b74	Allow FakeTensorProp to run on graphs traced with some None inputs (#94569 ) Without this tiny change in `torch/_subclasses/fake_tensor.py`, the added test may fail with ``` TypeError: cannot create weak reference to 'NoneType' object ``` Pull Request resolved: https://github.com/pytorch/pytorch/pull/94569 Approved by: https://github.com/ezyang	2023-02-10 20:38:22 +00:00
PyTorch MergeBot	f152a79be9	Revert "update aten op overload to not use `from` to avoid compile errors (#89797 )" This reverts commit `021d267694`. Reverted https://github.com/pytorch/pytorch/pull/89797 on behalf of https://github.com/jeanschmidt due to breaking internal builds - more details on https://fburl.com/sandcastle/bz8mgkil	2023-02-10 11:32:25 +00:00
Elias Ellison	021d267694	update aten op overload to not use `from` to avoid compile errors (#89797 ) Fix for https://github.com/pytorch/pytorch/issues/93591 by changing `random_.from` to `random_.from_int`. The previous signature would fail when printed in an fx graph, because `from` is a reserved python keyword. This change affects serialization but I have added an adapter. Pull Request resolved: https://github.com/pytorch/pytorch/pull/89797 Approved by: https://github.com/tugsbayasgalan	2023-02-08 22:04:59 +00:00
Yanbo Liang	605b661805	FakeTensor should constant propagate through ops that allow numbers as scalars (#94145 ) Fixes #92655 Thanks @eellison for the code change suggestion. Pull Request resolved: https://github.com/pytorch/pytorch/pull/94145 Approved by: https://github.com/eellison	2023-02-07 06:20:35 +00:00
Elias Ellison	e4f11e01bd	[Fake Tensor] Allow fake meta by default, delete unused ctor args (#93993 ) Two small changes that I'm bundling together because one of them needs to touch fbcode and I'm not sure how to do stacked diffs + internal changes + land before release cut. Remove allow_meta from ctor, and allow by default: we should be able to trace through meta with fake tensors, so in some senses it's a bit weird to expose to user to disallow this. However, it's still useful debug wise to error from time to time, so I've added an option to the config that will get back previous behavior. Remove `throw_on_data_dependent_ops=True`: this was intended as a temporary behavior as we were smoothing things turning on the erroring. There are no uses anywhere of `throw_on_data_dependent_ops=False` I could find. These are technically backward-incompatble, but fake tensor is new since the last release / in a private namespace, and I don't want to release it with baggage that would be hard to remove later. Fix for https://github.com/pytorch/pytorch/issues/92877. Pull Request resolved: https://github.com/pytorch/pytorch/pull/93993 Approved by: https://github.com/bdhirsh, https://github.com/ezyang	2023-02-03 09:23:38 +00:00
Edward Z. Yang	1237cf6b6c	Allow direct Tensor constructor to return preexisting PyObject (#92754 ) Signed-off-by: Edward Z. Yang <ezyang@meta.com> Pull Request resolved: https://github.com/pytorch/pytorch/pull/92754 Approved by: https://github.com/albanD, https://github.com/voznesenskym	2023-01-23 20:20:43 +00:00
Brian Hirsh	76cb2d0ede	fix incorrect _embedding_bag meta (#92549 ) Fixes https://github.com/pytorch/pytorch/issues/92286. See the issue for diagnosis. Pull Request resolved: https://github.com/pytorch/pytorch/pull/92549 Approved by: https://github.com/albanD, https://github.com/eellison	2023-01-18 22:50:31 +00:00
samdow	d8e795ecd5	[modes] make python arg parser also check for python key (#91573 ) Fixes #90652 Previously, we had assumed that the only way to call `handle_torch_function_no_python_arg_parser` was through the Python key. This is no longer true with FakeTensor. Specifically `_like` functions will call `.device()` on FakeTensors when the args list is being parsed. In order to respect that the mode stack shouldn't run when the python key is off, this just adds that a check that the python key is on/the torch_function equivalent to that function Pull Request resolved: https://github.com/pytorch/pytorch/pull/91573 Approved by: https://github.com/ezyang	2023-01-11 15:19:43 +00:00
Yanbo Liang	789b1437e9	Fix meta registration for aten._cudnn_rnn (#91333 ) Found this issue from [weekly running 7k github models](https://github.com/pytorch/torchdynamo/issues/1884). This caused regression on pass rate, there are 25 models failed due to this issue. The reason is argument ```cx``` of ```aten._cudnn_rnn``` can be ```None```, but it doesn't handle well in meta registration, so throws the following error: ``` Traceback (most recent call last): File "/scratch/ybliang/work/repos/pytorch/torch/_dynamo/utils.py", line 1059, in run_node return nnmodule(args, kwargs) File "/scratch/ybliang/work/repos/pytorch/torch/nn/modules/module.py", line 1482, in _call_impl return forward_call(args, *kwargs) File "/scratch/ybliang/work/repos/pytorch/torch/nn/modules/rnn.py", line 477, in forward result = _VF.rnn_tanh(input, hx, self._flat_weights, self.bias, self.num_layers, File "/scratch/ybliang/work/repos/pytorch/torch/_subclasses/fake_tensor.py", line 916, in __torch_dispatch__ r = func(args, *kwargs) File "/scratch/ybliang/work/repos/pytorch/torch/_ops.py", line 284, in __call__ return self._op(args, **kwargs or {}) File "/scratch/ybliang/work/repos/pytorch/torch/_meta_registrations.py", line 2108, in _cudnn_rnn cy = cx.new_empty(0 if cx is None else cell_shape) AttributeError: 'NoneType' object has no attribute 'new_empty' ``` Pull Request resolved: https://github.com/pytorch/pytorch/pull/91333 Approved by: https://github.com/ezyang	2022-12-23 22:59:31 +00:00
Joel Schlosser	3226209636	LSTM SymInt-aware changes & meta registration (cuDNN) (#90944 ) Pull Request resolved: https://github.com/pytorch/pytorch/pull/90944 Approved by: https://github.com/ezyang	2022-12-16 21:42:32 +00:00
Joel Schlosser	b0cda0b38c	LSTM SymInt-aware changes & meta registration (non-cuDNN CUDA) (#90701 ) Adds meta registrations for cuDNN and vanilla CUDA ops underneath `lstm()` and makes the logic SymInt-aware. TODO: * cuDNN side does some [nasty stuff](https://github.com/pytorch/pytorch/blob/master/aten/src/ATen/native/cudnn/RNN.cpp#L1567) with buffers; this needs larger redesign to figure out * Indicate that AOT Autograd can be used when an LSTM is present (remove the check for this once it's fully supported) Pull Request resolved: https://github.com/pytorch/pytorch/pull/90701 Approved by: https://github.com/ezyang	2022-12-16 18:08:45 +00:00
Edward Z. Yang	e686a442b4	If a torch.* returns non-Tensor, make this unimplemented rather than assert. (#89918 ) Signed-off-by: Edward Z. Yang <ezyang@fb.com> Pull Request resolved: https://github.com/pytorch/pytorch/pull/89918 Approved by: https://github.com/albanD	2022-12-15 21:53:54 +00:00
Tugsbayasgalan (Tugsuu) Manlaibaatar	1aab755320	Fakify params and weights under private config (#90417 ) Previously, we planned to lift the parameters and weights while exporting and implement our own transformer to "unlift" the lifted weights and params back to the graph as attributes. But this is bit challenging because: - We need to maintain correct ordering for weights and parameters that are passed as inputs so that we know how to map them back. - Some weights are unused in the graph, so our transformer needs to be aware of which weights and parameters are not used in the graph. And we need to distinguish which are real user input and which are parameters. - There can be more edge cases we haven't seen in other models yet. I am aware that @Chillee and @bdhirsh mentioned that functionalization won't work with fake-tensor attributes but this is fine for the short term as we don't expect users to be modifying weights and params in inference mode. In fact, we explicitly disable attribute mutation in torchdynamo export mode right now. Given above condition, it might be ok to just fakify params when we need. I use a flag to guard against this change. Differential Revision: [D41891201](https://our.internmc.facebook.com/intern/diff/D41891201) Pull Request resolved: https://github.com/pytorch/pytorch/pull/90417 Approved by: https://github.com/eellison	2022-12-14 09:33:18 +00:00
Elias Ellison	1a33b7cbfa	Make fake tensors preserve dense strides in type conversion (#89803 ) Pull Request resolved: https://github.com/pytorch/pytorch/pull/89803 Approved by: https://github.com/ngimel	2022-11-30 01:28:51 +00:00
Elias Ellison	72110d7833	Fix Upsample Decomp Striding For Small Channels (#89528 ) Fix for https://github.com/pytorch/torchdynamo/issues/623. Pull Request resolved: https://github.com/pytorch/pytorch/pull/89528 Approved by: https://github.com/ngimel, https://github.com/anijain2305	2022-11-23 20:47:39 +00:00
Wei-Sheng Chin	86b7aa26f0	Fix FakeTensorProp on Module with Parameters or Buffers (#88700 ) In `FakeTensorMode.__torch_dispatch__`, the output is now always computed by meta kernels in ```python try: with in_kernel_invocation_manager(self): r = func(args, *kwargs) # <----- "r" can be a real tensor. except NotImplementedError as not_implemented_error: # no meta kernel registered, fallback to kernel for the device if not self.allow_fallback_kernels: raise not_implemented_error return run_fallback_kernel(self, func, args, kwargs, not_implemented_error) return self.wrap_meta_outputs_with_default_device_logic(r, func, args, kwargs) ``` For example, I observed a CPU tensor is generated when executing `aten.addmm` when running `FakeTensorProp`. Therefore, I'd like to allow `FakeTensorMode` to wrap real tensor as `FakeTensor` during the computation. Does this PR look a good direction to fix this problem? If yes, I can go ahead and add some tests. Pull Request resolved: https://github.com/pytorch/pytorch/pull/88700 Approved by: https://github.com/eellison, https://github.com/ezyang	2022-11-11 03:49:29 +00:00
Elias Ellison	2ce2fc133d	Disable Current Modes when printing Tensor (#88344 ) Fix for https://github.com/pytorch/pytorch/issues/88087 Pull Request resolved: https://github.com/pytorch/pytorch/pull/88344 Approved by: https://github.com/ezyang, https://github.com/samdow	2022-11-04 00:45:35 +00:00
Edward Z. Yang	c2c269c10a	Convert MetaConverter's tensor memo into a weak value dictionary. (#87911 ) This is in preparation for unifying fake tensor converter and meta converter's memo tables. Signed-off-by: Edward Z. Yang <ezyang@fb.com> Pull Request resolved: https://github.com/pytorch/pytorch/pull/87911 Approved by: https://github.com/eellison	2022-10-28 21:05:13 +00:00
Edward Z. Yang	e72962a34d	Force people to call from_meta_and_device directly (#87903 ) It was pretty hard to tell at call site if I was doing device meta convert or not. This gets rid of the "dual" API and forces people to call the method manually for the device case. Signed-off-by: Edward Z. Yang <ezyang@fb.com> Pull Request resolved: https://github.com/pytorch/pytorch/pull/87903 Approved by: https://github.com/eellison, https://github.com/albanD	2022-10-28 21:05:13 +00:00
Elias Ellison	fc21b9db23	Use Eager Code To Determine Conv Layout (#87305 ) The logic for determine conv backend and therefore output striding is very complex. It depends on build settings, input striding/contiguity, sizes, etc. Eventually we should port that logic to the meta impl for dynamic shapes but that will require a lot more work and keeping the implementations in sync. See https://github.com/pytorch/torchdynamo/issues/1701 This is a prerequisite to removing the inductor conv stride propagation and more general fake tensor for inductor propagation. In that PR, the meta impls for cpu conv give incorrect striding which led to test failures (https://github.com/pytorch/pytorch/pull/87083). Pull Request resolved: https://github.com/pytorch/pytorch/pull/87305 Approved by: https://github.com/ezyang	2022-10-28 16:37:04 +00:00
Jagadish Krishnamoorthy	9efca7c085	[ROCm] [FakeTensorTest] Enable test_fallback_memory_prop (#85760 ) Signed-off-by: Jagadish Krishnamoorthy <jagdish.krishna@gmail.com> Pull Request resolved: https://github.com/pytorch/pytorch/pull/85760 Approved by: https://github.com/kit1980	2022-10-25 07:17:47 +00:00
Elias Ellison	d3f7c34cb3	Enable aten-aten decomps (#85921 ) Invokes aten-aten decomps with re-entrant FakeMode. These decomps are being used in other places, so it's good to unify the path static fake tensor takes / get additional testing etc. There is also an instance where we return different devices with cpu/cuda which this fixes ([batch_norm](https://github.com/pytorch/pytorch/blob/master/torch/_decomp/decompositions.py#L1374)) Pull Request resolved: https://github.com/pytorch/pytorch/pull/85921 Approved by: https://github.com/ezyang	2022-10-08 05:12:42 +00:00
PyTorch MergeBot	7ec12a559c	Revert "Enable aten-aten decomps (#85921 )" This reverts commit `62e4f51efd`. Reverted https://github.com/pytorch/pytorch/pull/85921 on behalf of https://github.com/huydhn due to Sorry for reverting your PR. I think it breaks a dynamo test in trunk `62e4f51efd`	2022-10-08 01:59:54 +00:00
Elias Ellison	62e4f51efd	Enable aten-aten decomps (#85921 ) Invokes aten-aten decomps with re-entrant FakeMode. These decomps are being used in other places, so it's good to unify the path static fake tensor takes / get additional testing etc. There is also an instance where we return different devices with cpu/cuda which this fixes ([batch_norm](https://github.com/pytorch/pytorch/blob/master/torch/_decomp/decompositions.py#L1374)) Pull Request resolved: https://github.com/pytorch/pytorch/pull/85921 Approved by: https://github.com/ezyang	2022-10-07 21:04:39 +00:00
Elias Ellison	9da5646cdb	Add device logic handling for functions which allow scalar inputs as tensors (#86149 ) Some functions allow scalars as tensor inputs. Add handling for them in device logic. Fix for https://github.com/pytorch/torchdynamo/issues/1445 Pull Request resolved: https://github.com/pytorch/pytorch/pull/86149 Approved by: https://github.com/ezyang, https://github.com/bdhirsh	2022-10-04 18:54:00 +00:00
Elias Ellison	e1859c0707	delete special fake tensor new handling (#86144 ) Delete the special-cased handling of `new` in FakeTensor. Ever since the dispatch keys were updated to reflect the FakeTensor's device, the special cased handling was not needed. Fixes https://github.com/pytorch/torchdynamo/issues/1448 Pull Request resolved: https://github.com/pytorch/pytorch/pull/86144 Approved by: https://github.com/ezyang	2022-10-04 16:08:52 +00:00
Elias Ellison	f183a989a2	Fix fake tensor kernel nesting (#85920 ) If you e.g. printed within a decomp which would call `in_kernel_invocation_manager`, on the exit from the manager it would unilaterally remove meta from the tls / set the tensor to return its real device. We should just restore what the existing state was. Pull Request resolved: https://github.com/pytorch/pytorch/pull/85920 Approved by: https://github.com/ezyang, https://github.com/bdhirsh, https://github.com/huydhn	2022-10-02 04:19:40 +00:00
PyTorch MergeBot	b562987c28	Revert "Fix fake tensor kernel nesting (#85920 )" This reverts commit `c2d9ea7f4b`. Reverted https://github.com/pytorch/pytorch/pull/85920 on behalf of https://github.com/huydhn due to Sorry for reverting your PR but I suspect that it causes a flaky memory leak issue in TestFakeTensorCUDA.test_fake_crossref_backward_amp_linalg_lstsq_cuda_float32	2022-10-01 19:30:21 +00:00
Elias Ellison	c2d9ea7f4b	Fix fake tensor kernel nesting (#85920 ) If you e.g. printed within a decomp which would call `in_kernel_invocation_manager`, on the exit from the manager it would unilaterally remove meta from the tls / set the tensor to return its real device. We should just restore what the existing state was. Pull Request resolved: https://github.com/pytorch/pytorch/pull/85920 Approved by: https://github.com/ezyang, https://github.com/bdhirsh	2022-09-30 23:11:20 +00:00
Edward Z. Yang	3b6588ab74	Consistent compute numel/contiguous strategy with SymInts (#85858 ) Previously, our handling for contiguity was inconsistent in the following ways: - is_strides_like 2d/3d and is_non_overlapping_and_dense always were computed based on sizes_and_strides_, even if you had symbolic ints - Furthermore, even if you set custom policy for strides, these quantities were not overridable by subclasses - Furthermore, we didn't even store these fields on ExtraMeta - We duplicate implementations of compute_contiguous (plain, channels last, channels last 3d) - We inconsistently called refresh_numel()/refresh_contiguous(), versus recomputing it ourselves This factor makes a consistent strategy for all of the boolean fields, and for numel computation. After this refactor: - All layout boolean fields are interposable via strides policy and can be overridden from Python; you will never access a garbage field - All layout boolean fields are on ExtraMeta - You can always call refresh_numel/contiguous, no matter if your Tensor is contiguous or not - The numel/layout boolean fields are always populated consistently with the sizes strides fields (either on Tensor or ExtraMeta), even if you have custom policy - There is only one implementation of the actual computation logic Signed-off-by: Edward Z. Yang <ezyang@fb.com> Differential Revision: [D39907696](https://our.internmc.facebook.com/intern/diff/D39907696) Pull Request resolved: https://github.com/pytorch/pytorch/pull/85858 Approved by: https://github.com/albanD	2022-09-30 21:26:34 +00:00
Elias Ellison	75db0225ad	Handle fake tensor in intlist (#85759 ) Previously, we were swallowing up the Fake Tensor Exception and throwing `TypeError`, which led to https://github.com/pytorch/torchdynamo/issues/1066. Now, we are propagating back the `DataDependentOutputException`. If this approach is accepted, I can go ahead and do doublelist, symintlist, afterward. Pull Request resolved: https://github.com/pytorch/pytorch/pull/85759 Approved by: https://github.com/ezyang	2022-09-28 21:58:54 +00:00
samdow	18d8c548f4	[Modes] remove enable and rewrite mode stack (squashed) (#84774 ) Based on @ezyang's suggestion, mode stack now has "one true mode" which is the _only_ mode that can ever be active at the C++ level. That mode's torch dispatch is just to take the top mode in the stack, reenable itself (if we aren't at the end of the mode stack), and run the top mode's torch_{dispatch\|function} This maintains that in the middle of a mode's torch dispatch, the mode itself will not be active. It changes the function the user has to call to see what the current mode is (no longer queries the C++, it's python only) but allows the user to also see the entire mode stack easily Removes `enable_torch_dispatch_mode` and `.restore()` since neither makes sense in this new setup ### Background Why do we want this? Well, a pretty common pattern that was coming up was that users had to do something like ```python ## PRE-PR UX def f(mode): with mode.restore(): # user needs to understand this restore thing? ... with Mode() as m: pass f(m) ``` Many users were getting error from forgetting to call `.restore` or from forgetting to add the (tbh weird) "mode instantiation" step where they use the mode as a context manager with an empty body. Really, they wanted to treat modes like context managers and just write ```python ## FROM FEEDBACK, USER DESIRED CODE. POSSIBLE POST-PR def f(mode): with mode: ... f(Mode()) ``` Technical Details With the old mode stack, we basically had a linked list so the mode itself could only be used once and had a fixed parent. In this new design, the mode stack is just a python list that we're pushing to and popping from. There's only one mode that's ever active at the C++ level and it runs the next mode in the Python list. The modes don't have state on them anymore Pull Request resolved: https://github.com/pytorch/pytorch/pull/84774 Approved by: https://github.com/ezyang, https://github.com/zou3519	2022-09-27 01:04:35 +00:00
Elias Ellison	90261945b7	Copy over non parameter grad (#85658 ) Wow, ugh silly mistake. Fix for https://github.com/pytorch/torchdynamo/issues/1291 not even sure how all the tests passed before this. Pull Request resolved: https://github.com/pytorch/pytorch/pull/85658 Approved by: https://github.com/Chillee, https://github.com/anijain2305	2022-09-26 23:04:01 +00:00
Edward Z. Yang	c5a8946e40	Revert "Revert "Redo how custom/python_custom methods on TensorImpl work (#84796 )" (#84806 ) This reverts commit `ca3b2bfbe3`. Pull Request resolved: https://github.com/pytorch/pytorch/pull/84806 Approved by: https://github.com/Chillee	2022-09-10 06:17:35 +00:00
Eli Uriegas	ca3b2bfbe3	Revert "Redo how custom/python_custom methods on TensorImpl work (#84796 ) This reverts commit `591b75bf98`. Manual revert of https://github.com/pytorch/pytorch/pull/84641 Fixes #ISSUE_NUMBER Pull Request resolved: https://github.com/pytorch/pytorch/pull/84796 Approved by: https://github.com/izaitsevfb	2022-09-10 00:18:13 +00:00
Edward Z. Yang	591b75bf98	Redo how custom/python_custom methods on TensorImpl work (#84641 ) A longstanding confusion in the implementation of fake tensor and proxy tensor is what to do about torch.ops.aten.sym_sizes and related calls. In particular, when you have a tensor that (1) has symbolic shapes and (2) has a `__torch_dispatch__` call, previously, you would always get `__torch_dispatch__` calls for sizes/strides query, even if you didn't request it via the dispatch kwargs in `make_wrapper_subclass`. The reason for this is because we were previously mixing several concepts: "I want to dispatch to Python", "I want to call a virtual method" and "I have dynamic shapes". A single boolean variable controlled all of these things, and so it was not possible to understand inside TensorImpl what the user had actually originally requested. In this PR, we track each of these concepts individually so that we can preserve user intent. Then, we combine these into a single "policy" variable that controls whether or not we can use the fastpath or not. For the policy to trigger, we only need one of the exceptional cases to be true. Billing of changes: * Rename `set_sizes_strides_policy` to `set_custom_sizes_strides`; in general, you cannot DIRECTLY set policy; you have to indirectly set it by the public functions. * Some helpers for sizes and strides, since it's more complicated (as it is an enum, rather than just bools as is the case for device and layout). `matches_python_custom` is used to test the Python dispatch user ask. `matches_policy` does the policy test (only used in the user facing functions.) * I reorged the accessor methods so that they are more logical. This makes the diff bad, so I recommend reading the final code directly. * The default custom implementations now more reliably call their default() implementations * As bonus refactor, I devirtualized some functions that don't need to be virtual * `set_sym_sizes_and_strides` is renamed to `set_sizes_and_strides` to make it easier to use in template contexts; it optionally takes a storage offset now so you can set all three values at the same time. If you use the SymInt overload but there are no symbolic integers, we give you a normal resize. * This adds `sym_storage_offset` since we had that in the symbolic shapes branch and there's no reason not to put it in (and it reduces merge conflicts) Signed-off-by: Edward Z. Yang <ezyang@fb.com> Pull Request resolved: https://github.com/pytorch/pytorch/pull/84641 Approved by: https://github.com/wconstab	2022-09-09 13:41:13 +00:00
Elias Ellison	15c5baf878	Throw on data dependent ops (#83567 ) Previously, we would trace through the following with no error: ``` from torch.fx.experimental.proxy_tensor import make_fx import torch def f(x, y): return x[0, y:] ``` Even though the output shape is dependent on the data of `y`. Now, throw on the conversion of `y` to an integer. It would be nice to not break on constant tensors but I'll do that as the next PR (Edit: done with https://github.com/pytorch/pytorch/pull/84387). Sketching out how that would work (and keep in mind this is applicable Dynamo tracing and not just AOT Autograd) I think to do that you would need to : - hold strong refs to a set of constant tensors, and only allow them to be captured from `lift_fresh.copy` - when you run a mutable op, either remove it from the set of constant tensors or run the operator for real - limit to small constant tensors Anything else ? Pull Request resolved: https://github.com/pytorch/pytorch/pull/83567 Approved by: https://github.com/ezyang	2022-09-07 02:37:00 +00:00
Elias Ellison	97b2dff600	Add Initial Support For Fake Tensor Constant Tracking (#84387 ) Adds support for constant tensor tracking within FakeTensors. Copy-pasta'ing from `proxy_tensor.py` why this is useful: ``` # In some circumstances, we will be tracing in a situation where a tensor # is statically known to be a constant (currently, this only happens if # you run torch.tensor; deterministic factory functions like torch.arange # don't get this treatment). When the tensor in question is small, it's # helpful to due constant propagation in case we call item() (in which # case we can return the constant value that is known, rather than give # an error.) ``` This PR only attempts to add support for the tracing scenarios where we run each operation linearly - aot autograd, torchdynamo. It does not yet handle how constant tensors should be handled as part of the persistent fx graph. Additionally, it does not yet attempt to de-duplicate or interact with ProxyMode's only constant tensor handling. Edit: plan is to rely on functionalization for fx graph Pull Request resolved: https://github.com/pytorch/pytorch/pull/84387 Approved by: https://github.com/ezyang	2022-09-02 02:43:04 +00:00
Elias Ellison	9c452abcf1	Use reentrant mode when invoking prims, delete global prim_fake_mode (#84090 ) Maybe I should be using the meta_impl instead of the prim_impl, but it's not terribly clear why, since the prim impl will be better tested and should work under the re-entrant FakeTensorMode. Fixes https://github.com/pytorch/pytorch/issues/78613 in the process Pull Request resolved: https://github.com/pytorch/pytorch/pull/84090 Approved by: https://github.com/ezyang, https://github.com/samdow	2022-08-31 01:58:44 +00:00
samdow	7532d5b125	[Modes] remove inner constructor kwarg (#83925 ) Pull Request resolved: https://github.com/pytorch/pytorch/pull/83925 Approved by: https://github.com/ezyang, https://github.com/zou3519	2022-08-31 00:05:56 +00:00
Elias Ellison	8a6b076196	lift numpy tensor, add randperm support (#83191 ) Couple changes needed to trace huggingface w fake tensors. Similar to https://github.com/pytorch/pytorch/pull/81927, need to call liftfresh for tensors created from numpy tensors. Also adds randperm for meta. Pull Request resolved: https://github.com/pytorch/pytorch/pull/83191 Approved by: https://github.com/bdhirsh	2022-08-10 22:27:51 +00:00
Elias Ellison	4f61aa79dd	Update nan_to_num prims (#82908 ) Pull Request resolved: https://github.com/pytorch/pytorch/pull/82908 Approved by: https://github.com/voznesenskym, https://github.com/ngimel	2022-08-08 14:28:28 +00:00
Edward Z. Yang	42fefd4403	Sparse fake tensor support (#82172 ) Add support for sparse fake tensors. - The testing strategy is to run a fake tensor cross ref test on `test_sparse.py`. This is necessary because OpInfo sparse coverage is completely nonexistent. We could have tried to turn on cross ref testing globally for all files, but that would be very time consuming and the tests I'm interested in are mostly in this file. There are some exclusions in testing for things that don't work. - I make fake tensor converter raise a UnsupportedFakeTensorException if the meta converter fails to do a conversion (which can happen in a relatively large number of situations). - I relax fake tensor invariants so that you can make a fake tensor from a meta tensor. This is useful because in the cross ref test sometimes we operate on meta tensors. - Fake tensor wrapping is improved to handle the case when a function doesn't return any tensors - Meta converter is taught how to convert sparse tensors to meta There's still a little more cleanup that needs to be done, but this is good for review. Signed-off-by: Edward Z. Yang <ezyang@fb.com> Pull Request resolved: https://github.com/pytorch/pytorch/pull/82172 Approved by: https://github.com/eellison	2022-08-03 14:29:36 +00:00
Elias Ellison	09a6d0b5bb	Fake: copy over grad attribute (#82593 ) Copy over the grad attribute... let me any suggestions of other tests/opinfos/crossrefs I should be doing to make this more comprehensive. Pull Request resolved: https://github.com/pytorch/pytorch/pull/82593 Approved by: https://github.com/ezyang	2022-08-02 04:08:42 +00:00
Elias Ellison	642aed8b99	Add Autocast Support for FakeTensors / use fake device dispatch keys (#82449 ) From PR: ``` Note: [Fake Tensor Dispatch Keys] In order to model the behavior of device-specific autocast and autograd logic, we update the dispatch keys of FakeTensors to reflect their fake device. This includes the BackendComponent (DispatchKey::Meta -> DispatchKey::CUDA), and also the BackendComponent related Autocast and Autograd keys. __torch__dispatch__ sits below Autocast and Autograd, and is only invoked when we are at the kernel for the BackendComponent. Then, we add Meta to the thread-local dispatch include set to hit the meta kernel instead of the kernel of the BackendComponent for the fake device. ``` Also adds the `conv1/2/3d.padding` operators to the Autocast rule set. Without that fix, the FakeTensor dtype would diverge. See: https://github.com/pytorch/pytorch/issues/81608 Pull Request resolved: https://github.com/pytorch/pytorch/pull/82449 Approved by: https://github.com/ezyang	2022-08-01 21:40:36 +00:00
Elias Ellison	452af0bc44	Call lift fresh in valueToTensor (#81927 ) `valueToTensor` is invoked in many python apis such as `x[0] = 0.5`, which lifts the 0.5 to a tensor. the problem is described similarly in https://github.com/pytorch/pytorch/pull/81609 (/s/scalar_to_tensor/valueToTensor) > scalar_to_tensor is not dispatched and thus there is no interposition point for modes to ensure that the resulting tensor is appropriately wrapped. lift_fresh introduces this interposition point. This prevents FakeTensorMode from erroring Pull Request resolved: https://github.com/pytorch/pytorch/pull/81927 Approved by: https://github.com/ezyang	2022-07-26 22:07:59 +00:00
Robert	47b4ec8aa7	Conv2d shape calculation for meta tensors (#79834 ) Fixes #79512 This PR adds support for convolutional meta modules and computes the output shape correctly for some meta input tensor. Currently in progress, no tests written so far. Feature implementations: - [x] `Conv1d` - [x] `Conv2d` - [x] `Conv3d` Tests: - [x] `Conv1d` - [x] `Conv2d` - [x] `Conv3d` cc @albanD @anjali411 Pull Request resolved: https://github.com/pytorch/pytorch/pull/79834 Approved by: https://github.com/ezyang, https://github.com/albanD	2022-07-23 05:58:56 +00:00
samdow	2ac24675cc	get rid of push_torch_{dispatch, function}_mode (#78215 ) Currently we have 2 ways of doing the same thing for torch dispatch and function modes: `with push_torch_dispatch_mode(X)` or `with X.push(...)` is now the equivalent of doing `with X()` This removes the first API (which is older and private so we don't need to go through a deprecation cycle) There is some risk here that this might land race with a PR that uses the old API but in general it seems like most are using the `with X()` API or `enable_torch_dispatch_mode(X())` which isn't getting removed. EDIT: left the `with X.push(...)` API since there were ~3 land races with that over the past day or so. But made it give a warning and ask users to use the other API Pull Request resolved: https://github.com/pytorch/pytorch/pull/78215 Approved by: https://github.com/ezyang	2022-07-22 18:56:37 +00:00
Elias Ellison	15fef3dc7e	normalize cuda device (#81739 ) Fix for https://github.com/pytorch/pytorch/issues/81068 Pull Request resolved: https://github.com/pytorch/pytorch/pull/81739 Approved by: https://github.com/ezyang	2022-07-21 19:48:56 +00:00
Elias Ellison	745739d8f3	reland of #80545 w skip rocm (#81738 ) Pull Request resolved: https://github.com/pytorch/pytorch/pull/81738 Approved by: https://github.com/ezyang	2022-07-21 19:46:57 +00:00
Horace He	a5fb41e3d3	Revert "Revert "Refactored prim utils into _prims_utils folder (#81746 ) Pull Request resolved: https://github.com/pytorch/pytorch/pull/81746 Approved by: https://github.com/anijain2305, https://github.com/Krovatkin	2022-07-20 23:43:57 +00:00
PyTorch MergeBot	e43a02c314	Revert "Refactored prim utils into _prims_utils folder (#81088 )" This reverts commit `80231d0a72`. Reverted https://github.com/pytorch/pytorch/pull/81088 on behalf of https://github.com/jeanschmidt due to breaking internal tests	2022-07-19 19:56:41 +00:00
Horace He	80231d0a72	Refactored prim utils into _prims_utils folder (#81088 ) Pull Request resolved: https://github.com/pytorch/pytorch/pull/81088 Approved by: https://github.com/ngimel	2022-07-19 03:55:51 +00:00
PyTorch MergeBot	d8cd15b00c	Revert "Allow returned tensor impls in fallback (#80545 )" This reverts commit `d7e4520d1d`. Reverted https://github.com/pytorch/pytorch/pull/80545 on behalf of https://github.com/malfet due to New test broke rocm, see `d7e4520d1d`	2022-07-01 01:30:40 +00:00
Elias Ellison	d7e4520d1d	Allow returned tensor impls in fallback (#80545 ) Previously, we had disallowed any reused storages in fallback kernels because the correct aliasing relationships weren't set up correctly. However, we can support the case of a reused TensorImpl by just returning the input FakeTensor it corresponds to. Additionally, we have `inplace_view` as a tag which will prevent us from running fallback kernels for kernels that mutate metadata. (thanks @ezyang who originally suggested I do this). Fix for https://github.com/pytorch/torchdynamo/issues/467 Pull Request resolved: https://github.com/pytorch/pytorch/pull/80545 Approved by: https://github.com/ezyang	2022-06-30 21:57:54 +00:00
Elias Ellison	1058b47562	Weak-ref-ify MetaConverter and FakeTensorConverter (#80544 ) Make `MetaConverter` and `FakeTensorConverter` hold weak references to their memoized tensors, and also have `MetaConverter` hold weak reference to Tensor storage. Otherwise it can be tricky for users to make sure all existing FakeTensors or FakeTensorModes are deleted which otherwise will leak memory. I ran into https://github.com/pytorch/pytorch/issues/7733 which I was able to get around with the following (see comment from code): ``` # torch.Tensors cannot be used as a key in a dictionary # because they define a custom __eq__ function which when used # to resolve hash collisions will throw when comparing tensors: # "RuntimeError: bool value of Tensor with more than one value is ambiguous." # To avoid that, we use an object which will hold a Tensor and use # its id for both hashing and equality. # In order to use this as a weak key reference, we cannot # simply use weakref.WeakKeyDictionary because the newly constructed # WeakTensorRefKey only use would be a dictionary so it would have no strong # references. # To get around this issue, we can use it as a normal key, and then set # `weakref.finalize` to delete the key when its contained tensor dies. ``` While for the tensor memo we can set a `weakref.finalize` callback that will remove the corresponding `WeakTensorRefKey` from the tensor memo, at the point that this callback is invoked the tensor storage is not yet deallocated.. See comment from code: ``` # [expired-storages] # NB: even though the tensor has died, # the deallocation of its storage can take longer, # even when the storage has no other uses/views. # In this case, the StorageWeakRef object will be kept alive # longer than it needs to be, however the storage itself # will be deallocated. We retain the possibly dead storages # and periodically check if any of them are expired and # can be freed. ``` partial fix for https://github.com/pytorch/torchdynamo/issues/468 Pull Request resolved: https://github.com/pytorch/pytorch/pull/80544 Approved by: https://github.com/ezyang	2022-06-29 23:36:35 +00:00
Elias Ellison	c1d11f40c9	[FakeTensor] Use the device of the meta tensor for fallback kernel (#80193 ) Otherwise, I would run into issues such as fp16 not supported for Conv Cpu, or cudnn_rnn not implemented for cpu. Pull Request resolved: https://github.com/pytorch/pytorch/pull/80193 Approved by: https://github.com/Chillee	2022-06-24 20:00:07 +00:00
PyTorch MergeBot	35268bdc2a	Revert "[FakeTensor] Use the device of the meta tensor for fallback kernel (#80193 )" This reverts commit `93e70c5973`. Reverted https://github.com/pytorch/pytorch/pull/80193 on behalf of https://github.com/b0noI due to broken test: https://github.com/pytorch/pytorch/runs/7035945243?check_suite_focus=true	2022-06-24 14:53:37 +00:00
Elias Ellison	93e70c5973	[FakeTensor] Use the device of the meta tensor for fallback kernel (#80193 ) Otherwise, I would run into issues such as fp16 not supported for Conv Cpu, or cudnn_rnn not implemented for cpu. Pull Request resolved: https://github.com/pytorch/pytorch/pull/80193 Approved by: https://github.com/Chillee	2022-06-24 03:53:21 +00:00
Elias Ellison	7e54959a8a	Add support for indexing cuda tensors with cpu (#80115 ) Pull Request resolved: https://github.com/pytorch/pytorch/pull/80115 Approved by: https://github.com/Chillee	2022-06-23 22:46:40 +00:00
Elias Ellison	9705fb03b3	Add support for a couple ops Pull Request resolved: https://github.com/pytorch/pytorch/pull/79581 Approved by: https://github.com/Chillee	2022-06-20 22:25:39 +00:00
Elias Ellison	268bbecf1c	Add option for allowing non-fake inputs, add deepcopy impl Pull Request resolved: https://github.com/pytorch/pytorch/pull/79580 Approved by: https://github.com/samdow	2022-06-17 19:36:26 +00:00
Elias Ellison	13a8867c01	Add Dynamic Output Shape Tagdfor ata-dependent ops, handle in FakeTensor Pull Request resolved: https://github.com/pytorch/pytorch/pull/79170 Approved by: https://github.com/ezyang	2022-06-09 22:16:16 +00:00
Elias Ellison	3c5a3ca9e8	Make FakeTensors return meta within kerenl invocation, add FakeTensor op tests Pull Request resolved: https://github.com/pytorch/pytorch/pull/78972 Approved by: https://github.com/ezyang	2022-06-09 01:39:27 +00:00
Elias Ellison	fe7a13496e	Add CPU Fallback Pull Request resolved: https://github.com/pytorch/pytorch/pull/78522 Approved by: https://github.com/ezyang	2022-06-08 22:35:13 +00:00
Elias Ellison	290d0979f1	Migrate FakeTensors to always call into FakeTensorMode and have them hold a reference Pull Request resolved: https://github.com/pytorch/pytorch/pull/78677 Approved by: https://github.com/ezyang	2022-06-08 22:30:34 +00:00
Elias Ellison	a711ce4a2a	add non-kwarg device and _like constructors Pull Request resolved: https://github.com/pytorch/pytorch/pull/78536 Approved by: https://github.com/ezyang	2022-06-07 20:36:41 +00:00
PyTorch MergeBot	0df77f320f	Revert "add non-kwarg device and _like constructors" This reverts commit `44937da6db`. Reverted https://github.com/pytorch/pytorch/pull/78536 on behalf of https://github.com/janeyx99 due to Broke meta tests on trunk and on PR https://github.com/pytorch/pytorch/runs/6765692797?check_suite_focus=true	2022-06-07 04:11:05 +00:00
Elias Ellison	44937da6db	add non-kwarg device and _like constructors Pull Request resolved: https://github.com/pytorch/pytorch/pull/78536 Approved by: https://github.com/ezyang	2022-06-07 00:28:59 +00:00
Elias Ellison	26d273959c	Add Caching of Conversion to Fake/Meta tensors in FakeTensorMode Pull Request resolved: https://github.com/pytorch/pytorch/pull/78090 Approved by: https://github.com/ezyang	2022-06-03 13:56:00 +00:00
Elias Ellison	6671b504f7	Modernize FakeTensorMode, throw on non-fake inputs Pull Request resolved: https://github.com/pytorch/pytorch/pull/78516 Approved by: https://github.com/samdow	2022-06-01 21:43:59 +00:00
Elias Ellison	cea7dd1646	Add FakeTensorMode Pull Request resolved: https://github.com/pytorch/pytorch/pull/77972 Approved by: https://github.com/ezyang	2022-05-31 16:27:06 +00:00
Elias Ellison	4c18f362a9	add support for type_as/_to_copy Pull Request resolved: https://github.com/pytorch/pytorch/pull/77971 Approved by: https://github.com/ezyang	2022-05-31 16:25:11 +00:00
Elias Ellison	98e0816986	Extend __new__ on subclasses to set custom_device and custom_strides Pull Request resolved: https://github.com/pytorch/pytorch/pull/77970 Approved by: https://github.com/Chillee	2022-05-31 16:23:18 +00:00
Elias Ellison	678213ead2	Fake Tensor Part 1 Pull Request resolved: https://github.com/pytorch/pytorch/pull/77969 Approved by: https://github.com/ezyang	2022-05-31 16:20:35 +00:00

1 2 3 4

185 Commits