pytorch

mirror of https://github.com/zebrajr/pytorch.git synced 2025-12-07 00:21:07 +01:00

Author	SHA1	Message	Date
Sam Gross	7c0b16c140	Add torch.take and Tensor.put_ (#3263 ) * Add torch.take and Tensor.put_ These are similar to numpy.take and numpy.put. The take function allows you to linearly index into a tensor without viewing it as a 1D tensor first. The output has the same shape as the indices. The put function copies value into a tensor also using linear indices.	2017-11-01 06:04:44 -04:00
Sam Gross	a65db4e956	Use ATen for torch.cat, torch.addmm, and friends on Variables. (#3286 ) This includes some changes to the dispatch code for torch.xxx functions: - Since Variable.addmm is an instance-method, the self argument has to come first. The dispatch code swaps the first two arguments if necessary to suppor the deprecated signatures where 'alpha' or 'beta' comes before the 'self' tensor. - Delete IMPLEMENT_STATELESS_REVERSED. These functions require output arguments to be passed in using the keyword 'out'. They were meant to handle torch.gt(out, a, b), but we haven't allowed that for a while.	2017-10-25 14:27:45 -04:00
Sam Gross	f1f64c8d07	Generate autograd functions for NN / more refactors (#3136 ) Generate autograd functions for NN and implement more derivatives in derivatives.yaml A big refactor of gen_variable_type.py	2017-10-19 15:03:26 -04:00
Adam Paszke	f9ee52efa9	Update DLPack bindings	2017-10-19 10:06:53 -04:00
Sam Gross	47beb64b5c	Use ATen generator as default CPU generator (#3135 ) ATen has it's own default CPU RNG. Use this as the default in PyTorch so that random functions called through ATen have the same behavior as random functions called through TensorMethods	2017-10-16 14:22:58 -04:00
Priya Goyal	756ab3f24f	Adding conversion from python tensor to dlpack tensor (#2933 )	2017-10-04 08:35:42 -04:00
Soumith Chintala	b3bc5fe302	refactor THCP method defs into cuda/Module.cpp	2017-09-30 13:14:35 -07:00
IraKorshunova	2b9765ad02	Erf and erfinv (#2799 )	2017-09-20 21:23:45 -04:00
Adam Paszke	8dae433de8	Move JIT passes to a separate directory	2017-09-19 10:53:32 -04:00
Edward Z. Yang	2e266837f5	Port TracingState to pybind11, new export() method. Along the way I added converters for Variable and TracingInput. Variable should probably be moved to a more widely known spot. Signed-off-by: Edward Z. Yang <ezyang@fb.com>	2017-09-05 17:48:55 -04:00
Adam Paszke	594f98ce16	Support multi-stage AutogradClosures	2017-09-05 17:48:55 -04:00
Zach DeVito	a3fdb281d1	Python wrapper for Node IR using pybind11 Supports almost all of the IR API.	2017-09-05 17:48:55 -04:00
Adam Paszke	bdcbbeaf68	Remove GlobalTracingState	2017-09-05 17:48:55 -04:00
Adam Paszke	e186d16e6b	Apply JIT optimizations form Python	2017-09-05 17:48:55 -04:00
Adam Paszke	ea05ac8f41	Move JIT-related files to jit dir. Remove IR interpreter	2017-09-05 17:48:55 -04:00
Edward Z. Yang	a797ab9343	Rewrite AST to a new, more functional representation. Previously, our AST was a DAG, where shared Nodes indicated a computation should be reused. This commit rewrites the IR into a new functional representation which represents sharing explicitly using variable bindings. We offer a few justifications for this new style: 1. The new representation is not all that different from the old one; it is about as easy to construct, and the lack of an explicit graph doesn't negatively impact our ability to interpret the graph, since we've chosen, as a matter of design, to NOT have the IR participate in the actual execution of a graph. 2. The new let-binding representation has an implicit ordering, which we can use to conveniently keep track of the original order the trace showed up as. This automatically gives us a topsort, and gives us an easier to read textual representation of our IR: %14 = Embedding %11, %0, -1, None, 2, False, False %15 = Dropout %14, 0.2, True, False %16 = Index %12, 0 %17 = Index %12, 1 %18 = Index %13, 0 %19 = Index %13, 1 %20 = Index %15, 0 %21 = Linear %20, %1, %3 %22 = Linear %16, %2, %4 3. It moves us closer to a Futhark style language (http://futhark-lang.org/publications/pldi17.pdf). Major aspects of the diff - Node is replaced with Expr and Arg, a pair of mutually recursive structures which represent our new language. In BNF, the language looks like this: a ::= c \| %i e ::= %i, ... = e \| PyOp e, ... \| Ret %i, ... Technically, Ret is not actually a return (no control flow is involved), it just tuples up a series of tensors (identified by variables). One important invariant is that locals are always tensors; they are never constants (this is asymmetric with Args.) - Arguments support Python constants. This is an important piece because many operators take extra Python literals like integers and tuples in order to specify extra parameters about how an operator operates. Adding this was essential to getting word_language_model to work. - As both Expr and Arg have multiple variants, there is new infrastructure for doing case on the variants using ExprVisitor and ArgVisitor. The strategy here is adapted from WebAssembly's visitors, although we have generalized to permit arbitrary argument forwarding, which is necessary to support tail-recursive visitor calls. TCO is important because our interpreter may recurse arbitrarily deep into a stack of nested lets. If users wish, they can also manually case on the type tag. - Tracing is now turned on and off using _tracer_enter/_tracer_exit in torch._C. _tracer_enter accepts a list of variables which are to be treated as arguments; _tracer_exit accepts the list of traced variables which should be returned when you reexecute the trace, and returns the trace expression which can be reexecuted. GlobalTracingState is a global variable which tracks whether or not we are tracing or not. - You use run_forward to execute a trace on some set of parameters. - When under tracing, variables keep track, via trace_local, what the name of their variables in the IR are. Here is a simple runner which leaks memory but can be used to JIT models: import torch.autograd.function as F import torch._C def jit(model): import types real_forward = model.forward def forward(self, args): def flatten(x): return tuple(F._iter_variables(x)) if not hasattr(self, "saved_trace"): torch._C._tracer_enter(tuple(self.parameters()) + flatten(args)) out = real_forward(args) self.saved_trace = torch._C._tracer_exit(flatten(out)) self.saved_outs = out return out else: flat_out = Variable._execution_engine.run_forward(self.saved_trace, tuple(self.parameters()) + flatten(args)) return F._unflatten(flat_out, self.saved_outs) Major problems: - Sanity checking is spotty at best, especially when users pass in variables. - The interpreter leaks tensor memory from the store. When we add back def-use we should be able to deallocate tensors as soon as we know they are no longer necessary. - The interpreter needs to reach feature parity with the old execution engine. From there, we need to see if backwards can be subsumed as well. - I still have no confidence in having memory managed everything correctly. This requires a close look. - Rather than return an open expression as a trace, we should return a lambda instead, which knows about how many formal parameters it requires. - The IR is not introspectable from Python at the moment, but this is simply a matter of implementing all the binding code. - The tracer is NOT reentrant (you can't trace while you're inside a trace.) Furthermore, no sanity checking is done if you try to incorrectly reuse things from one trace in another. Signed-off-by: Edward Z. Yang <ezyang@fb.com>	2017-09-05 17:48:55 -04:00
Edward Z. Yang	e1b7872fc2	Make it possible to access IR from Python. Also, add a new trace_fn field to attach forward IR to Variables. Signed-off-by: Edward Z. Yang <ezyang@fb.com>	2017-09-05 17:48:55 -04:00
Justin Johnson	94b5990201	Add torch.cuda.get_device_name function (#2540 )	2017-08-26 15:06:37 -04:00
Alykhan Tejani	eb58740651	add ones_like and zeros_like	2017-08-25 14:11:04 -04:00
gchanan	c000d15058	Properly use Py_RETURN_True, Py_RETURN_False in back compatibility warnings. (#2345 )	2017-08-08 21:54:20 -04:00
Zach DeVito	9d8cff9bc1	initialize aten and pytorch to share the same THCState	2017-07-11 10:35:03 -04:00
Adam Paszke	714351ff39	Officially enable process-group mode	2017-06-12 22:02:11 -04:00
Gregory Chanan	4f602a52b5	Use THPUtils_assert rather than THError in torch/csrc/Module.	2017-06-11 05:37:59 -04:00
Gregory Chanan	ffd808768e	Remove raiseErrors from THTensor functions, have THStorage functions take an error_buffer to return a proper error message while being able to handle memory management correctly from calling function.	2017-06-11 05:37:59 -04:00
Gregory Chanan	177785eecf	explicit Ptr constructors, fast transposed copy.	2017-06-11 05:37:59 -04:00
Gregory Chanan	be65f46c76	Add optional warning for backwards incompatible keepdim. Setting torch.utils.backcompat.keepdim.warning.enabled=True will cause Python warnings in the case where the default value of keepdim is used for 1-d reductions. Also specify keepdim via kwargs in library so these warnings have less noise.	2017-06-11 05:37:59 -04:00
Gregory Chanan	3556d1b8a3	Add optional warning for backwards incompatible broadcast. Setting torch.utils.backcompat.broadcast.warning.enabled=True will cause Python warnings in the case where broadcast occurs but previously 1-d view style pointwise ops occured.	2017-06-11 05:37:59 -04:00
Gregory Chanan	5af46cb352	Add broadcasting support for matmul.	2017-06-11 05:37:59 -04:00
Sam Gross	d81da41650	Make sure the number of MKL and OpenMP threads match Otherwise, on many machines, the size of the OpenMP thread pool will change between MKL and our OpenMP enabled functions. The constant thread creation and destruction results in worse performance and leaks memory on GCC 5.4	2017-06-07 14:53:29 -04:00
Adam Paszke	8ea7c87c29	Improve init methods	2017-06-02 23:42:11 +02:00
Adam Paszke	181d2f41bd	Add initial Python wrappers for THDTensors	2017-06-02 23:42:11 +02:00
Trevor Killeen	05bc877a05	make THPPointer have explicit constructors (#1636 )	2017-05-25 15:35:54 -04:00
ethanluoyc	d0504aa41d	Implement lgamma function.	2017-05-08 16:21:26 -07:00
Sam Gross	4c1cdb6148	Refactor Python string utility function	2017-04-28 21:25:26 +02:00
Sam Gross	27990fee54	Use fully qualified name as tp_name for tensors and storages (#1379 )	2017-04-27 16:26:44 -04:00
Martin Raison	cd3bbc9dfd	more operations and optimizations (hspmm, reorder, ...)	2017-04-18 12:46:54 -07:00
albanD	71303b8af4	Autograd deadlock for recent glibc fix (#1243 )	2017-04-12 22:24:31 +02:00
Adam Paszke	afeeb81e79	Add support for keyword arguments in torch.cat	2017-04-11 14:48:54 -07:00
Adam Paszke	91c4ba7980	Add torch.arange and deprecate torch.range	2017-04-03 10:38:58 -04:00
albanD	dfa2d26830	* make random_ range correct when both lower and upper are specified	2017-03-31 15:37:24 -04:00
Sergey Zagoruyko	8dc5d2a22e	export current_blas_handle	2017-03-23 23:32:45 +01:00
Brandon Amos	bb353ccc17	Add batch triangular factorization and solves, add IntegerTensor to cwrap (#903 )	2017-03-23 15:06:00 -04:00
Adam Paszke	faac0f5c25	Fix torch.cat bugs Always use PySequence API and disallow catting along inexistent dimensions.	2017-03-22 18:58:42 -04:00
Sam Gross	379ae6d865	Refactor out dispatchStateless (#1007 ) Some of the error messages were incorrect due to erroneous 'tensor == THPDefaultTensorClass' checks	2017-03-15 16:24:55 -04:00
Martin Raison	f17cfe4293	sparse tensor operations (#735 )	2017-03-03 18:37:03 +01:00
Zhou Chang	f366e5fc81	Support int16 numpy conversions issue #891	2017-03-02 09:15:57 -05:00
Sam Gross	fc6fcf23f7	Lock the cudaFree mutex. (#880 ) Prevents NCCL calls from overlapping with cudaFree() which can lead to deadlocks.	2017-03-01 11:29:25 -05:00
Adam Paszke	67f94557ff	Expose torch.HalfTensor	2017-02-27 19:35:47 -05:00
Sam Gross	bd5303010d	Refactor autograd package to separate Python dependencies. (#662 ) The core autograd Variable, Function, and Engine no longer depend on the Python API. This let's us implement functions in C++. In the future, we can also multithread engine and release the GIL for most of the non-Python backwards.	2017-02-13 16:00:16 -08:00
Sam Gross	712686ce91	Add cat, contiguous, squeeze, and unsqueeze to THPP Use unsqueeze and view from TH/THC	2017-02-11 17:49:31 +01:00

1 2 3

118 Commits