pytorch

mirror of https://github.com/zebrajr/pytorch.git synced 2025-12-07 12:21:27 +01:00

Author	SHA1	Message	Date
Jongsoo Park	fb8c3d62fe	removing quantization utility functions moved to fbgemm (#14301 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/14301 This diff removes quantization utility functions copied to fbgemm Reviewed By: Maratyszcza Differential Revision: D13159299 fbshipit-source-id: a7f3cd2af0aa241a8578d532a70a157da70d9289	2018-11-21 21:38:23 -08:00
Jongsoo Park	3c0ce51484	clang-format (#14160 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/14160 clang-format of C++ files Reviewed By: hx89 Differential Revision: D13115201 fbshipit-source-id: d2ad65f66209e00578ef90f87f41272de2d24aa9	2018-11-20 00:56:00 -08:00
Jongsoo Park	53c3a92a50	consistent rounding (#9 ) Summary: Pull Request resolved: https://github.com/pytorch/FBGEMM/pull/9 Pull Request resolved: https://github.com/pytorch/pytorch/pull/13960 The vectorized code was rounding to even in halfway cases with _mm256_round_ps + (_MM_FROUND_TO_NEAREST_INT \|_MM_FROUND_NO_EXC) (see more details in https://software.intel.com/en-us/node/523819), but we were still using std::round in a couple of places which does rounding away from zero in halfway cases. With this diff, we use std::nearbyint in all scalar code (except a few cases where we don't care exact rounding mode and uses rint which is the fastest in general) to be more consistent. nearbyint is the same as what the vectorized code does only when the current rounding mode is FE_TONEAREST but in practice this is OK because we almost always use the default rounding mode FE_TONEAREST. This is inspired by Marat's diff for mobile quantization. Reviewed By: dskhudia Differential Revision: D13017719 fbshipit-source-id: 6b8f99db7ea2e233aa2e3bd2adf622e03ed6258e	2018-11-14 10:21:42 -08:00
Jongsoo Park	90ea61800f	operators/quantized/server -> quantization/server (#13660 ) Summary: Pull Request resolved: https://github.com/pytorch/pytorch/pull/13660 Any change in server side quantized operator was triggering ios-sanity-check with more than 5 hours testing time. I suspect this was because the operator code was synced with xplat directory. This diff moves server side quantized operators to caffe2/caffe2/quantization/server to avoid this issue. Reviewed By: hx89 Differential Revision: D12955420 fbshipit-source-id: b6c824b9de5e2a696f8c748e1b2c77d81d46746b	2018-11-07 22:54:13 -08:00

4 Commits