eigen - C++ library for linear algebra

	Commit message (Collapse)	Author	Age
...
*	Removing executable bit from file mode	Christoph Hertzberg	2020-01-11
\|
*	Bug #1790: Make `areApprox` check `numext::isnan` instead of bitwise ↵	Christoph Hertzberg	2020-01-11
\| \| \| \|	equality (NaNs don't have to be bitwise equal).
*	Added special_packetmath test and tweaked bounds on tests.	Srinivas Vasudevan	2020-01-11
\| \| \| \| \|	Refactor shared packetmath code to header file. (Squashed from PR !38)
*	call Explicitly ::rint and ::rintf for targets without c++11. Without this, ↵	Rasmus Munk Larsen	2020-01-10
\| \| \| \|	the Windows build breaks when trying to compile numext::rint<double>.
*	Improvements to the tidiness and completeness of the NEON implementation	Joel Holdsworth	2020-01-10
\|
*	Fix for gcc build error when using Eigen headers with AVX512	Anuj Rawat	2020-01-10
\|
*	Adding RInt vector support for SYCL.	mehdi-goli	2020-01-10
\|
*	Properly initialize b vector in SplineFitting	Matthew Powelson	2020-01-09
\| \| \|	InterpolateWithDerivative does not initialize the be vector correctly. This issue is discussed In stackoverflow question 48382939.
*	Don't add EIGEN_DEVICE_FUNC to random() since ::rand is not available in Cuda.	Rasmus Munk Larsen	2020-01-09
\|
*	Add missing EIGEN_DEVICE_FUNC annotations in MathFunctions.h.	Rasmus Munk Larsen	2020-01-09
\|
*	Use data.data() instead of &data (since it is not obvious that Array is ↵	Christoph Hertzberg	2020-01-09
\| \| \| \|	trivially copyable)
*	Don't use the rational approximation to the logistic function on GPUs as it ↵	Rasmus Munk Larsen	2020-01-09
\| \| \| \|	appears to be slightly slower.
*	The upper limits for where to use the rational approximation to the logistic ↵	Rasmus Munk Larsen	2020-01-08
\| \| \| \|	function were not set carefully enough in the original commit, and some arguments would cause the function to return values greater than 1. This change set the versions found by scanning all floating point numbers (using std::nextafterf()).
*	Fix formatting	Christoph Hertzberg	2020-01-08
\|
*	Bug #1785: Introduce numext::rint.	Ilya Tokar	2020-01-07
\| \| \| \| \| \|	This provides a new op that matches std::rint and previous behavior of pround. Also adds corresponding unsupported/../Tensor op. Performance is the same as e. g. floor (tested SSE/AVX).
*	[SYCL Backend]	mehdi-goli	2020-01-07
\| \| \| \| \| \| \|	* Adding Missing operations for vector comparison in SYCL. This caused compiler error for vector comparison when compiling SYCL * Fixing the compiler error for placement new in TensorForcedEval.h This caused compiler error when compiling SYCL backend * Reducing the SYCL warning by removing the abort function inside the kernel * Adding Strong inline to functions inside SYCL interop.
*	Protecting integer_types's long long test with a check to see if we have ↵	Everton Constantino	2020-01-07
\| \| \| \|	CXX11 support.
*	Bug #1800: Guard against misleading indentation	Christoph Hertzberg	2020-01-03
\|
*	Fix -Werror -Wfloat-conversion warning.	Janek Kozicki	2019-12-23
\|
*	Fix for HIP breakage - 191220	Deven Desai	2019-12-20
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	The breakage was introduced by the following commit : https://gitlab.com/libeigen/eigen/commit/ae07801dd8d295657f28b006e1e4999edf835052 After the commit, HIPCC errors out on some tests with the following error ``` Building HIPCC object unsupported/test/CMakeFiles/cxx11_tensor_device_1.dir/cxx11_tensor_device_1_generated_cxx11_tensor_device.cu.o In file included from /home/rocm-user/eigen/unsupported/test/cxx11_tensor_device.cu:17: In file included from /home/rocm-user/eigen/unsupported/Eigen/CXX11/Tensor:100: /home/rocm-user/eigen/unsupported/Eigen/CXX11/src/Tensor/TensorBlock.h:129:12: error: no matching constructor for initialization of 'Eigen::internal::TensorBlockResourceRequirements' return {merge(lhs.shape_type, rhs.shape_type), // shape_type ^~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~ /home/rocm-user/eigen/unsupported/Eigen/CXX11/src/Tensor/TensorBlock.h:75:8: note: candidate constructor (the implicit copy constructor) not viable: requires 1 argument, but 3 were provided struct TensorBlockResourceRequirements { ^ /home/rocm-user/eigen/unsupported/Eigen/CXX11/src/Tensor/TensorBlock.h:75:8: note: candidate constructor (the implicit move constructor) not viable: requires 1 argument, but 3 were provided /home/rocm-user/eigen/unsupported/Eigen/CXX11/src/Tensor/TensorBlock.h:75:8: note: candidate constructor (the implicit copy constructor) not viable: requires 5 arguments, but 3 were provided /home/rocm-user/eigen/unsupported/Eigen/CXX11/src/Tensor/TensorBlock.h:75:8: note: candidate constructor (the implicit default constructor) not viable: requires 0 arguments, but 3 were provided ... ... ``` The fix is to explicitly decalre the (implicitly called) constructor as a device func
*	Bug #1796: Make matrix squareroot usable for Map and Ref types	Christoph Hertzberg	2019-12-20
\|
*	Reduce code duplication and avoid confusing Doxygen	Christoph Hertzberg	2019-12-19
\|
*	Hide recursive meta templates from Doxygen	Christoph Hertzberg	2019-12-19
\|
*	Use double-braces initialization (as everywhere else in the test-suite).	Christoph Hertzberg	2019-12-19
\|
*	Fix trivial shadow warning	Christoph Hertzberg	2019-12-19
\|
*	Bug #1788: Fix rule-of-three violations inside the stable modules.	Christoph Hertzberg	2019-12-19
\| \| \| \| \|	This fixes deprecated-copy warnings when compiling with GCC>=9 Also protect some additional Base-constructors from getting called by user code code (#1587)
*	Fix unit-test which I broke in previous fix	Christoph Hertzberg	2019-12-19
\|
*	Fix TensorPadding bug in squeezed reads from inner dimension	Eugene Zhulenev	2019-12-19
\|
*	Return const data pointer from TensorRef evaluator.data()	Eugene Zhulenev	2019-12-18
\|
*	Tensor block evaluation cost model	Eugene Zhulenev	2019-12-18
\|
*	Fix some maybe-unitialized warnings	Christoph Hertzberg	2019-12-18
\|
*	Workaround class-memaccess warnings on newer GCC versions	Christoph Hertzberg	2019-12-18
\|
*	fix compilation due to new HIP scalar accessor	Jeff Daily	2019-12-17
\|
*	Reduce block evaluation overhead for small tensor expressions	Eugene Zhulenev	2019-12-17
\|
*	Add default definition for EIGEN_PREDICT_*	Rasmus Munk Larsen	2019-12-16
\|
*	Improve accuracy of fast approximate tanh and the logistic functions in ↵	Rasmus Munk Larsen	2019-12-16
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Eigen, such that they preserve relative accuracy to within a few ULPs where their function values tend to zero (around x=0 for tanh, and for large negative x for the logistic function). This change re-instates the fast rational approximation of the logistic function for float32 in Eigen (removed in https://gitlab.com/libeigen/eigen/commit/66f07efeaed39d6a67005343d7e0caf7d9eeacdb), but uses the more accurate approximation 1/(1+exp(-1)) ~= exp(x) below -9. The exponential is only calculated on the vectorized path if at least one element in the SIMD input vector is less than -9. This change also contains a few improvements to speed up the original float specialization of logistic: - Introduce EIGEN_PREDICT_{FALSE,TRUE} for __builtin_predict and use it to predict that the logistic-only path is most likely (~2-3% speedup for the common case). - Carefully set the upper clipping point to the smallest x where the approximation evaluates to exactly 1. This saves the explicit clamping of the output (~7% speedup). The increased accuracy for tanh comes at a cost of 10-20% depending on instruction set. The benchmarks below repeated calls u = v.logistic() (u = v.tanh(), respectively) where u and v are of type Eigen::ArrayXf, have length 8k, and v contains random numbers in [-1,1]. Benchmark numbers for logistic: Before: Benchmark Time(ns) CPU(ns) Iterations ----------------------------------------------------------------- SSE BM_eigen_logistic_float 4467 4468 155835 model_time: 4827 AVX BM_eigen_logistic_float 2347 2347 299135 model_time: 2926 AVX+FMA BM_eigen_logistic_float 1467 1467 476143 model_time: 2926 AVX512 BM_eigen_logistic_float 805 805 858696 model_time: 1463 After: Benchmark Time(ns) CPU(ns) Iterations ----------------------------------------------------------------- SSE BM_eigen_logistic_float 2589 2590 270264 model_time: 4827 AVX BM_eigen_logistic_float 1428 1428 489265 model_time: 2926 AVX+FMA BM_eigen_logistic_float 1059 1059 662255 model_time: 2926 AVX512 BM_eigen_logistic_float 673 673 1000000 model_time: 1463 Benchmark numbers for tanh: Before: Benchmark Time(ns) CPU(ns) Iterations ----------------------------------------------------------------- SSE BM_eigen_tanh_float 2391 2391 292624 model_time: 4242 AVX BM_eigen_tanh_float 1256 1256 554662 model_time: 2633 AVX+FMA BM_eigen_tanh_float 823 823 866267 model_time: 1609 AVX512 BM_eigen_tanh_float 443 443 1578999 model_time: 805 After: Benchmark Time(ns) CPU(ns) Iterations ----------------------------------------------------------------- SSE BM_eigen_tanh_float 2588 2588 273531 model_time: 4242 AVX BM_eigen_tanh_float 1536 1536 452321 model_time: 2633 AVX+FMA BM_eigen_tanh_float 1007 1007 694681 model_time: 1609 AVX512 BM_eigen_tanh_float 471 471 1472178 model_time: 805
*	Resolve double-promotion warnings when compiling with clang.	Christoph Hertzberg	2019-12-13
\| \| \| \|	`sin` was calling `sin(double)` instead of `std::sin(float)`
*	Renamed .hgignore to .gitignore (removing hg-specific "syntax" line)	Christoph Hertzberg	2019-12-13
\|
*	Bug 1785: fix pround on x86 to use the same rounding mode as std::round.	Ilya Tokar	2019-12-12
\| \| \| \| \| \|	This also adds pset1frombits helper to Packet[24]d. Makes round ~45% slower for SSE: 1.65µs ± 1% before vs 2.45µs ± 2% after, stil an order of magnitude faster than scalar version: 33.8µs ± 2%.
*	Clamp tanh approximation outside [-c, c] where c is the smallest value where ↵	Rasmus Munk Larsen	2019-12-12
\| \| \| \|	the approximation is exactly +/-1. Without FMA, c = 7.90531110763549805, with FMA c = 7.99881172180175781.
*	Fix implementation of complex expm1. Add tests that fail with previous ↵	Srinivas Vasudevan	2019-12-12
\| \| \| \|	implementation, but pass with the current one.
*	Initialize non-trivially constructible types when allocating a temp buffer.	Eugene Zhulenev	2019-12-12
\|
*	Squeeze reads from two inner dimensions in TensorPadding	Eugene Zhulenev	2019-12-11
\|
*	Add back accidentally deleted default constructor to ↵	Eugene Zhulenev	2019-12-11
\| \| \| \|	TensorExecutorTilingContext.
*	Added io test	Joel Holdsworth	2019-12-11
\|
*	IO: Fixed printing of char and unsigned char matrices	Joel Holdsworth	2019-12-11
\|
*	Added Eigen::numext typedefs for uint8_t, int8_t, uint16_t and int16_t	Joel Holdsworth	2019-12-11
\|
*	Bug 1786: fix compilation with MSVC	Gael Guennebaud	2019-12-11
\|
*	Remove block memory allocation required by removed block evaluation API	Eugene Zhulenev	2019-12-10
\|
*	Remove V2 suffix from TensorBlock	Eugene Zhulenev	2019-12-10
\|