portelli/Grid - Grid - DiRAC Tursa git server

mirror of https://github.com/paboyle/Grid.git synced 2024-11-16 02:35:36 +00:00

Author	SHA1	Message	Date
Lanny91	a833f88c32	Added missing SIMD integer reduction implementation for AVX, AVX-512, SSE4, IMCI	2017-06-16 15:58:47 +01:00
paboyle	736bf3c866	Major rework of stencil. Half precision and MPI3 now working.	2017-04-22 11:33:50 +01:00
paboyle	3844bcf800	If no f16c instructions supported must use software half precision conversion. This will also become useful on BG/Q, so will move out from SSE4 into a general area. Lifted the Eigen half precision from web. Looks sensible, but not extensively regressed against the intrinsics implementation yet.	2017-04-20 15:30:52 +01:00
paboyle	db5ea001a3	Update to use Xcode 8.3 since -mfp16 causes SIGILL	2017-04-13 12:22:40 +01:00
paboyle	1d502e4ed6	FP16 optional compile time	2017-04-13 11:55:24 +01:00
paboyle	73cdf0fffe	Drop f16c from SSE because of a macos compile error on travis	2017-04-13 11:23:41 +01:00
paboyle	94eb829d08	Align cast fixed for __mm128i gcc complained	2017-04-13 08:40:44 +01:00
paboyle	68392ddb5b	Exchange in generic Precision change in AVX, SSE, AVX512, Generic. QPX still to do.	2017-04-13 08:38:12 +01:00
paboyle	cb6b81ae82	Half precision conversion	2017-04-12 19:32:37 +01:00
paboyle	bd600702cf	Vectorise the XYZT face gathering better. Hard coded for simd_layout <= 2 in any given spread out direction; full generality is inconsistent with efficiency.	2017-02-15 11:11:04 +00:00
paboyle	4b220972ac	Warning fix	2016-12-18 02:14:17 +00:00
paboyle	629f43e36c	Return statement needed	2016-12-18 02:09:37 +00:00
paboyle	a3172b3455	Precision error	2016-12-18 02:07:45 +00:00
Peter Boyle	69ae817d1c	Updates for supporting Mobius better	2016-12-08 16:43:28 +00:00
paboyle	836e929565	Divide handling improved	2016-09-26 09:42:22 +01:00
paboyle	e3f141f82f	Fixed SSE compile with typecasts	2016-04-22 10:30:30 -07:00
paboyle	528eb773ad	Merged. Merge branch 'master' of https://github.com/paboyle/Grid	2016-04-19 22:24:34 +01:00
paboyle	e5657510b0	Rotate support for Ls simd-ized	2016-04-19 22:24:18 +01:00
Christopher Kelly	ab56ccdd25	-Complete and working implementation of Grid_empty	2016-04-15 13:17:42 -04:00
paboyle	aae8bf31a7	Global edit adding copyright and license info to every source file.	2016-01-02 14:51:32 +00:00
Azusa Yamaguchi	24a5a81c53	SSE compile fix	2015-12-16 09:09:37 +00:00
Peter Boyle	814c79f38d	SIMD improvements for mac and madd use in complex for avx, sse	2015-10-09 00:38:52 +02:00
Peter Boyle	64d64d1ab6	Updating to modify non-inlining permute routines and hopefully get better reg use and enhance performance.	2015-09-25 08:55:04 -07:00
neo	6e5db0b1da	Corrected bug in integer multiplications for SSE4 and AVX2 Merge remote-tracking branch 'upstream/master' Conflicts: tests/Make.inc	2015-06-16 23:34:45 +09:00
neo	48bf4878c1	Experimental support for ARM	2015-06-09 15:46:21 +09:00
Peter Boyle	b72ca15bd2	Improving the reduction to go through our on permute. Must also do this for avx512	2015-05-27 16:07:17 +01:00
neo	64753ea633	Included Gpermute in the new Grid_simd.h file style. Now tested for SSE4. OK	2015-05-27 12:11:44 +09:00
neo	48cc816136	Merge remote-tracking branch 'upstream/master' Conflicts: lib/math/Grid_math_tensors.h lib/simd/Grid_vector_types.h	2015-05-26 13:14:06 +09:00
Peter Boyle	489b1b9633	Schur complement based red-black inversion working	2015-05-25 13:47:12 +01:00
neo	9e29ac6549	Completed implementation of new Grid_simd classes Tested performance for SSE4, Ok. AVX1/2, AVX512 yet untested	2015-05-22 17:33:15 +09:00
neo	cf7be0e461	Implemented all SSE4 functions. A test code Grid_simd_new.cc has been created to test the new class. Tests are all OK.	2015-05-20 17:22:40 +09:00
neo	74e91cd925	Partial implementation of the vector types SIMD Implementing SSE4 now A systematic series of tests must be written.	2015-05-19 17:21:17 +09:00

32 Commits