Peter Boyle
|
d0f3d525d5
|
Optimal block size for KNL
|
2017-08-25 19:33:54 +01:00 |
|
Peter Boyle
|
3a58217405
|
Updated
|
2017-08-25 14:29:53 +01:00 |
|
Peter Boyle
|
c289699d9a
|
updated from cambridge mpi3 shakeout
|
2017-08-25 11:41:01 +01:00 |
|
Peter Boyle
|
c3b1263e75
|
Benchmark prep
|
2017-08-25 09:25:54 +01:00 |
|
paboyle
|
ae56e556c6
|
finalise issue on new OPA revert
|
2017-08-20 02:53:12 +01:00 |
|
paboyle
|
a446d95c33
|
Trying to pass TeamCity and Travis
|
2017-08-20 01:10:50 +01:00 |
|
paboyle
|
be66e7dd95
|
Merge branch 'develop' into feature/multi-communicator
|
2017-08-19 23:12:38 +01:00 |
|
paboyle
|
bfef525ed2
|
New benchmark prep
|
2017-08-19 23:10:12 +01:00 |
|
Peter Boyle
|
7d88198387
|
Merge branch 'develop' into feature/multi-communicator
|
2017-08-19 13:03:35 -04:00 |
|
Peter Boyle
|
9e658de238
|
Use Vector
|
2017-08-19 12:52:44 -04:00 |
|
Peter Boyle
|
14d53e1c9e
|
Threaded MPI calls patches
|
2017-07-29 13:08:10 -04:00 |
|
Peter Boyle
|
40e119c61c
|
NUMA improvements worth preserving from AMD EPYC tests
|
2017-07-08 22:27:11 -04:00 |
|
Peter Boyle
|
b73bd151bb
|
Switch off counters by default
|
2017-06-30 10:16:35 +01:00 |
|
Peter Boyle
|
694b305cab
|
Update to reporting
|
2017-06-30 10:16:13 +01:00 |
|
paboyle
|
6f5a5cd9b3
|
Improved threaded comms benchmark
|
2017-06-28 23:27:02 +01:00 |
|
Peter Boyle
|
08e04b9676
|
Better benchmarks
|
2017-06-28 15:30:06 +01:00 |
|
paboyle
|
54e94360ad
|
Experimental: Multiple communicators to see if we can avoid thread locks in --enable-comms=mpit
|
2017-06-24 23:10:24 +01:00 |
|
paboyle
|
3bfd1f13e6
|
I/O improvements
|
2017-06-11 23:14:10 +01:00 |
|
Peter Boyle
|
725c513d94
|
Better MPI3 benchmarking
|
2017-05-29 16:47:32 -04:00 |
|
Guido Cossu
|
0ffc235741
|
Adding more statistics to the Benchmark_comms. Min and max
|
2017-05-19 10:55:04 +01:00 |
|
Guido Cossu
|
8e19c99c7d
|
Adding more statistical info in the Benchmark_comms
|
2017-05-18 19:07:35 +01:00 |
|
Guido Cossu
|
a0bc0ad06f
|
Reverting change in Bechmark_comms. Keeping 300 iterations
|
2017-05-18 17:48:11 +01:00 |
|
Guido Cossu
|
bc862ce3ab
|
Fixing an allocation issue in Benchmark_comms
|
2017-05-18 14:44:56 +01:00 |
|
paboyle
|
751f2b9703
|
Better check and benchmark driving
|
2017-05-05 19:54:38 +01:00 |
|
Guido Cossu
|
20999c1370
|
Merge branch 'develop' into feature/hmc_generalise
|
2017-05-05 12:47:17 +01:00 |
|
Peter Boyle
|
945767c6d8
|
More info
|
2017-05-03 20:26:35 -04:00 |
|
Peter Boyle
|
92e364a35f
|
Better reporting in benchmark for MPI3
|
2017-05-03 15:43:36 -04:00 |
|
Guido Cossu
|
4063238943
|
Adding HMC test file example for Mobius + smearing
|
2017-05-01 13:44:00 +01:00 |
|
Guido Cossu
|
3344788fa1
|
Merge branch 'develop' into feature/hmc_generalise
|
2017-05-01 12:13:56 +01:00 |
|
paboyle
|
738c1a11c2
|
longer nloop
|
2017-04-26 08:43:20 +01:00 |
|
paboyle
|
ab66bac4e6
|
Think I'm getting on top of the reduced cost exterior precomputed list of links
|
2017-04-25 08:50:26 +01:00 |
|
paboyle
|
c429ace748
|
Cleaner OpenMP use
|
2017-04-22 20:28:42 +01:00 |
|
Peter Boyle
|
1d1b225497
|
Hand unrolled Nc=3 kernels support split phase compute (on-node, off-node).
|
2017-04-22 09:05:28 -04:00 |
|
paboyle
|
fc4ab9ccd5
|
Working half precision comms
|
2017-04-20 11:20:26 +01:00 |
|
Guido Cossu
|
8c540333d5
|
Merge branch 'develop' into feature/hmc_generalise
|
2017-04-05 14:41:04 +01:00 |
|
paboyle
|
f18f5ed926
|
Drop random device
|
2017-04-02 00:26:26 +09:00 |
|
paboyle
|
4b17e8eba8
|
Merge branch 'develop' into feature/bgq-asm
Conflicts:
lib/qcd/action/fermion/Fermion.h
lib/qcd/action/fermion/WilsonFermion.cc
lib/util/Init.cc
tests/Test_cayley_even_odd_vec.cc
|
2017-03-28 04:49:30 -04:00 |
|
paboyle
|
18bde08d1b
|
Merge branch 'feature/staggering' into develop
|
2017-03-28 15:25:55 +09:00 |
|
paboyle
|
e099dcdae7
|
Merge branch 'develop' into feature/bgq-asm
|
2017-02-23 00:25:29 +00:00 |
|
azusayamaguchi
|
1c30e9a961
|
Verified
|
2017-02-21 23:01:25 +00:00 |
|
paboyle
|
3ae92fa2e6
|
Global changes to parallel_for structure.
Move the comms flags to more sensible names
|
2017-02-21 05:24:27 -05:00 |
|
paboyle
|
1a30455a10
|
1000 iters on bmark for more accurate timing
|
2017-02-20 17:47:01 -05:00 |
|
paboyle
|
aca7a3ef0a
|
Optimisation control improvements
|
2017-02-10 18:22:31 -05:00 |
|
Guido Cossu
|
8b6a6c8236
|
Resolving small merge conflict
|
2017-02-09 16:20:24 +00:00 |
|
Guido Cossu
|
e0571c872b
|
Merge branch 'develop' into feature/hmc_generalise
|
2017-02-09 16:12:00 +00:00 |
|
paboyle
|
2bf4688e83
|
Running on BNL KNL
|
2017-02-07 01:32:10 -05:00 |
|
paboyle
|
060da786e9
|
Comms benchmark improvements
|
2017-02-07 01:07:39 -05:00 |
|
Guido Cossu
|
17629b8d9e
|
Merge branch 'develop' into feature/hmc_generalise
|
2017-01-25 11:33:53 +00:00 |
|
|
a37e71f362
|
New automatic implementation of gamma matrices, Meson and SeqGamma are broken
|
2017-01-23 19:13:43 -08:00 |
|
azusayamaguchi
|
05c1924819
|
Timing loop change
|
2017-01-23 10:43:45 +00:00 |
|