#1023: Improve OpenMP parallelisation of SummationByParts -----------------------------------+---------------------------------------- Reporter: eschnett | Owner: Type: enhancement | Status: new Priority: major | Milestone: Component: EinsteinToolkit thorn | Version: Keywords: | -----------------------------------+---------------------------------------- The Intel compiler does not handle workshare constructs well. The attached patch replaces them by explicit loops, which execute faster. This makes a measurable difference on Hopper with 24 OpenMP threads.