User and Reference Guide for the Intel® C++ Compiler 14.0
To enable the auto-parallelizer, use the [Q]parallel option. This option detects parallel loops capable of being executed safely in parallel and automatically generates multi-threaded code for these loops.
You might need to set the KMP_STACKSIZE environment variable to an appropriately large size to enable parallelization with this option.
Using this option enables parallelization for both Intel® microprocessors and non-Intel microprocessors. The resulting executable may get additional performance gain on Intel® microprocessors than on non-Intel microprocessors. The parallelization can also be affected by certain options, such as /arch (Windows*), -m (Linux* and OS X*), or [Q]x.
An example of the command using auto-parallelization is as follows:
// (Linux* OS and OS X*) icc -c -parallel prog.cpp
// (Windows* OS) icl /c /Qparallel prog.cpp
Auto-parallelization uses two specific pragmas: #pragma parallel and #pragma noparallel.
The format of an auto-parallelization compiler pragma is:
|
Syntax |
|---|
<prefix> <pragma> |
where <prefix> indicates:
The <prefix> is followed by the pragma name, as in:
|
Syntax |
|---|
#pragma parallel |
The #pragma parallel pragma instructs the compiler to ignore dependencies that it assumes may exist and which would prevent correct parallelization in the immediately following loop. However, if dependencies are proven, they are not ignored. In addition, parallel [always] overrides the compiler heuristics that estimate the likelihood that parallelization of a loop increases performance. It allows a loop to be parallelized even if the compiler thinks parallelization may not improve performance. If the ASSERT keyword is added, as in #pragma parallel [always [assert]], the compiler generates an error-level assertion message saying that the compiler analysis and cost model indicate that the loop cannot be parallelized.
The #pragma noparallel pragma disables auto-parallelization.