Forum Discussion
Altera_Forum
Honored Contributor
8 years agoDifferent kernels of same algorithm give different throughputs
Hi, I'm trying to test the performance of my 385A card, using different OpenCL kernels, while all of them represent the same functionality. Here are two of my kernels: __attribute__((num...
Altera_Forum
Honored Contributor
8 years agoLoops iterations are NOT pipelined in NDRange kernels, but instead, different threads are scheduled into the same loop pipeline at runtime; obviously, the more threads you have, the more successful the runtime scheduler will be in keeping the pipeline, resulting in higher performance. Furthermore, I believe a major contributing factor to the performance difference you are seeing could be the difference in the operating frequency of the kernels.