IWOCL 2015 (International Workshop on OpenCL) presentations are available online for free download. It's a good thing that the organizers provide them not long after the conference takes place.
For more info about IWOCL: Link
![]() |
| The Zotac GTX-960 AMP! edition |
BYTEmark* Native Mode Benchmark ver. 2 (10/95)
Index-split by Andrew D. Balsa (11/97)
Linux/Unix* port by Uwe F. Mayer (12/96,11/97)
TEST : Iterations/sec. : Old Index : New Index
: : Pentium 90* : AMD K6/233*
--------------------:------------------:-------------:------------
NUMERIC SORT : 453.9 : 11.64 : 3.82
STRING SORT : 36.298 : 16.22 : 2.51
BITFIELD : 1.1028e+08 : 18.92 : 3.95
FP EMULATION : 82.381 : 39.53 : 9.12
FOURIER : 4877.8 : 5.55 : 3.12
ASSIGNMENT : 7.1713 : 27.29 : 7.08
IDEA : 1364.7 : 20.87 : 6.20
HUFFMAN : 663.8 : 18.41 : 5.88
NEURAL NET : 5.7769 : 9.28 : 3.90
LU DECOMPOSITION : 224.96 : 11.65 : 8.42
==========================ORIGINAL BYTEMARK RESULTS==========================
INTEGER INDEX : 20.419
FLOATING-POINT INDEX: 8.434
Baseline (MSDOS*) : Pentium* 90, 256 KB L2-cache, Watcom* compiler 10.0
==============================LINUX DATA BELOW===============================
CPU : 4 CPU ARMv7 Processor rev 5 (v7l)
L2 Cache :
OS : Linux 3.18.5-v7+
C compiler : gcc-4.7
libc : /lib/arm-linux-gnueabihf/libgcc_s.so.1
MEMORY INDEX : 4.125
INTEGER INDEX : 5.970
FLOATING-POINT INDEX: 4.678
Baseline (LINUX) : AMD K6/233*, 512 KB L2-cache, gcc 2.7.2.3, libc-5.4.38
* Trademarks are property of their respective holder.
Workgroup and sub-workgroup OpenCL 2.0 function evaluation test case Platform/Device selection Total platforms: 1 AMD Accelerated Parallel Processing 1. Bonaire/Advanced Micro Devices, Inc. 2. Intel(R) Pentium(R) 4 CPU 3.06GHz/GenuineIntel Select device index: Device info Platform: AMD Accelerated Parallel Processing Device: Bonaire Driver version: 1642.5 (VM) OpenCL version: OpenCL 2.0 AMD-APP (1642.5) Great! OpenCL 2.0 is supported :) Building kernel with options "-cl-std=CL2.0 -cl-uniform-work-group-size -DK3 -DK2 -DWAVEFRONT_SIZE=64" 1. Shared memory only kernel Executing...Done! Output: 2147450880 / Time: 0.089481 msecs (0.732401 billion elements/second) PASSED! 2. Hybrid kernel via subgroup functions Executing...Done! Output: 2147450880 / Time: 0.215851 msecs (0.303617 billion elements/second) Relative speed-up to kernel 1: 0.41455 PASSED! 3. Workgroup function kernel Executing...Done! Output: 2147450880 / Time: 0.475408 msecs (0.137852 billion elements/second) Relative speed-up to kernel 1: 0.188219 PASSED!