Name | de_modfit_85_bundle4_4s_south4s_bgset_4_1603804501_66072433_0 |
Workunit | 2136182185 |
Created | 17 Jan 2021, 21:40:38 UTC |
Sent | 17 Jan 2021, 21:50:28 UTC |
Report deadline | 29 Jan 2021, 21:50:28 UTC |
Received | 18 Jan 2021, 11:48:41 UTC |
Server state | Over |
Outcome | Validate error |
Client state | Done |
Exit status | 0 (0x00000000) |
Computer ID | 825027 |
Run time | 6 min 35 sec |
CPU time | 6 min 17 sec |
Validate state | Invalid |
Credit | 0.00 |
Device peak FLOPS | 1,143.60 GFLOPS |
Application version | Milkyway@home Separation v1.46 (opencl_nvidia_101) x86_64-pc-linux-gnu |
Peak working set size | 622.35 MB |
Peak swap size | 33.62 GB |
Peak disk usage | 0.03 MB |
<core_client_version>7.9.3</core_client_version> <![CDATA[ <stderr_txt> <search_application> milkyway_separation 1.46 Linux x86_64 double OpenCL </search_application> Reading preferences ended prematurely BOINC GPU type suggests using OpenCL vendor 'NVIDIA Corporation' Setting process priority to 0 (13): Permission denied Error loading Lua script 'astronomy_parameters.txt': [string "number_parameters: 4..."]:1: '<name>' expected near '4' Switching to Parameter File 'astronomy_parameters.txt' <number_WUs> 4 </number_WUs> <number_params_per_WU> 26 </number_params_per_WU> Using AVX path Found 1 platform Platform 0 information: Name: NVIDIA CUDA Version: OpenCL 1.2 CUDA 10.2.120 Vendor: NVIDIA Corporation Extensions: cl_khr_global_int32_base_atomics cl_khr_global_int32_extended_atomics cl_khr_local_int32_base_atomics cl_khr_local_int32_extended_atomics cl_khr_fp64 cl_khr_byte_addressable_store cl_khr_icd cl_khr_gl_sharing cl_nv_compiler_options cl_nv_device_attribute_query cl_nv_pragma_unroll cl_nv_copy_opts cl_nv_create_buffer Profile: FULL_PROFILE Using device 0 on platform 0 Found 2 CL devices Device 'Tesla P4' (NVIDIA Corporation:0x10de) (CL_DEVICE_TYPE_GPU) Board: Driver version: 430.26 Version: OpenCL 1.2 CUDA Compute capability: 6.1 Max compute units: 20 Clock frequency: 1113 Mhz Global mem size: 7981694976 Local mem size: 49152 Max const buf size: 65536 Double extension: cl_khr_fp64 Build log: -------------------------------------------------------------------------------- ptxas info : 0 bytes gmem ptxas info : Compiling entry function 'probabilities' for 'sm_61' ptxas info : Function properties for probabilities ptxas . 0 bytes stack frame, 0 bytes spill stores, 0 bytes spill loads ptxas info : Used 127 registers, 420 bytes cmem[0], 184 bytes cmem[2] -------------------------------------------------------------------------------- Build log: -------------------------------------------------------------------------------- -------------------------------------------------------------------------------- Estimated Nvidia GPU GFLOP/s: 1425 SP GFLOP/s, 712 DP FLOP/s Using a target frequency of 60.0 Using a block size of 10240 with 55 blocks/chunk Using clWaitForEvents() for polling with initial wait of 12 ms (mode 0) Range: { nu_steps = 320, mu_steps = 800, r_steps = 700 } Iteration area: 560000 Chunk estimate: 2 Num chunks: 1 Chunk size: 563200 Added area: 3200 Effective area: 563200 Initial wait: 12 ms Integration time: 86.797245 s. Average time per iteration = 271.241392 ms Integral 0 time = 87.125835 s Estimated Nvidia GPU GFLOP/s: 1425 SP GFLOP/s, 712 DP FLOP/s Using a target frequency of 60.0 Using a block size of 10240 with 4 blocks/chunk Using clWaitForEvents() for polling with initial wait of 0 ms (mode 0) Range: { nu_steps = 320, mu_steps = 58, r_steps = 700 } Iteration area: 40600 Chunk estimate: 1 Num chunks: 1 Chunk size: 40960 Added area: 360 Effective area: 40960 Initial wait: 0 ms Integration time: 7.131689 s. Average time per iteration = 22.286528 ms Integral 1 time = 7.207245 s Running likelihood with 31964 stars Likelihood time = 0.732647 s <background_integral> 0.000046777229099 </background_integral> <stream_integral> 137.487909052735063 21.778095435419647 0.000000000000000 115.487648946837339 </stream_integral> <background_likelihood> -3.178338604344296 </background_likelihood> <stream_only_likelihood> -3.739210525042044 -3.678998678444222 -224.263122464547564 -7.602050292913450 </stream_only_likelihood> <search_likelihood> -2.687157183748678 </search_likelihood> Using AVX path Found 1 platform Platform 0 information: Name: NVIDIA CUDA Version: OpenCL 1.2 CUDA 10.2.120 Vendor: NVIDIA Corporation Extensions: cl_khr_global_int32_base_atomics cl_khr_global_int32_extended_atomics cl_khr_local_int32_base_atomics cl_khr_local_int32_extended_atomics cl_khr_fp64 cl_khr_byte_addressable_store cl_khr_icd cl_khr_gl_sharing cl_nv_compiler_options cl_nv_device_attribute_query cl_nv_pragma_unroll cl_nv_copy_opts cl_nv_create_buffer Profile: FULL_PROFILE Using device 0 on platform 0 Found 2 CL devices Device 'Tesla P4' (NVIDIA Corporation:0x10de) (CL_DEVICE_TYPE_GPU) Board: Driver version: 430.26 Version: OpenCL 1.2 CUDA Compute capability: 6.1 Max compute units: 20 Clock frequency: 1113 Mhz Global mem size: 7981694976 Local mem size: 49152 Max const buf size: 65536 Double extension: cl_khr_fp64 Build log: -------------------------------------------------------------------------------- ptxas info : 0 bytes gmem ptxas info : Compiling entry function 'probabilities' for 'sm_61' ptxas info : Function properties for probabilities ptxas . 0 bytes stack frame, 0 bytes spill stores, 0 bytes spill loads ptxas info : Used 127 registers, 420 bytes cmem[0], 184 bytes cmem[2] -------------------------------------------------------------------------------- Build log: -------------------------------------------------------------------------------- -------------------------------------------------------------------------------- Estimated Nvidia GPU GFLOP/s: 1425 SP GFLOP/s, 712 DP FLOP/s Using a target frequency of 60.0 Using a block size of 10240 with 55 blocks/chunk Using clWaitForEvents() for polling with initial wait of 12 ms (mode 0) Range: { nu_steps = 320, mu_steps = 800, r_steps = 700 } Iteration area: 560000 Chunk estimate: 2 Num chunks: 1 Chunk size: 563200 Added area: 3200 Effective area: 563200 Initial wait: 12 ms Integration time: 90.333303 s. Average time per iteration = 282.291573 ms Integral 0 time = 90.705119 s Estimated Nvidia GPU GFLOP/s: 1425 SP GFLOP/s, 712 DP FLOP/s Using a target frequency of 60.0 Using a block size of 10240 with 4 blocks/chunk Using clWaitForEvents() for polling with initial wait of 0 ms (mode 0) Range: { nu_steps = 320, mu_steps = 58, r_steps = 700 } Iteration area: 40600 Chunk estimate: 1 Num chunks: 1 Chunk size: 40960 Added area: 360 Effective area: 40960 Initial wait: 0 ms Integration time: 7.142489 s. Average time per iteration = 22.320278 ms Integral 1 time = 7.213864 s Running likelihood with 31964 stars Likelihood time = 0.736509 s <background_integral1> 0.000046033878282 </background_integral1> <stream_integral1> 140.911234464605769 31.029826230193272 0.000000000000000 129.470425192973323 </stream_integral1> <background_likelihood1> -3.226641101951116 </background_likelihood1> <stream_only_likelihood1> -3.812907115999841 -3.343310579834471 -224.185529028257491 -7.571948593524957 </stream_only_likelihood1> <search_likelihood1> -2.686860807801480 </search_likelihood1> Using AVX path Found 1 platform Platform 0 information: Name: NVIDIA CUDA Version: OpenCL 1.2 CUDA 10.2.120 Vendor: NVIDIA Corporation Extensions: cl_khr_global_int32_base_atomics cl_khr_global_int32_extended_atomics cl_khr_local_int32_base_atomics cl_khr_local_int32_extended_atomics cl_khr_fp64 cl_khr_byte_addressable_store cl_khr_icd cl_khr_gl_sharing cl_nv_compiler_options cl_nv_device_attribute_query cl_nv_pragma_unroll cl_nv_copy_opts cl_nv_create_buffer Profile: FULL_PROFILE Using device 0 on platform 0 Found 2 CL devices Device 'Tesla P4' (NVIDIA Corporation:0x10de) (CL_DEVICE_TYPE_GPU) Board: Driver version: 430.26 Version: OpenCL 1.2 CUDA Compute capability: 6.1 Max compute units: 20 Clock frequency: 1113 Mhz Global mem size: 7981694976 Local mem size: 49152 Max const buf size: 65536 Double extension: cl_khr_fp64 Build log: -------------------------------------------------------------------------------- ptxas info : 0 bytes gmem ptxas info : Compiling entry function 'probabilities' for 'sm_61' ptxas info : Function properties for probabilities ptxas . 0 bytes stack frame, 0 bytes spill stores, 0 bytes spill loads ptxas info : Used 127 registers, 420 bytes cmem[0], 184 bytes cmem[2] -------------------------------------------------------------------------------- Build log: -------------------------------------------------------------------------------- -------------------------------------------------------------------------------- Estimated Nvidia GPU GFLOP/s: 1425 SP GFLOP/s, 712 DP FLOP/s Using a target frequency of 60.0 Using a block size of 10240 with 55 blocks/chunk Using clWaitForEvents() for polling with initial wait of 12 ms (mode 0) Range: { nu_steps = 320, mu_steps = 800, r_steps = 700 } Iteration area: 560000 Chunk estimate: 2 Num chunks: 1 Chunk size: 563200 Added area: 3200 Effective area: 563200 Initial wait: 12 ms Integration time: 90.416744 s. Average time per iteration = 282.552325 ms Integral 0 time = 90.744093 s Estimated Nvidia GPU GFLOP/s: 1425 SP GFLOP/s, 712 DP FLOP/s Using a target frequency of 60.0 Using a block size of 10240 with 4 blocks/chunk Using clWaitForEvents() for polling with initial wait of 0 ms (mode 0) Range: { nu_steps = 320, mu_steps = 58, r_steps = 700 } Iteration area: 40600 Chunk estimate: 1 Num chunks: 1 Chunk size: 40960 Added area: 360 Effective area: 40960 Initial wait: 0 ms Integration time: 7.105624 s. Average time per iteration = 22.205074 ms Integral 1 time = 7.178963 s Running likelihood with 31964 stars Likelihood time = 0.728008 s <background_integral2> 0.000046601054200 </background_integral2> <stream_integral2> 130.354562479232470 27.377681011887677 0.000000000000000 101.768483570054315 </stream_integral2> <background_likelihood2> -3.242319020557554 </background_likelihood2> <stream_only_likelihood2> -3.543352586531191 -3.356658708801019 -224.253352083388535 -11.314321928613278 </stream_only_likelihood2> <search_likelihood2> -2.686797974389359 </search_likelihood2> Using AVX path Found 1 platform Platform 0 information: Name: NVIDIA CUDA Version: OpenCL 1.2 CUDA 10.2.120 Vendor: NVIDIA Corporation Extensions: cl_khr_global_int32_base_atomics cl_khr_global_int32_extended_atomics cl_khr_local_int32_base_atomics cl_khr_local_int32_extended_atomics cl_khr_fp64 cl_khr_byte_addressable_store cl_khr_icd cl_khr_gl_sharing cl_nv_compiler_options cl_nv_device_attribute_query cl_nv_pragma_unroll cl_nv_copy_opts cl_nv_create_buffer Profile: FULL_PROFILE Using device 0 on platform 0 Found 2 CL devices Device 'Tesla P4' (NVIDIA Corporation:0x10de) (CL_DEVICE_TYPE_GPU) Board: Driver version: 430.26 Version: OpenCL 1.2 CUDA Compute capability: 6.1 Max compute units: 20 Clock frequency: 1113 Mhz Global mem size: 7981694976 Local mem size: 49152 Max const buf size: 65536 Double extension: cl_khr_fp64 Build log: -------------------------------------------------------------------------------- ptxas info : 0 bytes gmem ptxas info : Compiling entry function 'probabilities' for 'sm_61' ptxas info : Function properties for probabilities ptxas . 0 bytes stack frame, 0 bytes spill stores, 0 bytes spill loads ptxas info : Used 127 registers, 420 bytes cmem[0], 184 bytes cmem[2] -------------------------------------------------------------------------------- Build log: -------------------------------------------------------------------------------- -------------------------------------------------------------------------------- Estimated Nvidia GPU GFLOP/s: 1425 SP GFLOP/s, 712 DP FLOP/s Using a target frequency of 60.0 Using a block size of 10240 with 55 blocks/chunk Using clWaitForEvents() for polling with initial wait of 12 ms (mode 0) Range: { nu_steps = 320, mu_steps = 800, r_steps = 700 } Iteration area: 560000 Chunk estimate: 2 Num chunks: 1 Chunk size: 563200 Added area: 3200 Effective area: 563200 Initial wait: 12 ms Integration time: 90.201020 s. Average time per iteration = 281.878188 ms Integral 0 time = 90.572611 s Estimated Nvidia GPU GFLOP/s: 1425 SP GFLOP/s, 712 DP FLOP/s Using a target frequency of 60.0 Using a block size of 10240 with 4 blocks/chunk Using clWaitForEvents() for polling with initial wait of 0 ms (mode 0) Range: { nu_steps = 320, mu_steps = 58, r_steps = 700 } Iteration area: 40600 Chunk estimate: 1 Num chunks: 1 Chunk size: 40960 Added area: 360 Effective area: 40960 Initial wait: 0 ms Integration time: 7.126829 s. Average time per iteration = 22.271341 ms Integral 1 time = 7.199140 s Running likelihood with 31964 stars Likelihood time = 0.729755 s <background_integral3> 0.000045100841464 </background_integral3> <stream_integral3> 132.951555597940171 19.392716288785039 0.000000000000000 106.282806908493527 </stream_integral3> <background_likelihood3> -3.245869755468560 </background_likelihood3> <stream_only_likelihood3> -3.554967850474476 -3.638496660982142 -223.746101610143057 -9.637580646243606 </stream_only_likelihood3> <search_likelihood3> -2.688170493351780 </search_likelihood3> 11:48:32 (18100): called boinc_finish(0) </stderr_txt> ]]>
©2024 Astroinformatics Group