Name | de_modfit_84_bundle4_4s_south4s_bgset_4_1603804501_71959220_1 |
Workunit | 2142725079 |
Created | 26 Jan 2021, 3:38:18 UTC |
Sent | 26 Jan 2021, 3:47:46 UTC |
Report deadline | 7 Feb 2021, 3:47:46 UTC |
Received | 26 Jan 2021, 8:10:54 UTC |
Server state | Over |
Outcome | Success |
Client state | Done |
Exit status | 0 (0x00000000) |
Computer ID | 819576 |
Run time | 48 sec |
CPU time | 31 sec |
Validate state | Valid |
Credit | 0.00 |
Device peak FLOPS | 1,042.23 GFLOPS |
Application version | Milkyway@home Separation v1.46 (opencl_nvidia_101) x86_64-pc-linux-gnu |
Peak working set size | 354.68 MB |
Peak swap size | 128.90 GB |
Peak disk usage | 0.02 MB |
<core_client_version>7.9.3</core_client_version> <![CDATA[ <stderr_txt> <search_application> milkyway_separation 1.46 Linux x86_64 double OpenCL </search_application> BOINC GPU type suggests using OpenCL vendor 'NVIDIA Corporation' Setting process priority to 0 (13): Permission denied Error loading Lua script 'astronomy_parameters.txt': [string "number_parameters: 4..."]:1: '<name>' expected near '4' Switching to Parameter File 'astronomy_parameters.txt' <number_WUs> 4 </number_WUs> <number_params_per_WU> 26 </number_params_per_WU> Using AVX path Found 1 platform Platform 0 information: Name: NVIDIA CUDA Version: OpenCL 1.2 CUDA 11.1.102 Vendor: NVIDIA Corporation Extensions: cl_khr_global_int32_base_atomics cl_khr_global_int32_extended_atomics cl_khr_local_int32_base_atomics cl_khr_local_int32_extended_atomics cl_khr_fp64 cl_khr_byte_addressable_store cl_khr_icd cl_nv_compiler_options cl_nv_device_attribute_query cl_nv_pragma_unroll cl_nv_copy_opts cl_nv_create_buffer cl_khr_int64_base_atomics cl_khr_int64_extended_atomics cl_kernel_attribute_nv cl_khr_device_uuid Profile: FULL_PROFILE Using device 0 on platform 0 Found 2 CL devices Device 'Tesla K40c' (NVIDIA Corporation:0x10de) (CL_DEVICE_TYPE_GPU) Board: Driver version: 455.32.00 Version: OpenCL 1.2 CUDA Compute capability: 3.5 Max compute units: 15 Clock frequency: 745 Mhz Global mem size: 11996954624 Local mem size: 49152 Max const buf size: 65536 Double extension: cl_khr_fp64 Build log: -------------------------------------------------------------------------------- ptxas info : 0 bytes gmem ptxas info : Compiling entry function 'probabilities' for 'sm_35' ptxas info : Function properties for probabilities ptxas . 0 bytes stack frame, 0 bytes spill stores, 0 bytes spill loads ptxas info : Used 126 registers, 420 bytes cmem[0], 192 bytes cmem[2] -------------------------------------------------------------------------------- Build log: -------------------------------------------------------------------------------- -------------------------------------------------------------------------------- Estimated Nvidia GPU GFLOP/s: 715 SP GFLOP/s, 358 DP FLOP/s Using a target frequency of 60.0 Using a block size of 3840 with 146 blocks/chunk Using clWaitForEvents() for polling with initial wait of 12 ms (mode 0) Range: { nu_steps = 320, mu_steps = 800, r_steps = 700 } Iteration area: 560000 Chunk estimate: 4 Num chunks: 1 Chunk size: 560640 Added area: 640 Effective area: 560640 Initial wait: 12 ms Integration time: 9.583908 s. Average time per iteration = 29.949712 ms Integral 0 time = 9.703463 s Estimated Nvidia GPU GFLOP/s: 715 SP GFLOP/s, 358 DP FLOP/s Using a target frequency of 60.0 Using a block size of 3840 with 11 blocks/chunk Using clWaitForEvents() for polling with initial wait of 0 ms (mode 0) Range: { nu_steps = 320, mu_steps = 58, r_steps = 700 } Iteration area: 40600 Chunk estimate: 1 Num chunks: 1 Chunk size: 42240 Added area: 1640 Effective area: 42240 Initial wait: 0 ms Integration time: 0.798067 s. Average time per iteration = 2.493958 ms Integral 1 time = 0.813902 s Running likelihood with 31815 stars Likelihood time = 0.774009 s <background_integral> 0.000048193394294 </background_integral> <stream_integral> 2.667526920127586 28.398802504580061 140.185584412509002 0.000000000000000 </stream_integral> <background_likelihood> -3.279521075087734 </background_likelihood> <stream_only_likelihood> -56.446692566269093 -3.490615098018887 -3.659076214669599 -228.202342937315933 </stream_only_likelihood> <search_likelihood> -2.696926731499021 </search_likelihood> Using AVX path Found 1 platform Platform 0 information: Name: NVIDIA CUDA Version: OpenCL 1.2 CUDA 11.1.102 Vendor: NVIDIA Corporation Extensions: cl_khr_global_int32_base_atomics cl_khr_global_int32_extended_atomics cl_khr_local_int32_base_atomics cl_khr_local_int32_extended_atomics cl_khr_fp64 cl_khr_byte_addressable_store cl_khr_icd cl_nv_compiler_options cl_nv_device_attribute_query cl_nv_pragma_unroll cl_nv_copy_opts cl_nv_create_buffer cl_khr_int64_base_atomics cl_khr_int64_extended_atomics cl_kernel_attribute_nv cl_khr_device_uuid Profile: FULL_PROFILE Using device 0 on platform 0 Found 2 CL devices Device 'Tesla K40c' (NVIDIA Corporation:0x10de) (CL_DEVICE_TYPE_GPU) Board: Driver version: 455.32.00 Version: OpenCL 1.2 CUDA Compute capability: 3.5 Max compute units: 15 Clock frequency: 745 Mhz Global mem size: 11996954624 Local mem size: 49152 Max const buf size: 65536 Double extension: cl_khr_fp64 Build log: -------------------------------------------------------------------------------- ptxas info : 0 bytes gmem ptxas info : Compiling entry function 'probabilities' for 'sm_35' ptxas info : Function properties for probabilities ptxas . 0 bytes stack frame, 0 bytes spill stores, 0 bytes spill loads ptxas info : Used 126 registers, 420 bytes cmem[0], 192 bytes cmem[2] -------------------------------------------------------------------------------- Build log: -------------------------------------------------------------------------------- -------------------------------------------------------------------------------- Estimated Nvidia GPU GFLOP/s: 715 SP GFLOP/s, 358 DP FLOP/s Using a target frequency of 60.0 Using a block size of 3840 with 146 blocks/chunk Using clWaitForEvents() for polling with initial wait of 12 ms (mode 0) Range: { nu_steps = 320, mu_steps = 800, r_steps = 700 } Iteration area: 560000 Chunk estimate: 4 Num chunks: 1 Chunk size: 560640 Added area: 640 Effective area: 560640 Initial wait: 12 ms Integration time: 9.623274 s. Average time per iteration = 30.072733 ms Integral 0 time = 9.745822 s Estimated Nvidia GPU GFLOP/s: 715 SP GFLOP/s, 358 DP FLOP/s Using a target frequency of 60.0 Using a block size of 3840 with 11 blocks/chunk Using clWaitForEvents() for polling with initial wait of 0 ms (mode 0) Range: { nu_steps = 320, mu_steps = 58, r_steps = 700 } Iteration area: 40600 Chunk estimate: 1 Num chunks: 1 Chunk size: 42240 Added area: 1640 Effective area: 42240 Initial wait: 0 ms Integration time: 0.796915 s. Average time per iteration = 2.490358 ms Integral 1 time = 0.825617 s Running likelihood with 31815 stars Likelihood time = 0.545310 s <background_integral1> 0.000046379649939 </background_integral1> <stream_integral1> 3.709460019669782 38.177925595744391 105.504072991313819 0.000000000000223 </stream_integral1> <background_likelihood1> -3.312239350668512 </background_likelihood1> <stream_only_likelihood1> -82.452298620142074 -3.360482674094220 -3.637674143608495 -227.462137196712462 </stream_only_likelihood1> <search_likelihood1> -2.708438603183631 </search_likelihood1> Using AVX path Found 1 platform Platform 0 information: Name: NVIDIA CUDA Version: OpenCL 1.2 CUDA 11.1.102 Vendor: NVIDIA Corporation Extensions: cl_khr_global_int32_base_atomics cl_khr_global_int32_extended_atomics cl_khr_local_int32_base_atomics cl_khr_local_int32_extended_atomics cl_khr_fp64 cl_khr_byte_addressable_store cl_khr_icd cl_nv_compiler_options cl_nv_device_attribute_query cl_nv_pragma_unroll cl_nv_copy_opts cl_nv_create_buffer cl_khr_int64_base_atomics cl_khr_int64_extended_atomics cl_kernel_attribute_nv cl_khr_device_uuid Profile: FULL_PROFILE Using device 0 on platform 0 Found 2 CL devices Device 'Tesla K40c' (NVIDIA Corporation:0x10de) (CL_DEVICE_TYPE_GPU) Board: Driver version: 455.32.00 Version: OpenCL 1.2 CUDA Compute capability: 3.5 Max compute units: 15 Clock frequency: 745 Mhz Global mem size: 11996954624 Local mem size: 49152 Max const buf size: 65536 Double extension: cl_khr_fp64 Build log: -------------------------------------------------------------------------------- ptxas info : 0 bytes gmem ptxas info : Compiling entry function 'probabilities' for 'sm_35' ptxas info : Function properties for probabilities ptxas . 0 bytes stack frame, 0 bytes spill stores, 0 bytes spill loads ptxas info : Used 126 registers, 420 bytes cmem[0], 192 bytes cmem[2] -------------------------------------------------------------------------------- Build log: -------------------------------------------------------------------------------- -------------------------------------------------------------------------------- Estimated Nvidia GPU GFLOP/s: 715 SP GFLOP/s, 358 DP FLOP/s Using a target frequency of 60.0 Using a block size of 3840 with 146 blocks/chunk Using clWaitForEvents() for polling with initial wait of 12 ms (mode 0) Range: { nu_steps = 320, mu_steps = 800, r_steps = 700 } Iteration area: 560000 Chunk estimate: 4 Num chunks: 1 Chunk size: 560640 Added area: 640 Effective area: 560640 Initial wait: 12 ms Integration time: 9.596055 s. Average time per iteration = 29.987672 ms Integral 0 time = 9.805905 s Estimated Nvidia GPU GFLOP/s: 715 SP GFLOP/s, 358 DP FLOP/s Using a target frequency of 60.0 Using a block size of 3840 with 11 blocks/chunk Using clWaitForEvents() for polling with initial wait of 0 ms (mode 0) Range: { nu_steps = 320, mu_steps = 58, r_steps = 700 } Iteration area: 40600 Chunk estimate: 1 Num chunks: 1 Chunk size: 42240 Added area: 1640 Effective area: 42240 Initial wait: 0 ms Integration time: 0.806899 s. Average time per iteration = 2.521561 ms Integral 1 time = 0.839361 s Running likelihood with 31815 stars Likelihood time = 0.398861 s <background_integral2> 0.000048314518356 </background_integral2> <stream_integral2> 15.099201271227580 28.256700987275007 138.058276653898872 0.000000000000001 </stream_integral2> <background_likelihood2> -3.314110671522529 </background_likelihood2> <stream_only_likelihood2> -14.306789914381156 -3.434691892329881 -3.555445602549162 -229.430330564366727 </stream_only_likelihood2> <search_likelihood2> -2.708495064196514 </search_likelihood2> Using AVX path Found 1 platform Platform 0 information: Name: NVIDIA CUDA Version: OpenCL 1.2 CUDA 11.1.102 Vendor: NVIDIA Corporation Extensions: cl_khr_global_int32_base_atomics cl_khr_global_int32_extended_atomics cl_khr_local_int32_base_atomics cl_khr_local_int32_extended_atomics cl_khr_fp64 cl_khr_byte_addressable_store cl_khr_icd cl_nv_compiler_options cl_nv_device_attribute_query cl_nv_pragma_unroll cl_nv_copy_opts cl_nv_create_buffer cl_khr_int64_base_atomics cl_khr_int64_extended_atomics cl_kernel_attribute_nv cl_khr_device_uuid Profile: FULL_PROFILE Using device 0 on platform 0 Found 2 CL devices Device 'Tesla K40c' (NVIDIA Corporation:0x10de) (CL_DEVICE_TYPE_GPU) Board: Driver version: 455.32.00 Version: OpenCL 1.2 CUDA Compute capability: 3.5 Max compute units: 15 Clock frequency: 745 Mhz Global mem size: 11996954624 Local mem size: 49152 Max const buf size: 65536 Double extension: cl_khr_fp64 Build log: -------------------------------------------------------------------------------- ptxas info : 0 bytes gmem ptxas info : Compiling entry function 'probabilities' for 'sm_35' ptxas info : Function properties for probabilities ptxas . 0 bytes stack frame, 0 bytes spill stores, 0 bytes spill loads ptxas info : Used 126 registers, 420 bytes cmem[0], 192 bytes cmem[2] -------------------------------------------------------------------------------- Build log: -------------------------------------------------------------------------------- -------------------------------------------------------------------------------- Estimated Nvidia GPU GFLOP/s: 715 SP GFLOP/s, 358 DP FLOP/s Using a target frequency of 60.0 Using a block size of 3840 with 146 blocks/chunk Using clWaitForEvents() for polling with initial wait of 12 ms (mode 0) Range: { nu_steps = 320, mu_steps = 800, r_steps = 700 } Iteration area: 560000 Chunk estimate: 4 Num chunks: 1 Chunk size: 560640 Added area: 640 Effective area: 560640 Initial wait: 12 ms Integration time: 9.621092 s. Average time per iteration = 30.065912 ms Integral 0 time = 9.737510 s Estimated Nvidia GPU GFLOP/s: 715 SP GFLOP/s, 358 DP FLOP/s Using a target frequency of 60.0 Using a block size of 3840 with 11 blocks/chunk Using clWaitForEvents() for polling with initial wait of 0 ms (mode 0) Range: { nu_steps = 320, mu_steps = 58, r_steps = 700 } Iteration area: 40600 Chunk estimate: 1 Num chunks: 1 Chunk size: 42240 Added area: 1640 Effective area: 42240 Initial wait: 0 ms Integration time: 0.797727 s. Average time per iteration = 2.492896 ms Integral 1 time = 0.814019 s Running likelihood with 31815 stars Likelihood time = 0.751247 s Non-finite result: setting likelihood to -999 <background_integral3> 0.000048201403513 </background_integral3> <stream_integral3> 7.313471634939821 34.549255144347491 98.746010407555701 -0.000000000000000 </stream_integral3> <background_likelihood3> -3.325680575082707 </background_likelihood3> <stream_only_likelihood3> -31.015766432062751 -3.298418991760043 -3.846343977487686 nan </stream_only_likelihood3> <search_likelihood3> -999.000000000000000 </search_likelihood3> 23:28:56 (26204): called boinc_finish(0) </stderr_txt> ]]>
©2024 Astroinformatics Group