Name | de_modfit_84_bundle4_4s_south4s_bgset_4_1603804501_70431041_0 |
Workunit | 2141028862 |
Created | 23 Jan 2021, 9:10:11 UTC |
Sent | 23 Jan 2021, 9:21:24 UTC |
Report deadline | 4 Feb 2021, 9:21:24 UTC |
Received | 23 Jan 2021, 14:53:09 UTC |
Server state | Over |
Outcome | Success |
Client state | Done |
Exit status | 0 (0x00000000) |
Computer ID | 855352 |
Run time | 8 min 20 sec |
CPU time | 25 sec |
Validate state | Valid |
Credit | 0.00 |
Device peak FLOPS | 99.34 GFLOPS |
Application version | Milkyway@home Separation v1.46 (opencl_nvidia_101) x86_64-pc-linux-gnu |
Peak working set size | 199.18 MB |
Peak swap size | 16.97 GB |
Peak disk usage | 0.03 MB |
<core_client_version>7.16.6</core_client_version> <![CDATA[ <stderr_txt> <search_application> milkyway_separation 1.46 Linux x86_64 double OpenCL </search_application> Reading preferences ended prematurely BOINC GPU type suggests using OpenCL vendor 'NVIDIA Corporation' Setting process priority to 0 (13): Permission denied Error loading Lua script 'astronomy_parameters.txt': [string "number_parameters: 4..."]:1: '<name>' expected near '4' Switching to Parameter File 'astronomy_parameters.txt' <number_WUs> 4 </number_WUs> <number_params_per_WU> 26 </number_params_per_WU> Using SSE3 path Found 1 platform Platform 0 information: Name: NVIDIA CUDA Version: OpenCL 1.2 CUDA 9.1.84 Vendor: NVIDIA Corporation Extensions: cl_khr_global_int32_base_atomics cl_khr_global_int32_extended_atomics cl_khr_local_int32_base_atomics cl_khr_local_int32_extended_atomics cl_khr_fp64 cl_khr_byte_addressable_store cl_khr_icd cl_khr_gl_sharing cl_nv_compiler_options cl_nv_device_attribute_query cl_nv_pragma_unroll cl_nv_copy_opts cl_khr_gl_event cl_nv_create_buffer Profile: FULL_PROFILE Using device 0 on platform 0 Found 1 CL device Device 'Quadro 4000' (NVIDIA Corporation:0x10de) (CL_DEVICE_TYPE_GPU) Board: Driver version: 390.141 Version: OpenCL 1.1 CUDA Compute capability: 2.0 Max compute units: 8 Clock frequency: 950 Mhz Global mem size: 2080505856 Local mem size: 49152 Max const buf size: 65536 Double extension: cl_khr_fp64 Build log: -------------------------------------------------------------------------------- ptxas info : 0 bytes gmem ptxas info : Compiling entry function 'probabilities' for 'sm_20' ptxas info : Function properties for probabilities ptxas . 272 bytes stack frame, 272 bytes spill stores, 272 bytes spill loads ptxas info : Used 62 registers, 132 bytes cmem[0], 40 bytes cmem[16] -------------------------------------------------------------------------------- Build log: -------------------------------------------------------------------------------- -------------------------------------------------------------------------------- Estimated Nvidia GPU GFLOP/s: 486 SP GFLOP/s, 61 DP FLOP/s Using a target frequency of 60.0 Using a block size of 4096 with 5 blocks/chunk Using clWaitForEvents() for polling with initial wait of 12 ms (mode 0) Range: { nu_steps = 320, mu_steps = 800, r_steps = 700 } Iteration area: 560000 Chunk estimate: 23 Num chunks: 28 Chunk size: 20480 Added area: 13440 Effective area: 573440 Initial wait: 12 ms Integration time: 111.088651 s. Average time per iteration = 347.152033 ms Integral 0 time = 111.483127 s Estimated Nvidia GPU GFLOP/s: 486 SP GFLOP/s, 61 DP FLOP/s Using a target frequency of 60.0 Using a block size of 4096 with 9 blocks/chunk Using clWaitForEvents() for polling with initial wait of 21 ms (mode 0) Range: { nu_steps = 320, mu_steps = 58, r_steps = 700 } Iteration area: 40600 Chunk estimate: 1 Num chunks: 2 Chunk size: 36864 Added area: 33128 Effective area: 73728 Initial wait: 21 ms Integration time: 12.798862 s. Average time per iteration = 39.996445 ms Integral 1 time = 12.842225 s Running likelihood with 31815 stars Likelihood time = 0.873870 s <background_integral> 0.000047109630849 </background_integral> <stream_integral> 9.635604061443129 25.816988571343089 96.246289735899580 0.478368772192369 </stream_integral> <background_likelihood> -3.275493370357958 </background_likelihood> <stream_only_likelihood> -26.529255855240407 -3.490542208641087 -3.819215421013562 -132.415475444360993 </stream_only_likelihood> <search_likelihood> -2.705176801510927 </search_likelihood> Using SSE3 path Found 1 platform Platform 0 information: Name: NVIDIA CUDA Version: OpenCL 1.2 CUDA 9.1.84 Vendor: NVIDIA Corporation Extensions: cl_khr_global_int32_base_atomics cl_khr_global_int32_extended_atomics cl_khr_local_int32_base_atomics cl_khr_local_int32_extended_atomics cl_khr_fp64 cl_khr_byte_addressable_store cl_khr_icd cl_khr_gl_sharing cl_nv_compiler_options cl_nv_device_attribute_query cl_nv_pragma_unroll cl_nv_copy_opts cl_khr_gl_event cl_nv_create_buffer Profile: FULL_PROFILE Using device 0 on platform 0 Found 1 CL device Device 'Quadro 4000' (NVIDIA Corporation:0x10de) (CL_DEVICE_TYPE_GPU) Board: Driver version: 390.141 Version: OpenCL 1.1 CUDA Compute capability: 2.0 Max compute units: 8 Clock frequency: 950 Mhz Global mem size: 2080505856 Local mem size: 49152 Max const buf size: 65536 Double extension: cl_khr_fp64 Build log: -------------------------------------------------------------------------------- ptxas info : 0 bytes gmem ptxas info : Compiling entry function 'probabilities' for 'sm_20' ptxas info : Function properties for probabilities ptxas . 272 bytes stack frame, 272 bytes spill stores, 272 bytes spill loads ptxas info : Used 62 registers, 132 bytes cmem[0], 40 bytes cmem[16] -------------------------------------------------------------------------------- Build log: -------------------------------------------------------------------------------- -------------------------------------------------------------------------------- Estimated Nvidia GPU GFLOP/s: 486 SP GFLOP/s, 61 DP FLOP/s Using a target frequency of 60.0 Using a block size of 4096 with 5 blocks/chunk Using clWaitForEvents() for polling with initial wait of 12 ms (mode 0) Range: { nu_steps = 320, mu_steps = 800, r_steps = 700 } Iteration area: 560000 Chunk estimate: 23 Num chunks: 28 Chunk size: 20480 Added area: 13440 Effective area: 573440 Initial wait: 12 ms Integration time: 108.033726 s. Average time per iteration = 337.605395 ms Integral 0 time = 108.414522 s Estimated Nvidia GPU GFLOP/s: 486 SP GFLOP/s, 61 DP FLOP/s Using a target frequency of 60.0 Using a block size of 4096 with 9 blocks/chunk Using clWaitForEvents() for polling with initial wait of 21 ms (mode 0) Range: { nu_steps = 320, mu_steps = 58, r_steps = 700 } Iteration area: 40600 Chunk estimate: 1 Num chunks: 2 Chunk size: 36864 Added area: 33128 Effective area: 73728 Initial wait: 21 ms Integration time: 12.801436 s. Average time per iteration = 40.004487 ms Integral 1 time = 12.846066 s Running likelihood with 31815 stars Likelihood time = 0.864135 s <background_integral1> 0.000047254743277 </background_integral1> <stream_integral1> 3.647119430484928 29.465714119274828 99.673558566214410 0.000000000000000 </stream_integral1> <background_likelihood1> -3.291270176873298 </background_likelihood1> <stream_only_likelihood1> -33.205092579749667 -3.382216554968683 -3.824065075052576 -233.260525041871745 </stream_only_likelihood1> <search_likelihood1> -2.702262184515336 </search_likelihood1> Using SSE3 path Found 1 platform Platform 0 information: Name: NVIDIA CUDA Version: OpenCL 1.2 CUDA 9.1.84 Vendor: NVIDIA Corporation Extensions: cl_khr_global_int32_base_atomics cl_khr_global_int32_extended_atomics cl_khr_local_int32_base_atomics cl_khr_local_int32_extended_atomics cl_khr_fp64 cl_khr_byte_addressable_store cl_khr_icd cl_khr_gl_sharing cl_nv_compiler_options cl_nv_device_attribute_query cl_nv_pragma_unroll cl_nv_copy_opts cl_khr_gl_event cl_nv_create_buffer Profile: FULL_PROFILE Using device 0 on platform 0 Found 1 CL device Device 'Quadro 4000' (NVIDIA Corporation:0x10de) (CL_DEVICE_TYPE_GPU) Board: Driver version: 390.141 Version: OpenCL 1.1 CUDA Compute capability: 2.0 Max compute units: 8 Clock frequency: 950 Mhz Global mem size: 2080505856 Local mem size: 49152 Max const buf size: 65536 Double extension: cl_khr_fp64 Build log: -------------------------------------------------------------------------------- ptxas info : 0 bytes gmem ptxas info : Compiling entry function 'probabilities' for 'sm_20' ptxas info : Function properties for probabilities ptxas . 272 bytes stack frame, 272 bytes spill stores, 272 bytes spill loads ptxas info : Used 62 registers, 132 bytes cmem[0], 40 bytes cmem[16] -------------------------------------------------------------------------------- Build log: -------------------------------------------------------------------------------- -------------------------------------------------------------------------------- Estimated Nvidia GPU GFLOP/s: 486 SP GFLOP/s, 61 DP FLOP/s Using a target frequency of 60.0 Using a block size of 4096 with 5 blocks/chunk Using clWaitForEvents() for polling with initial wait of 12 ms (mode 0) Range: { nu_steps = 320, mu_steps = 800, r_steps = 700 } Iteration area: 560000 Chunk estimate: 23 Num chunks: 28 Chunk size: 20480 Added area: 13440 Effective area: 573440 Initial wait: 12 ms Integration time: 108.156540 s. Average time per iteration = 337.989186 ms Integral 0 time = 108.538275 s Estimated Nvidia GPU GFLOP/s: 486 SP GFLOP/s, 61 DP FLOP/s Using a target frequency of 60.0 Using a block size of 4096 with 9 blocks/chunk Using clWaitForEvents() for polling with initial wait of 21 ms (mode 0) Range: { nu_steps = 320, mu_steps = 58, r_steps = 700 } Iteration area: 40600 Chunk estimate: 1 Num chunks: 2 Chunk size: 36864 Added area: 33128 Effective area: 73728 Initial wait: 21 ms Integration time: 12.799583 s. Average time per iteration = 39.998696 ms Integral 1 time = 12.841172 s Running likelihood with 31815 stars Likelihood time = 0.864517 s <background_integral2> 0.000047166347437 </background_integral2> <stream_integral2> 3.987289225682057 29.088523585099605 81.998064961747261 0.142275828735448 </stream_integral2> <background_likelihood2> -3.271624746783580 </background_likelihood2> <stream_only_likelihood2> -56.195845295334721 -3.420381272316790 -3.936044172182415 -219.919296943336832 </stream_only_likelihood2> <search_likelihood2> -2.704182225760874 </search_likelihood2> Using SSE3 path Found 1 platform Platform 0 information: Name: NVIDIA CUDA Version: OpenCL 1.2 CUDA 9.1.84 Vendor: NVIDIA Corporation Extensions: cl_khr_global_int32_base_atomics cl_khr_global_int32_extended_atomics cl_khr_local_int32_base_atomics cl_khr_local_int32_extended_atomics cl_khr_fp64 cl_khr_byte_addressable_store cl_khr_icd cl_khr_gl_sharing cl_nv_compiler_options cl_nv_device_attribute_query cl_nv_pragma_unroll cl_nv_copy_opts cl_khr_gl_event cl_nv_create_buffer Profile: FULL_PROFILE Using device 0 on platform 0 Found 1 CL device Device 'Quadro 4000' (NVIDIA Corporation:0x10de) (CL_DEVICE_TYPE_GPU) Board: Driver version: 390.141 Version: OpenCL 1.1 CUDA Compute capability: 2.0 Max compute units: 8 Clock frequency: 950 Mhz Global mem size: 2080505856 Local mem size: 49152 Max const buf size: 65536 Double extension: cl_khr_fp64 Build log: -------------------------------------------------------------------------------- ptxas info : 0 bytes gmem ptxas info : Compiling entry function 'probabilities' for 'sm_20' ptxas info : Function properties for probabilities ptxas . 272 bytes stack frame, 272 bytes spill stores, 272 bytes spill loads ptxas info : Used 62 registers, 132 bytes cmem[0], 40 bytes cmem[16] -------------------------------------------------------------------------------- Build log: -------------------------------------------------------------------------------- -------------------------------------------------------------------------------- Estimated Nvidia GPU GFLOP/s: 486 SP GFLOP/s, 61 DP FLOP/s Using a target frequency of 60.0 Using a block size of 4096 with 5 blocks/chunk Using clWaitForEvents() for polling with initial wait of 12 ms (mode 0) Range: { nu_steps = 320, mu_steps = 800, r_steps = 700 } Iteration area: 560000 Chunk estimate: 23 Num chunks: 28 Chunk size: 20480 Added area: 13440 Effective area: 573440 Initial wait: 12 ms Integration time: 111.606955 s. Average time per iteration = 348.771733 ms Integral 0 time = 111.994839 s Estimated Nvidia GPU GFLOP/s: 486 SP GFLOP/s, 61 DP FLOP/s Using a target frequency of 60.0 Using a block size of 4096 with 9 blocks/chunk Using clWaitForEvents() for polling with initial wait of 21 ms (mode 0) Range: { nu_steps = 320, mu_steps = 58, r_steps = 700 } Iteration area: 40600 Chunk estimate: 1 Num chunks: 2 Chunk size: 36864 Added area: 33128 Effective area: 73728 Initial wait: 21 ms Integration time: 12.803254 s. Average time per iteration = 40.010168 ms Integral 1 time = 12.843004 s Running likelihood with 31815 stars Likelihood time = 0.871950 s <background_integral3> 0.000047345055487 </background_integral3> <stream_integral3> 4.554337844056457 28.703598425797267 112.136745853091185 0.964010208822804 </stream_integral3> <background_likelihood3> -3.296774467826794 </background_likelihood3> <stream_only_likelihood3> -29.170997615515681 -3.414254301832802 -3.767887236693591 -111.529618587495960 </stream_only_likelihood3> <search_likelihood3> -2.704172517667945 </search_likelihood3> 14:53:01 (123002): called boinc_finish(0) </stderr_txt> ]]>
©2024 Astroinformatics Group