Name | hadam3p_eu_e6ue_2013_1_008850657_1 |
Workunit | 8996586 |
Created | 7 Aug 2014, 9:57:43 UTC |
Sent | 7 Aug 2014, 9:58:20 UTC |
Report deadline | 20 Jul 2015, 15:18:20 UTC |
Received | 14 Aug 2014, 14:28:34 UTC |
Server state | Over |
Outcome | Computation error |
Client state | Compute error |
Exit status | 0 (0x00000000) |
Computer ID | 1084069 |
Run time | 2 days 8 hours 35 min 28 sec |
CPU time | 2 days 6 hours 24 min 30 sec |
Validate state | Invalid |
Credit | 1,194.02 |
Device peak FLOPS | 3.02 GFLOPS |
Application version | UK Met Office HadAM3P-HadRM3P Europe v6.09 windows_intelx86 |
Stderr | <core_client_version>7.2.42</core_client_version> <![CDATA[ <stderr_txt> CPDN Monitor - Quit request from BOINC... CPDN Monitor - Quit request from BOINC... Suspended CPDN Monitor - Suspend request from BOINC... CPDN Monitor - Quit request from BOINC... CPDN Monitor - Quit request from BOINC... CPDN Monitor - Quit request from BOINC... Regional Worker:: CPDN process is not running, exiting, bRetVal = 1, checkPID=9956, selfPID=9956, iMonCtr=2 CPDN Monitor - Quit request from BOINC... Suspended CPDN Monitor - Suspend request from BOINC... CPDN Monitor - Quit request from BOINC... Suspended CPDN Monitor - Suspend request from BOINC... CPDN Monitor - Quit request from BOINC... CPDN Monitor - Quit request from BOINC... CPDN Monitor - Quit request from BOINC... CPDN Monitor - Quit request from BOINC... Regional Worker:: CPDN process is not running, exiting, bRetVal = 1, checkPID=2568, selfPID=2568, iMonCtr=2 CPDN Monitor - Quit request from BOINC... Regional WoSuspended CPDN Monitor - Suspend request from BOINC... Suspended CPDN Monitor - Suspend request from BOINC... Suspended CPDN Monitor - Suspend request from BOINC... Suspended CPDN Monitor - Suspend request from BOINC... Suspended CPDN Monitor - Suspend request from BOINC... CPDN Monitor - Quit request from BOINC... CPDN Monitor - Quit request from BOINC... CPDN Monitor - Quit request from BOINC... CPDN Monitor - Quit request from BOINC... Regional Worker:: CPDN process is not running, exiting, bRetVal = 1, checkPID=18340, selfPID=18340, iMonCtr=2 CPDN Monitor - Quit request from BOINC... CPDN Monitor - Quit request from BOINC... 19:13:53 (19812): No heartbeat from core client for 30 sec - exiting CPDN Monitor - No 'heartbeat' from BOINC... Regional Worker:: CPDN process is not running, exiting, bRetVal = 1, checkPID=20948, selfPID=20948, iMonCtr=2 Suspended CPDN Monitor - Suspend request from BOINC... CPDN Monitor - Quit request from BOINC... CPDN Monitor - Quit request from BOINC... Regional Worker:: CPDN process is not running, exiting, bRetVal = 1, checkPID=22288, selfPID=22288, iMonCtr=2 CPDN Monitor - Quit request from BOINC... 23:48:30 (11304): No heartbeat from core client for 30 sec - exiting 23:48:31 (11304): No heartbeat from core client for 30 sec - exiting 23:48:32 (11304): No heartbeat from core client for 30 sec - exiting 23:48:33 (11304): No heartbeat from core client for 30 sec - exiting 23:48:34 (11304): No heartbeat from core client for 30 sec - exiting 23:48:35 (11304): No heartbeat from core client for 30 sec - exiting 23:48:36 (11304): No heartbeat from core client for 30 sec - exiting CPDN Monitor - No 'heartbeat' from BOINC... Suspended CPDN Monitor - Suspend request from BOINC... CPDN Monitor - Quit request from BOINC... CPDN Monitor - Quit request from BOINC... CPDN Monitor - Quit request from BOINC... 23:12:40 (8292): No heartbeat from core client for 30 sec - exiting 23:12:41 (8292): No heartbeat from core client for 30 sec - exiting 23:12:42 (8292): No heartbeat from core client for 30 sec - exiting 23:12:43 (8292): No heartbeat from core client for 30 sec - exiting 23:12:44 (8292): No heartbeat from core client for 30 sec - exiting 23:12:45 (8292): No heartbeat from core client for 30 sec - exiting 23:12:46 (8292): No heartbeat from core client for 30 sec - exiting 23:12:47 (8292): No heartbeat from core client for 30 sec - exiting 23:12:48 (8292): No heartbeat from core client for 30 sec - exiting 23:12:49 (8292): No heartbeat from core client for 30 sec - exiting 23:12:50 (8292): No heartbeat from core client for 30 sec - exiting 23:12:51 (8292): No heartbeat from core client for 30 sec - exiting 23:12:52 (8292): No heartbeat from core client for 30 sec - exiting 23:12:53 (8292): No heartbeat from core client for 30 sec - exiting 23:12:54 (8292): No heartbeat from core client for 30 sec - exiting 23:12:55 (8292): No heartbeat from core client for 30 sec - exiting CPDN Monitor - No 'heartbeat' from BOINC... Signal 11 received, exiting... Called boinc_finish Controller:: CPDN process is not running, exiting, bRetVal = 1, checkPID=0, selfPID=4172, iMonCtr=2 Model crash detected, will try to restart... Global Worker:: CPDN process is not running, exiting, bRetVal = 1, checkPID=0, selfPID=9188, iMonCtr=2 Leaving CPDN_Main::Monitor... Called boinc_finish </stderr_txt> <message> upload failure: <file_xfer_error> <file_name>hadam3p_eu_e6ue_2013_1_008850657_1_7.zip</file_name> <error_code>-161 (not found)</error_code> </file_xfer_error> <file_xfer_error> <file_name>hadam3p_eu_e6ue_2013_1_008850657_1_8.zip</file_name> <error_code>-161 (not found)</error_code> </file_xfer_error> <file_xfer_error> <file_name>hadam3p_eu_e6ue_2013_1_008850657_1_9.zip</file_name> <error_code>-161 (not found)</error_code> </file_xfer_error> <file_xfer_error> <file_name>hadam3p_eu_e6ue_2013_1_008850657_1_10.zip</file_name> <error_code>-161 (not found)</error_code> </file_xfer_error> <file_xfer_error> <file_name>hadam3p_eu_e6ue_2013_1_008850657_1_11.zip</file_name> <error_code>-161 (not found)</error_code> </file_xfer_error> <file_xfer_error> <file_name>hadam3p_eu_e6ue_2013_1_008850657_1_12.zip</file_name> <error_code>-161 (not found)</error_code> </file_xfer_error> </message> ]]> |
Latest Trickles Received | ||||||
---|---|---|---|---|---|---|
Time Sent (UTC) | Host ID | Result ID | Result Name | Timestep | CPU Time (sec) | Average (sec/TS) |
14 Aug 2014 14:36:13 | 1084069 | 16839575 | hadam3p_eu_e6ue_2013_1_008850657_1 | 69,216 | 172,538 | 2.4927 |
14 Aug 2014 14:36:13 | 1084069 | 16839575 | hadam3p_eu_e6ue_2013_1_008850657_1 | 57,696 | 143,596 | 2.4888 |
14 Aug 2014 14:36:13 | 1084069 | 16839575 | hadam3p_eu_e6ue_2013_1_008850657_1 | 46,176 | 115,632 | 2.5042 |
14 Aug 2014 14:36:13 | 1084069 | 16839575 | hadam3p_eu_e6ue_2013_1_008850657_1 | 34,656 | 87,449 | 2.5233 |
08 Aug 2014 08:31:30 | 1084069 | 16839575 | hadam3p_eu_e6ue_2013_1_008850657_1 | 23,136 | 58,749 | 2.5393 |
08 Aug 2014 00:37:49 | 1084069 | 16839575 | hadam3p_eu_e6ue_2013_1_008850657_1 | 11,616 | 29,416 | 2.5324 |
©2024 cpdn.org