• Some users have recently had their accounts hijacked. It seems that the now defunct EVGA forums might have compromised your password there and seems many are using the same PW here. We would suggest you UPDATE YOUR PASSWORD and TURN ON 2FA for your account here to further secure it. None of the compromised accounts had 2FA turned on.
    Once you have enabled 2FA, your account will be updated soon to show a badge, letting other members know that you use 2FA to protect your account. This should be beneficial for everyone that uses FSFT.

cannot upload the SMP work

Pocatello

DC Moderator and [H]ard DCOTM x7
Staff member
2FA
Joined
Jun 15, 2005
Messages
6,875
see the work log below; I cannot get the work to upload. I have tried the -send all flag. Sometimes it finishes (but nothing is uploaded) and sometimes it just sits there.

I am doing SMP WU's but they are not being accounted for.

I am connected to the internet. My other programs seem to be uploading okay right now. All my GPU work seems to be credited. I have tried Qfix, send all, qfix again, send all, but the units will not upload.


Code:
--- Opening Log file [November 12 04:02:44 UTC] 


# Windows SMP Console Edition #################################################
###############################################################################

                       Folding@Home Client Version 6.22 SMP Beta2

                          http://folding.stanford.edu

###############################################################################
###############################################################################

Launch directory: C:\Folding\smp
Executable: C:\Folding\smp\fah.exe
Arguments: -send all -smp 

[04:02:44] - Ask before connecting: No
[04:02:44] - User name: pocatello (Team 33)
[04:02:44] - User ID: /*-/*-/*-*-
[04:02:44] - Machine ID: 1
[04:02:44] 
[04:02:44] Loaded queue successfully.
[04:02:44] Attempting to return result(s) to server...
[04:02:44] Project: 2665 (Run 3, Clone 390, Gen 65)


[04:02:44] + Attempting to send results [November 12 04:02:44 UTC]
[04:50:16] - Unknown packet returned from server, expected ACK for results
[04:50:16] - Error: Could not transmit unit 04 (completed November 10) to work server.


[04:50:16] + Attempting to send results [November 12 04:50:16 UTC]
[05:17:36] - Unknown packet returned from server, expected ACK for results
[05:17:36]   Could not transmit unit 04 to Collection server; keeping in queue.
[05:17:36] - Failed to send all units to server

Folding@Home Client Shutdown.


--- Opening Log file [November 12 05:40:34 UTC] 


# Windows SMP Console Edition #################################################
###############################################################################

                       Folding@Home Client Version 6.22 SMP Beta2

                          http://folding.stanford.edu

###############################################################################
###############################################################################

Launch directory: C:\Folding\smp
Executable: C:\Folding\smp\fah.exe
Arguments: -send all -smp 

[05:40:34] - Ask before connecting: No
[05:40:34] - User name: pocatello (Team 33)
[05:40:34] - User ID: 3/-/*-/*-/*-
[05:40:34] - Machine ID: 1
[05:40:34] 
[05:40:34] Loaded queue successfully.
[05:40:34] Attempting to return result(s) to server...
[05:40:34] Project: 2665 (Run 3, Clone 390, Gen 65)


[05:40:34] + Attempting to send results [November 12 05:40:34 UTC]

Folding@Home Client Shutdown at user request.

Folding@Home Client Shutdown.


--- Opening Log file [November 12 06:01:27 UTC] 


# Windows SMP Console Edition #################################################
###############################################################################

                       Folding@Home Client Version 6.22 SMP Beta2

                          http://folding.stanford.edu

###############################################################################
###############################################################################

Launch directory: C:\Folding\smp
Executable: C:\Folding\smp\fah.exe
Arguments: -verbosity 9 -smp 

[06:01:27] - Ask before connecting: No
[06:01:27] - User name: pocatello (Team 33)
[06:01:27] - User ID: /*-/*-/*-/*-
[06:01:27] - Machine ID: 1
[06:01:27] 
[06:01:27] Loaded queue successfully.
[06:01:27] 
[06:01:27] + Processing work unit
[06:01:27] Work type a1 not eligible for variable processors
[06:01:27] Core required: FahCore_a1.exe
[06:01:27] Core found.[06:01:27] - Autosending finished units... [November 12 06:01:27 UTC]

[06:01:27] Trying to send all finished work units
[06:01:27] Using generic mpiexec calls
[06:01:27] Project: 2665 (Run 3, Clone 390, Gen 65)


[06:01:27] + Attempting to send results [November 12 06:01:27 UTC]
[06:01:27] - Reading file work/wuresults_04.dat from core
[06:01:27] Working on queue slot 05 [November 12 06:01:27 UTC]
[06:01:27]   (Read 22081724 bytes from disk)
[06:01:27] + Working ...
[06:01:27] Connecting to http://171.64.65.64:8080/
[06:01:27] - Calling 'mpiexec -np 4 -channel auto -host 127.0.0.1 FahCore_a1.exe -dir work/ -suffix 05 -cpu 88 -checkpoint 15 -verbose -lifeline 920 -version 622'

[06:01:28] 
[06:01:28] *------------------------------*
[06:01:28] Folding@Home Gromacs SMP Core
[06:01:28] Version 1.74 (March 10, 2007)
[06:01:28] 
[06:01:28] Preparing to commence simulation
[06:01:28] - Ensuring status. Please wait.
[06:01:45] - Looking at optimizations...
[06:01:45] - Working with standard loops on this execution.
[06:01:45] - Previous termination of core was improper.
[06:01:45] - Going to use standard loops.
[06:01:45] - Files status OK
[06:02:02] - Expanded 4724314 -> 24426905 (decompressed 517.0 percent)
[06:02:03] 
[06:02:03] Project: 2665 (Run 0, Clone 241, Gen 66)
[06:02:03] 
[06:02:04] Entering M.D.
[06:02:17] Calling FAH init
[06:02:20] Read topology
[06:02:20] ocal files
[06:02:20] rom checkpoint)
[06:02:20] Read checkpoint
[06:02:20] Protein: HGG in water
[06:02:21] Writing local files
[06:02:21] Completed 181697 out of 250000 steps  (72 percent)
[06:02:36] Extra SSE boost OK.
[06:10:12] Writing local files
[06:10:13] Completed 182500 out of 250000 steps  (73 percent)
[06:21:28] Posted data.
[06:25:13] Timered checkpoint triggered.
[06:33:07] Writing local files
[06:33:08] Completed 185000 out of 250000 steps  (74 percent)
[06:41:27] Initial: 0000; - Uploaded at ~8 kB/s
[06:44:36] - Averaged speed for that direction ~13 kB/s
[06:44:36] - Unknown packet returned from server, expected ACK for results
[06:44:36] - Error: Could not transmit unit 04 (completed November 10) to work server.
[06:44:36] - 9 failed uploads of this unit.


[06:44:36] + Attempting to send results [November 12 06:44:36 UTC]
[06:44:36] - Reading file work/wuresults_04.dat from core
[06:44:36]   (Read 22081724 bytes from disk)
[06:44:36] Connecting to http://171.67.108.25:8080/
[06:44:36] - Couldn't send HTTP request to server
[06:44:36]   (Got status 503)
[06:44:36] + Could not connect to Work Server (results)
[06:44:36]     (171.67.108.25:8080)
[06:44:36] + Retrying using alternative port
[06:44:36] Connecting to http://171.67.108.25:80/
[06:44:36] - Couldn't send HTTP request to server
[06:44:36]   (Got status 503)
[06:44:36] + Could not connect to Work Server (results)
[06:44:36]     (171.67.108.25:80)
[06:44:36]   Could not transmit unit 04 to Collection server; keeping in queue.
[06:44:36] + Sent 0 of 1 completed units to the server
[06:44:36] - Autosend completed
[06:48:08] Timered checkpoint triggered.
[06:56:32] Writing local files
[06:56:33] Completed 187500 out of 250000 steps  (75 percent)
[07:11:33] Timered checkpoint triggered.
[07:19:29] Writing local files
[07:19:30] Completed 190000 out of 250000 steps  (76 percent)
[07:34:31] Timered checkpoint triggered.
[07:43:17] Writing local files
[07:43:18] Completed 192500 out of 250000 steps  (77 percent)
[07:58:18] Timered checkpoint triggered.
[08:07:23] Writing local files
[08:07:24] Completed 195000 out of 250000 steps  (78 percent)
[08:22:24] Timered checkpoint triggered.
[08:30:59] Writing local files
[08:31:01] Completed 197500 out of 250000 steps  (79 percent)
[08:46:02] Timered checkpoint triggered.
[08:55:23] Writing local files
[08:55:24] Completed 200000 out of 250000 steps  (80 percent)
[09:10:25] Timered checkpoint triggered.
[09:19:11] Writing local files
[09:19:12] Completed 202500 out of 250000 steps  (81 percent)
[09:34:14] Timered checkpoint triggered.
[09:43:06] Writing local files
[09:43:07] Completed 205000 out of 250000 steps  (82 percent)
[09:58:09] Timered checkpoint triggered.
[10:06:46] Writing local files
[10:06:47] Completed 207500 out of 250000 steps  (83 percent)
[10:21:49] Timered checkpoint triggered.
[10:32:16] Writing local files
[10:32:17] Completed 210000 out of 250000 steps  (84 percent)
[10:47:19] Timered checkpoint triggered.
[10:55:38] Writing local files
[10:55:38] Completed 212500 out of 250000 steps  (85 percent)
[11:10:39] Timered checkpoint triggered.
[11:18:53] Writing local files
[11:18:55] Completed 215000 out of 250000 steps  (86 percent)
[11:33:55] Timered checkpoint triggered.
[11:42:24] Writing local files
[11:42:24] Completed 217500 out of 250000 steps  (87 percent)
[11:57:25] Timered checkpoint triggered.
[12:05:55] Writing local files
[12:05:55] Completed 220000 out of 250000 steps  (88 percent)
[12:20:56] Timered checkpoint triggered.
[12:29:53] Writing local files
[12:29:53] Completed 222500 out of 250000 steps  (89 percent)
[12:44:35] - Autosending finished units... [November 12 12:44:35 UTC]
[12:44:35] Trying to send all finished work units
[12:44:35] Project: 2665 (Run 3, Clone 390, Gen 65)


[12:44:35] + Attempting to send results [November 12 12:44:35 UTC]
[12:44:35] - Reading file work/wuresults_04.dat from core
[12:44:35]   (Read 22081724 bytes from disk)
[12:44:35] Connecting to http://171.64.65.64:8080/
[12:44:55] Timered checkpoint triggered.
[12:53:33] Writing local files
[12:53:34] Completed 225000 out of 250000 steps  (90 percent)
[13:04:38] Posted data.
[13:08:34] Timered checkpoint triggered.
[13:17:09] Writing local files
[13:17:09] Completed 227500 out of 250000 steps  (91 percent)
[13:24:38] Initial: 0000; Timered checkpoint triggered.
[13:40:50] Writing local files
[13:40:51] Completed 230000 out of 250000 steps  (92 percent)
[13:44:37] + Could not connect to Work Server (results)
[13:44:37]     (171.64.65.64:8080)
[13:44:37] + Retrying using alternative port
[13:44:37] Connecting to http://171.64.65.64:80/
[13:55:52] Timered checkpoint triggered.
[14:04:38] Writing local files
[14:04:38] Completed 232500 out of 250000 steps  (93 percent)
[14:04:39] Posted data.
[14:19:39] Timered checkpoint triggered.
[14:24:39] Initial: 001A; Writing local files
[14:28:29] Completed 235000 out of 250000 steps  (94 percent)
[14:43:30] Timered checkpoint triggered.
[14:44:39] + Could not connect to Work Server (results)
[14:44:39]     (171.64.65.64:80)
[14:44:39] - Error: Could not transmit unit 04 (completed November 10) to work server.
[14:44:39] - 10 failed uploads of this unit.


[14:44:39] + Attempting to send results [November 12 14:44:39 UTC]
[14:44:39] - Reading file work/wuresults_04.dat from core
[14:44:39]   (Read 22081724 bytes from disk)
[14:44:39] Connecting to http://171.67.108.25:8080/
[14:44:41] - Couldn't send HTTP request to server
[14:44:41]   (Got status 503)
[14:44:41] + Could not connect to Work Server (results)
[14:44:41]     (171.67.108.25:8080)
[14:44:41] + Retrying using alternative port
[14:44:41] Connecting to http://171.67.108.25:80/
[14:44:44] - Couldn't send HTTP request to server
[14:44:44]   (Got status 503)
[14:44:44] + Could not connect to Work Server (results)
[14:44:44]     (171.67.108.25:80)
[14:44:44]   Could not transmit unit 04 to Collection server; keeping in queue.
[14:44:44] + Sent 0 of 1 completed units to the server
[14:44:44] - Autosend completed
[14:52:36] Writing local files
[14:52:37] Completed 237500 out of 250000 steps  (95 percent)
[15:07:38] Timered checkpoint triggered.
[15:17:01] Writing local files
[15:17:02] Completed 240000 out of 250000 steps  (96 percent)
[15:32:03] Timered checkpoint triggered.
[15:41:06] Writing local files
[15:41:07] Completed 242500 out of 250000 steps  (97 percent)
[15:56:08] Timered checkpoint triggered.
[16:05:12] Writing local files
[16:05:13] Completed 245000 out of 250000 steps  (98 percent)
[16:20:15] Timered checkpoint triggered.
[16:29:38] Writing local files
[16:29:38] Completed 247500 out of 250000 steps  (99 percent)
[16:44:39] Timered checkpoint triggered.
[16:54:16] Writing local files
[16:54:17] Completed 250000 out of 250000 steps  (100 percent)
[16:54:17] Writing final coordinates.
[16:54:21] Past main M.D. loop
[16:54:21] Will end MPI now
[16:55:21] 
[16:55:21] Finished Work Unit:
[16:55:21] - Reading up to 21310704 from "work/wudata_05.arc": Read 21310704
[16:55:21] - Reading up to 558012 from "work/wudata_05.xtc": Read 558012
[16:55:21] goefile size: 0
[16:55:21] logfile size: 212416
[16:55:21] Leaving Run
[16:55:22] - Writing 22087504 bytes of core data to disk...
[16:55:22]   ... Done.
[16:55:23] - Failed to delete work/wudata_05.sas
[16:55:23] - Failed to delete work/wudata_05.goe
[16:55:23] Warning:  check for stray files
[16:55:23] - Shutting down core
[16:57:23] 
[16:57:23] Folding@home Core Shutdown: FINISHED_UNIT
[16:57:23] 
[16:57:23] Folding@home Core Shutdown: FINISHED_UNIT
[16:57:28] CoreStatus = 64 (100)
[16:57:28] Unit 5 finished with 71 percent of time to deadline remaining.
[16:57:28] Updated performance fraction: 0.707505
[16:57:28] Sending work to server
[16:57:28] Project: 2665 (Run 0, Clone 241, Gen 66)


[16:57:28] + Attempting to send results [November 12 16:57:28 UTC]
[16:57:28] - Reading file work/wuresults_05.dat from core
[16:57:28]   (Read 22087504 bytes from disk)
[16:57:28] Connecting to http://171.64.65.64:8080/
[17:17:32] Posted data.
[17:37:32] Initial: 0000; + Could not connect to Work Server (results)
[17:57:32]     (171.64.65.64:8080)
[17:57:32] + Retrying using alternative port
[17:57:32] Connecting to http://171.64.65.64:80/
[18:17:34] Posted data.
[18:37:33] Initial: 001A; + Could not connect to Work Server (results)
[18:57:33]     (171.64.65.64:80)
[18:57:33] - Error: Could not transmit unit 05 (completed November 12) to work server.
[18:57:33] - 1 failed uploads of this unit.
[18:57:33]   Keeping unit 05 in queue.
[18:57:33] Trying to send all finished work units
[18:57:33] Project: 2665 (Run 3, Clone 390, Gen 65)


[18:57:33] + Attempting to send results [November 12 18:57:33 UTC]
[18:57:33] - Reading file work/wuresults_04.dat from core
[18:57:34]   (Read 22081724 bytes from disk)
[18:57:34] Connecting to http://171.64.65.64:8080/
[19:17:37] Posted data.
[19:37:37] Initial: 0000; + Could not connect to Work Server (results)
[19:57:37]     (171.64.65.64:8080)
[19:57:37] + Retrying using alternative port
[19:57:37] Connecting to http://171.64.65.64:80/
[20:17:39] Posted data.
[20:37:39] Initial: 001A; - Autosending finished units... [November 12 20:44:42 UTC]
[20:44:42] Trying to send all finished work units
[20:44:42] - Already sending work
[20:44:42] - Already sending work
[20:44:42] + Sent 0 of 2 completed units to the server
[20:44:42] - Autosend completed
[20:57:39] + Could not connect to Work Server (results)
[20:57:39]     (171.64.65.64:80)
[20:57:39] - Error: Could not transmit unit 04 (completed November 10) to work server.
[20:57:39] - 11 failed uploads of this unit.

 
I uninstalled FAH6 and reinstalled attempting to upload my finished SMP wu's after seeing the same error you're seeing.
Current status: Uninstalled until I hear it's not a waste of my resources again.
 
folding the SMP is really a waste when I cannot return the work units.
 
Pocatello, you are using the SMP client version which doesn't include the ACK fix and you have issues with this precisely. Get the latest from http://foldingforum.org/viewtopic.php?f=46&t=6143 and replace the executable then retry.


Thanks

I did not know anything about it. :cool:

I will give it a try. Another try. :)

Damn! Too much time and effort on my part! "Spare CPU cycles" my ass! What about my spare time being wasted for the Panda group?
 
Do you want a cake for your spare time wasted ? :D

 
okay

seems to have worked for me

It uploaded 3 WU's .... but I got zero credit for them. :mad: :confused:
 
This just might be the key:

I snagged that from the log file posted.

During client setup NEVER use “internet explorer settings”




[21:27:04] + Attempting to send results [November 10 21:27:04 UTC]
[21:28:00] + Results successfully sent
[21:28:00] Thank you for your contribution to Folding@Home.
[21:28:00] + Number of Units Completed: 11

[21:28:01] Using generic mpiexec calls
[21:30:22] - Preparing to get new work unit...
[21:30:22] + Attempting to get work packet
[21:30:22] - Connecting to assignment server
[21:30:22] - Successful: assigned to (171.64.65.64).
[21:30:22] + News From Folding@Home: Welcome to Folding@Home

note the lack of the "http//" prefix. IE settings should be turned off.;)

 
Ok, so as of last night my SMPs will no longer send either so it’s not all in your heads;)

Server status shows a number of servers down and this is one of those little things that as far as Stanford goes just pisses me off.

Nobody from Stanford can be bothered to check at least once a weekend to see if there is a problem?

Now, if this is some way related to the fires, ok, that I could understand but logic would suggest that since GPU units are sending just fine…..and SMP is sending just fine….ohoh, I used the “L” word….:rolleyes:

 
BillR

I just went through the config only setup for the SMP and I did not notice anywhere that I could disable the IE connection settings like you mentioned. I remember this from the older clients, but I did not see it. Does the newer client that does both regular and SMP, and that requires the -SMP flag, still have IE settings?
 
Pocatello, you are correct. The IE setting option is gone in the v6 clients.

 
There are some SMP servers issues yesterday and this might explain this.

 
Ive been having the same problem but instead of just not being able to send the WU's back, it encounters a fatal error and shuts down. Ive lost at least 8 projects the last 3 days and I'm really getting tired of all the BS this thing has caused!!!!
 
Back
Top