• Some users have recently had their accounts hijacked. It seems that the now defunct EVGA forums might have compromised your password there and seems many are using the same PW here. We would suggest you UPDATE YOUR PASSWORD and TURN ON 2FA for your account here to further secure it. None of the compromised accounts had 2FA turned on.
    Once you have enabled 2FA, your account will be updated soon to show a badge, letting other members know that you use 2FA to protect your account. This should be beneficial for everyone that uses FSFT.

EUE always on one unit...

Vaulter98c

[H]ard|DCer of the Month - October 2009
2FA
Joined
May 21, 2008
Messages
5,840
OK, so until tomorrow my tri box is down to 2, and they are both 96's. BOth have run fine before, but I had power issues all along so I couldnt run too many at the same time. Anyways, Ive got 2 in there, and the OC's were moderate but not heavy (only shader, nothing else). I keep having a problem now, my damn cards always EUE on 353's (or after getting 353's, not sure, but FAHmon always says 353 beside it when it crashes) These things can run stable for a day or two, but everytime they get one of those unit's, BAM, EUE at 100%. I keep lowering the clocks each time thinking thats it, but its not, because I'm almost stock now and each card has had one today alone. ANy ideas? Is it power related? If so, I'm changing out PSU's tomorrow, so I'll try that path, but all 3 cards would do this, and now all 2 cards do it (RIP buddy)

HELP! I have no clue and its bugging the hell out of me. Its Win 7-64 on a 750i board, P4 cpu, and 3 gigs ram on a 450w PSU. Everything is up to date and the install is less than 30 days old.

Here's the log
Code:
[00:25:45] Completed 98%
[00:27:04] Completed 99%
[00:28:23] Completed 100%
[00:28:23] Successful run
[00:28:23] DynamicWrapper: Finished Work Unit: sleep=10000
[00:28:33] Reserved 75944 bytes for xtc file; Cosm status=0
[00:28:33] Allocated 75944 bytes for xtc file
[00:28:33] - Reading up to 75944 from "work/wudata_07.xtc": Read 75944
[00:28:33] Read 75944 bytes from xtc file; available packet space=786354520
[00:28:33] xtc file hash check passed.
[00:28:33] Reserved 15168 15168 786354520 bytes for arc file=<work/wudata_07.trr> Cosm status=0
[00:28:33] Allocated 15168 bytes for arc file
[00:28:33] - Reading up to 15168 from "work/wudata_07.trr": Read 15168
[00:28:33] Read 15168 bytes from arc file; available packet space=786339352
[00:28:33] trr file hash check passed.
[00:28:33] Allocated 560 bytes for edr file
[00:28:33] Read bedfile
[00:28:33] edr file hash check passed.
[00:28:33] Allocated 26158 bytes for logfile
[00:28:33] Read logfile
[00:28:33] GuardedRun: success in DynamicWrapper
[00:28:33] GuardedRun: done
[00:28:33] Run: GuardedRun completed.
[00:28:35] - Writing 118342 bytes of core data to disk...
[00:28:35] Done: 117830 -> 97816 (compressed to 83.0 percent)
[00:28:35]   ... Done.
[00:28:35] - Shutting down core 
[00:28:35] 
[00:28:35] Folding@home Core Shutdown: FINISHED_UNIT
[00:28:39] CoreStatus = 64 (100)
[00:28:39] Sending work to server
[00:28:39] Project: 5769 (Run 4, Clone 102, Gen 403)
[00:28:39] - Read packet limit of 540015616... Set to 524286976.


[00:28:39] + Attempting to send results [May 14 00:28:39 UTC]
[00:28:42] + Results successfully sent
[00:28:42] Thank you for your contribution to Folding@Home.
[00:28:42] + Number of Units Completed: 41

[00:28:46] - Preparing to get new work unit...
[00:28:46] + Attempting to get work packet
[00:28:46] - Connecting to assignment server
[00:28:47] - Successful: assigned to (171.67.108.11).
[00:28:47] + News From Folding@Home: GPU folding beta
[00:28:47] Loaded queue successfully.
[00:28:49] + Closed connections
[00:28:49] 
[00:28:49] + Processing work unit
[00:28:49] Core required: FahCore_11.exe
[00:28:49] Core found.
[00:28:49] Working on queue slot 08 [May 14 00:28:49 UTC]
[00:28:49] + Working ...
[00:28:49] 
[00:28:49] *------------------------------*
[00:28:49] Folding@Home GPU Core - Beta
[00:28:49] Version 1.19 (Mon Nov 3 09:34:13 PST 2008)
[00:28:49] 
[00:28:49] Compiler  : Microsoft (R) 32-bit C/C++ Optimizing Compiler Version 14.00.50727.762 for 80x86 
[00:28:49] Build host: amoeba
[00:28:49] Board Type: Nvidia
[00:28:49] Core      : 
[00:28:49] Preparing to commence simulation
[00:28:49] - Looking at optimizations...
[00:28:49] - Created dyn
[00:28:49] - Files status OK
[00:28:49] - Expanded 46667 -> 252912 (decompressed 541.9 percent)
[00:28:49] Called DecompressByteArray: compressed_data_size=46667 data_size=252912, decompressed_data_size=252912 diff=0
[00:28:49] - Digital signature verified
[00:28:49] 
[00:28:49] Project: 5768 (Run 10, Clone 41, Gen 530)
[00:28:49] 
[00:28:49] Assembly optimizations on if available.
[00:28:49] Entering M.D.
[00:28:55] Working on Protein
[00:28:57] Client config found, loading data.
[00:28:57] mdrun_gpu returned 
[00:28:57] NANs detected on GPU
[00:28:57] 
[00:28:57] Folding@home Core Shutdown: UNSTABLE_MACHINE
[00:28:59] CoreStatus = 7A (122)
[00:28:59] Sending work to server
[00:28:59] Project: 5768 (Run 10, Clone 41, Gen 530)
[00:28:59] - Read packet limit of 540015616... Set to 524286976.
[00:28:59] - Error: Could not get length of results file work/wuresults_08.dat
[00:28:59] - Error: Could not read unit 08 file. Removing from queue.
[00:28:59] - Preparing to get new work unit...
[00:28:59] + Attempting to get work packet
[00:28:59] - Connecting to assignment server
[00:28:59] - Successful: assigned to (171.67.108.11).
[00:28:59] + News From Folding@Home: GPU folding beta
[00:29:00] Loaded queue successfully.
[00:29:01] + Closed connections
[00:29:06] 
[00:29:06] + Processing work unit
[00:29:06] Core required: FahCore_11.exe
[00:29:06] Core found.
[00:29:06] Working on queue slot 09 [May 14 00:29:06 UTC]
[00:29:06] + Working ...
[00:29:06] 
[00:29:06] *------------------------------*
[00:29:06] Folding@Home GPU Core - Beta
[00:29:06] Version 1.19 (Mon Nov 3 09:34:13 PST 2008)
[00:29:06] 
[00:29:06] Compiler  : Microsoft (R) 32-bit C/C++ Optimizing Compiler Version 14.00.50727.762 for 80x86 
[00:29:06] Build host: amoeba
[00:29:06] Board Type: Nvidia
[00:29:06] Core      : 
[00:29:06] Preparing to commence simulation
[00:29:06] - Looking at optimizations...
[00:29:06] - Created dyn
[00:29:06] - Files status OK
[00:29:06] - Expanded 46667 -> 252912 (decompressed 541.9 percent)
[00:29:06] Called DecompressByteArray: compressed_data_size=46667 data_size=252912, decompressed_data_size=252912 diff=0
[00:29:06] - Digital signature verified
[00:29:06] 
[00:29:06] Project: 5768 (Run 10, Clone 41, Gen 530)
[00:29:06] 
[00:29:06] Assembly optimizations on if available.
[00:29:06] Entering M.D.
[00:29:12] Working on Protein
[00:29:14] Client config found, loading data.
[00:29:14] mdrun_gpu returned 
[00:29:14] NANs detected on GPU
[00:29:14] 
[00:29:14] Folding@home Core Shutdown: UNSTABLE_MACHINE
[00:29:16] CoreStatus = 7A (122)
[00:29:16] Sending work to server
[00:29:16] Project: 5768 (Run 10, Clone 41, Gen 530)
[00:29:16] - Read packet limit of 540015616... Set to 524286976.
[00:29:16] - Error: Could not get length of results file work/wuresults_09.dat
[00:29:16] - Error: Could not read unit 09 file. Removing from queue.
[00:29:16] - Preparing to get new work unit...
[00:29:16] + Attempting to get work packet
[00:29:16] - Connecting to assignment server
[00:29:16] - Successful: assigned to (171.67.108.11).
[00:29:16] + News From Folding@Home: GPU folding beta
[00:29:17] Loaded queue successfully.
[00:29:18] + Closed connections
[00:29:23] 
[00:29:23] + Processing work unit
[00:29:23] Core required: FahCore_11.exe
[00:29:23] Core found.
[00:29:23] Working on queue slot 00 [May 14 00:29:23 UTC]
[00:29:23] + Working ...
[00:29:23] 
[00:29:23] *------------------------------*
[00:29:23] Folding@Home GPU Core - Beta
[00:29:23] Version 1.19 (Mon Nov 3 09:34:13 PST 2008)
[00:29:23] 
[00:29:23] Compiler  : Microsoft (R) 32-bit C/C++ Optimizing Compiler Version 14.00.50727.762 for 80x86 
[00:29:23] Build host: amoeba
[00:29:23] Board Type: Nvidia
[00:29:23] Core      : 
[00:29:23] Preparing to commence simulation
[00:29:23] - Looking at optimizations...
[00:29:23] - Created dyn
[00:29:23] - Files status OK
[00:29:23] - Expanded 46667 -> 252912 (decompressed 541.9 percent)
[00:29:23] Called DecompressByteArray: compressed_data_size=46667 data_size=252912, decompressed_data_size=252912 diff=0
[00:29:23] - Digital signature verified
[00:29:23] 
[00:29:23] Project: 5768 (Run 10, Clone 41, Gen 530)
[00:29:23] 
[00:29:23] Assembly optimizations on if available.
[00:29:23] Entering M.D.
[00:29:29] Working on Protein
[00:29:31] Client config found, loading data.
[00:29:31] Starting GUI Server
[00:29:31] mdrun_gpu returned 
[00:29:31] NANs detected on GPU
[00:29:31] 
[00:29:31] Folding@home Core Shutdown: UNSTABLE_MACHINE
[00:29:33] CoreStatus = 7A (122)
[00:29:33] Sending work to server
[00:29:33] Project: 5768 (Run 10, Clone 41, Gen 530)
[00:29:33] - Read packet limit of 540015616... Set to 524286976.
[00:29:33] - Error: Could not get length of results file work/wuresults_00.dat
[00:29:33] - Error: Could not read unit 00 file. Removing from queue.
[00:29:33] - Preparing to get new work unit...
[00:29:33] + Attempting to get work packet
[00:29:33] - Connecting to assignment server
[00:29:33] - Successful: assigned to (171.67.108.11).
[00:29:33] + News From Folding@Home: GPU folding beta
[00:29:33] Loaded queue successfully.
[00:29:34] + Closed connections
[00:29:39] 
[00:29:39] + Processing work unit
[00:29:39] Core required: FahCore_11.exe
[00:29:39] Core found.
[00:29:39] Working on queue slot 01 [May 14 00:29:39 UTC]
[00:29:39] + Working ...
[00:29:40] 
[00:29:40] *------------------------------*
[00:29:40] Folding@Home GPU Core - Beta
[00:29:40] Version 1.19 (Mon Nov 3 09:34:13 PST 2008)
[00:29:40] 
[00:29:40] Compiler  : Microsoft (R) 32-bit C/C++ Optimizing Compiler Version 14.00.50727.762 for 80x86 
[00:29:40] Build host: amoeba
[00:29:40] Board Type: Nvidia
[00:29:40] Core      : 
[00:29:40] Preparing to commence simulation
[00:29:40] - Looking at optimizations...
[00:29:40] - Created dyn
[00:29:40] - Files status OK
[00:29:40] - Expanded 46667 -> 252912 (decompressed 541.9 percent)
[00:29:40] Called DecompressByteArray: compressed_data_size=46667 data_size=252912, decompressed_data_size=252912 diff=0
[00:29:40] - Digital signature verified
[00:29:40] 
[00:29:40] Project: 5768 (Run 10, Clone 41, Gen 530)
[00:29:40] 
[00:29:40] Assembly optimizations on if available.
[00:29:40] Entering M.D.
[00:29:46] Working on Protein
[00:29:48] Client config found, loading data.
[00:29:48] mdrun_gpu returned 
[00:29:48] NANs detected on GPU
[00:29:48] 
[00:29:48] Folding@home Core Shutdown: UNSTABLE_MACHINE
[00:29:50] CoreStatus = 7A (122)
[00:29:50] Sending work to server
[00:29:50] Project: 5768 (Run 10, Clone 41, Gen 530)
[00:29:50] - Read packet limit of 540015616... Set to 524286976.
[00:29:50] - Error: Could not get length of results file work/wuresults_01.dat
[00:29:50] - Error: Could not read unit 01 file. Removing from queue.
[00:29:50] - Preparing to get new work unit...
[00:29:50] + Attempting to get work packet
[00:29:50] - Connecting to assignment server
[00:29:50] - Successful: assigned to (171.67.108.11).
[00:29:50] + News From Folding@Home: GPU folding beta
[00:29:50] Loaded queue successfully.
[00:29:52] + Closed connections
[00:29:57] 
[00:29:57] + Processing work unit
[00:29:57] Core required: FahCore_11.exe
[00:29:57] Core found.
[00:29:57] Working on queue slot 02 [May 14 00:29:57 UTC]
[00:29:57] + Working ...
[00:29:57] 
[00:29:57] *------------------------------*
[00:29:57] Folding@Home GPU Core - Beta
[00:29:57] Version 1.19 (Mon Nov 3 09:34:13 PST 2008)
[00:29:57] 
[00:29:57] Compiler  : Microsoft (R) 32-bit C/C++ Optimizing Compiler Version 14.00.50727.762 for 80x86 
[00:29:57] Build host: amoeba
[00:29:57] Board Type: Nvidia
[00:29:57] Core      : 
[00:29:57] Preparing to commence simulation
[00:29:57] - Looking at optimizations...
[00:29:57] - Created dyn
[00:29:57] - Files status OK
[00:29:57] - Expanded 46667 -> 252912 (decompressed 541.9 percent)
[00:29:57] Called DecompressByteArray: compressed_data_size=46667 data_size=252912, decompressed_data_size=252912 diff=0
[00:29:57] - Digital signature verified
[00:29:57] 
[00:29:57] Project: 5768 (Run 10, Clone 41, Gen 530)
[00:29:57] 
[00:29:57] Assembly optimizations on if available.
[00:29:57] Entering M.D.
[00:30:03] Working on Protein
[00:30:05] Client config found, loading data.
[00:30:05] mdrun_gpu returned 
[00:30:05] NANs detected on GPU
[00:30:05] 
[00:30:05] Folding@home Core Shutdown: UNSTABLE_MACHINE
[00:30:07] CoreStatus = 7A (122)
[00:30:07] Sending work to server
[00:30:07] Project: 5768 (Run 10, Clone 41, Gen 530)
[00:30:07] - Read packet limit of 540015616... Set to 524286976.
[00:30:07] - Error: Could not get length of results file work/wuresults_02.dat
[00:30:07] - Error: Could not read unit 02 file. Removing from queue.
[00:30:07] EUE limit exceeded. Pausing 24 hours.
 
its not your system.. its a bad string of 5768 WU's.. if it was your system it would do 1 or 2 percent then fail.. not fail instantly.. reminds me of the problems the gpu client had when the newer 5XXXX WU's came out.. a bunch of them were failing and was fixed a little later.. might be something you want to post in the F@H forums.. but i bet they will say the same thing..
 
It looks like it tried to do that same WU several times. It happens sometimes and caused me to EUE left and right eventually causing me to exceed EUE limits.

Nothing I did made it better. Manual intervention is usually required. I had to do the following (doing these from memory so bear with me):
Shut down GPU2 client
Delete work folder
Delete queue.dat
Delete machineindependent.dat
Delete unitinfo.txt
Delete fahcore11.exe
Restart GPU2 client

Most of the times it will pick up a new WU and continue chugging along. There are occassions when it will STILL pickup the bad WU. Part of it may be timing. Give it an hour and then repeat the above steps.

 
It looks like it tried to do that same WU several times. It happens sometimes and caused me to EUE left and right eventually causing me to exceed EUE limits.

Nothing I did made it better. Manual intervention is usually required. I had to do the following (doing these from memory so bear with me):
Shut down GPU2 client
Delete work folder
Delete queue.dat
Delete machineindependent.dat
Delete unitinfo.txt
Delete fahcore11.exe
Restart GPU2 client
Actually, the only thing you really need to delete is the queue.dat file. The client reads its work assignment from that file alone, so if it's missing it will automatically request and download a new workunit, even if the work folder and other files are still present.
 
Interesting. I'll keep that in mind. I normally go thru the above routine to completely flush things and hopefully get a different WU.
 
Interesting. I'll keep that in mind. I normally go thru the above routine to completely flush things and hopefully get a different WU.
I used to do that as well, but then I figured out that you only need to delete the one file. Saves a bit of time and effort, especially when you're doing it quite a few times ;).
 
Back
Top