|
1)
Message boards :
Graphics cards (GPUs) :
LLM GPU utilization
(Message 62458)
Posted 22 Jun 2025 by FritzB Post: Since June 19th there is no load with LLM on the GPU (4090) using Linux. Just 1-2 minutes CPU load at the beginning and then running without progress. After aboarding it, the resent to a Windows PC seems to work fine. https://gpugrid.net/gpugrid/workunit.php?wuid=31499403 https://gpugrid.net/gpugrid/workunit.php?wuid=31499413 I've aboarded 8 other WUs after seeing no load after some minutes. |
|
2)
Message boards :
Number crunching :
LLM
(Message 62369)
Posted 11 Apr 2025 by FritzB Post: I have no information abou LLM but received some of them. https://gpugrid.net/gpugrid/results.php?userid=143331&offset=0&show_names=0&state=0&appid=49 |
|
3)
Message boards :
Graphics cards (GPUs) :
erreur
(Message 62355)
Posted 5 Apr 2025 by FritzB Post: Seems to work fine again. At the moment I am getting new work which is running 1-2 hours. Nope, there have just been a few good WUs.Back to errors. |
|
4)
Message boards :
Graphics cards (GPUs) :
erreur
(Message 62354)
Posted 5 Apr 2025 by FritzB Post: Seems to work fine again. At the moment I am getting new work which is running 1-2 hours. The broken ones errored out <2 min. But about 15% of those even finished successfully in the same time. |
|
5)
Message boards :
Graphics cards (GPUs) :
erreur
(Message 62348)
Posted 4 Apr 2025 by FritzB Post: I have about ~100 errors. 2 PCs, 4070 Ti and 4090. It began around early afternoon today. |
|
6)
Message boards :
Graphics cards (GPUs) :
gpu dédié a une tache
(Message 61759)
Posted 3 Sep 2024 by FritzB Post: You'll find some by asking Google for "boinc multiple instance linux" |
|
7)
Message boards :
Graphics cards (GPUs) :
AMD GPU
(Message 60640)
Posted 6 Aug 2023 by FritzB Post: Is it possible to translate the CUDA code to AMD by using this: https://github.com/ROCm-Developer-Tools/HIPIFY/blob/master/README.md |
|
8)
Message boards :
Graphics cards (GPUs) :
rtx 40xx
(Message 60337)
Posted 17 Apr 2023 by FritzB Post: First one finishing successfully on RTX 4090: http://www.gpugrid.net/workunit.php?wuid=27474364 |
|
9)
Message boards :
Graphics cards (GPUs) :
rtx 40xx
(Message 60280)
Posted 5 Apr 2023 by FritzB Post: ATMbeta still crashing on 4070Ti |
|
10)
Message boards :
Graphics cards (GPUs) :
rtx 40xx
(Message 60110)
Posted 17 Mar 2023 by FritzB Post: They get but crash after a few seconds https://www.gpugrid.net/show_host_detail.php?hostid=605125 |
|
11)
Message boards :
Number crunching :
ATM: Free Energy Calculations new application
(Message 59988)
Posted 26 Feb 2023 by FritzB Post: I just aborted the upload (not the workunit) and then it was reported as valid. https://www.gpugrid.net/results.php?hostid=604029 |
|
12)
Message boards :
Number crunching :
ATM: Free Energy Calculations new application
(Message 59967)
Posted 23 Feb 2023 by FritzB Post: I've also finished one: https://www.gpugrid.net/workunit.php?wuid=27410166 We're both using Linux Mint. It seems to crash on Win 10 machines (computer #600532 is mine, too). |
|
13)
Message boards :
Number crunching :
ATM: Free Energy Calculations new application
(Message 59899)
Posted 10 Feb 2023 by FritzB Post: This one https://www.gpugrid.net/workunit.php?wuid=27399736 is runnig for about 11 hours and it is stuck at 66,666% for at least 4 hours now. There is almost no load on the GPU. Just a few percent (3-5) once in a while, but constantly some load on the memory controller (10-30). Hope it will finish some day :) |
|
14)
Message boards :
News :
Experimental Python tasks (beta) - task description
(Message 58995)
Posted 10 Jul 2022 by FritzB Post: Sorry for OT, but some people need admin help and I've seen one beeing active here :) Password reset doesn't work and there seems to be an alternative method some years ago. Maybe this can be done again? Please have a look in this thread: http://www.gpugrid.net/forum_thread.php?id=2587&nowrap=true#58958 Thanks! Fritz |
|
15)
Message boards :
Server and website :
Account - password forgotten and impossible to get it back
(Message 58958)
Posted 21 Jun 2022 by FritzB Post: Unfortunately, the problem exists again. Not with me but with another user ("X_FISH"). He has reinstalled Boinc on a computer and now the password is no longer accepted. The password cannot be reset because sending it to his e-mail address is being refused. Can he therefore contact someone directly by e-mail? Or should I send his email address, which is linked to his account, to someone via PN? Many thanks in advance on behalf of X_FISH |
|
16)
Message boards :
News :
Experimental Python tasks (beta) - task description
(Message 58300)
Posted 19 Jan 2022 by FritzB Post: it seems to work better now but I've reached time limit after 1800sec https://www.gpugrid.net/result.php?resultid=32734648 19:39:23 (6124): task /usr/bin/flock reached time limit 1800 application ./gpugridpy/bin/python missing |
|
17)
Message boards :
News :
Experimental Python tasks (beta) - task description
(Message 58265)
Posted 10 Jan 2022 by FritzB Post: I got 20 bad WU's today on this host: https://www.gpugrid.net/results.php?hostid=520456
Stderr Ausgabe
<core_client_version>7.16.6</core_client_version>
<![CDATA[
<message>
process exited with code 195 (0xc3, -61)</message>
<stderr_txt>
13:25:53 (6392): wrapper (7.7.26016): starting
13:25:53 (6392): wrapper (7.7.26016): starting
13:25:53 (6392): wrapper: running /usr/bin/flock (/var/lib/boinc-client/projects/www.gpugrid.net/miniconda.lock -c "/bin/bash ./miniconda-installer.sh -b -u -p /var/lib/boinc-client/projects/www.gpugrid.net/miniconda &&
/var/lib/boinc-client/projects/www.gpugrid.net/miniconda/bin/conda install -m -y -p gpugridpy --file requirements.txt ")
0%| | 0/45 [00:00<?, ?it/s]
concurrent.futures.process._RemoteTraceback:
'''
Traceback (most recent call last):
File "concurrent/futures/process.py", line 368, in _queue_management_worker
File "multiprocessing/connection.py", line 251, in recv
TypeError: __init__() missing 1 required positional argument: 'msg'
'''
The above exception was the direct cause of the following exception:
Traceback (most recent call last):
File "entry_point.py", line 69, in <module>
File "concurrent/futures/process.py", line 484, in _chain_from_iterable_of_lists
File "concurrent/futures/_base.py", line 611, in result_iterator
File "concurrent/futures/_base.py", line 439, in result
File "concurrent/futures/_base.py", line 388, in __get_result
concurrent.futures.process.BrokenProcessPool: A process in the process pool was terminated abruptly while the future was running or pending.
[6689] Failed to execute script entry_point
13:25:58 (6392): /usr/bin/flock exited; CPU time 3.906269
13:25:58 (6392): app exit status: 0x1
13:25:58 (6392): called boinc_finish(195)
</stderr_txt>
]]>
|
©2026 Universitat Pompeu Fabra