| 3 |
0 |
6384
|
26 |
The job attribute OnExitHold expression '(ExitBySignal == true) || (ExitCode != 0)' evaluated to TRUE |
6 mins ago
|
| 34 |
102 |
5606
|
35 |
Error from slot2_4@gpu2005.chtc.wisc.edu: Job has gone over cgroup memory limit of 8192 megabytes. Last measured usage: 8192 megabytes. Consider resubmitting with a higher request_memory. |
6 mins ago
|
| 12 |
2 |
4258
|
71 |
Transfer output files failure at the execution point while sending files to access point submit06. Details: reading from file /tmp/glide_IXyz2d/execute/dir_403997/output_368342.pairs: (errno 2) No such file or directory |
6 mins ago
|
| 14 |
13 |
3324
|
1 |
Cannot access initial working directory /nfs/direct/jonesrt/simsamples/particle_gun.d: Permission denied |
2138 hours ago
|
| 47 |
0 |
3161
|
30 |
The job exceeded allowed execute duration of 14+00:00:00 |
6 mins ago
|
| 1 |
0 |
3127
|
21 |
User requested pause new work; let current running jobs finish; no automatic restart (by user anzheng.li) |
6 mins ago
|
| 21 |
102 |
3071
|
19 |
Error from slot1_2@glidein_3845102_138587700@execute-130.mortimer.hpc.uwm.edu: memory usage exceeded request_memory |
6 mins ago
|
| 12 |
122 |
2637
|
7 |
Transfer output files failure at access point ap2001 while receiving files from execution point slot3_5@hsharma-chtcgpu4000.chtc.wisc.edu. Details: writing to file /home/ychen878/PALite-CHTC/uv-setuptools-bb603b59abf98e69.lock: (errno 122) Disk quota exceeded |
6 mins ago
|
| 13 |
256 |
2464
|
22 |
Transfer input files failure at the execution point using protocol osdf. Details: Pelican Client Error: Attempt #3: from mghpcc-cache.nationalresearchplatform.org:8443: Specification.FileNotFound Error: Error code 5011: server returned 404 Not Found (100ms elapsed, 200ms since start); Attempt #2: from osdf-cache-01.gwave.ics.psu.edu:8443: Specification.FileNotFound Error: Error code 5011: server returned 404 Not Found (100ms elapsed, 100ms since start); Attempt #1: from osdf1.newy32aoa.nrp.internet2.edu:8443: Specification.FileNotFound Error: Error code 5011: server returned 404 Not Found (0s since start) (Version: 7.26.0; Site: UConn) ( URL file = jlab.gluex+osdf:///jlab-osdf/gluex/test/testfile1 )| |
6 mins ago
|
| 13 |
2 |
2298
|
30 |
Transfer input files failure at access point ap1 while sending files to the execution point. Details: 6 total failures: first failure: reading from file /path-facility/projects/Morgridge_Bockelman/Protein-DB/APO-sys/jessica/1n1j/topol.top: (errno 2) No such file or directory |
6 mins ago
|
| 21 |
103 |
1154
|
25 |
Error from slot1_12@e2599.chtc.wisc.edu: disk usage exceeded request_disk |
6 mins ago
|
| 9 |
122 |
1094
|
1 |
Error from slot1_6@glidein_721931_23784321@compute04: StreamHandler: stdout: write of 134 bytes to output.out.41978559-54 failed: Disk quota exceeded |
84 hours ago
|
| 21 |
0 |
953
|
46 |
Error from slot1_1@e2013.chtc.wisc.edu: Job failed to complete in 72 hrs |
6 mins ago
|
| 26 |
100 |
928
|
15 |
The job restarted too many times |
6 mins ago
|
| 9 |
0 |
664
|
1 |
Error from slot1_3@glidein_1476404_800798700@comphy.fsu.edu: unable to establish standard output stream |
84 hours ago
|
| 21 |
104 |
462
|
17 |
Error from slot1_1@glidein_3477228_359097200@scigrid4.physics.fsu.edu: disk usage exceeded allocated disk |
6 mins ago
|
| 13 |
107 |
233
|
1 |
Transfer input files failure at access point ap40 while sending files to execution point slot1_2@glidein_605854_595494167@rails14. Details: 3 total failures: first failure: reading from file /home/emily.ostrow/../../ospool/ap40/data/emily.ostrow/sratoolkit.tar.gz: (errno 107) Transport endpoint is not connected |
982 hours ago
|
| 14 |
19 |
224
|
1 |
Cannot access initial working directory /home/tier3/cmsprod/cms/data/nanohr/D08/EGamma2+Run2025G-PromptReco-v1+MINIAOD: No such device |
172 hours ago
|
| 12 |
256 |
159
|
24 |
Transfer output files failure at execution point slot1_1@glidein_49861_376592796@gpu05 using protocol osdf. Details: Pelican Client Error: local file "/tmp/glide_JgmMuF/execute/dir_80012/scratch/output_data.tar.gz" does not exist: stat /tmp/glide_JgmMuF/execute/dir_80012/scratch/output_data.tar.gz: no such file or directory (Version: 7.26.0; Site: PDX-Coeus) ( URL file = osdf:///ospool/ap41/data/bakary/results/md/output_data_12317421_0.tar.gz )| |
6 mins ago
|
| 13 |
2816 |
139
|
9 |
Transfer input files failure at the execution point using protocol osdf. Details: Pelican Client Error: Attempt #3: from dtn-pas.cinc.nrp.internet2.edu:8443: Transfer.SlowTransfer Error: Error code 6002: cancelled transfer, too slow; detected speed=97.4 KB/s, total transferred=14.6 MB, total transfer time=2m28.001s, cache hit (2m28s elapsed, 31m27.5s since start); Attempt #2: from osdf-cache-01.gwave.ics.psu.edu:8443: Transfer.SlowTransfer Error: Error code 6002: cancelled transfer, too slow; detected speed=69.0 KB/s, total transferred=9.3 MB, total transfer time=2m10.501s (2m10.5s elapsed, 28m59.5s since start); Attempt #1: from osdf1.newy32aoa.nrp.internet2.edu:8443: Transfer.SlowTransfer Error: Error code 6002: cancelled transfer, too slow; detected speed=99.6 KB/s, total transferred=218.0 MB, total transfer time=26m49.001s, cache hit (26m49s since start) (Version: 7.26.0; Site: VU-AUGIE) ( URL file = osdf:///ospool/ap40/data/zitao.zhang/spcbridge_v3.sif ) |
6 mins ago
|
| 35 |
0 |
100
|
4 |
Error from slot1_5@e4061.chtc.wisc.edu: Cannot pull image pytorch/pytorch:2.5.1-cpu Error response from daemon: manifest for pytorch/pytorch:2.5.1-cpu not found: manifest unknown: manifest unknown |
6 mins ago
|
| 4 |
0 |
95
|
5 |
Job credentials are not available |
6 mins ago
|
| 34 |
104 |
70
|
10 |
Error from slot2_11@oconnor2007.chtc.wisc.edu: Job has exceeded allocated disk (15.68 GB). Consider increasing the value of request_disk. |
6 mins ago
|
| 13 |
40 |
45
|
1 |
Transfer input files failure at access point uclhc-2 while sending files to execution point slot1_9@gpatlas2.ps.uci.edu. Details: 1 total failures: first failure: reading from file /export/nfs0home/rmgleaso/Continue_Rose_v2PtThrCalc/run/paramHistos_merged_Riley.root: (errno 40) Too many levels of symbolic links |
148 hours ago
|
| 26 |
102 |
44
|
11 |
Excessive CPU usage. Job used 3 CPUs, while request_cpus=1. Please verify that the code is configured to use a limited number of cpus/threads, and matches request_cpus. |
6 mins ago
|
| 20 |
0 |
19
|
1 |
Job missed deferred execution time |
1722 hours ago
|
| 3 |
1 |
10
|
1 |
Nonzero exit-code |
368 hours ago
|
| 15 |
0 |
8
|
5 |
submitted on hold at user's request |
6 mins ago
|
| 14 |
107 |
7
|
1 |
Cannot access initial working directory /mnt/htc-cephfs/fuse/root/ospool/ap40/data/oz.arie/MPT_AKLT/Fusion: Transport endpoint is not connected |
963 hours ago
|
| 6 |
13 |
6
|
1 |
Error from slot1_15@e2548.chtc.wisc.edu: Failed to execute '/var/lib/condor/execute/slot1/dir_3740134/scratch/integrate_uv.sh' with arguments uv_spectrum: (errno=13: 'Permission denied') |
19 hours ago
|
| 6 |
0 |
5
|
4 |
Error from slot2_5@gpu4003.chtc.wisc.edu: Error running docker job: failed to create task for container: failed to create shim task: OCI runtime create failed: runc create failed: unable to start container process: error during container init: error running createContainer hook #1: exit status 1, stdout: , stderr: time=\'2026-08-20T23:38:49-05:00\' level=info msg=\'Symlinking /var/lib/docker/overlay2/4498b5f6ec5f1817a19c2dce50d3eb043a56ecc46275dd8625a86caf92471849/merged/usr/lib64/libcuda.so to libcuda.so.1\'\ntime=\'2026-08-20T23:38:49-05:00\' level=error msg=\'failed to create link [libcuda.so.1 /usr/lib64/libcuda.so]: failed to create temporary symlink: bad file descriptor\' |
6 mins ago
|
| 19 |
1 |
5
|
3 |
Error from slot1_8@glidein_3926248_485583295@node2045.palmetto.clemson.edu: PREPARE_JOB (prepare-hook) failed (reported status 001): Unable to download or build singularity image /cvmfs/oasis.opensciencegrid.org/jlab/halla/solid/soft/container/jeffersonlab_jlabce_tag1.5_digest:sha156:9b9a9ec8c793035d5bfe6651150b54ac198f5ad17dca490a8039c530d0301008_10110413_s3.9.5.sif |
6 mins ago
|
| 12 |
36 |
4
|
1 |
Transfer output files failure at execution point slot1_12@e2010.chtc.wisc.edu while sending files to access point ap2001. Details: 1 total failures: first failure: reading from file /var/lib/condor/execute/slot1/dir_1868296/scratch/H9_D0_SD500_CHIR7_inefficient_scdblfinder.h5ad H9_D0_SD500_CHIR8_efficient_scdblfinder.h5ad H9_D0_SD500_CHIR12_intermediate_scdblfinder.h5ad H9_D1_SD500_CHIR7_inefficient_scdblfinder.h5ad H9_D1_SD500_CHIR8_efficient_scdblfinder.h5ad H9_D1_SD500_CHIR12_intermediate_scdblfinder.h5ad H9_D2_SD500_CHIR7_inefficient_scdblfinder.h5ad H9_D2_SD500_CHIR8_efficient_scdblfinder.h5ad H9_D2_SD500_CHIR12_intermediate_scdblfinder.h5ad H9_D3_SD500_CHIR7_inefficient_scdblfinder.h5ad H9_D3_SD500_CHIR8_efficient_scdblfinder.h5ad H9_D3_SD500_CHIR12_intermediate_scdblfinder.h5ad H9_D4_SD500_CHIR7_inefficient_scdblfinder.h5ad H9_D4_SD500_CHIR8_efficient_scdblfinder.h5ad H9_D4_SD500_CHIR12_intermediate_scdblfinder.h5ad: (errno 36) File name too long |
729 hours ago
|
| 12 |
21 |
4
|
2 |
Transfer output files failure at execution point slot2_4@vetsigian0001.chtc.wisc.edu while sending files to access point ap2002. Details: 1 total failures: first failure: reading from file /var/lib/condor/execute/slot2/dir_1087596/scratch/wandb/latest-run: (errno 21) Is a directory- Transfer of symlinks to directories is not supported. |
6 mins ago
|
| 32 |
0 |
4
|
3 |
TransferInputSizeMB (5506) is greater than MAX_TRANSFER_INPUT_MB (5000) at submit time |
6 mins ago
|
| 14 |
2 |
4
|
1 |
Cannot access initial working directory /ospool/ap41/data/amanjalingal/fireFLY/TIPSPc/overlap_condor: No such file or directory |
283 hours ago
|
| 12 |
28 |
3
|
1 |
Transfer output files failure at access point ap41 while receiving files from the execution point. Details: writing to file /var/lib/condor/spool/7044/0/cluster12327044.proc0.subproc0/job.12327044.0.out: (errno 28) No space left on device |
6 mins ago
|
| 26 |
10001 |
3
|
1 |
Policy violation. CPU consumption limit exceeded: used 2.081378178835111E+01 usr and 2.879409351927810E-01 sys > requested 1. |
1197 hours ago
|
| 13 |
20 |
3
|
1 |
Transfer input files failure at access point ap2001 while sending files to execution point slot1_5@e2607.chtc.wisc.edu. Details: 1 total failures: first failure: reading from file /home/groups/keller_group/fsgDB_KellerDrott_Analysis/blabri1/blabri1_hmmer_dom.out.gz: (errno 20) Not a directory |
669 hours ago
|
| 13 |
122 |
3
|
1 |
Transfer input files failure at execution point slot1_1@Purdue-Anvil.929c709587.g006 while receiving files from access point grid-submitter. Details: Error from slot1_1@Purdue-Anvil.929c709587.g006: Failed to transfer files: STARTER at 172.18.92.44 failed to create directory /pilot/osgvo-pilot-td1amn/execute/dir_10392/scratch/spice_3.2_lumi: Disk quota exceeded (errno 122) |
67 hours ago
|
| 6 |
2 |
3
|
1 |
Error from slot1_5@e2456.chtc.wisc.edu: Failed to execute '/home/skim2324/venvs/kineticmodel/bin/python' with arguments test.py: (errno=2: 'No such file or directory') |
499 hours ago
|
| 46 |
0 |
3
|
1 |
The job exceeded allowed job duration of 1+00:00:00 |
575 hours ago
|
| 13 |
13 |
2
|
1 |
Transfer input files failure at execution point slot1_25@e4040.chtc.wisc.edu while receiving files from access point ap2002. Details: writing to file /var/lib/condor/execute/slot1/dir_1372467/scratch/data/breast_trans_prob.csv: (errno 13) Permission denied |
6 mins ago
|
| 3 |
5 |
2
|
1 |
crop63 run_in_job.sh exit code 5: see <tag>.manifest.json and <tag>.err in the run dir |
84 hours ago
|
| 13 |
28 |
2
|
1 |
Transfer input files failure at execution point slot2_2@gpu4003.chtc.wisc.edu while receiving files from access point ap2002. Details: writing to file /var/lib/condor/execute/slot2/dir_628230/scratch/realistic_v31.bundle: (errno 28) No space left on device |
6 mins ago
|
| 13 |
3 |
2
|
1 |
Transfer input files failure at access point ap2002 while sending files to execution point slot2_8@gpu2011.chtc.wisc.edu. Details: Error from slot2_8@gpu2011.chtc.wisc.edu: Failed to transfer files: SHADOW at 128.105.68.113 failed to send file(s) to <128.105.68.103:45893>: DoUpload: Failure when signing URL 's3://web.s3.wisc.edu/chestnut/chestnut-imagery/UAV/raw-images/50ft_single_images/DJI_20250827125515_0001_point0.JPG': unable to read from access key file; STARTER failed to receive file(s) from <128.105.68.113:9618> |
829 hours ago
|
| 26 |
100001 |
2
|
1 |
Policy violation. Execution time limit exceeded, job was evicted and held to prevent rematching |
6 mins ago
|
| 19 |
0 |
2
|
1 |
Error from slot1_12@glidein_755926_510265136@spark-a011.chtc.wisc.edu: failed to execute PREPARE_JOB (/var/lib/condor/execute/osg01/glide_v6PW5t/client_group_main/prepare-hook) |
17 hours ago
|
| 36 |
-1 |
2
|
1 |
Error from slot1_16@glidein_114319_643782136@CRUSH-OSG-C7-10-5-188-217: Starter failed to upload checkpoint |
39 hours ago
|
| 16 |
0 |
2
|
1 |
Spooling input data files |
7 hours ago
|
| 12 |
2816 |
2
|
1 |
Transfer output files failure at execution point backfill1_62@build4000.chtc.wisc.edu using protocol osdf. Details: Contact.ConnectionReset Error: Error code 3005: Error occurred when querying for metadata: Get "https://osg-htc.org/.well-known/pelican-configuration": read tcp [2607:f388:2200:100:1270:fdff:fe56:41cc]:39188->[2606:4700:3030::ac43:ab13]:443: read: connection reset by peer ( URL file = pelican.peglegel9+osdf:///icecube/wipac/data/user/syan079/pegleg_grid/2025/0109/Run00140339_00000206.tgz )| |
6 mins ago
|
| 23 |
0 |
1
|
1 |
Unable to switch to user: avkekane@chtc.wisc.edu |
6 mins ago
|
| 13 |
12 |
1
|
1 |
Transfer input files failure at the execution point while receiving files from access point submit06. Details: writing to file /tmp/glide_iHCGKu/execute/dir_994278/kraken_15_0_14.tgz: (errno 12) Cannot allocate memory |
98 hours ago
|
| 13 |
1 |
1
|
1 |
Transfer input files failure at execution point backfill1_5@mrudolphgpu4000.chtc.wisc.edu while receiving files from access point ap2002. Details: Error from backfill1_5@mrudolphgpu4000.chtc.wisc.edu: Failed to transfer files: Attempt to write to illegal sandbox path: |
15 hours ago
|
| 6 |
8 |
1
|
1 |
Error from slot1_1@e4020.chtc.wisc.edu: Failed to execute '/var/lib/condor/execute/slot1/dir_3485002/scratch/run.sh': (errno=8: 'Exec format error') |
6 mins ago
|
| 13 |
21 |
1
|
1 |
Transfer input files failure at access point ap2001 while sending files to execution point slot1_17@e2010.chtc.wisc.edu. Details: 10 total failures: first failure: reading from file /home/mkaur32/MK_Liver_RNAseq_2026/16_BooteJTK/16.03_CHTC/bootejtk_py2_env/lib/icu/current: (errno 21) Is a directory- Transfer of symlinks to directories is not supported. |
167 hours ago
|
| 7 |
122 |
1
|
1 |
Error from slot1_1@Purdue-Anvil.929c709587.g006: Failed to open '/pilot/osgvo-pilot-td1amn/execute/dir_10396/scratch/_condor_stdout' as standard output: Disk quota exceeded (errno 122) |
67 hours ago
|
| 13 |
-1004 |
1
|
1 |
Transfer input files failure at the execution point while receiving files from access point grid-submitter. Details: file transfer plugin /usr/libexec/condor/stash_plugin exited (exit code 3), no valid classads in output file /pilot/osgvo-pilot-naOf6n/execute/dir_383987/scratch/.stash_plugin.out |
1373 hours ago
|
| 19 |
7 |
1
|
1 |
Error from slot1@glidein_1126977_268142376@c-5-1.aglt2.org: PREPARE_JOB (prepare-hook) failed (reported status 007): Unable to download or build singularity image /cvmfs/singularity.opensciencegrid.org/jeffersonlab/clas12software:production |
49 hours ago
|
| 12 |
107 |
1
|
1 |
Transfer output files failure at access point ap40 while receiving files from execution point slot1_16@UA-LR-ITS-EP.e8fe1928b823. Details: writing to file /ospool/ap40/data/ka784/magic_msr_rawpt_tf4l_pm0205_20260820_v1/results/magic_msr_rawpt_tf4l_pm0205_20260820_v1_random_dual_clifford_pm0p205_L64_pt0p05_batch000485.tar.gz: (errno 107) Transport endpoint is not connected |
984 hours ago
|
| 34 |
0 |
1
|
1 |
Error from slot2_3@vetsigian0000.chtc.wisc.edu: Job has gone over cgroup memory limit of 2048 megabytes. Last measured usage: 1680 megabytes. Consider resubmitting with a higher request_memory. |
1553 hours ago
|
| 13 |
11 |
1
|
1 |
Transfer input files failure at the execution point while receiving files from access point ap40. Details: receiving file /var/lib/condor/execute/dir_24207/glide_dbv9LW/execute/dir_128405/database.tar.gz: FILETRANSFER:1:FILETRANSFER: plugin for type osdf not found!|FILETRANSFER:110:No output from /var/lib/condor/execute/dir_24207/glide_dbv9LW/main/condor/libexec/stash_plugin -classad, ignoring |
803 hours ago
|
| 12 |
20 |
1
|
1 |
Transfer output files failure at execution point slot1_1@glidein_2869077_470054144@gpu06 while sending files to access point ap40. Details: 1 total failures: first failure: reading from file /tmp/glide_UrVfrY/execute/dir_2897476/scratch/raw_image_denoising/checkpoints/improved: (errno 20) Not a directory |
566 hours ago
|
| 0 |
0 |
1
|
1 |
|
9 hours ago
|
| 45 |
-1000 |
1
|
1 |
Error from slot1_33@UA-LR-ITS-EP.4c50b455fe13: Singularity test failed:FATAL: context deadline exceeded |
6 mins ago
|
| 3 |
4 |
1
|
1 |
crop63 run_in_job.sh exit code 4: see <tag>.manifest.json and <tag>.err in the run dir |
109 hours ago
|
| 3 |
6 |
1
|
1 |
crop63 run_in_job.sh exit code 6: see <tag>.manifest.json and <tag>.err in the run dir |
89 hours ago
|
| 26 |
0 |
1
|
1 |
Job in status 2 put on hold by SYSTEM_PERIODIC_HOLD due to memory usage 49296480. |
1958 hours ago
|
| 13 |
-1001 |
1
|
1 |
Transfer input files failure at the execution point while receiving files from access point grid-submitter. Details: receiving file /scratch/glide_GEOlHi/execute/dir_3052258/scratch/decay_charge1_beta0.0003_lambda0.001_nevents10000_4942.i3.zst: FILETRANSFER:1:FILETRANSFER: plugin for type osdf not found!|FILETRANSFER:110:No output from /scratch/glide_GEOlHi/main/condor/libexec/stash_plugin -classad, ignoring |
1211 hours ago
|
| 26 |
1001 |
1
|
1 |
Policy violation. Memory limit exceeded: 14006 MB resident > 8192 MB requested. |
6 mins ago
|
| 9 |
22 |
1
|
1 |
Error from slot2_1@gpu4005.chtc.wisc.edu: StreamHandler: stdout: couldn't write to ssh_4945036_0.out: Invalid argument (-1!=0) |
6 mins ago
|
| 3 |
9001 |
1
|
1 |
Retrying after high memory usage |
353 hours ago
|
| 13 |
5 |
1
|
1 |
Transfer input files failure at execution point slot1_1@glidein_14723_512829888@voh1 while receiving files from access point ap2002. Details: writing to file /wsu/tmp/glide_xGl9x1/execute/dir_1827/./debug/check_size.py: (errno 5) Input/output error |
522 hours ago
|
| 13 |
512 |
1
|
1 |
Transfer input files failure at execution point slot1_4@UWL-Test-EP.072ca3c51e51 while receiving files from access point ap41. Details: file transfer plugin /usr/libexec/condor/stash_plugin exited (exit code 2), no valid classads in output file /pilot/osgvo-pilot-uoIAqx/execute/dir_325633/.stash_plugin.out |
2508 hours ago
|
| 12 |
11 |
1
|
1 |
Transfer output files failure at the execution point while sending files to access point ap40. Details: sending file /var/lib/condor/execute/dir_7673/glide_469FaF/execute/dir_28312/result/PBC/L=14/p=0.365/tau=0/mi_r3847324908.dat: FILETRANSFER:1:FILETRANSFER: plugin for type osdf not found!|FILETRANSFER:110:No output from /var/lib/condor/execute/dir_7673/glide_469FaF/main/condor/libexec/stash_plugin -classad, ignoring |
1075 hours ago
|
| 3 |
42 |
1
|
1 |
Job ran for more than two hours |
97 hours ago
|
| 12 |
1 |
1
|
1 |
Transfer output files failure at access point ap2002 while receiving files from execution point backfill1_2@dsigpu4000.chtc.wisc.edu. Details: Error from backfill1_2@dsigpu4000.chtc.wisc.edu: STARTER at 128.105.68.119 failed to send file(s) to <128.105.68.113:9618>; Remap of output file resulted in a URL: osdf:///chtc/staging/s/srusso6/smaxi-normalize-gpu-20260925-staged/ID_7075_013_normalized.mp4 |
6 mins ago
|
| 26 |
101 |
1
|
1 |
The job (shadow) restarted too many times |
31 hours ago
|
| 13 |
9 |
1
|
1 |
Transfer input files failure at the execution point while receiving files from access point ap40. Details: Error from slot1_3@glidein_3994351_105468480@warlock23.beocat.ksu.edu: Failed to transfer files: FILETRANSFER:1:File transfer plugin /tmp/glide_bq781D/main/condor/libexec/stash_plugin failed unexpectedly with exit code 0, did not report a TransferError message. |
6 mins ago
|
| 12 |
-1001 |
1
|
1 |
Transfer output files failure at the execution point while sending files to access point grid-submitter. Details: 1 total failures: first failure: sending file /var/lib/condor/execute/slot1/dir_2886709/scratch/glide_0s8QmM/execute/dir_180262/scratch/batch99992.npz: FILETRANSFER:1:FILETRANSFER: plugin for type osdf not found!|FILETRANSFER:110:No output from /var/lib/condor/execute/slot1/dir_2886709/scratch/glide_0s8QmM/main/condor/libexec/stash_plugin -classad, ignoring |
766 hours ago
|
| 19 |
254 |
1
|
1 |
Error from slot1@glidein_862497_473595615@talon05.cm.cluster: PREPARE_JOB (prepare-hook) failed (exited with status 254): <no message> |
46 hours ago
|
| 12 |
512 |
1
|
1 |
Transfer output files failure at execution point slot2_1@gpu2005.chtc.wisc.edu using protocol pelican. Details: Pelican Client Error: failed upload to researchdrive-origin.uwdf-prod.chtc.io:8443: request failed (HTTP status 500): Transfer Error: Error code 6000: server returned 500 Internal Server Error (0s since start) (Version: 7.25.0) ( URL file = pelican://chtc.wisc.edu/researchdrive/dli55/CHTC/62741_0/a100_62741_petal/masks/batch_00067/Rudbeckia_hirta_547563350_instance_2.npy )||FILETRANSFER:1:non-zero exit (2) from /usr/libexec/condor/stash_plugin. |Error: Pelican Client Error: failed upload to researchdrive-origin.uwdf-prod.chtc.io:8443: request failed (HTTP status 500): Transfer Error: Error code 6000: server returned 500 Internal Server Error (0s since start) (Version: 7.25.0) ( URL file = pelican://chtc.wisc.edu/researchdrive/dli55/CHTC/62741_0/a100_62741_petal/masks/batch_00067/Rudbeckia_hirta_549531719_instance_11.npy )||FILETRANSFER:1:non-zero exit (2) from /usr/libexec/condor/stash_plugin. |Error: Pelican Client Error: failed upload to researchdrive-origin.uwdf-prod.chtc.io:8443: request failed (HTTP status 500): Transfer Error: Error code 6000: server returned 500 Internal Server Error (0s since start) (Version: 7.25.0) ( URL file = pelican://chtc.wisc.edu/researchdrive/dli55/CHTC/62741_0/a100_62741_petal/errors.txt )||FILETRANSFER:1:non-zero exit (2) from /usr/libexec/condor/stash_plugin. |Error: Pelican Client Error: failed upload to researchdrive-origin.uwdf-prod.chtc.io:8443: request failed (HTTP status 403): Authorization Error: Error code 4000: server returned 403 Forbidden (0s since start) (Version: 7.25.0) ( URL file = pelican://chtc.wisc.edu/researchdrive/dli55/CHTC/62741_0/a100_62741_petal/masks/batch_00067/Rudbeckia_hirta_545092662_instance_3.npy )||FILETRANSFER:1:non-zero exit (2) from /usr/libexec/condor/stash_plugin. |Error: Pelican Client Error: failed upload to researchdrive-origin.uwdf-prod.chtc.io:8443: request failed (HTTP status 500): Transfer Error: Error code 6000: server returned 500 Internal Server Error (0s since start) (Version: 7.25.0) ( URL file = pelican://chtc.wisc.edu/researchdrive/dli55/CHTC/62741_0/a100_62741_petal/masks/batch_00067/Rudbeckia_hirta_548406666_instance_1.npy )||FILETRANSFER:1:non-zero exit (2) from /usr/libexec/condor/stash_plugin. |Error: Pelican Client Error: failed upload to researchdrive-origin.uwdf-prod.chtc.io:8443: request failed (HTTP status 500): Transfer Error: Error code 6000: server returned 500 Internal Server Error (0s since start) (Version: 7.25.0) ( URL file = pelican://chtc.wisc.edu/researchdrive/dli55/CHTC/62741_0/a100_62741_petal/masks/batch_00067/Rudbeckia_hirta_549325675_instance_4.npy )| |
6 mins ago
|
| 13 |
0 |
1
|
1 |
Transfer input files failure at execution point slot1_2@USC-CARC-Artemis-Backfill.cef453c84752 while receiving files from access point ap41. Details: Error from slot1_2@USC-CARC-Artemis-Backfill.cef453c84752: Failed to transfer files: reason unknown. |
647 hours ago
|
| 12 |
13 |
1
|
1 |
Transfer output files failure at access point ap40 while receiving files from the execution point. Details: Error from slot1_3@glidein_3555484_587001160@execute-136.mortimer.hpc.uwm.edu: STARTER at 172.20.32.64 failed to send file(s) to <128.105.68.62:9618>; SHADOW at 128.105.68.62 failed to create output directory (/home/test_results): Permission denied (errno 13) |
19 hours ago
|
| 19 |
127 |
1
|
1 |
Error from slot1_7@Harvard.64cf75365d.holygpu8a31504.rc.fas.harvard.edu: PREPARE_JOB (prepare-hook) failed (exited with status 127): <no message> |
2179 hours ago
|
| 26 |
101001 |
1
|
1 |
Policy violation. Memory limit exceeded: 3745 MB resident > 2500 MB requested. Execution time limit exceeded, job was evicted and held to prevent rematching |
1858 hours ago
|