Jobsub ID 43078.87@dunegpschedd01.fnal.gov
Jobsub ID | 43078.87@dunegpschedd01.fnal.gov |
Workflow ID | 2323 |
Stage ID | 1 |
User name | ykermaid@fnal.gov |
HTCondor Group | group_dune.prod.mcsim |
Requested | Processors | 1 |
GPU | No |
RSS bytes | 4193255424 (3999 MiB) |
Wall seconds limit | 18000 (5 hours) |
Submitted time | 2025-09-08 12:07:53 |
Site | NL_NIKHEF |
Entry | VIRGO_NL_NIKHEF_dissel |
Last heartbeat | 2025-09-08 15:52:47 |
From worker node | Hostname | wn-lot-031.farm.nikhef.nl |
cpuinfo | AMD EPYC 7702P 64-Core Processor |
OS release | Scientific Linux release 7.9 (Nitrogen) |
Processors | 1 |
RSS bytes | 4194304000 (4000 MiB) |
Wall seconds limit | 129600 (36 hours) |
GPU | |
Inner Apptainer? | True |
Job state | outputting_failed |
Started | 2025-09-08 14:44:21 |
Input files | vd-protodune:np02vd_raw_run039324_1716_df-s02-d2_dw_0_20250907T042330.hdf5
|
Jobscript | Exit code | 1 |
Real time | 0m (0s) |
CPU time | 0m (0s = 0%) |
Max RSS bytes | 0 (0 MiB) |
Outputting started | |
Output files | |
Finished | 2025-09-08 15:52:47 |
List job events Cached HTCondor job logs |
Jobscript log (last 10,000 characters)
aTrackCreation 0.284791 0.724204 2.75935 0.575328 0.506423 26
produce:pandoraGnocalo:GnocchiCalorimetry 0.0176897 0.0281584 0.0560241 0.0259394 0.00890109 26
[art]:TriggerResults:TriggerResultInserter 1.564e-05 1.9152e-05 4.807e-05 1.73725e-05 6.18928e-06 26
end_path:out1:RootOutput 3.486e-06 4.44135e-06 1.7723e-05 3.787e-06 2.67913e-06 26
end_path:out1:RootOutput(write) 3.83408 4.28204 6.62732 4.07118 0.628479 26
==================================================================================================================================
====================================================================================================
MemoryTracker summary (base-10 MB units used)
Peak virtual memory usage (VmPeak) : 4905.93 MB
Peak resident set size usage (VmHWM): 3023.86 MB
Details saved in: 'mem.db'
====================================================================================================
Art has completed and will exit with status 0.
Output files:
\tReco: np02vd_raw_run039324_1716_df-s02-d2_dw_0_20250907T042330_reco_stage1_20250908T154727_keepup.root
\tHists: np02vd_raw_run039324_1716_df-s02-d2_dw_0_20250907T042330_reco_stage1_20250908T154727_keepup_hists.root
Forming reco metadata
Successfully opened file np02vd_raw_run039324_1716_df-s02-d2_dw_0_20250907T042330_reco_stage1_20250908T154727_keepup.root
Ran successfully
{
"name": "np02vd_raw_run039324_1716_df-s02-d2_dw_0_20250907T042330_reco_stage1_20250908T154727_keepup.root",
"namespace": "vd-protodune-det-reco",
"metadata": {
"core.file_format": "artroot",
"core.application.name": "reco",
"core.application.family": "dunesw",
"core.application.version": "v10_10_00d00",
"core.data_tier": "full-reconstructed",
"dune.config_file": "standard_reco_stage1_protodunevd_keepup_all.fcl",
"dune.campaign": "vd-protodune-reco-keepup-v0",
"core.start_time": 1757346448.0,
"core.end_time": 1757346448.0,
"core.events": [
899204,
899224,
899244,
899264,
899284,
899304,
899324,
899344,
899364,
899384,
899404,
899424,
899444,
899464,
899484,
899504,
899524,
899544,
899564,
899584,
899604,
899624,
899644,
899664,
899684,
899704
],
"core.event_count": 26,
"core.first_event_number": 899204,
"core.last_event_number": 899704,
"core.data_stream": "physics",
"core.file_content_status": "good",
"core.file_type": "detector",
"core.run_type": "vd-protodune",
"core.runs": [
39324
],
"core.runs_subruns": [
3932400001
],
"dune.daq_test": false,
"retention.status": "active",
"retention.class": "physics"
},
"parents": [
{
"did": "vd-protodune:np02vd_raw_run039324_1716_df-s02-d2_dw_0_20250907T042330.hdf5"
}
]
}Forming hist metadata
Traceback (most recent call last):
File "/cvmfs/fermilab.opensciencegrid.org/products/common/prd/python_future_six_request/v1_3_1/Linux64bit-3-10-2-17-python3-9/urllib3/connectionpool.py", line 699, in urlopen
httplib_response = self._make_request(
File "/cvmfs/fermilab.opensciencegrid.org/products/common/prd/python_future_six_request/v1_3_1/Linux64bit-3-10-2-17-python3-9/urllib3/connectionpool.py", line 445, in _make_request
six.raise_from(e, None)
File "<string>", line 3, in raise_from
File "/cvmfs/fermilab.opensciencegrid.org/products/common/prd/python_future_six_request/v1_3_1/Linux64bit-3-10-2-17-python3-9/urllib3/connectionpool.py", line 440, in _make_request
httplib_response = conn.getresponse()
File "/cvmfs/larsoft.opensciencegrid.org/products/python/v3_9_15/Linux64bit+3.10-2.17/lib/python3.9/http/client.py", line 1377, in getresponse
response.begin()
File "/cvmfs/larsoft.opensciencegrid.org/products/python/v3_9_15/Linux64bit+3.10-2.17/lib/python3.9/http/client.py", line 320, in begin
version, status, reason = self._read_status()
File "/cvmfs/larsoft.opensciencegrid.org/products/python/v3_9_15/Linux64bit+3.10-2.17/lib/python3.9/http/client.py", line 289, in _read_status
raise RemoteDisconnected("Remote end closed connection without"
http.client.RemoteDisconnected: Remote end closed connection without response
During handling of the above exception, another exception occurred:
Traceback (most recent call last):
File "/cvmfs/fermilab.opensciencegrid.org/products/common/prd/python_future_six_request/v1_3_1/Linux64bit-3-10-2-17-python3-9/requests/adapters.py", line 439, in send
resp = conn.urlopen(
File "/cvmfs/fermilab.opensciencegrid.org/products/common/prd/python_future_six_request/v1_3_1/Linux64bit-3-10-2-17-python3-9/urllib3/connectionpool.py", line 755, in urlopen
retries = retries.increment(
File "/cvmfs/fermilab.opensciencegrid.org/products/common/prd/python_future_six_request/v1_3_1/Linux64bit-3-10-2-17-python3-9/urllib3/util/retry.py", line 532, in increment
raise six.reraise(type(error), error, _stacktrace)
File "/cvmfs/fermilab.opensciencegrid.org/products/common/prd/python_future_six_request/v1_3_1/Linux64bit-3-10-2-17-python3-9/urllib3/packages/six.py", line 769, in reraise
raise value.with_traceback(tb)
File "/cvmfs/fermilab.opensciencegrid.org/products/common/prd/python_future_six_request/v1_3_1/Linux64bit-3-10-2-17-python3-9/urllib3/connectionpool.py", line 699, in urlopen
httplib_response = self._make_request(
File "/cvmfs/fermilab.opensciencegrid.org/products/common/prd/python_future_six_request/v1_3_1/Linux64bit-3-10-2-17-python3-9/urllib3/connectionpool.py", line 445, in _make_request
six.raise_from(e, None)
File "<string>", line 3, in raise_from
File "/cvmfs/fermilab.opensciencegrid.org/products/common/prd/python_future_six_request/v1_3_1/Linux64bit-3-10-2-17-python3-9/urllib3/connectionpool.py", line 440, in _make_request
httplib_response = conn.getresponse()
File "/cvmfs/larsoft.opensciencegrid.org/products/python/v3_9_15/Linux64bit+3.10-2.17/lib/python3.9/http/client.py", line 1377, in getresponse
response.begin()
File "/cvmfs/larsoft.opensciencegrid.org/products/python/v3_9_15/Linux64bit+3.10-2.17/lib/python3.9/http/client.py", line 320, in begin
version, status, reason = self._read_status()
File "/cvmfs/larsoft.opensciencegrid.org/products/python/v3_9_15/Linux64bit+3.10-2.17/lib/python3.9/http/client.py", line 289, in _read_status
raise RemoteDisconnected("Remote end closed connection without"
urllib3.exceptions.ProtocolError: ('Connection aborted.', RemoteDisconnected('Remote end closed connection without response'))
During handling of the above exception, another exception occurred:
Traceback (most recent call last):
File "/cvmfs/larsoft.opensciencegrid.org/products/python/v3_9_15/Linux64bit+3.10-2.17/lib/python3.9/runpy.py", line 197, in _run_module_as_main
return _run_code(code, main_globals, None,
File "/cvmfs/larsoft.opensciencegrid.org/products/python/v3_9_15/Linux64bit+3.10-2.17/lib/python3.9/runpy.py", line 87, in _run_code
exec(code, run_globals)
File "/cvmfs/fifeuser2.opensciencegrid.org/sw/dune/20943968d0063a8f11ac6a4062b5ad0c8392b138/meta_maker.py", line 48, in <module>
results = inherit_metadata.inherit(args.parent)
File "/cvmfs/fifeuser2.opensciencegrid.org/sw/dune/20943968d0063a8f11ac6a4062b5ad0c8392b138/inherit_metadata.py", line 72, in inherit
'metadata':get_parent_md(parent_name),
File "/cvmfs/fifeuser2.opensciencegrid.org/sw/dune/20943968d0063a8f11ac6a4062b5ad0c8392b138/inherit_metadata.py", line 27, in get_parent_md
parent_file = mc.get_file(did=parent_name, with_metadata=True,
File "/cvmfs/dune.opensciencegrid.org/products/dune/metacat/v4_0_2/NULL/lib/python3.9/site-packages/metacat/webapi/webapi.py", line 1259, in get_file
return self.get_json(url)
File "/cvmfs/dune.opensciencegrid.org/products/dune/metacat/v4_0_2/NULL/lib/python3.9/site-packages/metacat/webapi/webapi.py", line 206, in get_json
return self.unpack_json_data(self.send_request("get", uri_suffix, headers=headers, stream=True))
File "/cvmfs/dune.opensciencegrid.org/products/dune/metacat/v4_0_2/NULL/lib/python3.9/site-packages/metacat/webapi/webapi.py", line 154, in send_request
self.LastResponse = response = self.retry_request(method, url, headers=headers, **args)
File "/cvmfs/dune.opensciencegrid.org/products/dune/metacat/v4_0_2/NULL/lib/python3.9/site-packages/metacat/webapi/webapi.py", line 131, in retry_request
response = requests.get(url, timeout=self.Timeout, **args)
File "/cvmfs/fermilab.opensciencegrid.org/products/common/prd/python_future_six_request/v1_3_1/Linux64bit-3-10-2-17-python3-9/requests/api.py", line 75, in get
return request('get', url, params=params, **kwargs)
File "/cvmfs/fermilab.opensciencegrid.org/products/common/prd/python_future_six_request/v1_3_1/Linux64bit-3-10-2-17-python3-9/requests/api.py", line 61, in request
return session.request(method=method, url=url, **kwargs)
File "/cvmfs/fermilab.opensciencegrid.org/products/common/prd/python_future_six_request/v1_3_1/Linux64bit-3-10-2-17-python3-9/requests/sessions.py", line 542, in request
resp = self.send(prep, **send_kwargs)
File "/cvmfs/fermilab.opensciencegrid.org/products/common/prd/python_future_six_request/v1_3_1/Linux64bit-3-10-2-17-python3-9/requests/sessions.py", line 655, in send
r = adapter.send(request, **kwargs)
File "/cvmfs/fermilab.opensciencegrid.org/products/common/prd/python_future_six_request/v1_3_1/Linux64bit-3-10-2-17-python3-9/requests/adapters.py", line 498, in send
raise ConnectionError(err, request=request)
requests.exceptions.ConnectionError: ('Connection aborted.', RemoteDisconnected('Remote end closed connection without response'))
Error in hist metadata