Lmod Warning:
-------------------------------------------------------------------------------
The following dependent module(s) are not currently loaded: curl/8.4.0
(required by: htslib/1.16)
-------------------------------------------------------------------------------



Lmod Warning:
-------------------------------------------------------------------------------
The following dependent module(s) are not currently loaded: curl/8.17.0
(required by: ucsc-utils/v489), openssl/3.0.7 (required by: curl/8.4.0)
-------------------------------------------------------------------------------




The following have been reloaded with a version change:
  1) curl/8.17.0 => curl/8.4.0

Lmod Warning:
-------------------------------------------------------------------------------
The following dependent module(s) are not currently loaded: curl/8.4.0
(required by: htslib/1.16)
-------------------------------------------------------------------------------




The following have been reloaded with a version change:
  1) curl/8.4.0 => curl/8.17.0

2026-07-10 14:45:04.461680: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 14:45:31.476906: I tensorflow/compiler/jit/xla_cpu_device.cc:41] Not creating XLA devices, tf_xla_enable_xla_devices not set
2026-07-10 14:45:31.491005: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcuda.so.1
2026-07-10 14:45:32.111971: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1720] Found device 0 with properties: 
pciBusID: 0000:bb:00.0 name: NVIDIA H100 80GB HBM3 computeCapability: 9.0
coreClock: 1.98GHz coreCount: 132 deviceMemorySize: 79.10GiB deviceMemoryBandwidth: 3.05TiB/s
2026-07-10 14:45:32.112058: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 14:45:32.673100: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublas.so.11
2026-07-10 14:45:32.673207: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublasLt.so.11
2026-07-10 14:45:33.037360: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcufft.so.10
2026-07-10 14:45:33.615175: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcurand.so.10
2026-07-10 14:45:34.212455: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusolver.so.10
2026-07-10 14:45:34.439763: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusparse.so.11
2026-07-10 14:45:34.610128: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudnn.so.8
2026-07-10 14:45:34.618895: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1862] Adding visible gpu devices: 0
2026-07-10 14:45:34.619338: I tensorflow/core/platform/cpu_feature_guard.cc:142] This TensorFlow binary is optimized with oneAPI Deep Neural Network Library (oneDNN) to use the following CPU instructions in performance-critical operations:  AVX2 AVX512F FMA
To enable them in other operations, rebuild TensorFlow with the appropriate compiler flags.
2026-07-10 14:45:34.619408: I tensorflow/compiler/jit/xla_gpu_device.cc:99] Not creating XLA devices, tf_xla_enable_xla_devices not set
2026-07-10 14:45:34.622013: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1720] Found device 0 with properties: 
pciBusID: 0000:bb:00.0 name: NVIDIA H100 80GB HBM3 computeCapability: 9.0
coreClock: 1.98GHz coreCount: 132 deviceMemorySize: 79.10GiB deviceMemoryBandwidth: 3.05TiB/s
2026-07-10 14:45:34.622038: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 14:45:34.622053: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublas.so.11
2026-07-10 14:45:34.622063: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublasLt.so.11
2026-07-10 14:45:34.622089: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcufft.so.10
2026-07-10 14:45:34.622098: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcurand.so.10
2026-07-10 14:45:34.622106: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusolver.so.10
2026-07-10 14:45:34.622115: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusparse.so.11
2026-07-10 14:45:34.622123: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudnn.so.8
2026-07-10 14:45:34.627376: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1862] Adding visible gpu devices: 0
2026-07-10 14:45:34.627416: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 14:47:39.771877: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1261] Device interconnect StreamExecutor with strength 1 edge matrix:
2026-07-10 14:47:39.771977: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1267]      0 
2026-07-10 14:47:39.771989: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1280] 0:   N 
2026-07-10 14:47:39.776361: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1406] Created TensorFlow device (/job:localhost/replica:0/task:0/device:GPU:0 with 75496 MB memory) -> physical GPU (device: 0, name: NVIDIA H100 80GB HBM3, pci bus id: 0000:bb:00.0, compute capability: 9.0)
2026-07-10 14:47:47.721686: I tensorflow/compiler/mlir/mlir_graph_optimization_pass.cc:116] None of the MLIR optimization passes are enabled (registered 2)
2026-07-10 14:47:47.722182: I tensorflow/core/platform/profile_utils/cpu_utils.cc:112] CPU Frequency: 2800000000 Hz
2026-07-10 14:47:49.743446: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublas.so.11
2026-07-10 14:49:45.537642: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublasLt.so.11
2026-07-10 14:49:45.542556: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudnn.so.8
2026-07-10 14:55:16.129445: W tensorflow/stream_executor/gpu/asm_compiler.cc:63] Running ptxas --version returned 256
2026-07-10 14:55:16.271226: W tensorflow/stream_executor/gpu/redzone_allocator.cc:314] Internal: ptxas exited with non-zero error code 256, output: 
Relying on driver to perform ptx compilation. 
Modify $PATH to customize ptxas location.
This message will be only logged once.
2026-07-10 14:55:37.216065: I tensorflow/stream_executor/cuda/cuda_blas.cc:1838] TensorFloat-32 will be used for the matrix multiplication. This will only be logged once.
/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/keras/engine/functional.py:595: UserWarning: Input dict contained keys ['coordinates', 'jitters', 'index', 'status', 'rev_comp'] which did not match any model input. They will be ignored by the model.
  [n for n in tensors.keys() if n not in ref_input_names])
/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/keras/engine/functional.py:595: UserWarning: Input dict contained keys ['coordinates'] which did not match any model input. They will be ignored by the model.
  [n for n in tensors.keys() if n not in ref_input_names])
2026-07-10 15:00:58.515220: W tensorflow/python/util/util.cc:348] Sets are not currently considered sequences, but this may change in the future, so consider avoiding using them.
2026-07-10 15:01:03.162925: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 15:02:19.413239: I tensorflow/compiler/jit/xla_cpu_device.cc:41] Not creating XLA devices, tf_xla_enable_xla_devices not set
2026-07-10 15:02:19.423866: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcuda.so.1
2026-07-10 15:02:19.934496: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1720] Found device 0 with properties: 
pciBusID: 0000:bb:00.0 name: NVIDIA H100 80GB HBM3 computeCapability: 9.0
coreClock: 1.98GHz coreCount: 132 deviceMemorySize: 79.10GiB deviceMemoryBandwidth: 3.05TiB/s
2026-07-10 15:02:19.934581: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 15:02:19.947670: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublas.so.11
2026-07-10 15:02:19.947757: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublasLt.so.11
2026-07-10 15:02:19.954160: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcufft.so.10
2026-07-10 15:02:19.960863: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcurand.so.10
2026-07-10 15:02:19.967538: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusolver.so.10
2026-07-10 15:02:19.973035: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusparse.so.11
2026-07-10 15:02:19.977778: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudnn.so.8
2026-07-10 15:02:19.984015: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1862] Adding visible gpu devices: 0
2026-07-10 15:02:19.984350: I tensorflow/core/platform/cpu_feature_guard.cc:142] This TensorFlow binary is optimized with oneAPI Deep Neural Network Library (oneDNN) to use the following CPU instructions in performance-critical operations:  AVX2 AVX512F FMA
To enable them in other operations, rebuild TensorFlow with the appropriate compiler flags.
2026-07-10 15:02:19.984414: I tensorflow/compiler/jit/xla_gpu_device.cc:99] Not creating XLA devices, tf_xla_enable_xla_devices not set
2026-07-10 15:02:19.987026: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1720] Found device 0 with properties: 
pciBusID: 0000:bb:00.0 name: NVIDIA H100 80GB HBM3 computeCapability: 9.0
coreClock: 1.98GHz coreCount: 132 deviceMemorySize: 79.10GiB deviceMemoryBandwidth: 3.05TiB/s
2026-07-10 15:02:19.987050: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 15:02:19.987062: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublas.so.11
2026-07-10 15:02:19.987071: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublasLt.so.11
2026-07-10 15:02:19.987080: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcufft.so.10
2026-07-10 15:02:19.987089: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcurand.so.10
2026-07-10 15:02:19.987098: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusolver.so.10
2026-07-10 15:02:19.987107: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusparse.so.11
2026-07-10 15:02:19.987116: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudnn.so.8
2026-07-10 15:02:19.991946: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1862] Adding visible gpu devices: 0
2026-07-10 15:02:19.991971: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 15:03:37.779017: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1261] Device interconnect StreamExecutor with strength 1 edge matrix:
2026-07-10 15:03:37.779123: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1267]      0 
2026-07-10 15:03:37.779142: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1280] 0:   N 
2026-07-10 15:03:37.784142: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1406] Created TensorFlow device (/job:localhost/replica:0/task:0/device:GPU:0 with 75496 MB memory) -> physical GPU (device: 0, name: NVIDIA H100 80GB HBM3, pci bus id: 0000:bb:00.0, compute capability: 9.0)
/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/keras/layers/core.py:1059: UserWarning: bpnet.model.arch is not loaded, but a Lambda layer uses it. It may cause errors.
  , UserWarning)
batch:   0%|          | 0/14 [00:00<?, ?it/s]2026-07-10 15:03:39.970026: I tensorflow/compiler/mlir/mlir_graph_optimization_pass.cc:116] None of the MLIR optimization passes are enabled (registered 2)
2026-07-10 15:03:39.970563: I tensorflow/core/platform/profile_utils/cpu_utils.cc:112] CPU Frequency: 2800000000 Hz
2026-07-10 15:03:40.616084: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublas.so.11
2026-07-10 15:05:15.466277: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublasLt.so.11
2026-07-10 15:05:15.467565: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudnn.so.8
2026-07-10 15:11:19.536615: W tensorflow/stream_executor/gpu/asm_compiler.cc:63] Running ptxas --version returned 256
2026-07-10 15:11:19.664124: W tensorflow/stream_executor/gpu/redzone_allocator.cc:314] Internal: ptxas exited with non-zero error code 256, output: 
Relying on driver to perform ptx compilation. 
Modify $PATH to customize ptxas location.
This message will be only logged once.
/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/keras/engine/functional.py:595: UserWarning: Input dict contained keys ['coordinates', 'true_profiles', 'true_logcounts', 'rev_comp'] which did not match any model input. They will be ignored by the model.
  [n for n in tensors.keys() if n not in ref_input_names])
batch:   7%|▋         | 1/14 [08:08<1:45:56, 488.98s/it]batch:  14%|█▍        | 2/14 [08:09<40:17, 201.48s/it]  batch:  21%|██▏       | 3/14 [08:09<20:05, 109.58s/it]batch:  29%|██▊       | 4/14 [08:09<11:04, 66.41s/it] batch:  36%|███▌      | 5/14 [08:09<06:22, 42.54s/it]batch:  43%|████▎     | 6/14 [08:10<03:45, 28.16s/it]batch:  50%|█████     | 7/14 [08:10<02:13, 19.02s/it]batch:  57%|█████▋    | 8/14 [08:10<01:18, 13.04s/it]batch:  64%|██████▍   | 9/14 [08:10<00:45,  9.03s/it]batch:  71%|███████▏  | 10/14 [08:11<00:25,  6.31s/it]batch:  79%|███████▊  | 11/14 [08:11<00:13,  4.45s/it]batch:  86%|████████▌ | 12/14 [08:11<00:06,  3.16s/it]batch:  93%|█████████▎| 13/14 [08:11<00:02,  2.27s/it]batch: 100%|██████████| 14/14 [08:11<00:00,  1.65s/it]batch: 100%|██████████| 14/14 [08:11<00:00, 35.13s/it]
  0%|          | 0/886 [00:00<?, ?it/s] 15%|█▍        | 131/886 [00:00<00:00, 1289.33it/s] 29%|██▉       | 260/886 [00:00<00:00, 1199.23it/s] 43%|████▎     | 384/886 [00:00<00:00, 1216.96it/s] 57%|█████▋    | 508/886 [00:00<00:00, 1221.06it/s] 71%|███████   | 631/886 [00:00<00:00, 1218.78it/s] 85%|████████▍ | 753/886 [00:00<00:00, 1211.41it/s] 99%|█████████▉| 875/886 [00:00<00:00, 1195.48it/s]100%|██████████| 886/886 [00:00<00:00, 1209.02it/s]
  0%|          | 0/886 [00:00<?, ?it/s] 15%|█▍        | 131/886 [00:00<00:00, 1293.55it/s] 29%|██▉       | 261/886 [00:00<00:00, 1283.57it/s] 44%|████▍     | 390/886 [00:00<00:00, 1271.32it/s] 58%|█████▊    | 518/886 [00:00<00:00, 1257.57it/s] 73%|███████▎  | 644/886 [00:00<00:00, 1188.94it/s] 86%|████████▌ | 764/886 [00:00<00:00, 1191.90it/s]100%|█████████▉| 884/886 [00:00<00:00, 1185.63it/s]100%|██████████| 886/886 [00:00<00:00, 1212.81it/s]
2026-07-10 15:12:54.163957: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 15:13:54.941576: I tensorflow/compiler/jit/xla_cpu_device.cc:41] Not creating XLA devices, tf_xla_enable_xla_devices not set
2026-07-10 15:13:54.955911: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcuda.so.1
2026-07-10 15:13:55.471422: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1720] Found device 0 with properties: 
pciBusID: 0000:bb:00.0 name: NVIDIA H100 80GB HBM3 computeCapability: 9.0
coreClock: 1.98GHz coreCount: 132 deviceMemorySize: 79.10GiB deviceMemoryBandwidth: 3.05TiB/s
2026-07-10 15:13:55.471529: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 15:13:55.494463: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublas.so.11
2026-07-10 15:13:55.494592: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublasLt.so.11
2026-07-10 15:13:55.505762: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcufft.so.10
2026-07-10 15:13:55.516126: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcurand.so.10
2026-07-10 15:13:55.529009: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusolver.so.10
2026-07-10 15:13:55.539919: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusparse.so.11
2026-07-10 15:13:55.550375: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudnn.so.8
2026-07-10 15:13:55.555729: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1862] Adding visible gpu devices: 0
2026-07-10 15:13:55.556130: I tensorflow/core/platform/cpu_feature_guard.cc:142] This TensorFlow binary is optimized with oneAPI Deep Neural Network Library (oneDNN) to use the following CPU instructions in performance-critical operations:  AVX2 AVX512F FMA
To enable them in other operations, rebuild TensorFlow with the appropriate compiler flags.
2026-07-10 15:13:55.556205: I tensorflow/compiler/jit/xla_gpu_device.cc:99] Not creating XLA devices, tf_xla_enable_xla_devices not set
2026-07-10 15:13:55.558815: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1720] Found device 0 with properties: 
pciBusID: 0000:bb:00.0 name: NVIDIA H100 80GB HBM3 computeCapability: 9.0
coreClock: 1.98GHz coreCount: 132 deviceMemorySize: 79.10GiB deviceMemoryBandwidth: 3.05TiB/s
2026-07-10 15:13:55.558857: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 15:13:55.558872: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublas.so.11
2026-07-10 15:13:55.558884: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublasLt.so.11
2026-07-10 15:13:55.558896: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcufft.so.10
2026-07-10 15:13:55.558908: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcurand.so.10
2026-07-10 15:13:55.558920: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusolver.so.10
2026-07-10 15:13:55.558931: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusparse.so.11
2026-07-10 15:13:55.558943: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudnn.so.8
2026-07-10 15:13:55.565014: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1862] Adding visible gpu devices: 0
2026-07-10 15:13:55.565052: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 15:18:41.955200: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1261] Device interconnect StreamExecutor with strength 1 edge matrix:
2026-07-10 15:18:41.955303: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1267]      0 
2026-07-10 15:18:41.955313: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1280] 0:   N 
2026-07-10 15:18:41.960323: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1406] Created TensorFlow device (/job:localhost/replica:0/task:0/device:GPU:0 with 75496 MB memory) -> physical GPU (device: 0, name: NVIDIA H100 80GB HBM3, pci bus id: 0000:bb:00.0, compute capability: 9.0)
2026-07-10 15:18:42.005849: I tensorflow/compiler/mlir/mlir_graph_optimization_pass.cc:196] None of the MLIR optimization passes are enabled (registered 0 passes)
2026-07-10 15:18:42.077136: I tensorflow/core/platform/profile_utils/cpu_utils.cc:112] CPU Frequency: 2800000000 Hz
2026-07-10 15:18:43.859368: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublas.so.11
2026-07-10 15:21:34.256487: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublasLt.so.11
2026-07-10 15:21:34.263113: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudnn.so.8
2026-07-10 15:27:10.232564: W tensorflow/stream_executor/gpu/asm_compiler.cc:63] Running ptxas --version returned 256
2026-07-10 15:27:10.370655: W tensorflow/stream_executor/gpu/redzone_allocator.cc:314] Internal: ptxas exited with non-zero error code 256, output: 
Relying on driver to perform ptx compilation. 
Modify $PATH to customize ptxas location.
This message will be only logged once.
2026-07-10 15:27:12.555888: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:27:14.728617: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:27:16.631290: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:27:18.363364: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:27:20.679234: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:27:23.244202: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:27:26.257747: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:27:33.535186: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:27:38.752897: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:27:38.753626: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:27:38.753720: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
RuntimeError: module compiled against API version 0xe but this version of numpy is 0xd
/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/keras/layers/core.py:1059: UserWarning: bpnet.model.arch is not loaded, but a Lambda layer uses it. It may cause errors.
  , UserWarning)
Traceback (most recent call last):
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/client/session.py", line 1375, in _do_call
    return fn(*args)
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/client/session.py", line 1360, in _run_fn
    target_list, run_metadata)
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/client/session.py", line 1453, in _call_tf_sessionrun
    run_metadata)
tensorflow.python.framework.errors_impl.InternalError: 2 root error(s) found.
  (0) Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
	 [[{{node gradients/main_conv_0_relu/Relu_grad/Abs}}]]
	 [[gradients/main_logsumexp_counts_bias_0/ReduceLogSumExp/Sub_grad/Reshape/_65]]
  (1) Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
	 [[{{node gradients/main_conv_0_relu/Relu_grad/Abs}}]]
0 successful operations.
0 derived errors ignored.

During handling of the above exception, another exception occurred:

Traceback (most recent call last):
  File "/home/users/shouvikm/miniconda3/envs/bpnet/bin/bpnet-shap", line 8, in <module>
    sys.exit(shap_scores_main())
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/bpnet/cli/shap_scores.py", line 448, in shap_scores_main
    shap_scores(args, shap_scores_dir)
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/bpnet/cli/shap_scores.py", line 326, in shap_scores
    counts_shap_inputs, progress_message=100)
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py", line 294, in shap_values
    sample_phis = self.run(self.phi_symbolic(feature_ind), self.model_inputs, joint_input)
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py", line 322, in run
    return self.session.run(out, feed_dict)
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/client/session.py", line 968, in run
    run_metadata_ptr)
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/client/session.py", line 1191, in _run
    feed_dict_tensor, options, run_metadata)
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/client/session.py", line 1369, in _do_run
    run_metadata)
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/client/session.py", line 1394, in _do_call
    raise type(e)(node_def, op, message)
tensorflow.python.framework.errors_impl.InternalError: 2 root error(s) found.
  (0) Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
	 [[node gradients/main_conv_0_relu/Relu_grad/Abs (defined at /lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py:494) ]]
	 [[gradients/main_logsumexp_counts_bias_0/ReduceLogSumExp/Sub_grad/Reshape/_65]]
  (1) Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
	 [[node gradients/main_conv_0_relu/Relu_grad/Abs (defined at /lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py:494) ]]
0 successful operations.
0 derived errors ignored.

Errors may have originated from an input operation.
Input Source operations connected to node gradients/main_conv_0_relu/Relu_grad/Abs:
 gradients/main_conv_0_relu/Relu_grad/sub (defined at /lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py:489)

Input Source operations connected to node gradients/main_conv_0_relu/Relu_grad/Abs:
 gradients/main_conv_0_relu/Relu_grad/sub (defined at /lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py:489)

Original stack trace for 'gradients/main_conv_0_relu/Relu_grad/Abs':
  File "/bin/bpnet-shap", line 8, in <module>
    sys.exit(shap_scores_main())
  File "/lib/python3.7/site-packages/bpnet/cli/shap_scores.py", line 448, in shap_scores_main
    shap_scores(args, shap_scores_dir)
  File "/lib/python3.7/site-packages/bpnet/cli/shap_scores.py", line 326, in shap_scores
    counts_shap_inputs, progress_message=100)
  File "/lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py", line 294, in shap_values
    sample_phis = self.run(self.phi_symbolic(feature_ind), self.model_inputs, joint_input)
  File "/lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py", line 229, in phi_symbolic
    self.phi_symbolics[i] = tf.gradients(out, self.model_inputs)
  File "/lib/python3.7/site-packages/tensorflow/python/ops/gradients_impl.py", line 318, in gradients_v2
    unconnected_gradients)
  File "/lib/python3.7/site-packages/tensorflow/python/ops/gradients_util.py", line 684, in _GradientsHelper
    lambda: grad_fn(op, *out_grads))
  File "/lib/python3.7/site-packages/tensorflow/python/ops/gradients_util.py", line 340, in _MaybeCompile
    return grad_fn()  # Exit early
  File "/lib/python3.7/site-packages/tensorflow/python/ops/gradients_util.py", line 684, in <lambda>
    lambda: grad_fn(op, *out_grads))
  File "/lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py", line 327, in custom_grad
    return op_handlers[op.type](self, op, *grads)
  File "/lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py", line 477, in handler
    return nonlinearity_1d_handler(input_ind, explainer, op, *grads)
  File "/lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py", line 494, in nonlinearity_1d_handler
    tf.tile(tf.abs(delta_in0), dup0) < 1e-6,
  File "/lib/python3.7/site-packages/tensorflow/python/util/dispatch.py", line 201, in wrapper
    return target(*args, **kwargs)
  File "/lib/python3.7/site-packages/tensorflow/python/ops/math_ops.py", line 401, in abs
    return gen_math_ops._abs(x, name=name)
  File "/lib/python3.7/site-packages/tensorflow/python/ops/gen_math_ops.py", line 56, in _abs
    "Abs", x=x, name=name)
  File "/lib/python3.7/site-packages/tensorflow/python/framework/op_def_library.py", line 750, in _apply_op_helper
    attrs=attr_protos, op_def=op_def)
  File "/lib/python3.7/site-packages/tensorflow/python/framework/ops.py", line 3536, in _create_op_internal
    op_def=op_def)
  File "/lib/python3.7/site-packages/tensorflow/python/framework/ops.py", line 1990, in __init__
    self._traceback = tf_stack.extract_stack()

...which was originally created as op 'main_conv_0_relu/Relu', defined at:
  File "/bin/bpnet-shap", line 8, in <module>
    sys.exit(shap_scores_main())
[elided 0 identical lines from previous traceback]
  File "/lib/python3.7/site-packages/bpnet/cli/shap_scores.py", line 448, in shap_scores_main
    shap_scores(args, shap_scores_dir)
  File "/lib/python3.7/site-packages/bpnet/cli/shap_scores.py", line 96, in shap_scores
    model = load_model(args.model, compile=False)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/saving/save.py", line 212, in load_model
    return saved_model_load.load(filepath, compile, options)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/saving/saved_model/load.py", line 138, in load
    keras_loader.load_layers(compile=compile)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/saving/saved_model/load.py", line 376, in load_layers
    node_metadata.metadata)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/saving/saved_model/load.py", line 417, in _load_layer
    obj, setter = self._revive_from_config(identifier, metadata, node_id)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/saving/saved_model/load.py", line 435, in _revive_from_config
    self._revive_layer_from_config(metadata, node_id))
  File "/lib/python3.7/site-packages/tensorflow/python/keras/saving/saved_model/load.py", line 495, in _revive_layer_from_config
    generic_utils.serialize_keras_class_and_config(class_name, config))
  File "/lib/python3.7/site-packages/tensorflow/python/keras/layers/serialization.py", line 177, in deserialize
    printable_module_name='layer')
  File "/lib/python3.7/site-packages/tensorflow/python/keras/utils/generic_utils.py", line 358, in deserialize_keras_object
    list(custom_objects.items())))
  File "/lib/python3.7/site-packages/tensorflow/python/keras/engine/training.py", line 2262, in from_config
    config, custom_objects=custom_objects)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/engine/functional.py", line 669, in from_config
    config, custom_objects)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/engine/functional.py", line 1285, in reconstruct_from_config
    process_node(layer, node_data)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/engine/functional.py", line 1233, in process_node
    output_tensors = layer(input_tensors, **kwargs)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/engine/base_layer_v1.py", line 786, in __call__
    outputs = call_fn(cast_inputs, *args, **kwargs)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/layers/advanced_activations.py", line 420, in call
    threshold=self.threshold)
  File "/lib/python3.7/site-packages/tensorflow/python/util/dispatch.py", line 201, in wrapper
    return target(*args, **kwargs)

