Lmod Warning:
-------------------------------------------------------------------------------
The following dependent module(s) are not currently loaded: curl/8.4.0
(required by: htslib/1.16)
-------------------------------------------------------------------------------



Lmod Warning:
-------------------------------------------------------------------------------
The following dependent module(s) are not currently loaded: curl/8.17.0
(required by: ucsc-utils/v489), openssl/3.0.7 (required by: curl/8.4.0)
-------------------------------------------------------------------------------




The following have been reloaded with a version change:
  1) curl/8.17.0 => curl/8.4.0

Lmod Warning:
-------------------------------------------------------------------------------
The following dependent module(s) are not currently loaded: curl/8.4.0
(required by: htslib/1.16)
-------------------------------------------------------------------------------




The following have been reloaded with a version change:
  1) curl/8.4.0 => curl/8.17.0

2026-07-10 14:45:04.461683: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 14:45:35.470762: I tensorflow/compiler/jit/xla_cpu_device.cc:41] Not creating XLA devices, tf_xla_enable_xla_devices not set
2026-07-10 14:45:35.472080: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcuda.so.1
2026-07-10 14:45:35.515224: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1720] Found device 0 with properties: 
pciBusID: 0000:5d:00.0 name: NVIDIA H100 80GB HBM3 computeCapability: 9.0
coreClock: 1.98GHz coreCount: 132 deviceMemorySize: 79.10GiB deviceMemoryBandwidth: 3.05TiB/s
2026-07-10 14:45:35.515319: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 14:45:35.521268: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublas.so.11
2026-07-10 14:45:35.521373: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublasLt.so.11
2026-07-10 14:45:35.523773: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcufft.so.10
2026-07-10 14:45:35.525443: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcurand.so.10
2026-07-10 14:45:35.529364: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusolver.so.10
2026-07-10 14:45:35.531245: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusparse.so.11
2026-07-10 14:45:35.533148: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudnn.so.8
2026-07-10 14:45:35.542426: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1862] Adding visible gpu devices: 0
2026-07-10 14:45:35.542872: I tensorflow/core/platform/cpu_feature_guard.cc:142] This TensorFlow binary is optimized with oneAPI Deep Neural Network Library (oneDNN) to use the following CPU instructions in performance-critical operations:  AVX2 AVX512F FMA
To enable them in other operations, rebuild TensorFlow with the appropriate compiler flags.
2026-07-10 14:45:35.542936: I tensorflow/compiler/jit/xla_gpu_device.cc:99] Not creating XLA devices, tf_xla_enable_xla_devices not set
2026-07-10 14:45:35.545557: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1720] Found device 0 with properties: 
pciBusID: 0000:5d:00.0 name: NVIDIA H100 80GB HBM3 computeCapability: 9.0
coreClock: 1.98GHz coreCount: 132 deviceMemorySize: 79.10GiB deviceMemoryBandwidth: 3.05TiB/s
2026-07-10 14:45:35.545581: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 14:45:35.545595: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublas.so.11
2026-07-10 14:45:35.545604: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublasLt.so.11
2026-07-10 14:45:35.545631: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcufft.so.10
2026-07-10 14:45:35.545640: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcurand.so.10
2026-07-10 14:45:35.545648: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusolver.so.10
2026-07-10 14:45:35.545657: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusparse.so.11
2026-07-10 14:45:35.545666: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudnn.so.8
2026-07-10 14:45:35.550234: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1862] Adding visible gpu devices: 0
2026-07-10 14:45:35.550266: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 14:47:40.985934: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1261] Device interconnect StreamExecutor with strength 1 edge matrix:
2026-07-10 14:47:40.986116: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1267]      0 
2026-07-10 14:47:40.986141: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1280] 0:   N 
2026-07-10 14:47:40.991199: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1406] Created TensorFlow device (/job:localhost/replica:0/task:0/device:GPU:0 with 75496 MB memory) -> physical GPU (device: 0, name: NVIDIA H100 80GB HBM3, pci bus id: 0000:5d:00.0, compute capability: 9.0)
2026-07-10 14:47:54.953124: I tensorflow/compiler/mlir/mlir_graph_optimization_pass.cc:116] None of the MLIR optimization passes are enabled (registered 2)
2026-07-10 14:47:54.953619: I tensorflow/core/platform/profile_utils/cpu_utils.cc:112] CPU Frequency: 2800000000 Hz
2026-07-10 14:47:58.131826: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublas.so.11
2026-07-10 14:49:39.436579: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublasLt.so.11
2026-07-10 14:49:39.441765: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudnn.so.8
2026-07-10 14:55:10.284184: W tensorflow/stream_executor/gpu/asm_compiler.cc:63] Running ptxas --version returned 256
2026-07-10 14:55:10.435853: W tensorflow/stream_executor/gpu/redzone_allocator.cc:314] Internal: ptxas exited with non-zero error code 256, output: 
Relying on driver to perform ptx compilation. 
Modify $PATH to customize ptxas location.
This message will be only logged once.
2026-07-10 14:55:33.075687: I tensorflow/stream_executor/cuda/cuda_blas.cc:1838] TensorFloat-32 will be used for the matrix multiplication. This will only be logged once.
/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/keras/engine/functional.py:595: UserWarning: Input dict contained keys ['coordinates', 'jitters', 'index', 'status', 'rev_comp'] which did not match any model input. They will be ignored by the model.
  [n for n in tensors.keys() if n not in ref_input_names])
/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/keras/engine/functional.py:595: UserWarning: Input dict contained keys ['coordinates'] which did not match any model input. They will be ignored by the model.
  [n for n in tensors.keys() if n not in ref_input_names])
2026-07-10 15:15:32.649612: W tensorflow/python/util/util.cc:348] Sets are not currently considered sequences, but this may change in the future, so consider avoiding using them.
2026-07-10 15:15:47.333139: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 15:17:10.016504: I tensorflow/compiler/jit/xla_cpu_device.cc:41] Not creating XLA devices, tf_xla_enable_xla_devices not set
2026-07-10 15:17:10.047958: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcuda.so.1
2026-07-10 15:17:10.519089: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1720] Found device 0 with properties: 
pciBusID: 0000:5d:00.0 name: NVIDIA H100 80GB HBM3 computeCapability: 9.0
coreClock: 1.98GHz coreCount: 132 deviceMemorySize: 79.10GiB deviceMemoryBandwidth: 3.05TiB/s
2026-07-10 15:17:10.519180: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 15:17:10.528573: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublas.so.11
2026-07-10 15:17:10.528692: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublasLt.so.11
2026-07-10 15:17:10.533958: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcufft.so.10
2026-07-10 15:17:10.538252: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcurand.so.10
2026-07-10 15:17:10.545726: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusolver.so.10
2026-07-10 15:17:10.550767: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusparse.so.11
2026-07-10 15:17:10.554845: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudnn.so.8
2026-07-10 15:17:10.566540: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1862] Adding visible gpu devices: 0
2026-07-10 15:17:10.566890: I tensorflow/core/platform/cpu_feature_guard.cc:142] This TensorFlow binary is optimized with oneAPI Deep Neural Network Library (oneDNN) to use the following CPU instructions in performance-critical operations:  AVX2 AVX512F FMA
To enable them in other operations, rebuild TensorFlow with the appropriate compiler flags.
2026-07-10 15:17:10.566951: I tensorflow/compiler/jit/xla_gpu_device.cc:99] Not creating XLA devices, tf_xla_enable_xla_devices not set
2026-07-10 15:17:10.569546: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1720] Found device 0 with properties: 
pciBusID: 0000:5d:00.0 name: NVIDIA H100 80GB HBM3 computeCapability: 9.0
coreClock: 1.98GHz coreCount: 132 deviceMemorySize: 79.10GiB deviceMemoryBandwidth: 3.05TiB/s
2026-07-10 15:17:10.569570: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 15:17:10.569584: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublas.so.11
2026-07-10 15:17:10.569594: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublasLt.so.11
2026-07-10 15:17:10.569603: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcufft.so.10
2026-07-10 15:17:10.569613: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcurand.so.10
2026-07-10 15:17:10.569622: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusolver.so.10
2026-07-10 15:17:10.569631: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusparse.so.11
2026-07-10 15:17:10.569641: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudnn.so.8
2026-07-10 15:17:10.574199: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1862] Adding visible gpu devices: 0
2026-07-10 15:17:10.574226: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 15:19:08.427091: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1261] Device interconnect StreamExecutor with strength 1 edge matrix:
2026-07-10 15:19:08.427204: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1267]      0 
2026-07-10 15:19:08.427222: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1280] 0:   N 
2026-07-10 15:19:08.432386: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1406] Created TensorFlow device (/job:localhost/replica:0/task:0/device:GPU:0 with 75496 MB memory) -> physical GPU (device: 0, name: NVIDIA H100 80GB HBM3, pci bus id: 0000:5d:00.0, compute capability: 9.0)
/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/keras/layers/core.py:1059: UserWarning: bpnet.model.arch is not loaded, but a Lambda layer uses it. It may cause errors.
  , UserWarning)
batch:   0%|          | 0/82 [00:00<?, ?it/s]2026-07-10 15:19:10.690881: I tensorflow/compiler/mlir/mlir_graph_optimization_pass.cc:116] None of the MLIR optimization passes are enabled (registered 2)
2026-07-10 15:19:10.691405: I tensorflow/core/platform/profile_utils/cpu_utils.cc:112] CPU Frequency: 2800000000 Hz
2026-07-10 15:19:11.406594: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublas.so.11
2026-07-10 15:21:50.450465: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublasLt.so.11
2026-07-10 15:21:50.452159: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudnn.so.8
2026-07-10 15:26:57.582127: W tensorflow/stream_executor/gpu/asm_compiler.cc:63] Running ptxas --version returned 256
2026-07-10 15:26:57.736898: W tensorflow/stream_executor/gpu/redzone_allocator.cc:314] Internal: ptxas exited with non-zero error code 256, output: 
Relying on driver to perform ptx compilation. 
Modify $PATH to customize ptxas location.
This message will be only logged once.
/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/keras/engine/functional.py:595: UserWarning: Input dict contained keys ['coordinates', 'true_profiles', 'true_logcounts', 'rev_comp'] which did not match any model input. They will be ignored by the model.
  [n for n in tensors.keys() if n not in ref_input_names])
batch:   1%|          | 1/82 [08:22<11:17:42, 502.01s/it]batch:   2%|▏         | 2/82 [08:22<4:35:48, 206.85s/it] batch:   4%|▎         | 3/82 [08:22<2:28:09, 112.52s/it]batch:   5%|▍         | 4/82 [08:22<1:28:38, 68.19s/it] batch:   6%|▌         | 5/82 [08:22<56:03, 43.68s/it]  batch:   7%|▋         | 6/82 [08:23<36:37, 28.91s/it]batch:   9%|▊         | 7/82 [08:23<24:24, 19.53s/it]batch:  10%|▉         | 8/82 [08:23<16:30, 13.39s/it]batch:  11%|█         | 9/82 [08:23<11:17,  9.28s/it]batch:  12%|█▏        | 10/82 [08:24<07:47,  6.50s/it]batch:  13%|█▎        | 11/82 [08:24<05:25,  4.58s/it]batch:  15%|█▍        | 12/82 [08:24<03:47,  3.26s/it]batch:  16%|█▌        | 13/82 [08:24<02:41,  2.34s/it]batch:  17%|█▋        | 14/82 [08:25<01:55,  1.70s/it]batch:  18%|█▊        | 15/82 [08:25<01:24,  1.26s/it]batch:  20%|█▉        | 16/82 [08:25<01:02,  1.06it/s]batch:  21%|██        | 17/82 [08:25<00:48,  1.35it/s]batch:  22%|██▏       | 18/82 [08:26<00:37,  1.71it/s]batch:  23%|██▎       | 19/82 [08:26<00:30,  2.05it/s]batch:  24%|██▍       | 20/82 [08:26<00:25,  2.44it/s]batch:  26%|██▌       | 21/82 [08:26<00:21,  2.82it/s]batch:  27%|██▋       | 22/82 [08:26<00:19,  3.15it/s]batch:  28%|██▊       | 23/82 [08:27<00:17,  3.42it/s]batch:  29%|██▉       | 24/82 [08:27<00:25,  2.30it/s]batch:  30%|███       | 25/82 [08:28<00:21,  2.68it/s]batch:  32%|███▏      | 26/82 [08:28<00:18,  3.03it/s]batch:  33%|███▎      | 27/82 [08:28<00:16,  3.34it/s]batch:  34%|███▍      | 28/82 [08:28<00:14,  3.60it/s]batch:  35%|███▌      | 29/82 [08:29<00:14,  3.77it/s]batch:  37%|███▋      | 30/82 [08:29<00:13,  3.79it/s]batch:  38%|███▊      | 31/82 [08:29<00:12,  3.93it/s]batch:  39%|███▉      | 32/82 [08:29<00:12,  4.03it/s]batch:  40%|████      | 33/82 [08:30<00:11,  4.14it/s]batch:  41%|████▏     | 34/82 [08:30<00:11,  4.15it/s]batch:  43%|████▎     | 35/82 [08:30<00:11,  4.22it/s]batch:  44%|████▍     | 36/82 [08:30<00:11,  4.10it/s]batch:  45%|████▌     | 37/82 [08:31<00:10,  4.09it/s]batch:  46%|████▋     | 38/82 [08:31<00:10,  4.16it/s]batch:  48%|████▊     | 39/82 [08:31<00:10,  4.22it/s]batch:  49%|████▉     | 40/82 [08:31<00:09,  4.28it/s]batch:  50%|█████     | 41/82 [08:31<00:09,  4.31it/s]batch:  51%|█████     | 42/82 [08:32<00:09,  4.34it/s]batch:  52%|█████▏    | 43/82 [08:32<00:09,  4.21it/s]batch:  54%|█████▎    | 44/82 [08:32<00:08,  4.25it/s]batch:  55%|█████▍    | 45/82 [08:33<00:13,  2.81it/s]batch:  56%|█████▌    | 46/82 [08:33<00:11,  3.14it/s]batch:  57%|█████▋    | 47/82 [08:33<00:10,  3.44it/s]batch:  59%|█████▊    | 48/82 [08:34<00:09,  3.58it/s]batch:  60%|█████▉    | 49/82 [08:34<00:08,  3.79it/s]batch:  61%|██████    | 50/82 [08:34<00:08,  3.92it/s]batch:  62%|██████▏   | 51/82 [08:34<00:07,  4.05it/s]batch:  63%|██████▎   | 52/82 [08:34<00:07,  4.12it/s]batch:  65%|██████▍   | 53/82 [08:35<00:06,  4.18it/s]batch:  66%|██████▌   | 54/82 [08:35<00:06,  4.26it/s]batch:  67%|██████▋   | 55/82 [08:35<00:06,  4.30it/s]batch:  68%|██████▊   | 56/82 [08:35<00:06,  4.20it/s]batch:  70%|██████▉   | 57/82 [08:36<00:06,  4.13it/s]batch:  71%|███████   | 58/82 [08:36<00:05,  4.20it/s]batch:  72%|███████▏  | 59/82 [08:36<00:05,  4.27it/s]batch:  73%|███████▎  | 60/82 [08:36<00:05,  4.29it/s]batch:  74%|███████▍  | 61/82 [08:37<00:04,  4.32it/s]batch:  76%|███████▌  | 62/82 [08:37<00:04,  4.34it/s]batch:  77%|███████▋  | 63/82 [08:37<00:04,  4.15it/s]batch:  78%|███████▊  | 64/82 [08:37<00:04,  4.16it/s]batch:  79%|███████▉  | 65/82 [08:38<00:04,  4.21it/s]batch:  80%|████████  | 66/82 [08:38<00:03,  4.27it/s]batch:  82%|████████▏ | 67/82 [08:38<00:04,  3.66it/s]batch:  83%|████████▎ | 68/82 [08:38<00:03,  3.84it/s]batch:  84%|████████▍ | 69/82 [08:39<00:03,  3.80it/s]batch:  85%|████████▌ | 70/82 [08:39<00:03,  3.96it/s]batch:  87%|████████▋ | 71/82 [08:39<00:02,  4.07it/s]batch:  88%|████████▊ | 72/82 [08:39<00:02,  4.17it/s]batch:  89%|████████▉ | 73/82 [08:40<00:02,  4.22it/s]batch:  90%|█████████ | 74/82 [08:40<00:01,  4.28it/s]batch:  91%|█████████▏| 75/82 [08:40<00:01,  4.19it/s]batch:  93%|█████████▎| 76/82 [08:40<00:01,  4.11it/s]batch:  94%|█████████▍| 77/82 [08:40<00:01,  4.19it/s]batch:  95%|█████████▌| 78/82 [08:41<00:00,  4.25it/s]batch:  96%|█████████▋| 79/82 [08:41<00:00,  4.32it/s]batch:  98%|█████████▊| 80/82 [08:41<00:00,  4.32it/s]batch:  99%|█████████▉| 81/82 [08:41<00:00,  4.85it/s]batch: 100%|██████████| 82/82 [08:41<00:00,  6.36s/it]
  0%|          | 0/5156 [00:00<?, ?it/s]  2%|▏         | 110/5156 [00:00<00:04, 1061.11it/s]  4%|▍         | 231/5156 [00:00<00:04, 1146.26it/s]  7%|▋         | 346/5156 [00:00<00:04, 1093.17it/s]  9%|▉         | 472/5156 [00:00<00:04, 1155.81it/s] 12%|█▏        | 600/5156 [00:00<00:03, 1196.31it/s] 14%|█▍        | 728/5156 [00:00<00:03, 1220.00it/s] 17%|█▋        | 857/5156 [00:00<00:03, 1237.30it/s] 19%|█▉        | 985/5156 [00:00<00:03, 1244.83it/s] 22%|██▏       | 1113/5156 [00:00<00:03, 1250.75it/s] 24%|██▍       | 1239/5156 [00:01<00:03, 1253.48it/s] 26%|██▋       | 1365/5156 [00:01<00:03, 1252.58it/s] 29%|██▉       | 1491/5156 [00:01<00:02, 1253.19it/s] 31%|███▏      | 1617/5156 [00:01<00:02, 1246.82it/s] 34%|███▍      | 1742/5156 [00:01<00:02, 1237.08it/s] 36%|███▌      | 1866/5156 [00:01<00:02, 1235.68it/s] 39%|███▊      | 1990/5156 [00:01<00:02, 1204.83it/s] 41%|████      | 2111/5156 [00:01<00:02, 1151.45it/s] 43%|████▎     | 2233/5156 [00:01<00:02, 1165.54it/s] 46%|████▌     | 2354/5156 [00:01<00:02, 1177.48it/s] 48%|████▊     | 2475/5156 [00:02<00:02, 1182.42it/s] 50%|█████     | 2596/5156 [00:02<00:02, 1188.26it/s] 53%|█████▎    | 2716/5156 [00:02<00:02, 1190.34it/s] 55%|█████▌    | 2836/5156 [00:02<00:01, 1192.05it/s] 57%|█████▋    | 2956/5156 [00:02<00:01, 1194.20it/s] 60%|█████▉    | 3076/5156 [00:02<00:01, 1189.77it/s] 62%|██████▏   | 3196/5156 [00:02<00:01, 1187.59it/s] 64%|██████▍   | 3315/5156 [00:02<00:01, 1179.41it/s] 67%|██████▋   | 3433/5156 [00:02<00:01, 1177.77it/s] 69%|██████▉   | 3551/5156 [00:02<00:01, 1174.19it/s] 71%|███████   | 3669/5156 [00:03<00:01, 1163.42it/s] 73%|███████▎  | 3786/5156 [00:03<00:01, 1163.25it/s] 76%|███████▌  | 3903/5156 [00:03<00:01, 1159.82it/s] 78%|███████▊  | 4019/5156 [00:03<00:00, 1158.34it/s] 80%|████████  | 4135/5156 [00:03<00:00, 1152.83it/s] 82%|████████▏ | 4251/5156 [00:03<00:00, 1150.11it/s] 85%|████████▍ | 4367/5156 [00:03<00:00, 1151.77it/s] 87%|████████▋ | 4483/5156 [00:03<00:00, 1145.35it/s] 89%|████████▉ | 4598/5156 [00:03<00:00, 1146.13it/s] 91%|█████████▏| 4713/5156 [00:03<00:00, 1144.28it/s] 94%|█████████▎| 4828/5156 [00:04<00:00, 1141.48it/s] 96%|█████████▌| 4943/5156 [00:04<00:00, 1135.02it/s] 98%|█████████▊| 5057/5156 [00:04<00:00, 1136.24it/s]100%|██████████| 5156/5156 [00:04<00:00, 1179.28it/s]
  0%|          | 0/5156 [00:00<?, ?it/s]  3%|▎         | 129/5156 [00:00<00:03, 1286.40it/s]  5%|▌         | 258/5156 [00:00<00:03, 1288.19it/s]  8%|▊         | 388/5156 [00:00<00:03, 1284.92it/s] 10%|█         | 518/5156 [00:00<00:03, 1272.02it/s] 13%|█▎        | 646/5156 [00:00<00:03, 1176.25it/s] 15%|█▌        | 775/5156 [00:00<00:03, 1206.62it/s] 18%|█▊        | 903/5156 [00:00<00:03, 1223.00it/s] 20%|█▉        | 1031/5156 [00:00<00:03, 1240.44it/s] 22%|██▏       | 1156/5156 [00:00<00:03, 1167.37it/s] 25%|██▍       | 1281/5156 [00:01<00:03, 1190.82it/s] 27%|██▋       | 1406/5156 [00:01<00:03, 1207.78it/s] 30%|██▉       | 1530/5156 [00:01<00:02, 1214.40it/s] 32%|███▏      | 1654/5156 [00:01<00:02, 1219.85it/s] 34%|███▍      | 1777/5156 [00:01<00:02, 1218.37it/s] 37%|███▋      | 1900/5156 [00:01<00:02, 1220.60it/s] 39%|███▉      | 2023/5156 [00:01<00:02, 1219.06it/s] 42%|████▏     | 2146/5156 [00:01<00:02, 1219.83it/s] 44%|████▍     | 2269/5156 [00:01<00:02, 1211.18it/s] 46%|████▋     | 2391/5156 [00:01<00:02, 1210.82it/s] 49%|████▊     | 2513/5156 [00:02<00:02, 1207.94it/s] 51%|█████     | 2634/5156 [00:02<00:02, 1200.33it/s] 53%|█████▎    | 2755/5156 [00:02<00:01, 1200.76it/s] 56%|█████▌    | 2876/5156 [00:02<00:01, 1199.23it/s] 58%|█████▊    | 2996/5156 [00:02<00:01, 1196.29it/s] 60%|██████    | 3116/5156 [00:02<00:01, 1150.62it/s] 63%|██████▎   | 3232/5156 [00:02<00:01, 1131.03it/s] 65%|██████▍   | 3350/5156 [00:02<00:01, 1143.02it/s] 67%|██████▋   | 3465/5156 [00:02<00:01, 1078.13it/s] 69%|██████▉   | 3582/5156 [00:03<00:01, 1098.71it/s] 72%|███████▏  | 3699/5156 [00:03<00:01, 1115.78it/s] 74%|███████▍  | 3816/5156 [00:03<00:01, 1130.66it/s] 76%|███████▋  | 3932/5156 [00:03<00:01, 1138.73it/s] 78%|███████▊  | 4047/5156 [00:03<00:00, 1135.24it/s] 81%|████████  | 4162/5156 [00:03<00:00, 1139.39it/s] 83%|████████▎ | 4278/5156 [00:03<00:00, 1141.20it/s] 85%|████████▌ | 4393/5156 [00:03<00:00, 1142.02it/s] 87%|████████▋ | 4508/5156 [00:03<00:00, 1142.47it/s] 90%|████████▉ | 4623/5156 [00:03<00:00, 1143.08it/s] 92%|█████████▏| 4738/5156 [00:04<00:00, 1141.87it/s] 94%|█████████▍| 4853/5156 [00:04<00:00, 1138.09it/s] 96%|█████████▋| 4967/5156 [00:04<00:00, 1055.00it/s] 98%|█████████▊| 5074/5156 [00:04<00:00, 1044.32it/s]100%|██████████| 5156/5156 [00:04<00:00, 1160.03it/s]
2026-07-10 15:28:24.871691: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 15:28:51.157094: I tensorflow/compiler/jit/xla_cpu_device.cc:41] Not creating XLA devices, tf_xla_enable_xla_devices not set
2026-07-10 15:28:51.158578: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcuda.so.1
2026-07-10 15:28:51.598009: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1720] Found device 0 with properties: 
pciBusID: 0000:5d:00.0 name: NVIDIA H100 80GB HBM3 computeCapability: 9.0
coreClock: 1.98GHz coreCount: 132 deviceMemorySize: 79.10GiB deviceMemoryBandwidth: 3.05TiB/s
2026-07-10 15:28:51.598121: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 15:28:51.606901: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublas.so.11
2026-07-10 15:28:51.607005: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublasLt.so.11
2026-07-10 15:28:51.611249: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcufft.so.10
2026-07-10 15:28:51.614380: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcurand.so.10
2026-07-10 15:28:51.620006: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusolver.so.10
2026-07-10 15:28:51.625537: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusparse.so.11
2026-07-10 15:28:51.629386: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudnn.so.8
2026-07-10 15:28:51.634687: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1862] Adding visible gpu devices: 0
2026-07-10 15:28:51.635250: I tensorflow/core/platform/cpu_feature_guard.cc:142] This TensorFlow binary is optimized with oneAPI Deep Neural Network Library (oneDNN) to use the following CPU instructions in performance-critical operations:  AVX2 AVX512F FMA
To enable them in other operations, rebuild TensorFlow with the appropriate compiler flags.
2026-07-10 15:28:51.635350: I tensorflow/compiler/jit/xla_gpu_device.cc:99] Not creating XLA devices, tf_xla_enable_xla_devices not set
2026-07-10 15:28:51.637973: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1720] Found device 0 with properties: 
pciBusID: 0000:5d:00.0 name: NVIDIA H100 80GB HBM3 computeCapability: 9.0
coreClock: 1.98GHz coreCount: 132 deviceMemorySize: 79.10GiB deviceMemoryBandwidth: 3.05TiB/s
2026-07-10 15:28:51.638024: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 15:28:51.638049: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublas.so.11
2026-07-10 15:28:51.638072: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublasLt.so.11
2026-07-10 15:28:51.638114: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcufft.so.10
2026-07-10 15:28:51.638138: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcurand.so.10
2026-07-10 15:28:51.638159: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusolver.so.10
2026-07-10 15:28:51.638181: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcusparse.so.11
2026-07-10 15:28:51.638202: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudnn.so.8
2026-07-10 15:28:51.643056: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1862] Adding visible gpu devices: 0
2026-07-10 15:28:51.643104: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudart.so.11.0
2026-07-10 15:29:26.262021: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1261] Device interconnect StreamExecutor with strength 1 edge matrix:
2026-07-10 15:29:26.262108: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1267]      0 
2026-07-10 15:29:26.262118: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1280] 0:   N 
2026-07-10 15:29:26.267195: I tensorflow/core/common_runtime/gpu/gpu_device.cc:1406] Created TensorFlow device (/job:localhost/replica:0/task:0/device:GPU:0 with 75496 MB memory) -> physical GPU (device: 0, name: NVIDIA H100 80GB HBM3, pci bus id: 0000:5d:00.0, compute capability: 9.0)
2026-07-10 15:29:26.309114: I tensorflow/compiler/mlir/mlir_graph_optimization_pass.cc:196] None of the MLIR optimization passes are enabled (registered 0 passes)
2026-07-10 15:29:26.325302: I tensorflow/core/platform/profile_utils/cpu_utils.cc:112] CPU Frequency: 2800000000 Hz
2026-07-10 15:29:30.537854: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublas.so.11
2026-07-10 15:30:38.823374: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcublasLt.so.11
2026-07-10 15:30:38.828908: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:30:38.829085: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:30:38.829235: I tensorflow/stream_executor/platform/default/dso_loader.cc:49] Successfully opened dynamic library libcudnn.so.8
2026-07-10 15:34:26.798379: W tensorflow/stream_executor/gpu/asm_compiler.cc:63] Running ptxas --version returned 256
2026-07-10 15:34:26.943665: W tensorflow/stream_executor/gpu/redzone_allocator.cc:314] Internal: ptxas exited with non-zero error code 256, output: 
Relying on driver to perform ptx compilation. 
Modify $PATH to customize ptxas location.
This message will be only logged once.
2026-07-10 15:34:28.661427: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:34:30.512118: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:34:32.534019: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:34:34.381089: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:34:36.756307: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:34:39.216440: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:34:41.695430: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:34:44.174504: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
2026-07-10 15:34:46.658376: W tensorflow/core/framework/op_kernel.cc:1763] OP_REQUIRES failed at cwise_op_gpu_base.cc:89 : Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
RuntimeError: module compiled against API version 0xe but this version of numpy is 0xd
/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/keras/layers/core.py:1059: UserWarning: bpnet.model.arch is not loaded, but a Lambda layer uses it. It may cause errors.
  , UserWarning)
Traceback (most recent call last):
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/client/session.py", line 1375, in _do_call
    return fn(*args)
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/client/session.py", line 1360, in _run_fn
    target_list, run_metadata)
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/client/session.py", line 1453, in _call_tf_sessionrun
    run_metadata)
tensorflow.python.framework.errors_impl.InternalError: 2 root error(s) found.
  (0) Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
	 [[{{node gradients/main_logsumexp_counts_bias_0/ReduceLogSumExp/Exp_grad/Abs}}]]
	 [[gradients/main_logsumexp_counts_bias_0/ReduceLogSumExp/Sub_grad/Reshape/_65]]
  (1) Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
	 [[{{node gradients/main_logsumexp_counts_bias_0/ReduceLogSumExp/Exp_grad/Abs}}]]
0 successful operations.
0 derived errors ignored.

During handling of the above exception, another exception occurred:

Traceback (most recent call last):
  File "/home/users/shouvikm/miniconda3/envs/bpnet/bin/bpnet-shap", line 8, in <module>
    sys.exit(shap_scores_main())
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/bpnet/cli/shap_scores.py", line 448, in shap_scores_main
    shap_scores(args, shap_scores_dir)
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/bpnet/cli/shap_scores.py", line 326, in shap_scores
    counts_shap_inputs, progress_message=100)
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py", line 294, in shap_values
    sample_phis = self.run(self.phi_symbolic(feature_ind), self.model_inputs, joint_input)
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py", line 322, in run
    return self.session.run(out, feed_dict)
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/client/session.py", line 968, in run
    run_metadata_ptr)
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/client/session.py", line 1191, in _run
    feed_dict_tensor, options, run_metadata)
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/client/session.py", line 1369, in _do_run
    run_metadata)
  File "/home/users/shouvikm/miniconda3/envs/bpnet/lib/python3.7/site-packages/tensorflow/python/client/session.py", line 1394, in _do_call
    raise type(e)(node_def, op, message)
tensorflow.python.framework.errors_impl.InternalError: 2 root error(s) found.
  (0) Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
	 [[node gradients/main_logsumexp_counts_bias_0/ReduceLogSumExp/Exp_grad/Abs (defined at /lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py:494) ]]
	 [[gradients/main_logsumexp_counts_bias_0/ReduceLogSumExp/Sub_grad/Reshape/_65]]
  (1) Internal: Failed to load in-memory CUBIN: CUDA_ERROR_NO_BINARY_FOR_GPU: no kernel image is available for execution on the device
	 [[node gradients/main_logsumexp_counts_bias_0/ReduceLogSumExp/Exp_grad/Abs (defined at /lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py:494) ]]
0 successful operations.
0 derived errors ignored.

Errors may have originated from an input operation.
Input Source operations connected to node gradients/main_logsumexp_counts_bias_0/ReduceLogSumExp/Exp_grad/Abs:
 gradients/main_logsumexp_counts_bias_0/ReduceLogSumExp/Exp_grad/sub (defined at /lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py:489)

Input Source operations connected to node gradients/main_logsumexp_counts_bias_0/ReduceLogSumExp/Exp_grad/Abs:
 gradients/main_logsumexp_counts_bias_0/ReduceLogSumExp/Exp_grad/sub (defined at /lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py:489)

Original stack trace for 'gradients/main_logsumexp_counts_bias_0/ReduceLogSumExp/Exp_grad/Abs':
  File "/bin/bpnet-shap", line 8, in <module>
    sys.exit(shap_scores_main())
  File "/lib/python3.7/site-packages/bpnet/cli/shap_scores.py", line 448, in shap_scores_main
    shap_scores(args, shap_scores_dir)
  File "/lib/python3.7/site-packages/bpnet/cli/shap_scores.py", line 326, in shap_scores
    counts_shap_inputs, progress_message=100)
  File "/lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py", line 294, in shap_values
    sample_phis = self.run(self.phi_symbolic(feature_ind), self.model_inputs, joint_input)
  File "/lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py", line 229, in phi_symbolic
    self.phi_symbolics[i] = tf.gradients(out, self.model_inputs)
  File "/lib/python3.7/site-packages/tensorflow/python/ops/gradients_impl.py", line 318, in gradients_v2
    unconnected_gradients)
  File "/lib/python3.7/site-packages/tensorflow/python/ops/gradients_util.py", line 684, in _GradientsHelper
    lambda: grad_fn(op, *out_grads))
  File "/lib/python3.7/site-packages/tensorflow/python/ops/gradients_util.py", line 340, in _MaybeCompile
    return grad_fn()  # Exit early
  File "/lib/python3.7/site-packages/tensorflow/python/ops/gradients_util.py", line 684, in <lambda>
    lambda: grad_fn(op, *out_grads))
  File "/lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py", line 327, in custom_grad
    return op_handlers[op.type](self, op, *grads)
  File "/lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py", line 477, in handler
    return nonlinearity_1d_handler(input_ind, explainer, op, *grads)
  File "/lib/python3.7/site-packages/shap/explainers/deep/deep_tf.py", line 494, in nonlinearity_1d_handler
    tf.tile(tf.abs(delta_in0), dup0) < 1e-6,
  File "/lib/python3.7/site-packages/tensorflow/python/util/dispatch.py", line 201, in wrapper
    return target(*args, **kwargs)
  File "/lib/python3.7/site-packages/tensorflow/python/ops/math_ops.py", line 401, in abs
    return gen_math_ops._abs(x, name=name)
  File "/lib/python3.7/site-packages/tensorflow/python/ops/gen_math_ops.py", line 56, in _abs
    "Abs", x=x, name=name)
  File "/lib/python3.7/site-packages/tensorflow/python/framework/op_def_library.py", line 750, in _apply_op_helper
    attrs=attr_protos, op_def=op_def)
  File "/lib/python3.7/site-packages/tensorflow/python/framework/ops.py", line 3536, in _create_op_internal
    op_def=op_def)
  File "/lib/python3.7/site-packages/tensorflow/python/framework/ops.py", line 1990, in __init__
    self._traceback = tf_stack.extract_stack()

...which was originally created as op 'main_logsumexp_counts_bias_0/ReduceLogSumExp/Exp', defined at:
  File "/bin/bpnet-shap", line 8, in <module>
    sys.exit(shap_scores_main())
[elided 0 identical lines from previous traceback]
  File "/lib/python3.7/site-packages/bpnet/cli/shap_scores.py", line 448, in shap_scores_main
    shap_scores(args, shap_scores_dir)
  File "/lib/python3.7/site-packages/bpnet/cli/shap_scores.py", line 96, in shap_scores
    model = load_model(args.model, compile=False)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/saving/save.py", line 212, in load_model
    return saved_model_load.load(filepath, compile, options)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/saving/saved_model/load.py", line 138, in load
    keras_loader.load_layers(compile=compile)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/saving/saved_model/load.py", line 376, in load_layers
    node_metadata.metadata)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/saving/saved_model/load.py", line 417, in _load_layer
    obj, setter = self._revive_from_config(identifier, metadata, node_id)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/saving/saved_model/load.py", line 435, in _revive_from_config
    self._revive_layer_from_config(metadata, node_id))
  File "/lib/python3.7/site-packages/tensorflow/python/keras/saving/saved_model/load.py", line 495, in _revive_layer_from_config
    generic_utils.serialize_keras_class_and_config(class_name, config))
  File "/lib/python3.7/site-packages/tensorflow/python/keras/layers/serialization.py", line 177, in deserialize
    printable_module_name='layer')
  File "/lib/python3.7/site-packages/tensorflow/python/keras/utils/generic_utils.py", line 358, in deserialize_keras_object
    list(custom_objects.items())))
  File "/lib/python3.7/site-packages/tensorflow/python/keras/engine/training.py", line 2262, in from_config
    config, custom_objects=custom_objects)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/engine/functional.py", line 669, in from_config
    config, custom_objects)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/engine/functional.py", line 1285, in reconstruct_from_config
    process_node(layer, node_data)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/engine/functional.py", line 1233, in process_node
    output_tensors = layer(input_tensors, **kwargs)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/engine/base_layer_v1.py", line 786, in __call__
    outputs = call_fn(cast_inputs, *args, **kwargs)
  File "/lib/python3.7/site-packages/tensorflow/python/keras/layers/core.py", line 917, in call
    result = self.function(inputs, **kwargs)
  File "/lib/python3.7/site-packages/bpnet/model/arch.py", line 446, in <lambda>
    lambda x: tf.math.reduce_logsumexp(x, axis=-1, keepdims=True),

