Primus-SaFE(Stability and Fault Endurance)
-
Updated
Sep 18, 2026 - Go
Primus-SaFE(Stability and Fault Endurance)
Code and data of the EMNLP 2022 Main Conference paper "Reduce Catastrophic Forgetting of Dense Retrieval Training with Teleportation Negatives".
Global L2 norm adaptive gradient clipping engine to mitigate exploding gradients during deep neural network training.
Global L2 norm adaptive gradient clipping engine to mitigate exploding gradients during deep neural network training.
PyTorch NaNs are silent killers. This hook catches them at the exact layer and batch — with ~3 ms overhead vs ~7 ms for set_detect_anomaly.
Drift-Aware Adaptive Aggregation (DAA) for federated learning on CIFAR-10 under heterogeneous client partitions.
RVAV: a physics-informed PyTorch optimizer for energy-stable, high-LR training—with a simple closure API, tests, CI, and quickstart.
An unofficial extended version for ICLR 2021"Gradient Descent on Neural Networks Typically Occurs at the Edge of Stability"
Benchmarking GAN optimizers (Adam, RMSprop, SGD, Lookahead) on CIFAR-10 using WGAN-GP and FID evaluation.
To associate your repository with the training-stability topic, visit your repo's landing page and select "manage topics."