The is the official implementation of ICML 2023 paper "Revisiting Weighted Aggregation in Federated Learning with Neural Networks".
-
Updated
Aug 28, 2023 - Python
The is the official implementation of ICML 2023 paper "Revisiting Weighted Aggregation in Federated Learning with Neural Networks".
[NeurIPS 2023] The PyTorch Implementation of Scheduled (Stable) Weight Decay.
AdamW optimizer engine decoupling L2 weight decay regularization from adaptive gradient first and second moments.
AdamW optimizer engine decoupling L2 weight decay regularization from adaptive gradient first and second moments.
Investigating the Role of Weight Decay in Enhancing Nonconvex SGD, CVPR 2025
Advanced CIFAR-10 image classification using ResNet-inspired CNN with residual blocks, achieving 92%+ accuracy through comprehensive regularization, data augmentation, and professional ML engineering practices.
Code accompanying "Learning to Forget: Continual Learning with Adaptive Weight Decay"
Cheap online diagnostics for grokking transformers: weight-decay regimes, attention-head order parameters, data, and Lean 4 verification
We proposed an order parameter for grokking. Then we killed it. A pre-registered falsification, with primary data, code, and a documented retraction.
Folder contains implementation of Multi layer feed forward networks, Autoencoders, Sparse Autoencoders and many..
Super-Convergence on CIFAR10
Implementation of some new techniques from fastai and other papers which works with keras models
Machine Learning university project
Code and figures for To Grok Grokking: Provable Grokking in Ridge Regression (ICML 2026), with ridge-regression, random-feature, NTK-style, and fully trained ReLU experiments.
To associate your repository with the weight-decay topic, visit your repo's landing page and select "manage topics."