View sprakashdash's full-sized avatar
🎯
Focusing
Satya Prakash Dash sprakashdash
🎯
Focusing
My current interest lies in training dynamics of LLMs using Natural Gradient Descent and using randomized linear algebra to make them scalable.
- Manchester, UK
-
19:37
(UTC +01:00) - https://sprakashdash.github.io/
- https://orcid.org/0009-0000-4822-823X
- in/sprakashdash
Highlights
- Pro
Pinned Loading
-
reinforcement-learning-algorithms
reinforcement-learning-algorithms PublicForked from TianhongDai/reinforcement-learning-algorithms
This repository contains most of pytorch implementation based classic deep reinforcement learning algorithms, including - DQN, DDQN, Dueling Network, DDPG, SAC, A2C, PPO, TRPO. (More algorithms are...
Python
-
RL.Fun.Do
RL.Fun.Do PublicA repository for easy understanding of codes in Deep Reinforcement Learning
Python
-
RotEqNet
RotEqNet PublicForked from COGMAR/RotEqNet
Rotational Equivariant Networks for PyTorch/Python
Python
-
strassen-torch
strassen-torch PublicStrassen's algorithm implemented for Neural Networks in PyTorch.
Python
-
robinhenry/gym-anm
robinhenry/gym-anm PublicA framework to design Reinforcement Learning environments that model Active Network Management (ANM) tasks in electricity distribution networks.
-
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.