Always interesting how much the optimizer alone can change behavior, even with everything else fixed. Nice experiment by Gradient Thoughts. When did the differences start showing up most clearly during training?
Training the Same Neural Network with Different Optimizers
2 Comments
Gradient Thoughts
•
@[Andrew Mewborn] Thank you for your views !
Predominant changes started showing up as I moved towards more dynamic and adaptive optimizers like Adam and RMSProp. Again this is subject to the conditions and constraints that I have imposed on the neural net. Varying other parameters may result in totally different behaviour in the convergence plots.
The primary focus of this experiment is to observe optimizer behaviour under tightly controlled conditions.
Please log in to add a comment.
🔥 Join developers growing publicly
Share your knowledge, build in public, and grow your developer presence with a global community.
Please log in to comment on this post.
More Posts
- © 2026 Coder Legion
- Feedback / Bug
- Privacy
- About Us
- Contacts
- You Tube
- Premium Subscription
- Terms of Service
- Early Builders
chevron_left
Related Jobs
- Network & Infrastructure IntegratorAcestack · Full time · Canada
- Network Security EngineerBitdeer · Full time · Singapore
- Network Automation EngineerBitdeer · Full time · Singapore
Commenters (This Week)
hetptis
3 comments
SpaceShaman
1 comment
jomynn
1 comment
Contribute meaningful comments to climb the leaderboard and earn badges!