Talks and workshops:
third-crowd-c77.notion.site/NeurIPS2024-...
Curated reading list
fracturedplane.notion.site/NeurIPS2024-...
#Holidayreading
Talks and workshops:
third-crowd-c77.notion.site/NeurIPS2024-...
Curated reading list
fracturedplane.notion.site/NeurIPS2024-...
#Holidayreading
Link: arxiv.org/pdf/2401.023...
#ReinforcementLearning #ICLR2025 #ACL2025 #NAACL2025 #NeurIPS2024 #ICML2025 #DeepRL #DeepReinforcementLearning
Link: arxiv.org/pdf/2401.023...
#ReinforcementLearning #ICLR2025 #ACL2025 #NAACL2025 #NeurIPS2024 #ICML2025 #DeepRL #DeepReinforcementLearning
www.cell.com/trends/neuro...
Very short thread below to summarize our review
#neuroscience #neuroskyence #compneurosky #PsychSciSky
Link: proceedings.mlr.press/v235/korkmaz...
#ReinforcementLearning #ICLR2025 #ACL2025 #NAACL2025 #NeurIPS2024 #ICML2025 #DeepRL #DeepReinforcementLearning
Link: proceedings.mlr.press/v235/korkmaz...
#ReinforcementLearning #ICLR2025 #ACL2025 #NAACL2025 #NeurIPS2024 #ICML2025 #DeepRL #DeepReinforcementLearning
Highly recommended read.
https://arxiv.org/abs/2407.00695
Highly recommended read.
https://arxiv.org/abs/2407.00695
Tips here: neo-x.github.io/blog/2023/09...
Tips here: neo-x.github.io/blog/2023/09...
bryantmcgill.blogspot.com/2025/07/rewa...
This investigation was inspired by Lex's (@LexFridman) @MIT 6.S091: Introduction to Deep RL.
Soundcloud:
soundcloud.com/bryantmcgill...
bryantmcgill.blogspot.com/2025/07/rewa...
This investigation was inspired by Lex's (@LexFridman) @MIT 6.S091: Introduction to Deep RL.
Soundcloud:
soundcloud.com/bryantmcgill...
youtu.be/1ZuvCWvj0HM
youtu.be/1ZuvCWvj0HM
We propose gradient interventions that enable stable, scalable learning, unlocking significant performance gains across agents and environments!
Details below 👇
Link: github.com/EzgiKorkmaz/...
#ReinforcementLearning #SafeAI #Adversarial #Robust #DeepRL #robustRL #LanguageModels #AdversarialRL #AISafety #ExplainableAI #TrustworthyAI #ResponsibleAI #DeepReinforcementLearning
Link: github.com/EzgiKorkmaz/...
#ReinforcementLearning #SafeAI #Adversarial #Robust #DeepRL #robustRL #LanguageModels #AdversarialRL #AISafety #ExplainableAI #TrustworthyAI #ResponsibleAI #DeepReinforcementLearning
#NeurIPS2024 @neuripsconf.bsky.social #NeurIPS24
#reinforcementlearning #AIsafety #AISecurity #ResponsibleAI #TrustworthyAI #RobustAI #DeepRL
bsky.app/profile/ezgi...
Adversarial Robust Deep Reinforcement Learning is Neither Robust Nor Safe
Link: openreview.net/pdf?id=EPa0u...
#NeurIPS2024
neuripsconf.bsky.social
#NeurIPS24
#NeurIPS2024 @neuripsconf.bsky.social #NeurIPS24
#reinforcementlearning #AIsafety #AISecurity #ResponsibleAI #TrustworthyAI #RobustAI #DeepRL
bsky.app/profile/ezgi...
Link: neurips2023-enlsp.github.io/papers/paper...
#ReinforcementLearning #FoundationModels #DeepRL #DeepReinforcementLearning #ResponsibleAI #AIBias #LLMs #LanguageModels
Link: neurips2023-enlsp.github.io/papers/paper...
#ReinforcementLearning #FoundationModels #DeepRL #DeepReinforcementLearning #ResponsibleAI #AIBias #LLMs #LanguageModels
#DeepRL #Robotics #WarehouseAutomation #AI
#DeepRL #Robotics #WarehouseAutomation #AI
#BigData #Analytics #DataScience #AI #MachineLearning #ReinforcementLearning #RL #DeepRL #PyTorch #Python #RStats #TensorFlow #JavaScript #CloudComputing
Amazon: geni.us/RL-ML-Comput...
Free Book: incompleteideas.net/book/the-boo...
#BigData #Analytics #DataScience #AI #MachineLearning #ReinforcementLearning #RL #DeepRL #PyTorch #Python #RStats #TensorFlow #JavaScript #CloudComputing
Amazon: geni.us/RL-ML-Comput...
Free Book: incompleteideas.net/book/the-boo...
✨See my new paper on scaling, capacity and complexity of reinforcement learning published at #AAAI2026 ! @aaai.org
#AAAI #AAAI26 #ReinforcementLearning #DeepRL
"Deep reinforcement learning without experience replay, target networks, or batch updates"
Deep RL networks in the streaming setting without replay buffers thanks to signal normalization & step-size bounding 🤯
📄Paper: openreview.net/pdf?id=yqQJG...
"Deep reinforcement learning without experience replay, target networks, or batch updates"
Deep RL networks in the streaming setting without replay buffers thanks to signal normalization & step-size bounding 🤯
📄Paper: openreview.net/pdf?id=yqQJG...