René's URL Explorer Experiment


Title: GitHub - USTCPCS/reinforcement-learning: Implementation of Reinforcement Learning Algorithms. Python, OpenAI Gym, Tensorflow. Exercises and Solutions to accompany Sutton's Book and David Silver's course. · GitHub

Open Graph Title: GitHub - USTCPCS/reinforcement-learning: Implementation of Reinforcement Learning Algorithms. Python, OpenAI Gym, Tensorflow. Exercises and Solutions to accompany Sutton's Book and David Silver's course.

X Title: GitHub - USTCPCS/reinforcement-learning: Implementation of Reinforcement Learning Algorithms. Python, OpenAI Gym, Tensorflow. Exercises and Solutions to accompany Sutton's Book and David Silver's course.

Description: Implementation of Reinforcement Learning Algorithms. Python, OpenAI Gym, Tensorflow. Exercises and Solutions to accompany Sutton's Book and David Silver's course. - USTCPCS/reinforcement-learning

Open Graph Description: Implementation of Reinforcement Learning Algorithms. Python, OpenAI Gym, Tensorflow. Exercises and Solutions to accompany Sutton's Book and David Silver's course. - USTCPCS/reinforcement-le...

X Description: Implementation of Reinforcement Learning Algorithms. Python, OpenAI Gym, Tensorflow. Exercises and Solutions to accompany Sutton's Book and David Silver's course. - USTCPCS/reinforc...

Opengraph URL: https://github.com/USTCPCS/reinforcement-learning

X: @github

direct link

Domain: github.com

route-pattern/:user_id/:repository
route-controllerfiles
route-actiondisambiguate
fetch-noncev2:f898db00-0f4e-f0c5-544e-b3a32bd26b9b
current-catalog-service-hashf3abb0cc802f3d7b95fc8762b94bdcb13bf39634c40c357301c4aa1d67a256fb
request-idB076:37941:182816F:20DCF30:6A65289D
html-safe-nonceb6e9cc4d8ae6ad6f2e5250bb42dbb3d2ed9f2da3d65bb2059501264f67e23758
visitor-payloadeyJyZWZlcnJlciI6IiIsInJlcXVlc3RfaWQiOiJCMDc2OjM3OTQxOjE4MjgxNkY6MjBEQ0YzMDo2QTY1Mjg5RCIsInZpc2l0b3JfaWQiOiI3MDM4MjA0MjA1ODUwNzY1NDY5IiwicmVnaW9uX2VkZ2UiOiJpYWQiLCJyZWdpb25fcmVuZGVyIjoiaWFkIn0=
visitor-hmace3fee420b0101c159c8a6a997bee36d7e83544755a93c72a1e2c7b7702f73db9
hovercard-subject-tagrepository:111353298
github-keyboard-shortcutsrepository,copilot
google-site-verificationApib7-x98H0j5cPqHWwSMm6dNU4GmODRoqxLiDzdx9I
octolytics-urlhttps://collector.github.com/github/collect
analytics-location//
fb:app_id1401488693436528
apple-itunes-appapp-id=1477376905, app-argument=https://github.com/USTCPCS/reinforcement-learning
twitter:imagehttps://opengraph.githubassets.com/c6ef215055f943111379430bf2569b7d71a1993d9d4c2439157e42fdfe5b9917/USTCPCS/reinforcement-learning
twitter:cardsummary_large_image
og:imagehttps://opengraph.githubassets.com/c6ef215055f943111379430bf2569b7d71a1993d9d4c2439157e42fdfe5b9917/USTCPCS/reinforcement-learning
og:image:altImplementation of Reinforcement Learning Algorithms. Python, OpenAI Gym, Tensorflow. Exercises and Solutions to accompany Sutton's Book and David Silver's course. - USTCPCS/reinforcement-le...
og:image:width1200
og:image:height600
og:site_nameGitHub
og:typeobject
hostnamegithub.com
expected-hostnamegithub.com
None52c76df668885aaff23b50bdca1fa1ea44ac9c1553e888ebc70ff1e4daa4625b
turbo-cache-controlno-cache
go-importgithub.com/USTCPCS/reinforcement-learning git https://github.com/USTCPCS/reinforcement-learning.git
octolytics-dimension-user_id17957381
octolytics-dimension-user_loginUSTCPCS
octolytics-dimension-repository_id111353298
octolytics-dimension-repository_nwoUSTCPCS/reinforcement-learning
octolytics-dimension-repository_publictrue
octolytics-dimension-repository_is_forktrue
octolytics-dimension-repository_parent_id66483240
octolytics-dimension-repository_parent_nwodennybritz/reinforcement-learning
octolytics-dimension-repository_network_root_id66483240
octolytics-dimension-repository_network_root_nwodennybritz/reinforcement-learning
turbo-body-classeslogged-out env-production page-responsive
disable-turbofalse
browser-stats-urlhttps://api.github.com/_private/browser/stats
browser-errors-urlhttps://api.github.com/_private/browser/errors
release309153364422b3c499922d1a2a6404910a58ed8e
ui-targetfull
theme-color#1e2327
color-schemelight dark

Links:

Skip to contenthttps://github.com/USTCPCS/reinforcement-learning#start-of-content
https://github.com/
Sign in https://github.com/login?return_to=https%3A%2F%2Fgithub.com%2FUSTCPCS%2Freinforcement-learning
GitHub CopilotWrite better code with AIhttps://github.com/features/copilot
GitHub Copilot appDirect agents from issue to mergehttps://github.com/features/ai/github-app
MCP RegistryNewIntegrate external toolshttps://github.com/mcp
ActionsAutomate any workflowhttps://github.com/features/actions
CodespacesInstant dev environmentshttps://github.com/features/codespaces
IssuesPlan and track workhttps://github.com/features/issues
Code ReviewManage code changeshttps://github.com/features/code-review
Code QualityEnforce quality at mergehttps://github.com/features/code-quality
GitHub Advanced SecurityFind and fix vulnerabilitieshttps://github.com/security/advanced-security
Code securitySecure your code as you buildhttps://github.com/security/advanced-security/code-security
Secret protectionStop leaks before they starthttps://github.com/security/advanced-security/secret-protection
Why GitHubhttps://github.com/why-github
Documentationhttps://docs.github.com
Bloghttps://github.blog
Changeloghttps://github.blog/changelog
Marketplacehttps://github.com/marketplace
View all featureshttps://github.com/features
Enterpriseshttps://github.com/enterprise
Small and medium teamshttps://github.com/team
Startupshttps://github.com/enterprise/startups
Nonprofitshttps://github.com/solutions/industry/nonprofits
App Modernizationhttps://github.com/solutions/use-case/app-modernization
DevSecOpshttps://github.com/solutions/use-case/devsecops
DevOpshttps://github.com/solutions/use-case/devops
CI/CDhttps://github.com/solutions/use-case/ci-cd
View all use caseshttps://github.com/solutions/use-case
Healthcarehttps://github.com/solutions/industry/healthcare
Financial serviceshttps://github.com/solutions/industry/financial-services
Manufacturinghttps://github.com/solutions/industry/manufacturing
Governmenthttps://github.com/solutions/industry/government
View all industrieshttps://github.com/solutions/industry
View all solutionshttps://github.com/solutions
AIhttps://github.com/resources/articles?topic=ai
Software Developmenthttps://github.com/resources/articles?topic=software-development
DevOpshttps://github.com/resources/articles?topic=devops
Securityhttps://github.com/resources/articles?topic=security
View all topicshttps://github.com/resources/articles
Customer storieshttps://github.com/customer-stories
Events & webinarshttps://github.com/resources/events
Ebooks & reportshttps://github.com/resources/whitepapers
Business insightshttps://github.com/solutions/executive-insights
GitHub Skillshttps://skills.github.com
Documentationhttps://docs.github.com
Customer supporthttps://support.github.com
Community forumhttps://github.com/orgs/community/discussions
Trust centerhttps://github.com/trust-center
Partnershttps://github.com/partners
View all resourceshttps://github.com/resources
GitHub SponsorsFund open source developershttps://github.com/open-source/sponsors
Security Labhttps://securitylab.github.com
Maintainer Communityhttps://maintainers.github.com
Acceleratorhttps://github.com/open-source/accelerator
GitHub Starshttps://stars.github.com
Archive Programhttps://archiveprogram.github.com
Topicshttps://github.com/topics
Trendinghttps://github.com/trending
Collectionshttps://github.com/collections
Enterprise platformAI-powered developer platformhttps://github.com/enterprise
GitHub Advanced SecurityEnterprise-grade security featureshttps://github.com/security/advanced-security
Copilot for BusinessEnterprise-grade AI featureshttps://github.com/features/copilot/copilot-business
Premium SupportEnterprise-grade 24/7 supporthttps://github.com/enterprise/premium-support
Pricinghttps://github.com/pricing
Search syntax tipshttps://docs.github.com/search-github/github-code-search/understanding-github-code-search-syntax
documentationhttps://docs.github.com/search-github/github-code-search/understanding-github-code-search-syntax
Sign in https://github.com/login?return_to=https%3A%2F%2Fgithub.com%2FUSTCPCS%2Freinforcement-learning
Sign up https://github.com/signup?ref_cta=Sign+up&ref_loc=header+logged+out&ref_page=%2F%3Cuser-name%3E%2F%3Crepo-name%3E&source=header-repo&source_repo=USTCPCS%2Freinforcement-learning
Reloadhttps://github.com/USTCPCS/reinforcement-learning
Reloadhttps://github.com/USTCPCS/reinforcement-learning
Reloadhttps://github.com/USTCPCS/reinforcement-learning
USTCPCS https://github.com/USTCPCS
reinforcement-learninghttps://github.com/USTCPCS/reinforcement-learning
dennybritz/reinforcement-learninghttps://github.com/dennybritz/reinforcement-learning
Notifications https://github.com/login?return_to=%2FUSTCPCS%2Freinforcement-learning
Fork 1 https://github.com/login?return_to=%2FUSTCPCS%2Freinforcement-learning
Star 2 https://github.com/login?return_to=%2FUSTCPCS%2Freinforcement-learning
Code https://github.com/USTCPCS/reinforcement-learning
Pull requests 0 https://github.com/USTCPCS/reinforcement-learning/pulls
Actions https://github.com/USTCPCS/reinforcement-learning/actions
Projects https://github.com/USTCPCS/reinforcement-learning/projects
Wiki https://github.com/USTCPCS/reinforcement-learning/wiki
Security and quality 0 https://github.com/USTCPCS/reinforcement-learning/security
Insights https://github.com/USTCPCS/reinforcement-learning/pulse
Code https://github.com/USTCPCS/reinforcement-learning
Pull requests https://github.com/USTCPCS/reinforcement-learning/pulls
Actions https://github.com/USTCPCS/reinforcement-learning/actions
Projects https://github.com/USTCPCS/reinforcement-learning/projects
Wiki https://github.com/USTCPCS/reinforcement-learning/wiki
Security and quality https://github.com/USTCPCS/reinforcement-learning/security
Insights https://github.com/USTCPCS/reinforcement-learning/pulse
https://github.com/USTCPCS/reinforcement-learning
Brancheshttps://github.com/USTCPCS/reinforcement-learning/branches
Tagshttps://github.com/USTCPCS/reinforcement-learning/tags
https://github.com/USTCPCS/reinforcement-learning/branches
https://github.com/USTCPCS/reinforcement-learning/tags
185 Commitshttps://github.com/USTCPCS/reinforcement-learning/commits/master/
https://github.com/USTCPCS/reinforcement-learning/commits/master/
DPhttps://github.com/USTCPCS/reinforcement-learning/tree/master/DP
DPhttps://github.com/USTCPCS/reinforcement-learning/tree/master/DP
DQNhttps://github.com/USTCPCS/reinforcement-learning/tree/master/DQN
DQNhttps://github.com/USTCPCS/reinforcement-learning/tree/master/DQN
FAhttps://github.com/USTCPCS/reinforcement-learning/tree/master/FA
FAhttps://github.com/USTCPCS/reinforcement-learning/tree/master/FA
Introductionhttps://github.com/USTCPCS/reinforcement-learning/tree/master/Introduction
Introductionhttps://github.com/USTCPCS/reinforcement-learning/tree/master/Introduction
MChttps://github.com/USTCPCS/reinforcement-learning/tree/master/MC
MChttps://github.com/USTCPCS/reinforcement-learning/tree/master/MC
MDPhttps://github.com/USTCPCS/reinforcement-learning/tree/master/MDP
MDPhttps://github.com/USTCPCS/reinforcement-learning/tree/master/MDP
PolicyGradienthttps://github.com/USTCPCS/reinforcement-learning/tree/master/PolicyGradient
PolicyGradienthttps://github.com/USTCPCS/reinforcement-learning/tree/master/PolicyGradient
TDhttps://github.com/USTCPCS/reinforcement-learning/tree/master/TD
TDhttps://github.com/USTCPCS/reinforcement-learning/tree/master/TD
libhttps://github.com/USTCPCS/reinforcement-learning/tree/master/lib
libhttps://github.com/USTCPCS/reinforcement-learning/tree/master/lib
.gitignorehttps://github.com/USTCPCS/reinforcement-learning/blob/master/.gitignore
.gitignorehttps://github.com/USTCPCS/reinforcement-learning/blob/master/.gitignore
LICENSEhttps://github.com/USTCPCS/reinforcement-learning/blob/master/LICENSE
LICENSEhttps://github.com/USTCPCS/reinforcement-learning/blob/master/LICENSE
README.mdhttps://github.com/USTCPCS/reinforcement-learning/blob/master/README.md
README.mdhttps://github.com/USTCPCS/reinforcement-learning/blob/master/README.md
__init__.pyhttps://github.com/USTCPCS/reinforcement-learning/blob/master/__init__.py
__init__.pyhttps://github.com/USTCPCS/reinforcement-learning/blob/master/__init__.py
READMEhttps://github.com/USTCPCS/reinforcement-learning
MIT licensehttps://github.com/USTCPCS/reinforcement-learning
https://github.com/USTCPCS/reinforcement-learning#overview
Reinforcement Learning: An Introduction (2nd Edition)http://incompleteideas.net/sutton/book/bookdraft2017june.pdf
David Silver's Reinforcement Learning Coursehttp://www0.cs.ucl.ac.uk/staff/d.silver/web/Teaching.html
OpenAI Gymhttps://gym.openai.com/
Tensorflowhttps://www.tensorflow.org/
https://github.com/USTCPCS/reinforcement-learning#table-of-contents
Introduction to RL problems & OpenAI Gymhttps://github.com/USTCPCS/reinforcement-learning/blob/master/Introduction
MDPs and Bellman Equationshttps://github.com/USTCPCS/reinforcement-learning/blob/master/MDP
Dynamic Programming: Model-Based RL, Policy Iteration and Value Iterationhttps://github.com/USTCPCS/reinforcement-learning/blob/master/DP
Monte Carlo Model-Free Prediction & Controlhttps://github.com/USTCPCS/reinforcement-learning/blob/master/MC
Temporal Difference Model-Free Prediction & Controlhttps://github.com/USTCPCS/reinforcement-learning/blob/master/TD
Function Approximationhttps://github.com/USTCPCS/reinforcement-learning/blob/master/FA
Deep Q Learninghttps://github.com/USTCPCS/reinforcement-learning/blob/master/DQN
Policy Gradient Methodshttps://github.com/USTCPCS/reinforcement-learning/blob/master/PolicyGradient
https://github.com/USTCPCS/reinforcement-learning#list-of-implemented-algorithms
Asynchronous Advantage Actor Critic (A3C)https://github.com/USTCPCS/reinforcement-learning/blob/master/PolicyGradient/a3c
https://github.com/USTCPCS/reinforcement-learning#resources
Reinforcement Learning: An Introduction (2nd Edition)http://incompleteideas.net/sutton/book/bookdraft2017june.pdf
David Silver's Reinforcement Learning Course (UCL, 2015)http://www0.cs.ucl.ac.uk/staff/d.silver/web/Teaching.html
CS294 - Deep Reinforcement Learning (Berkeley, Fall 2015)http://rll.berkeley.edu/deeprlcourse/
CS 8803 - Reinforcement Learning (Georgia Tech)https://www.udacity.com/course/reinforcement-learning--ud600
Introduction to Reinforcement Learning (Joelle Pineau @ Deep Learning Summer School 2016)http://videolectures.net/deeplearning2016_pineau_reinforcement_learning/
Deep Reinforcement Learning (Pieter Abbeel @ Deep Learning Summer School 2016)http://videolectures.net/deeplearning2016_abbeel_deep_reinforcement/
Deep Reinforcement Learning ICML 2016 Tutorial (David Silver)http://techtalks.tv/talks/deep-reinforcement-learning/62360/
Tutorial: Introduction to Reinforcement Learning with Function Approximationhttps://www.youtube.com/watch?v=ggqnxyjaKe4
John Schulman - Deep Reinforcement Learning (4 Lectures)https://www.youtube.com/playlist?list=PLjKEIQlKCTZYN3CYBlj8r58SbNorobqcp
Deep Reinforcement Learning Slides @ NIPS 2016http://people.eecs.berkeley.edu/~pabbeel/nips-tutorial-policy-optimization-Schulman-Abbeel.pdf
carpedm20/deep-rl-tensorflowhttps://github.com/carpedm20/deep-rl-tensorflow
matthiasplappert/keras-rlhttps://github.com/matthiasplappert/keras-rl
Human-Level Control through Deep Reinforcement Learning (2015-02)http://www.readcube.com/articles/10.1038/nature14236
Deep Reinforcement Learning with Double Q-learning (2015-09)http://arxiv.org/abs/1509.06461
Continuous control with deep reinforcement learning (2015-09)https://arxiv.org/abs/1509.02971
Prioritized Experience Replay (2015-11)http://arxiv.org/abs/1511.05952
Dueling Network Architectures for Deep Reinforcement Learning (2015-11)http://arxiv.org/abs/1511.06581
Asynchronous Methods for Deep Reinforcement Learning (2016-02)http://arxiv.org/abs/1602.01783
Deep Reinforcement Learning from Self-Play in Imperfect-Information Games (2016-03)http://arxiv.org/abs/1603.01121
Mastering the game of Go with deep neural networks and tree searchhttps://gogameguru.com/i/2016/03/deepmind-mastering-go.pdf
www.wildml.com/2016/10/learning-reinforcement-learning/http://www.wildml.com/2016/10/learning-reinforcement-learning/
Readmehttps://github.com/USTCPCS/reinforcement-learning#readme-ov-file
MIT licensehttps://github.com/USTCPCS/reinforcement-learning#MIT-1-ov-file
Activityhttps://github.com/USTCPCS/reinforcement-learning/activity
2 starshttps://github.com/USTCPCS/reinforcement-learning/stargazers
1 watchinghttps://github.com/USTCPCS/reinforcement-learning/watchers
1 forkhttps://github.com/USTCPCS/reinforcement-learning/forks
Report repositoryhttps://github.com/contact/report-content?content_url=https%3A%2F%2Fgithub.com%2FUSTCPCS%2Freinforcement-learning&report=USTCPCS+%28user%29
https://github.com
Termshttps://docs.github.com/site-policy/github-terms/github-terms-of-service
Privacyhttps://docs.github.com/site-policy/privacy-policies/github-privacy-statement
Securityhttps://github.com/security
Statushttps://www.githubstatus.com/
Communityhttps://github.community/
Docshttps://docs.github.com/
Contacthttps://support.github.com?tags=dotcom-footer

Viewport: width=device-width


URLs of crawlers that visited me.