René's URL Explorer Experiment


Title: tutorials/Reinforcement_learning_TUT at master · LN512/tutorials · GitHub

Open Graph Title: tutorials/Reinforcement_learning_TUT at master · LN512/tutorials

X Title: tutorials/Reinforcement_learning_TUT at master · LN512/tutorials

Description: 机器学习相关教程. Contribute to LN512/tutorials development by creating an account on GitHub.

Open Graph Description: 机器学习相关教程. Contribute to LN512/tutorials development by creating an account on GitHub.

X Description: 机器学习相关教程. Contribute to LN512/tutorials development by creating an account on GitHub.

Opengraph URL: https://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT

X: @github

direct link

Domain: github.com

route-pattern/:user_id/:repository/tree/*name(/*path)
route-controllerfiles
route-actiondisambiguate
fetch-noncev2:e6b2e306-9f92-32e7-727c-5a319063c34d
current-catalog-service-hashf3abb0cc802f3d7b95fc8762b94bdcb13bf39634c40c357301c4aa1d67a256fb
request-idA028:17A025:2F5C0EA:3F43FAE:6A652C55
html-safe-nonce2ff1af2f90f23a358ef0103517893984f269eee3dc9b1e005ca323afe16d1687
visitor-payloadeyJyZWZlcnJlciI6IiIsInJlcXVlc3RfaWQiOiJBMDI4OjE3QTAyNToyRjVDMEVBOjNGNDNGQUU6NkE2NTJDNTUiLCJ2aXNpdG9yX2lkIjoiODU2MTI0MjgwNzc5MzQyOTU4OSIsInJlZ2lvbl9lZGdlIjoiaWFkIiwicmVnaW9uX3JlbmRlciI6ImlhZCJ9
visitor-hmac78c45b01b6c8caf2533825273ccb6ab1922539b6c9f199ae7bf4d3b3ef5a04fb
hovercard-subject-tagrepository:90147225
github-keyboard-shortcutsrepository,source-code,file-tree,copilot
google-site-verificationApib7-x98H0j5cPqHWwSMm6dNU4GmODRoqxLiDzdx9I
octolytics-urlhttps://collector.github.com/github/collect
analytics-location///files/disambiguate
fb:app_id1401488693436528
apple-itunes-appapp-id=1477376905, app-argument=https://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT
twitter:imagehttps://opengraph.githubassets.com/e321c045a1a54b64c5a54c456c12550d7dda0f4a70da4ed555976189259c3d0a/LN512/tutorials
twitter:cardsummary_large_image
og:imagehttps://opengraph.githubassets.com/e321c045a1a54b64c5a54c456c12550d7dda0f4a70da4ed555976189259c3d0a/LN512/tutorials
og:image:alt机器学习相关教程. Contribute to LN512/tutorials development by creating an account on GitHub.
og:image:width1200
og:image:height600
og:site_nameGitHub
og:typeobject
hostnamegithub.com
expected-hostnamegithub.com
None52c76df668885aaff23b50bdca1fa1ea44ac9c1553e888ebc70ff1e4daa4625b
turbo-cache-controlno-cache
go-importgithub.com/LN512/tutorials git https://github.com/LN512/tutorials.git
octolytics-dimension-user_id26219357
octolytics-dimension-user_loginLN512
octolytics-dimension-repository_id90147225
octolytics-dimension-repository_nwoLN512/tutorials
octolytics-dimension-repository_publictrue
octolytics-dimension-repository_is_forktrue
octolytics-dimension-repository_parent_id59944455
octolytics-dimension-repository_parent_nwoMorvanZhou/tutorials
octolytics-dimension-repository_network_root_id59944455
octolytics-dimension-repository_network_root_nwoMorvanZhou/tutorials
turbo-body-classeslogged-out env-production page-responsive
disable-turbofalse
browser-stats-urlhttps://api.github.com/_private/browser/stats
browser-errors-urlhttps://api.github.com/_private/browser/errors
release309153364422b3c499922d1a2a6404910a58ed8e
ui-targetfull
theme-color#1e2327
color-schemelight dark

Links:

Skip to contenthttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT#start-of-content
https://github.com/
Sign in https://github.com/login?return_to=https%3A%2F%2Fgithub.com%2FLN512%2Ftutorials%2Ftree%2Fmaster%2FReinforcement_learning_TUT
GitHub CopilotWrite better code with AIhttps://github.com/features/copilot
GitHub Copilot appDirect agents from issue to mergehttps://github.com/features/ai/github-app
MCP RegistryNewIntegrate external toolshttps://github.com/mcp
ActionsAutomate any workflowhttps://github.com/features/actions
CodespacesInstant dev environmentshttps://github.com/features/codespaces
IssuesPlan and track workhttps://github.com/features/issues
Code ReviewManage code changeshttps://github.com/features/code-review
Code QualityEnforce quality at mergehttps://github.com/features/code-quality
GitHub Advanced SecurityFind and fix vulnerabilitieshttps://github.com/security/advanced-security
Code securitySecure your code as you buildhttps://github.com/security/advanced-security/code-security
Secret protectionStop leaks before they starthttps://github.com/security/advanced-security/secret-protection
Why GitHubhttps://github.com/why-github
Documentationhttps://docs.github.com
Bloghttps://github.blog
Changeloghttps://github.blog/changelog
Marketplacehttps://github.com/marketplace
View all featureshttps://github.com/features
Enterpriseshttps://github.com/enterprise
Small and medium teamshttps://github.com/team
Startupshttps://github.com/enterprise/startups
Nonprofitshttps://github.com/solutions/industry/nonprofits
App Modernizationhttps://github.com/solutions/use-case/app-modernization
DevSecOpshttps://github.com/solutions/use-case/devsecops
DevOpshttps://github.com/solutions/use-case/devops
CI/CDhttps://github.com/solutions/use-case/ci-cd
View all use caseshttps://github.com/solutions/use-case
Healthcarehttps://github.com/solutions/industry/healthcare
Financial serviceshttps://github.com/solutions/industry/financial-services
Manufacturinghttps://github.com/solutions/industry/manufacturing
Governmenthttps://github.com/solutions/industry/government
View all industrieshttps://github.com/solutions/industry
View all solutionshttps://github.com/solutions
AIhttps://github.com/resources/articles?topic=ai
Software Developmenthttps://github.com/resources/articles?topic=software-development
DevOpshttps://github.com/resources/articles?topic=devops
Securityhttps://github.com/resources/articles?topic=security
View all topicshttps://github.com/resources/articles
Customer storieshttps://github.com/customer-stories
Events & webinarshttps://github.com/resources/events
Ebooks & reportshttps://github.com/resources/whitepapers
Business insightshttps://github.com/solutions/executive-insights
GitHub Skillshttps://skills.github.com
Documentationhttps://docs.github.com
Customer supporthttps://support.github.com
Community forumhttps://github.com/orgs/community/discussions
Trust centerhttps://github.com/trust-center
Partnershttps://github.com/partners
View all resourceshttps://github.com/resources
GitHub SponsorsFund open source developershttps://github.com/open-source/sponsors
Security Labhttps://securitylab.github.com
Maintainer Communityhttps://maintainers.github.com
Acceleratorhttps://github.com/open-source/accelerator
GitHub Starshttps://stars.github.com
Archive Programhttps://archiveprogram.github.com
Topicshttps://github.com/topics
Trendinghttps://github.com/trending
Collectionshttps://github.com/collections
Enterprise platformAI-powered developer platformhttps://github.com/enterprise
GitHub Advanced SecurityEnterprise-grade security featureshttps://github.com/security/advanced-security
Copilot for BusinessEnterprise-grade AI featureshttps://github.com/features/copilot/copilot-business
Premium SupportEnterprise-grade 24/7 supporthttps://github.com/enterprise/premium-support
Pricinghttps://github.com/pricing
Search syntax tipshttps://docs.github.com/search-github/github-code-search/understanding-github-code-search-syntax
documentationhttps://docs.github.com/search-github/github-code-search/understanding-github-code-search-syntax
Sign in https://github.com/login?return_to=https%3A%2F%2Fgithub.com%2FLN512%2Ftutorials%2Ftree%2Fmaster%2FReinforcement_learning_TUT
Sign up https://github.com/signup?ref_cta=Sign+up&ref_loc=header+logged+out&ref_page=%2F%3Cuser-name%3E%2F%3Crepo-name%3E%2Ffiles%2Fdisambiguate&source=header-repo&source_repo=LN512%2Ftutorials
Reloadhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT
Reloadhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT
Reloadhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT
LN512 https://github.com/LN512
tutorialshttps://github.com/LN512/tutorials
MorvanZhou/tutorialshttps://github.com/MorvanZhou/tutorials
Notifications https://github.com/login?return_to=%2FLN512%2Ftutorials
Fork 1 https://github.com/login?return_to=%2FLN512%2Ftutorials
Star 0 https://github.com/login?return_to=%2FLN512%2Ftutorials
Code https://github.com/LN512/tutorials
Pull requests 0 https://github.com/LN512/tutorials/pulls
Actions https://github.com/LN512/tutorials/actions
Projects https://github.com/LN512/tutorials/projects
Security and quality 0 https://github.com/LN512/tutorials/security
Insights https://github.com/LN512/tutorials/pulse
Code https://github.com/LN512/tutorials
Pull requests https://github.com/LN512/tutorials/pulls
Actions https://github.com/LN512/tutorials/actions
Projects https://github.com/LN512/tutorials/projects
Security and quality https://github.com/LN512/tutorials/security
Insights https://github.com/LN512/tutorials/pulse
https://github.com/LN512/tutorials
tutorialshttps://github.com/LN512/tutorials/tree/master
Historyhttps://github.com/LN512/tutorials/commits/master/Reinforcement_learning_TUT
https://github.com/LN512/tutorials/commits/master/Reinforcement_learning_TUT
tutorialshttps://github.com/LN512/tutorials/tree/master
..https://github.com/LN512/tutorials/tree/master
10_A3Chttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/10_A3C
10_A3Chttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/10_A3C
11_Dyna_Qhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/11_Dyna_Q
11_Dyna_Qhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/11_Dyna_Q
1_command_line_reinforcement_learninghttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/1_command_line_reinforcement_learning
1_command_line_reinforcement_learninghttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/1_command_line_reinforcement_learning
2_Q_Learning_mazehttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/2_Q_Learning_maze
2_Q_Learning_mazehttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/2_Q_Learning_maze
3_Sarsa_mazehttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/3_Sarsa_maze
3_Sarsa_mazehttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/3_Sarsa_maze
4_Sarsa_lambda_mazehttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/4_Sarsa_lambda_maze
4_Sarsa_lambda_mazehttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/4_Sarsa_lambda_maze
5.1_Double_DQNhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/5.1_Double_DQN
5.1_Double_DQNhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/5.1_Double_DQN
5.2_Prioritized_Replay_DQNhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/5.2_Prioritized_Replay_DQN
5.2_Prioritized_Replay_DQNhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/5.2_Prioritized_Replay_DQN
5.3_Dueling_DQNhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/5.3_Dueling_DQN
5.3_Dueling_DQNhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/5.3_Dueling_DQN
5_Deep_Q_Networkhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/5_Deep_Q_Network
5_Deep_Q_Networkhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/5_Deep_Q_Network
6_OpenAI_gymhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/6_OpenAI_gym
6_OpenAI_gymhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/6_OpenAI_gym
7_Policy_gradient_softmaxhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/7_Policy_gradient_softmax
7_Policy_gradient_softmaxhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/7_Policy_gradient_softmax
8_Actor_Critic_Advantagehttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/8_Actor_Critic_Advantage
8_Actor_Critic_Advantagehttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/8_Actor_Critic_Advantage
9_Deep_Deterministic_Policy_Gradient_DDPGhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/9_Deep_Deterministic_Policy_Gradient_DDPG
9_Deep_Deterministic_Policy_Gradient_DDPGhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/9_Deep_Deterministic_Policy_Gradient_DDPG
experimentshttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/experiments
experimentshttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT/experiments
README.mdhttps://github.com/LN512/tutorials/blob/master/Reinforcement_learning_TUT/README.md
README.mdhttps://github.com/LN512/tutorials/blob/master/Reinforcement_learning_TUT/README.md
RL_cover.jpghttps://github.com/LN512/tutorials/blob/master/Reinforcement_learning_TUT/RL_cover.jpg
RL_cover.jpghttps://github.com/LN512/tutorials/blob/master/Reinforcement_learning_TUT/RL_cover.jpg
README.mdhttps://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT#readme
https://www.youtube.com/watch?v=pieI7rOXELI&list=PLXO45tsB95cIplu-fLMpUEEZTwrDNh6Ba
https://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT#reinforcement-learning-methods-and-tutorials
莫烦 Pythonhttps://morvanzhou.github.io/tutorials/
Youtube channelhttps://www.youtube.com/channel/UCdyjiB5H8Pu7aDTNVXTTpcg
https://www.youtube.com/playlist?list=PLXO45tsB95cIplu-fLMpUEEZTwrDNh6Bahttps://www.youtube.com/playlist?list=PLXO45tsB95cIplu-fLMpUEEZTwrDNh6Ba
Simple entry examplehttps://github.com/MorvanZhou/tutorials/tree/master/Reinforcement_learning_TUT/1_command_line_reinforcement_learning
Q-learninghttps://github.com/MorvanZhou/tutorials/tree/master/Reinforcement_learning_TUT/2_Q_Learning_maze
Sarsahttps://github.com/MorvanZhou/tutorials/tree/master/Reinforcement_learning_TUT/3_Sarsa_maze
Sarsa(lambda)https://github.com/MorvanZhou/tutorials/tree/master/Reinforcement_learning_TUT/4_Sarsa_lambda_maze
Deep Q Networkhttps://github.com/MorvanZhou/tutorials/tree/master/Reinforcement_learning_TUT/5_Deep_Q_Network
Using OpenAI Gymhttps://github.com/MorvanZhou/tutorials/tree/master/Reinforcement_learning_TUT/6_OpenAI_gym
Double DQNhttps://github.com/MorvanZhou/tutorials/tree/master/Reinforcement_learning_TUT/5.1_Double_DQN
DQN with Prioitized Experience Replayhttps://github.com/MorvanZhou/tutorials/tree/master/Reinforcement_learning_TUT/5.2_Prioritized_Replay_DQN
Dueling DQNhttps://github.com/MorvanZhou/tutorials/tree/master/Reinforcement_learning_TUT/5.3_Dueling_DQN
Policy Gradientshttps://github.com/MorvanZhou/tutorials/tree/master/Reinforcement_learning_TUT/7_Policy_gradient_softmax
Actor Critichttps://github.com/MorvanZhou/tutorials/tree/master/Reinforcement_learning_TUT/8_Actor_Critic_Advantage
Deep Deterministic Policy Gradienthttps://github.com/MorvanZhou/tutorials/tree/master/Reinforcement_learning_TUT/9_Deep_Deterministic_Policy_Gradient_DDPG
A3Chttps://github.com/MorvanZhou/tutorials/tree/master/Reinforcement_learning_TUT/10_A3C
Dyna-Qhttps://github.com/MorvanZhou/tutorials/tree/master/Reinforcement_learning_TUT/11_Dyna_Q
https://github.com/LN512/tutorials/tree/master/Reinforcement_learning_TUT#donation
https://www.paypal.com/cgi-bin/webscr?cmd=_donations&business=morvanzhou%40gmail%2ecom&lc=C2&item_name=MorvanPython¤cy_code=AUD&bn=PP%2dDonationsBF%3abtn_donateCC_LG%2egif%3aNonHosted
https://github.com
Termshttps://docs.github.com/site-policy/github-terms/github-terms-of-service
Privacyhttps://docs.github.com/site-policy/privacy-policies/github-privacy-statement
Securityhttps://github.com/security
Statushttps://www.githubstatus.com/
Communityhttps://github.community/
Docshttps://docs.github.com/
Contacthttps://support.github.com?tags=dotcom-footer

Viewport: width=device-width


URLs of crawlers that visited me.