{"id":8264,"date":"2024-01-19T12:03:42","date_gmt":"2024-01-19T17:03:42","guid":{"rendered":"https:\/\/www.clearobject.com\/?post_type=glossary&#038;p=8264"},"modified":"2024-01-19T12:03:42","modified_gmt":"2024-01-19T17:03:42","slug":"reinforcement-learning","status":"publish","type":"glossary","link":"https:\/\/www.clearobject.com\/?glossary=reinforcement-learning","title":{"rendered":"Reinforcement Learning"},"content":{"rendered":"<p><span data-sheets-root=\"1\" data-sheets-value=\"{&quot;1&quot;:2,&quot;2&quot;:&quot;Reinforcement Learning&quot;}\" data-sheets-userformat=\"{&quot;2&quot;:513,&quot;3&quot;:{&quot;1&quot;:0},&quot;12&quot;:0}\">Reinforcement Learning is a\u00a0field of study within artificial intelligence focused on training agents via a pre-defined reward systems. The models are able to interact with a system however they see fit in oder to maximize the reward function. Common examples are models allowed to play simple video games to learn how to get the most points, or using models to determine the best advertisement to show a certain customer.<\/span><\/p>\n","protected":false},"excerpt":{"rendered":"<p>Reinforcement Learning is a\u00a0field of study within artificial intelligence focused on training agents via a pre-defined reward systems. The models are able to interact with a system however they see fit in oder to maximize the reward function. Common examples are models allowed to play simple video games to learn how to get the most [&hellip;]<\/p>\n","protected":false},"author":11,"featured_media":0,"menu_order":0,"template":"","meta":{"_et_pb_use_builder":"","_et_pb_old_content":"","_et_gb_content_width":"","content-type":"","_monsterinsights_skip_tracking":false,"_uf_show_specific_survey":0,"_uf_disable_surveys":false,"footnotes":""},"class_list":["post-8264","glossary","type-glossary","status-publish","hentry"],"yoast_head":"<!-- This site is optimized with the Yoast SEO Premium plugin v28.1 (Yoast SEO v28.1) - https:\/\/yoast.com\/product\/yoast-seo-premium-wordpress\/ -->\n<title>Reinforcement Learning - Clear Object<\/title>\n<meta name=\"description\" content=\"AI focused on training agents via a pre-defined reward systems. The models are able to interact to maximize results.\" \/>\n<meta name=\"robots\" content=\"index, follow, max-snippet:-1, max-image-preview:large, max-video-preview:-1\" \/>\n<link rel=\"canonical\" href=\"https:\/\/www.clearobject.com\/?glossary=reinforcement-learning\" \/>\n<meta property=\"og:locale\" content=\"en_US\" \/>\n<meta property=\"og:type\" content=\"article\" \/>\n<meta property=\"og:title\" content=\"Reinforcement Learning\" \/>\n<meta property=\"og:description\" content=\"AI focused on training agents via a pre-defined reward systems. The models are able to interact to maximize results.\" \/>\n<meta property=\"og:url\" content=\"https:\/\/www.clearobject.com\/?glossary=reinforcement-learning\" \/>\n<meta property=\"og:site_name\" content=\"Clear Object\" \/>\n<meta property=\"article:publisher\" content=\"https:\/\/www.facebook.com\/ClearObjectInc\" \/>\n<meta name=\"twitter:card\" content=\"summary_large_image\" \/>\n<meta name=\"twitter:site\" content=\"@ClearObject\" \/>\n<meta name=\"twitter:label1\" content=\"Est. reading time\" \/>\n\t<meta name=\"twitter:data1\" content=\"1 minute\" \/>\n<script type=\"application\/ld+json\" class=\"yoast-schema-graph\">{\"@context\":\"https:\\\/\\\/schema.org\",\"@graph\":[{\"@type\":\"WebPage\",\"@id\":\"https:\\\/\\\/www.clearobject.com\\\/?glossary=reinforcement-learning\",\"url\":\"https:\\\/\\\/www.clearobject.com\\\/?glossary=reinforcement-learning\",\"name\":\"Reinforcement Learning - Clear Object\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/www.clearobject.com\\\/#website\"},\"datePublished\":\"2024-01-19T17:03:42+00:00\",\"description\":\"AI focused on training agents via a pre-defined reward systems. The models are able to interact to maximize results.\",\"breadcrumb\":{\"@id\":\"https:\\\/\\\/www.clearobject.com\\\/?glossary=reinforcement-learning#breadcrumb\"},\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"ReadAction\",\"target\":[\"https:\\\/\\\/www.clearobject.com\\\/?glossary=reinforcement-learning\"]}]},{\"@type\":\"BreadcrumbList\",\"@id\":\"https:\\\/\\\/www.clearobject.com\\\/?glossary=reinforcement-learning#breadcrumb\",\"itemListElement\":[{\"@type\":\"ListItem\",\"position\":1,\"name\":\"Home\",\"item\":\"https:\\\/\\\/www.clearobject.com\\\/\"},{\"@type\":\"ListItem\",\"position\":2,\"name\":\"Reinforcement Learning\"}]},{\"@type\":\"WebSite\",\"@id\":\"https:\\\/\\\/www.clearobject.com\\\/#website\",\"url\":\"https:\\\/\\\/www.clearobject.com\\\/\",\"name\":\"ClearObject\",\"description\":\"Transforming data into valuable outcomes\",\"publisher\":{\"@id\":\"https:\\\/\\\/www.clearobject.com\\\/#organization\"},\"potentialAction\":[{\"@type\":\"SearchAction\",\"target\":{\"@type\":\"EntryPoint\",\"urlTemplate\":\"https:\\\/\\\/www.clearobject.com\\\/?s={search_term_string}\"},\"query-input\":{\"@type\":\"PropertyValueSpecification\",\"valueRequired\":true,\"valueName\":\"search_term_string\"}}],\"inLanguage\":\"en-US\"},{\"@type\":\"Organization\",\"@id\":\"https:\\\/\\\/www.clearobject.com\\\/#organization\",\"name\":\"ClearObject\",\"url\":\"https:\\\/\\\/www.clearobject.com\\\/\",\"logo\":{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/www.clearobject.com\\\/#\\\/schema\\\/logo\\\/image\\\/\",\"url\":\"https:\\\/\\\/www.clearobject.com\\\/wp-content\\\/uploads\\\/2023\\\/09\\\/co_teal_logo.svg\",\"contentUrl\":\"https:\\\/\\\/www.clearobject.com\\\/wp-content\\\/uploads\\\/2023\\\/09\\\/co_teal_logo.svg\",\"width\":1,\"height\":1,\"caption\":\"ClearObject\"},\"image\":{\"@id\":\"https:\\\/\\\/www.clearobject.com\\\/#\\\/schema\\\/logo\\\/image\\\/\"},\"sameAs\":[\"https:\\\/\\\/www.facebook.com\\\/ClearObjectInc\",\"https:\\\/\\\/x.com\\\/ClearObject\",\"https:\\\/\\\/www.linkedin.com\\\/company\\\/clearobject\\\/\"]}]}<\/script>\n<!-- \/ Yoast SEO Premium plugin. -->","yoast_head_json":{"title":"Reinforcement Learning - Clear Object","description":"AI focused on training agents via a pre-defined reward systems. The models are able to interact to maximize results.","robots":{"index":"index","follow":"follow","max-snippet":"max-snippet:-1","max-image-preview":"max-image-preview:large","max-video-preview":"max-video-preview:-1"},"canonical":"https:\/\/www.clearobject.com\/?glossary=reinforcement-learning","og_locale":"en_US","og_type":"article","og_title":"Reinforcement Learning","og_description":"AI focused on training agents via a pre-defined reward systems. The models are able to interact to maximize results.","og_url":"https:\/\/www.clearobject.com\/?glossary=reinforcement-learning","og_site_name":"Clear Object","article_publisher":"https:\/\/www.facebook.com\/ClearObjectInc","twitter_card":"summary_large_image","twitter_site":"@ClearObject","twitter_misc":{"Est. reading time":"1 minute"},"schema":{"@context":"https:\/\/schema.org","@graph":[{"@type":"WebPage","@id":"https:\/\/www.clearobject.com\/?glossary=reinforcement-learning","url":"https:\/\/www.clearobject.com\/?glossary=reinforcement-learning","name":"Reinforcement Learning - Clear Object","isPartOf":{"@id":"https:\/\/www.clearobject.com\/#website"},"datePublished":"2024-01-19T17:03:42+00:00","description":"AI focused on training agents via a pre-defined reward systems. The models are able to interact to maximize results.","breadcrumb":{"@id":"https:\/\/www.clearobject.com\/?glossary=reinforcement-learning#breadcrumb"},"inLanguage":"en-US","potentialAction":[{"@type":"ReadAction","target":["https:\/\/www.clearobject.com\/?glossary=reinforcement-learning"]}]},{"@type":"BreadcrumbList","@id":"https:\/\/www.clearobject.com\/?glossary=reinforcement-learning#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https:\/\/www.clearobject.com\/"},{"@type":"ListItem","position":2,"name":"Reinforcement Learning"}]},{"@type":"WebSite","@id":"https:\/\/www.clearobject.com\/#website","url":"https:\/\/www.clearobject.com\/","name":"ClearObject","description":"Transforming data into valuable outcomes","publisher":{"@id":"https:\/\/www.clearobject.com\/#organization"},"potentialAction":[{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https:\/\/www.clearobject.com\/?s={search_term_string}"},"query-input":{"@type":"PropertyValueSpecification","valueRequired":true,"valueName":"search_term_string"}}],"inLanguage":"en-US"},{"@type":"Organization","@id":"https:\/\/www.clearobject.com\/#organization","name":"ClearObject","url":"https:\/\/www.clearobject.com\/","logo":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/www.clearobject.com\/#\/schema\/logo\/image\/","url":"https:\/\/www.clearobject.com\/wp-content\/uploads\/2023\/09\/co_teal_logo.svg","contentUrl":"https:\/\/www.clearobject.com\/wp-content\/uploads\/2023\/09\/co_teal_logo.svg","width":1,"height":1,"caption":"ClearObject"},"image":{"@id":"https:\/\/www.clearobject.com\/#\/schema\/logo\/image\/"},"sameAs":["https:\/\/www.facebook.com\/ClearObjectInc","https:\/\/x.com\/ClearObject","https:\/\/www.linkedin.com\/company\/clearobject\/"]}]}},"_links":{"self":[{"href":"https:\/\/www.clearobject.com\/index.php?rest_route=\/wp\/v2\/glossary\/8264","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.clearobject.com\/index.php?rest_route=\/wp\/v2\/glossary"}],"about":[{"href":"https:\/\/www.clearobject.com\/index.php?rest_route=\/wp\/v2\/types\/glossary"}],"author":[{"embeddable":true,"href":"https:\/\/www.clearobject.com\/index.php?rest_route=\/wp\/v2\/users\/11"}],"version-history":[{"count":0,"href":"https:\/\/www.clearobject.com\/index.php?rest_route=\/wp\/v2\/glossary\/8264\/revisions"}],"wp:attachment":[{"href":"https:\/\/www.clearobject.com\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=8264"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}