{"id":183139,"date":"2024-03-18T02:16:00","date_gmt":"2024-03-18T01:16:00","guid":{"rendered":"https:\/\/liora.io\/en\/?p=183139"},"modified":"2026-08-09T19:43:28","modified_gmt":"2026-08-09T18:43:28","slug":"sarsa-how-does-machine-learning-work","status":"publish","type":"post","link":"https:\/\/liora.io\/en\/sarsa-how-does-machine-learning-work","title":{"rendered":"SARSA: How does Machine Learning work?"},"content":{"rendered":"\n<p><strong>Reinforcement learning is, along with supervised and unsupervised learning, one of the three major machine learning techniques.<\/strong><\/p>\n\n\n<p>This family of algorithms has been creating a lot of buzz in recent years, with innovative products from the <a href=\"https:\/\/liora.io\/en\/unveiling-the-future-a-comprehensive-guide-to-the-open-ai-api\">OpenAI<\/a> company such as<strong> OpenAI Five,<\/strong> an AI that managed to beat a team of professional players on the Dota 2 video game, or the<a href=\"https:\/\/liora.io\/en\/chatgpt-how-does-this-nlp-algorithm-work\"> famous ChatGPT<\/a>, which uses this technique to adjust its parameters.<\/p>\n\n\n<div class=\"wp-block-buttons is-layout-flex wp-block-buttons-is-layout-flex is-content-justification-center wp-container-core-buttons-is-layout-5ee10de4\" style=\"margin-top:32px;margin-bottom:32px\"><div class=\"wp-block-button\"><a class=\"wp-block-button__link wp-element-button\" href=\"\/en\/courses\/data-ai\/data-scientist\">Learn all about reinforcement learning<\/a><\/div><\/div>\n\n\n<h2 class=\"wp-block-heading\" id=\"what-is-reinforcement-learning\">What is reinforcement learning?<\/h2>\n\n\n<p>Reinforcement learning is a field of <a href=\"https:\/\/liora.io\/en\/machine-learning-engineer-bootcamp-why-is-it-interesting\">machine learning<\/a> in which an agent (virtual entity: robot, program, etc.) is placed in an interactive environment in which it must learn to perform actions that maximize quantitative rewards.<\/p>\n\n\n<div class=\"wp-block-group has-background has-global-padding is-layout-constrained wp-container-core-group-is-layout-7441ce17 wp-block-group-is-layout-constrained\" style=\"border-left-color:#ff5c43;border-left-width:4px;border-radius:12px;background-color:#fff5f2;margin-top:32px;margin-bottom:32px;padding-top:24px;padding-right:28px;padding-bottom:24px;padding-left:28px\">\n\n<p style=\"margin-top:0;margin-bottom:16px;font-size:clamp(14.642px, 0.915rem + ((1vw - 3.2px) * 0.575), 22px);font-style:normal;font-weight:600\">Related articles<\/p>\n\n\n<ul class=\"wp-block-list\" style=\"margin-top:0;margin-bottom:0;padding-left:20px\">\n<li><a href=\"https:\/\/liora.io\/en\/image-processing-fundamental-principles-and-practical-uses\" rel=\"noopener\" target=\"_blank\">Image Processing<\/a><\/li>\n<li><a href=\"https:\/\/liora.io\/en\/all-about-deep-learning\" rel=\"noopener\" target=\"_blank\">Deep Learning ,  All you need to know<\/a><\/li>\n<li><a href=\"https:\/\/liora.io\/en\/mushroom-recognition\" rel=\"noopener\" target=\"_blank\">Mushroom Recognition<\/a><\/li>\n<li><a href=\"https:\/\/liora.io\/en\/tensor-flow-all-about-googles-machine-learning-framework\" rel=\"noopener\" target=\"_blank\">Tensor Flow ,  Google&#8217;s ML<\/a><\/li>\n<li><a href=\"https:\/\/liora.io\/en\/unlock-your-future-dive-into-machine-learning-engineer-training\" rel=\"noopener\" target=\"_blank\">Dive into ML<\/a><\/li>\n<\/ul>\n\n<\/div>\n\n\n<h2 class=\"wp-block-heading\" id=\"what-is-the-sarsa-algorithm\">What is the SARSA algorithm?<\/h2>\n\n\n<p><strong>SARSA<\/strong> is a learning algorithm whose name comes from State-Action-Reward-State-Action, meaning State-Action-Reward-State-Action, and refers to the sequence of elements that make up the algorithm. It is an algorithm based on a table of action values (or Q-table, Q representing the measure of the quality of an action performed) which assigns to each state-action pair a value representing the expected reward.<\/p>\n\n\n<h2 class=\"wp-block-heading\" id=\"conclusion\">Conclusion<\/h2>\n\n\n<p>In summary, SARSA is a reinforcement learning algorithm that aims to teach an agent the decisions to be made in an environment by means of an iteratively updated Q-table. It follows a policy of exploration and exploitation while interacting with the environment, and is used in various fields such as video games, decision-making in robotics, or solving path planning problems.<\/p>\n\n\n<figure class=\"wp-block-image size-full\" style=\"margin-top:32px;margin-bottom:32px\"><img alt=\"Illustration for Conclusion\" decoding=\"async\" height=\"600\" loading=\"lazy\" src=\"https:\/\/liora.io\/app\/uploads\/2023\/11\/SARSA-2.jpg\" style=\"width:100%;height:auto\" width=\"762\"\/><\/figure>\n\n\n<p>If you&#8217;d like to learn more about this field, take a look at our Data Scientist training course.<\/p>\n\n\n<div class=\"wp-block-buttons is-layout-flex wp-block-buttons-is-layout-flex is-content-justification-center wp-container-core-buttons-is-layout-5ee10de4\" style=\"margin-top:32px;margin-bottom:32px\"><div class=\"wp-block-button\"><a class=\"wp-block-button__link wp-element-button\" href=\"\/en\/courses\/data-ai\/data-scientist\">Discover our Data Scientist training<\/a><\/div><\/div>\n\n","protected":false},"excerpt":{"rendered":"<p>Reinforcement learning is, along with supervised and unsupervised learning, one of the three major machine learning techniques. This family of algorithms has been creating a lot of buzz in recent years, with innovative products from the OpenAI company such as OpenAI Five, an AI that managed to beat a team of professional players on the [&hellip;]<\/p>\n","protected":false},"author":76,"featured_media":183141,"comment_status":"open","ping_status":"open","sticky":false,"template":"elementor_theme","format":"standard","meta":{"_acf_changed":false,"editor_notices":[],"footnotes":""},"categories":[2433],"class_list":["post-183139","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-data-ai"],"acf":[],"_links":{"self":[{"href":"https:\/\/liora.io\/en\/wp-json\/wp\/v2\/posts\/183139","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/liora.io\/en\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/liora.io\/en\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/liora.io\/en\/wp-json\/wp\/v2\/users\/76"}],"replies":[{"embeddable":true,"href":"https:\/\/liora.io\/en\/wp-json\/wp\/v2\/comments?post=183139"}],"version-history":[{"count":3,"href":"https:\/\/liora.io\/en\/wp-json\/wp\/v2\/posts\/183139\/revisions"}],"predecessor-version":[{"id":211074,"href":"https:\/\/liora.io\/en\/wp-json\/wp\/v2\/posts\/183139\/revisions\/211074"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/liora.io\/en\/wp-json\/wp\/v2\/media\/183141"}],"wp:attachment":[{"href":"https:\/\/liora.io\/en\/wp-json\/wp\/v2\/media?parent=183139"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/liora.io\/en\/wp-json\/wp\/v2\/categories?post=183139"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}