{"id":13501,"date":"2025-03-16T13:21:32","date_gmt":"2025-03-16T13:21:32","guid":{"rendered":"https:\/\/www.yiaho.com\/reinforcement-learning-when-machines-learn-through-trial-and-error\/"},"modified":"2025-03-16T13:21:32","modified_gmt":"2025-03-16T13:21:32","slug":"reinforcement-learning-when-machines-learn-through-trial-and-error","status":"publish","type":"post","link":"https:\/\/www.yiaho.com\/en\/reinforcement-learning-when-machines-learn-through-trial-and-error\/","title":{"rendered":"Reinforcement Learning: When machines learn through trial and error"},"content":{"rendered":"<p>Artificial intelligence keeps pushing the limits of what machines can do. Among the approaches shaping this revolution, Reinforcement Learning (or reinforcement learning) stands out for its ability to mimic a very human process: learning by experimenting. <\/p>\n<p>But what exactly is it? How does it work? And why is it so powerful? This article, written by the Yiaho team, takes you on a deep dive into this fascinating field, with clear explanations and concrete examples.   <\/p>\n<h2>What is Reinforcement Learning in AI?<\/h2>\n<p>Reinforcement Learning (RL) is a branch of <a href=\"https:\/\/www.yiaho.com\/en\/machine-learning-what-is-machine-learning-definition-and-explanation\/\" target=\"_blank\" rel=\"noopener\">Machine Learning<\/a> where a machine, called an agent, learns to make decisions by interacting with an environment. Unlike other methods such as supervised learning (where AI is fed labeled data) or unsupervised learning (where it finds patterns without guidance), RL is based on a simple principle: the agent acts, observes the results of its actions, and adjusts its behavior based on the rewards or penalties it receives. <\/p>\n<p>Imagine a child learning to ride a bike. They pedal, fall, get back up, adjust their balance, and eventually ride without help. RL follows a similar logic: the machine learns through trial and error, guided by a feedback system.  <\/p>\n<h2>Key elements of Reinforcement Learning<\/h2>\n<p>To understand reinforcement learning, you need to grasp its core components:<\/p>\n<ul>\n<li><strong>The agent<\/strong>: The entity that makes decisions (for example, a robot, a computer program).<\/li>\n<li><strong>The environment<\/strong>: The world in which the agent operates (a video game, an automated factory).<\/li>\n<li><strong>Actions<\/strong>: The choices the agent can make (turn left, accelerate).<\/li>\n<li><strong>Rewards<\/strong>: A numerical signal the environment sends back to the agent to evaluate its actions (+1 for a good decision, -1 for a mistake).<\/li>\n<li><strong>Policy<\/strong>: The strategy the agent uses to decide its actions based on situations.<\/li>\n<li><strong>State<\/strong>: The current situation of the environment as perceived by the agent (for example, its position in a maze).<\/li>\n<\/ul>\n<p><strong>The agent\u2019s goal?<\/strong> Maximize the total rewards over the long term, even if that sometimes means sacrificing immediate gains for future benefits.<\/p>\n<h3>How does it work? A simple example<\/h3>\n<p>Let\u2019s take a concrete example: a robot learning to get out of a maze.<\/p>\n<ul>\n<li><strong>Initial situation<\/strong>: The robot is placed at the entrance of the maze (initial state).<\/li>\n<li><strong>Possible actions<\/strong>: Go left, right, straight ahead, or back up.<\/li>\n<li><strong>Rewards<\/strong>: +10 if it reaches the exit, -1 if it hits a wall, 0 if it moves forward without incident.<\/li>\n<li><strong>Process<\/strong>: At first, the robot tries actions at random. If it hits a wall, it gets -1 and adjusts its strategy. If it moves toward the exit, it earns positive points. Over time, thanks to an algorithm like Q-Learning (a popular RL method), the robot learns to favor paths that lead to the exit.   <\/li>\n<\/ul>\n<p>As trials go on, the agent doesn\u2019t just fumble around anymore: it develops an optimal policy, almost as if it were mentally drawing a map of the maze.<\/p>\n<h2>Algorithms behind reinforcement learning<\/h2>\n<p>RL relies on sophisticated <a href=\"https:\/\/www.yiaho.com\/en\/what-is-an-algorithm-definition-origin-and-explanation\/\" target=\"_blank\" rel=\"noopener\">algorithms<\/a> that balance exploration (trying new actions) and exploitation (using what already works). Some of the best known include: <\/p>\n<ul>\n<li><strong>Q-Learning<\/strong>: The agent builds a table (Q-table) that assigns values to each state-action pair to estimate future rewards.<\/li>\n<li><strong>Deep Reinforcement Learning<\/strong>: When the environment is too complex (like a video game with millions of possible states), RL is combined with deep neural networks. This is what DeepMind used to create AlphaGo, which beat the best human Go players. <\/li>\n<li><strong>Policy Gradient<\/strong>: Rather than evaluating individual actions, these algorithms directly optimize the agent\u2019s policy.<\/li>\n<\/ul>\n<p>Also read on this topic: <a href=\"https:\/\/www.yiaho.com\/une-intelligence-artificielle-hors-de-controle-elle-triche-pour-gagner-aux-echecs\/\" target=\"_blank\" rel=\"noopener\">An out-of-control AI: it cheats to win at chess<\/a><\/p>\n<h2>Real-world applications of Reinforcement Learning<\/h2>\n<p>Reinforcement learning shines in areas where decisions are sequential and results aren\u2019t immediate. Here are a few examples: <\/p>\n<ul>\n<li><strong>Video games<\/strong>: In 2013, DeepMind developed an AI capable of playing Atari games (like Breakout) by learning only from on-screen pixels and the score. It outperformed humans after a few hours of training. <\/li>\n<li><strong>Robotics<\/strong>: Robots learn to grasp objects or walk by adjusting their movements using RL.<\/li>\n<li><strong>Finance<\/strong>: RL algorithms optimize investment portfolios by testing strategies on market data.<\/li>\n<li><strong>Self-driving cars<\/strong>: A car can learn to navigate traffic by maximizing safety and smooth driving.<\/li>\n<\/ul>\n<p>Also read: <a href=\"https:\/\/www.yiaho.com\/en\/ai-hallucination-why-does-chatgpt-sometimes-invent-answers\/\" target=\"_blank\" rel=\"noopener\">AI hallucinations: Why does ChatGPT sometimes make up answers?<\/a><\/p>\n<h2>Advantages and limitations<\/h2>\n<h3>Advantages:<\/h3>\n<ul>\n<li>RL is flexible and doesn\u2019t require pre-labeled data.<\/li>\n<li>It excels in dynamic and uncertain environments.<\/li>\n<\/ul>\n<h3>Limitations:<\/h3>\n<ul>\n<li>It requires a lot of time and computation, because the agent has to experiment extensively.<\/li>\n<li>Designing a relevant reward system is tricky: poor design can lead to unexpected behaviors (for example, an agent that cheats to maximize points instead of solving the problem).<\/li>\n<\/ul>\n<h2>Why is Reinforcement Learning revolutionary in AI?<\/h2>\n<p>Reinforcement Learning pushes the boundaries of AI by enabling it to adapt to unpredictable situations without explicit instructions. It\u2019s a step toward <a href=\"https:\/\/www.yiaho.com\/en\/chatgpt-5-free-unlimited\/\" target=\"_blank\" rel=\"noopener\">artificial general intelligence<\/a> (AGI), which Yiaho or OpenAI are trying to develop, where machines could learn like humans, through experience. <\/p>\n<p>Whether it\u2019s beating champions at Go or optimizing production lines, RL shows that machines can not only perform tasks, but also learn how to learn.<br \/>\nConclusion<\/p>\n<p>Reinforcement Learning is a powerful approach that illustrates machines\u2019 ability to improve on their own. By combining trials, errors, and rewards, it opens up incredible possibilities, from video games to robotics and everyday life. If you\u2019re curious about AI, keep an eye on this field: it could well be at the heart of the next major technological breakthroughs.  <\/p>\n","protected":false},"excerpt":{"rendered":"<p>Artificial intelligence keeps pushing the limits of what machines can do. Among the approaches shaping this revolution, Reinforcement Learning (or reinforcement learning) stands out for its ability to mimic a very human process: learning by experimenting. But what exactly is it? How does it work? And why is it so powerful? This article, written by&hellip;&nbsp;<a href=\"https:\/\/www.yiaho.com\/en\/reinforcement-learning-when-machines-learn-through-trial-and-error\/\" rel=\"bookmark\">Read More &raquo;<span class=\"screen-reader-text\">Reinforcement Learning: When machines learn through trial and error<\/span><\/a><\/p>\n","protected":false},"author":4,"featured_media":13504,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"neve_meta_sidebar":"","neve_meta_container":"","neve_meta_enable_content_width":"off","neve_meta_content_width":70,"neve_meta_title_alignment":"","neve_meta_author_avatar":"","neve_post_elements_order":"","neve_meta_disable_header":"","neve_meta_disable_footer":"","neve_meta_disable_title":"","neve_meta_reading_time":"","footnotes":""},"categories":[50],"tags":[],"class_list":["post-13501","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-ai-glossary"],"yoast_head":"<!-- This site is optimized with the Yoast SEO plugin v28.3 - https:\/\/yoast.com\/product\/yoast-seo-wordpress\/ -->\n<title>What is &quot;Reinforcement Learning&quot;? Explanation and examples<\/title>\n<meta name=\"description\" content=\"Discover Reinforcement Learning: how machines learn through trial and error, from AI and robotics to video games.\" \/>\n<meta name=\"robots\" content=\"index, follow, max-snippet:-1, max-image-preview:large, max-video-preview:-1\" \/>\n<link rel=\"canonical\" href=\"https:\/\/www.yiaho.com\/en\/reinforcement-learning-when-machines-learn-through-trial-and-error\/\" \/>\n<meta property=\"og:locale\" content=\"en_US\" \/>\n<meta property=\"og:type\" content=\"article\" \/>\n<meta property=\"og:title\" content=\"What is &quot;Reinforcement Learning&quot;? Explanation and examples\" \/>\n<meta property=\"og:description\" content=\"Discover Reinforcement Learning: how machines learn through trial and error, from AI and robotics to video games.\" \/>\n<meta property=\"og:url\" content=\"https:\/\/www.yiaho.com\/en\/reinforcement-learning-when-machines-learn-through-trial-and-error\/\" \/>\n<meta property=\"og:site_name\" content=\"YIAHO\" \/>\n<meta property=\"article:publisher\" content=\"https:\/\/www.facebook.com\/yiaho.ia.gratuite\" \/>\n<meta property=\"article:published_time\" content=\"2025-03-16T13:21:32+00:00\" \/>\n<meta property=\"og:image\" content=\"https:\/\/www.yiaho.com\/wp-content\/uploads\/2025\/03\/Reinforcement-Learning-definition-exemple.webp\" \/>\n\t<meta property=\"og:image:width\" content=\"1200\" \/>\n\t<meta property=\"og:image:height\" content=\"800\" \/>\n\t<meta property=\"og:image:type\" content=\"image\/webp\" \/>\n<meta name=\"author\" content=\"G. de Yiaho\" \/>\n<meta name=\"twitter:card\" content=\"summary_large_image\" \/>\n<meta name=\"twitter:creator\" content=\"@Yiaho_AI\" \/>\n<meta name=\"twitter:site\" content=\"@Yiaho_AI\" \/>\n<meta name=\"twitter:label1\" content=\"Written by\" \/>\n\t<meta name=\"twitter:data1\" content=\"G. de Yiaho\" \/>\n\t<meta name=\"twitter:label2\" content=\"Est. reading time\" \/>\n\t<meta name=\"twitter:data2\" content=\"4 minutes\" \/>\n<script type=\"application\/ld+json\" class=\"yoast-schema-graph\">{\"@context\":\"https:\\\/\\\/schema.org\",\"@graph\":[{\"@type\":\"Article\",\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/reinforcement-learning-when-machines-learn-through-trial-and-error\\\/#article\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/reinforcement-learning-when-machines-learn-through-trial-and-error\\\/\"},\"author\":{\"name\":\"G. de Yiaho\",\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/#\\\/schema\\\/person\\\/09fcc2462849b463e2b2511013897d80\"},\"headline\":\"Reinforcement Learning: When machines learn through trial and error\",\"datePublished\":\"2025-03-16T13:21:32+00:00\",\"mainEntityOfPage\":{\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/reinforcement-learning-when-machines-learn-through-trial-and-error\\\/\"},\"wordCount\":904,\"commentCount\":0,\"publisher\":{\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/#organization\"},\"image\":{\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/reinforcement-learning-when-machines-learn-through-trial-and-error\\\/#primaryimage\"},\"thumbnailUrl\":\"https:\\\/\\\/www.yiaho.com\\\/wp-content\\\/uploads\\\/2025\\\/03\\\/Reinforcement-Learning-definition-exemple.webp\",\"articleSection\":[\"AI Glossary\"],\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"CommentAction\",\"name\":\"Comment\",\"target\":[\"https:\\\/\\\/www.yiaho.com\\\/en\\\/reinforcement-learning-when-machines-learn-through-trial-and-error\\\/#respond\"]}]},{\"@type\":\"WebPage\",\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/reinforcement-learning-when-machines-learn-through-trial-and-error\\\/\",\"url\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/reinforcement-learning-when-machines-learn-through-trial-and-error\\\/\",\"name\":\"What is \\\"Reinforcement Learning\\\"? Explanation and examples\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/#website\"},\"primaryImageOfPage\":{\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/reinforcement-learning-when-machines-learn-through-trial-and-error\\\/#primaryimage\"},\"image\":{\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/reinforcement-learning-when-machines-learn-through-trial-and-error\\\/#primaryimage\"},\"thumbnailUrl\":\"https:\\\/\\\/www.yiaho.com\\\/wp-content\\\/uploads\\\/2025\\\/03\\\/Reinforcement-Learning-definition-exemple.webp\",\"datePublished\":\"2025-03-16T13:21:32+00:00\",\"description\":\"Discover Reinforcement Learning: how machines learn through trial and error, from AI and robotics to video games.\",\"breadcrumb\":{\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/reinforcement-learning-when-machines-learn-through-trial-and-error\\\/#breadcrumb\"},\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"ReadAction\",\"target\":[\"https:\\\/\\\/www.yiaho.com\\\/en\\\/reinforcement-learning-when-machines-learn-through-trial-and-error\\\/\"]}]},{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/reinforcement-learning-when-machines-learn-through-trial-and-error\\\/#primaryimage\",\"url\":\"https:\\\/\\\/www.yiaho.com\\\/wp-content\\\/uploads\\\/2025\\\/03\\\/Reinforcement-Learning-definition-exemple.webp\",\"contentUrl\":\"https:\\\/\\\/www.yiaho.com\\\/wp-content\\\/uploads\\\/2025\\\/03\\\/Reinforcement-Learning-definition-exemple.webp\",\"width\":1200,\"height\":800,\"caption\":\"What is reinforcement learning in artificial intelligence? Definition and examples\"},{\"@type\":\"BreadcrumbList\",\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/reinforcement-learning-when-machines-learn-through-trial-and-error\\\/#breadcrumb\",\"itemListElement\":[{\"@type\":\"ListItem\",\"position\":1,\"name\":\"Accueil\",\"item\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/\"},{\"@type\":\"ListItem\",\"position\":2,\"name\":\"Reinforcement Learning: When machines learn through trial and error\"}]},{\"@type\":\"WebSite\",\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/#website\",\"url\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/\",\"name\":\"YIAHO\",\"description\":\"L&#039;intelligence Artificielle gratuite en ligne et fran\u00e7aise\",\"publisher\":{\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/#organization\"},\"potentialAction\":[{\"@type\":\"SearchAction\",\"target\":{\"@type\":\"EntryPoint\",\"urlTemplate\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/?s={search_term_string}\"},\"query-input\":{\"@type\":\"PropertyValueSpecification\",\"valueRequired\":true,\"valueName\":\"search_term_string\"}}],\"inLanguage\":\"en-US\"},{\"@type\":\"Organization\",\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/#organization\",\"name\":\"Yiaho\",\"url\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/\",\"logo\":{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/#\\\/schema\\\/logo\\\/image\\\/\",\"url\":\"https:\\\/\\\/www.yiaho.com\\\/wp-content\\\/uploads\\\/2025\\\/11\\\/YIAHO-logo.webp\",\"contentUrl\":\"https:\\\/\\\/www.yiaho.com\\\/wp-content\\\/uploads\\\/2025\\\/11\\\/YIAHO-logo.webp\",\"width\":417,\"height\":424,\"caption\":\"Yiaho\"},\"image\":{\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/#\\\/schema\\\/logo\\\/image\\\/\"},\"sameAs\":[\"https:\\\/\\\/www.facebook.com\\\/yiaho.ia.gratuite\",\"https:\\\/\\\/x.com\\\/Yiaho_AI\"]},{\"@type\":\"Person\",\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/en\\\/#\\\/schema\\\/person\\\/09fcc2462849b463e2b2511013897d80\",\"name\":\"G. de Yiaho\",\"image\":{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/www.yiaho.com\\\/wp-content\\\/uploads\\\/2025\\\/11\\\/YIAHO-logo-96x96.webp\",\"url\":\"https:\\\/\\\/www.yiaho.com\\\/wp-content\\\/uploads\\\/2025\\\/11\\\/YIAHO-logo-96x96.webp\",\"contentUrl\":\"https:\\\/\\\/www.yiaho.com\\\/wp-content\\\/uploads\\\/2025\\\/11\\\/YIAHO-logo-96x96.webp\",\"caption\":\"G. de Yiaho\"}}]}<\/script>\n<!-- \/ Yoast SEO plugin. -->","yoast_head_json":{"title":"What is \"Reinforcement Learning\"? Explanation and examples","description":"Discover Reinforcement Learning: how machines learn through trial and error, from AI and robotics to video games.","robots":{"index":"index","follow":"follow","max-snippet":"max-snippet:-1","max-image-preview":"max-image-preview:large","max-video-preview":"max-video-preview:-1"},"canonical":"https:\/\/www.yiaho.com\/en\/reinforcement-learning-when-machines-learn-through-trial-and-error\/","og_locale":"en_US","og_type":"article","og_title":"What is \"Reinforcement Learning\"? Explanation and examples","og_description":"Discover Reinforcement Learning: how machines learn through trial and error, from AI and robotics to video games.","og_url":"https:\/\/www.yiaho.com\/en\/reinforcement-learning-when-machines-learn-through-trial-and-error\/","og_site_name":"YIAHO","article_publisher":"https:\/\/www.facebook.com\/yiaho.ia.gratuite","article_published_time":"2025-03-16T13:21:32+00:00","og_image":[{"width":1200,"height":800,"url":"https:\/\/www.yiaho.com\/wp-content\/uploads\/2025\/03\/Reinforcement-Learning-definition-exemple.webp","type":"image\/webp"}],"author":"G. de Yiaho","twitter_card":"summary_large_image","twitter_creator":"@Yiaho_AI","twitter_site":"@Yiaho_AI","twitter_misc":{"Written by":"G. de Yiaho","Est. reading time":"4 minutes"},"schema":{"@context":"https:\/\/schema.org","@graph":[{"@type":"Article","@id":"https:\/\/www.yiaho.com\/en\/reinforcement-learning-when-machines-learn-through-trial-and-error\/#article","isPartOf":{"@id":"https:\/\/www.yiaho.com\/en\/reinforcement-learning-when-machines-learn-through-trial-and-error\/"},"author":{"name":"G. de Yiaho","@id":"https:\/\/www.yiaho.com\/en\/#\/schema\/person\/09fcc2462849b463e2b2511013897d80"},"headline":"Reinforcement Learning: When machines learn through trial and error","datePublished":"2025-03-16T13:21:32+00:00","mainEntityOfPage":{"@id":"https:\/\/www.yiaho.com\/en\/reinforcement-learning-when-machines-learn-through-trial-and-error\/"},"wordCount":904,"commentCount":0,"publisher":{"@id":"https:\/\/www.yiaho.com\/en\/#organization"},"image":{"@id":"https:\/\/www.yiaho.com\/en\/reinforcement-learning-when-machines-learn-through-trial-and-error\/#primaryimage"},"thumbnailUrl":"https:\/\/www.yiaho.com\/wp-content\/uploads\/2025\/03\/Reinforcement-Learning-definition-exemple.webp","articleSection":["AI Glossary"],"inLanguage":"en-US","potentialAction":[{"@type":"CommentAction","name":"Comment","target":["https:\/\/www.yiaho.com\/en\/reinforcement-learning-when-machines-learn-through-trial-and-error\/#respond"]}]},{"@type":"WebPage","@id":"https:\/\/www.yiaho.com\/en\/reinforcement-learning-when-machines-learn-through-trial-and-error\/","url":"https:\/\/www.yiaho.com\/en\/reinforcement-learning-when-machines-learn-through-trial-and-error\/","name":"What is \"Reinforcement Learning\"? Explanation and examples","isPartOf":{"@id":"https:\/\/www.yiaho.com\/en\/#website"},"primaryImageOfPage":{"@id":"https:\/\/www.yiaho.com\/en\/reinforcement-learning-when-machines-learn-through-trial-and-error\/#primaryimage"},"image":{"@id":"https:\/\/www.yiaho.com\/en\/reinforcement-learning-when-machines-learn-through-trial-and-error\/#primaryimage"},"thumbnailUrl":"https:\/\/www.yiaho.com\/wp-content\/uploads\/2025\/03\/Reinforcement-Learning-definition-exemple.webp","datePublished":"2025-03-16T13:21:32+00:00","description":"Discover Reinforcement Learning: how machines learn through trial and error, from AI and robotics to video games.","breadcrumb":{"@id":"https:\/\/www.yiaho.com\/en\/reinforcement-learning-when-machines-learn-through-trial-and-error\/#breadcrumb"},"inLanguage":"en-US","potentialAction":[{"@type":"ReadAction","target":["https:\/\/www.yiaho.com\/en\/reinforcement-learning-when-machines-learn-through-trial-and-error\/"]}]},{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/www.yiaho.com\/en\/reinforcement-learning-when-machines-learn-through-trial-and-error\/#primaryimage","url":"https:\/\/www.yiaho.com\/wp-content\/uploads\/2025\/03\/Reinforcement-Learning-definition-exemple.webp","contentUrl":"https:\/\/www.yiaho.com\/wp-content\/uploads\/2025\/03\/Reinforcement-Learning-definition-exemple.webp","width":1200,"height":800,"caption":"What is reinforcement learning in artificial intelligence? Definition and examples"},{"@type":"BreadcrumbList","@id":"https:\/\/www.yiaho.com\/en\/reinforcement-learning-when-machines-learn-through-trial-and-error\/#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Accueil","item":"https:\/\/www.yiaho.com\/en\/"},{"@type":"ListItem","position":2,"name":"Reinforcement Learning: When machines learn through trial and error"}]},{"@type":"WebSite","@id":"https:\/\/www.yiaho.com\/en\/#website","url":"https:\/\/www.yiaho.com\/en\/","name":"YIAHO","description":"L&#039;intelligence Artificielle gratuite en ligne et fran\u00e7aise","publisher":{"@id":"https:\/\/www.yiaho.com\/en\/#organization"},"potentialAction":[{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https:\/\/www.yiaho.com\/en\/?s={search_term_string}"},"query-input":{"@type":"PropertyValueSpecification","valueRequired":true,"valueName":"search_term_string"}}],"inLanguage":"en-US"},{"@type":"Organization","@id":"https:\/\/www.yiaho.com\/en\/#organization","name":"Yiaho","url":"https:\/\/www.yiaho.com\/en\/","logo":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/www.yiaho.com\/en\/#\/schema\/logo\/image\/","url":"https:\/\/www.yiaho.com\/wp-content\/uploads\/2025\/11\/YIAHO-logo.webp","contentUrl":"https:\/\/www.yiaho.com\/wp-content\/uploads\/2025\/11\/YIAHO-logo.webp","width":417,"height":424,"caption":"Yiaho"},"image":{"@id":"https:\/\/www.yiaho.com\/en\/#\/schema\/logo\/image\/"},"sameAs":["https:\/\/www.facebook.com\/yiaho.ia.gratuite","https:\/\/x.com\/Yiaho_AI"]},{"@type":"Person","@id":"https:\/\/www.yiaho.com\/en\/#\/schema\/person\/09fcc2462849b463e2b2511013897d80","name":"G. de Yiaho","image":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/www.yiaho.com\/wp-content\/uploads\/2025\/11\/YIAHO-logo-96x96.webp","url":"https:\/\/www.yiaho.com\/wp-content\/uploads\/2025\/11\/YIAHO-logo-96x96.webp","contentUrl":"https:\/\/www.yiaho.com\/wp-content\/uploads\/2025\/11\/YIAHO-logo-96x96.webp","caption":"G. de Yiaho"}}]}},"_links":{"self":[{"href":"https:\/\/www.yiaho.com\/en\/wp-json\/wp\/v2\/posts\/13501","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.yiaho.com\/en\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.yiaho.com\/en\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.yiaho.com\/en\/wp-json\/wp\/v2\/users\/4"}],"replies":[{"embeddable":true,"href":"https:\/\/www.yiaho.com\/en\/wp-json\/wp\/v2\/comments?post=13501"}],"version-history":[{"count":0,"href":"https:\/\/www.yiaho.com\/en\/wp-json\/wp\/v2\/posts\/13501\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.yiaho.com\/en\/wp-json\/wp\/v2\/media\/13504"}],"wp:attachment":[{"href":"https:\/\/www.yiaho.com\/en\/wp-json\/wp\/v2\/media?parent=13501"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.yiaho.com\/en\/wp-json\/wp\/v2\/categories?post=13501"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.yiaho.com\/en\/wp-json\/wp\/v2\/tags?post=13501"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}