Conference paper Restricted Access

Practical Resolution Methods for MDPs in Robotics Exemplified With Disassembly Planning

Suárez-Hernández, Alejandro; Torras, Carme; Alenyà, Guillem


JSON Export

{
  "owners": [
    78418
  ], 
  "doi": "10.1109/LRA.2019.2901905", 
  "stats": {
    "version_unique_downloads": 1.0, 
    "unique_views": 39.0, 
    "views": 49.0, 
    "version_views": 49.0, 
    "unique_downloads": 1.0, 
    "version_unique_views": 39.0, 
    "volume": 9324760.0, 
    "version_downloads": 5.0, 
    "downloads": 5.0, 
    "version_volume": 9324760.0
  }, 
  "links": {
    "latest_html": "https://zenodo.org/record/3463404", 
    "doi": "https://doi.org/10.1109/LRA.2019.2901905", 
    "badge": "https://zenodo.org/badge/doi/10.1109/LRA.2019.2901905.svg", 
    "html": "https://zenodo.org/record/3463404", 
    "latest": "https://zenodo.org/api/records/3463404"
  }, 
  "created": "2019-09-27T14:21:53.322237+00:00", 
  "updated": "2019-09-27T14:25:09.951308+00:00", 
  "conceptrecid": "3463403", 
  "revision": 2, 
  "id": 3463404, 
  "metadata": {
    "access_right_category": "danger", 
    "doi": "10.1109/LRA.2019.2901905", 
    "description": "<p>In this letter, we focus on finding practical resolution methods for Markov decision processes (MDPs) in robotics. Some of the main difficulties of applying MDPs to real-world robotics problems are: first, having to deal with huge state spaces; and second, designing a method that is robust enough to dead ends. These complications restrict or make more difficult the application of methods, such as value iteration, policy iteration, or labeled real-time dynamic programming (LRTDP). We see in determinization and heuristic search a way to successfully work around these problems. In addition, we believe that many practical use cases offer the opportunity to identify hierarchies of subtasks and solve smaller, simplified problems. We propose a decision-making unit that operates in a probabilistic planning setting through stochastic shortest path problems, which generalize the most common types of MDPs. Our decision-making unit combines: first, automatic hierarchical organization of subtasks; and second, on-line resolution via determinization. We argue that several applications of planning benefit from these two strategies. We exemplify our approach with a robotized disassembly application. The disassembly problem is modeled in probabilistic planning definition language, and serves to define our experiments. Our results show many advantages of our method over LRTDP, such as a better capability to handle problems with large state spaces and state definitions that change when new fluents are discovered.</p>", 
    "language": "eng", 
    "title": "Practical Resolution Methods for MDPs in Robotics Exemplified With Disassembly Planning", 
    "journal": {
      "volume": "4", 
      "issue": "3", 
      "pages": "2282-2288", 
      "title": "IEEE Robotics and Automation Letters"
    }, 
    "relations": {
      "version": [
        {
          "count": 1, 
          "index": 0, 
          "parent": {
            "pid_type": "recid", 
            "pid_value": "3463403"
          }, 
          "is_last": true, 
          "last_child": {
            "pid_type": "recid", 
            "pid_value": "3463404"
          }
        }
      ]
    }, 
    "access_conditions": "<p>RA-Letters are Hybrid Open Access: papers are free of charge to authors, and are freely available to <strong>IEEE RAS members</strong>.</p>", 
    "grants": [
      {
        "code": "731761", 
        "links": {
          "self": "https://zenodo.org/api/grants/10.13039/501100000780::731761"
        }, 
        "title": "Robots Understanding Their Actions by Imagining Their Effects", 
        "acronym": "IMAGINE", 
        "program": "H2020", 
        "funder": {
          "doi": "10.13039/501100000780", 
          "acronyms": [], 
          "name": "European Commission", 
          "links": {
            "self": "https://zenodo.org/api/funders/10.13039/501100000780"
          }
        }
      }
    ], 
    "keywords": [
      "lanning,  scheduling  and  coordination,  hybridlogical/dynamical planning and verification, task planning"
    ], 
    "publication_date": "2019-02-27", 
    "creators": [
      {
        "orcid": "0000-0003-1611-614X", 
        "affiliation": "IRI, CSIC-UPC", 
        "name": "Su\u00e1rez-Hern\u00e1ndez, Alejandro"
      }, 
      {
        "orcid": "0000-0002-2933-398X", 
        "affiliation": "IRI, CSIC-UPC", 
        "name": "Torras, Carme"
      }, 
      {
        "orcid": "0000-0002-6018-154X", 
        "affiliation": "IRI, CSIC-UPC", 
        "name": "Aleny\u00e0, Guillem"
      }
    ], 
    "access_right": "restricted", 
    "resource_type": {
      "subtype": "conferencepaper", 
      "type": "publication", 
      "title": "Conference paper"
    }
  }
}
49
5
views
downloads
Views 49
Downloads 5
Data volume 9.3 MB
Unique views 39
Unique downloads 1

Share

Cite as