kth.sePublikationer KTH
Ändra sökning
RefereraExporteraLänk till posten
Permanent länk

Direktlänk
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annat format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annat språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf
Reducing the Learning Time of Reinforcement Learning for the Supervisory Control of Discrete Event Systems
School of Electro-Mechanical Engineering, Xidian University, Xi’an, China.ORCID-id: 0000-0002-5988-0335
KTH, Skolan för industriell teknik och management (ITM), Maskinkonstruktion, Mekatronik och inbyggda styrsystem.ORCID-id: 0000-0003-4535-3849
KTH, Skolan för industriell teknik och management (ITM), Maskinkonstruktion, Mekatronik och inbyggda styrsystem.ORCID-id: 0000-0001-5703-5923
Department of Industrial Engineering, College of Engineering, King Saud University, Riyadh, Saudi Arabia.ORCID-id: 0000-0003-3559-6249
Visa övriga samt affilieringar
2023 (Engelska)Ingår i: IEEE Access, E-ISSN 2169-3536, Vol. 11, s. 59840-59853Artikel i tidskrift (Refereegranskat) Published
Abstract [en]

Reinforcement learning (RL) can obtain the supervisory controller for discrete-event systems modeled by finite automata and temporal logic. The published methods often have two limitations. First, a large number of training data are required to learn the RL controller. Second, the RL algorithms do not consider uncontrollable events, which are essential for supervisory control theory (SCT). To address the limitations, we first apply SCT to find the supervisors for the specifications modeled by automata. These supervisors remove illegal training data violating these specifications and hence reduce the exploration space of the RL algorithm. For the remaining specifications modeled by temporal logic, the RL algorithm is applied to search for the optimal control decision within the confined exploration space. Uncontrollable events are considered by the RL algorithm as uncertainties in the plant model. The proposed method can obtain a nonblocking supervisor for all specifications with less learning time than the published methods.

Ort, förlag, år, upplaga, sidor
IEEE, 2023. Vol. 11, s. 59840-59853
Nyckelord [en]
Discrete event system, linear temporal logic, supervisory control theory, reinforcement learning
Nationell ämneskategori
Reglerteknik
Forskningsämne
Tillämpad matematik och beräkningsmatematik, Optimeringslära och systemteori; Datalogi; Industriella informations- och styrsystem
Identifikatorer
URN: urn:nbn:se:kth:diva-330695DOI: 10.1109/access.2023.3285432ISI: 001018594800001Scopus ID: 2-s2.0-85163172875OAI: oai:DiVA.org:kth-330695DiVA, id: diva2:1778207
Projekt
XPRES
Forskningsfinansiär
XPRES - Initiative for excellence in production research
Anmärkning

QC 20230704

Tillgänglig från: 2023-06-30 Skapad: 2023-06-30 Senast uppdaterad: 2023-07-13Bibliografiskt granskad

Open Access i DiVA

Fulltext saknas i DiVA

Övriga länkar

Förlagets fulltextScopushttps://ieeexplore.ieee.org/abstract/document/10149832/authors#authors

Person

Tan, KaigeFeng, Lei

Sök vidare i DiVA

Av författaren/redaktören
Yang, JunjunTan, KaigeFeng, LeiEl-Sherbeeny, Ahmed M.Li, Zhiwu
Av organisationen
Mekatronik och inbyggda styrsystem
I samma tidskrift
IEEE Access
Reglerteknik

Sök vidare utanför DiVA

GoogleGoogle Scholar

doi
urn-nbn

Altmetricpoäng

doi
urn-nbn
Totalt: 324 träffar
RefereraExporteraLänk till posten
Permanent länk

Direktlänk
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annat format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annat språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf