Skip to main navigation Skip to search Skip to main content

Enabling cooperative behavior for building demand response based on extended joint action learning

  • L.A. Hurtado Munoz
  • , E. Mocanu
  • , H.P. Nguyen
  • , M. Gibescu
  • , I.G. Kamphuis

Research output: Contribution to journalArticleAcademicpeer-review

598 Downloads (Pure)

Abstract

This paper explores the use of distributed intelligence to assist the integration of the demand as a flexible resource, to mitigate the emerging uncertainty in the power system, while fulfilling the customer's local needs, i.e., comfort management. More exactly, our contribution is twofold. First, we propose a novel cooperative and decentralized reinforcement learning method, dubbed extended joint action learning (eJAL). Second, we perform a comparison between eJAL to noncooperative decentralized decision making strategies, i.e., Q-learning, and a centralized game theoretic approach, i.e., Nash n-player game. This comparison has been conducted on the basis of grid support effectiveness and the loss of comfort for each customer. Various metrics were used to analyze the advantages and disadvantages of each method. We demonstrated that a range of flexibility requests can be met by providing an optimal energy portfolio of buildings without substantially violating comfort constraints. Moreover, we showed that the proposed eJAL method achieves the highest fairness index.

Original languageEnglish
Article number8039194
Pages (from-to)127-136
Number of pages10
JournalIEEE Transactions on Industrial Informatics
Volume14
Issue number1
DOIs
Publication statusPublished - 1 Jan 2018

UN SDGs

This output contributes to the following UN Sustainable Development Goals (SDGs)

  1. SDG 7 - Affordable and Clean Energy
    SDG 7 Affordable and Clean Energy

Keywords

  • Cooperation
  • decision-making
  • demand flexibility
  • demand response (DR)
  • distributed control
  • multiagent systems (MAS)
  • smart grid

Fingerprint

Dive into the research topics of 'Enabling cooperative behavior for building demand response based on extended joint action learning'. Together they form a unique fingerprint.

Cite this