跳到主要导航 跳到搜索 跳到主要内容

SC-COO: A feedback-based service composition algorithm combining offline and online reinforcement learning

  • Xiaoming Yu*
  • , Wenjun Wu
  • , Jiadong Wang
  • , Xin Ji
  • *此作品的通讯作者
  • Beijing Wuzi University
  • Beihang University
  • State Grid Corporation of China

科研成果: 期刊稿件文章同行评审

摘要

Faced with the current dynamic service environment, rapid and efficient service composition has attracted much attention in recent years. The service composition could complete the reuse of existing services and its ultimate goal is to better satisfy users. However, it is challenging to interact with the service environment to collect data in practical applications due to factors such as high cost and risk. To overcome this limitation, this paper proposes the SC-COO method: A feedback-based service composition algorithm combining offline and online reinforcement learning. The SC-COO method mainly consists of two stages: the offline training module (SC-COO-offline) is the main stage, and the online update module (SC-COO-online) is the auxiliary stage. The SC-COO-offline model is trained through collected offline data, avoiding the drawback of online learning requiring multiple iterations to converge. And online training (SC-COO-online) serves as an auxiliary stage to jointly make decisions and recommend services to users to better adapt to dynamic environments. Furthermore, our SC-COO method offers users’ score preferences in service composition by designing a feedback-based reward mechanism. Continuous interactive feedback with humans can significantly improve the robustness of the service composition system. Finally, some experiments on the RapidAPI dataset demonstrate that SC-COO outperforms other baselines in accuracy, scalability, and convergence. And some results of the ablation experiment also verify the efficiency and applicability of SC-COO.

源语言英语
期刊论文编号806
期刊Applied Intelligence
55
11
DOI
出版状态已出版 - 7月 2025

学术指纹

探究 'SC-COO: A feedback-based service composition algorithm combining offline and online reinforcement learning' 的科研主题。它们共同构成独一无二的学术指纹。

引用此