首页> 外国专利> Delayed processing for arm policy determination for content management system messaging

Delayed processing for arm policy determination for content management system messaging

机译:用于内容管理系统消息传递的ARM政策确定的延迟处理

摘要

Techniques are provided for delayed processing for arm policy determination for content management system messaging, including, during a delayed processing window, receiving reward data for arm actions taken, where the arm actions were chosen based on a previous version of an arm choice policy, and the previous version of the arm choice policy was determined based on a previous set of reward data for a previous set of arm actions taken. When the delayed processing window has closed, a new arm choice policy is determined based at least in part on the action-reward data, and the previous set of reward data and/or the previous arm choice policy. After a request to choose an arm choice is received, a particular arm action to take is determined based on the new arm choice policy. This chosen arm is provided in response to the request.
机译:提供了用于内容管理系统消息传递的ARM策略确定的延迟处理的技术,包括在延迟处理窗口期间接收用于拍摄的ARM动作的奖励数据,其中基于先前版本的ARM选择策略选择ARM动作,以及 以前的ARM选择策略版本是根据采取的前一组ARM行动的上一组奖励数据确定的。 当延迟处理窗口已关闭时,至少部分地基于Action-right数据和上一组奖励数据和/或先前的ARM选择策略确定新的ARM选择策略。 在收到选择ARM选择之后,基于新的ARM选择政策来确定要采取的特定ARM动作。 根据请求提供此选择的臂。

著录项

相似文献

  • 专利
  • 外文文献
  • 中文文献
获取专利

客服邮箱:kefu@zhangqiaokeyan.com

京公网安备:11010802029741号 ICP备案号:京ICP备15016152号-6 六维联合信息科技 (北京) 有限公司©版权所有
  • 客服微信

  • 服务号