• Home
  • Search
  • Risk-aware multi-armed bandit problem with application to portfolio selection
  • Cite Icon72
  • https://doi.org/10.1098/rsos.171377Copy DOI Icon

Risk-aware multi-armed bandit problem with application to portfolio selection

Show More
  • Abstract
  • Highlights & Summary
  • PDF
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Sequential portfolio selection has attracted increasing interest in the machine learning and quantitative finance communities in recent years. As a mathematical framework for reinforcement learning policies, the stochastic multi-armed bandit problem addresses the primary difficulty in sequential decision-making under uncertainty, namely the exploration versus exploitation dilemma, and therefore provides a natural connection to portfolio selection. In this paper, we incorporate risk awareness into the classic multi-armed bandit setting and introduce an algorithm to construct portfolio. Through filtering assets based on the topological structure of the financial market and combining the optimal multi-armed bandit policy with the minimization of a coherent risk measure, we achieve a balance between risk and return.

Loading PDF

Similar Papers
  • Conference Article
  • Citations11

Approximation Algorithms for Restless Bandit Problems

  • Jan 04, 2009
  • Sudipto Guha +2
  • Research Article

Multi-Armed Bandits: Algorithms, Applications, and Future Directions

  • Jan 29, 2026
  • Academic Journal of Science and Technology
  • Yurong Zheng
  • Research Article
  • Citations1

An optimal selection for ensembles of influential projects

  • Feb 13, 2020
  • Annals of Operations Research
  • Kemal Gürsoy
  • Research Article

Understanding the stochastic dynamics of sequential decision-making processes: A path-integral analysis of multi-armed bandits.

  • Jun 01, 2023
  • Chaos (Woodbury, N.Y.)
  • Bo Li +1
  • Research Article
  • Citations5

Adaptive Cyber Defense Technique Based on Multiagent Reinforcement Learning Strategies

  • Jan 01, 2023
  • Intelligent Automation & Soft Computing
  • Adel Alshamrani +1
  • Single Book
  • Citations19

Mathematical Analysis of Machine Learning Algorithms

  • Jul 20, 2023
  • Tong Zhang
  • Conference Article
  • Citations13

Thompson Sampling on Symmetric Alpha-Stable Bandits

  • Aug 01, 2019
  • Abhimanyu Dubey +1
  • Research Article

Modified Index Policies for Multi-Armed Bandits with Network-like Markovian Dependencies

  • Jan 29, 2025
  • Network
  • Abdalaziz Sawwan +1
  • PDF
  • Research Article
  • Citations6

Optimal Policy for Bernoulli Bandits: Computation and Algorithm Gauge

  • Feb 01, 2021
  • IEEE Transactions on Artificial Intelligence
  • Sebastian Pilarski +2
  • PDF
  • Single Report

Creative Cognition as a Bandit Problem

  • May 01, 2023
  • Noémie Berlin +4
  • Research Article
  • Citations8

Dynamic Pricing and Placing for Distributed Machine Learning Jobs: An Online Learning Approach

  • Apr 01, 2023
  • IEEE Journal on Selected Areas in Communications
  • Ruiting Zhou +3
  • Research Article
  • Citations70

On Optimality of Myopic Policy for Restless Multi-Armed Bandit Problem: An Axiomatic Approach

  • Jan 01, 2012
  • IEEE Transactions on Signal Processing
  • Kehao Wang +1
  • Conference Article
  • Citations16

Performance Evaluation of Machine Learning Based Channel Selection Algorithm Implemented on IoT Sensor Devices in Coexisting IoT Networks

  • Jan 01, 2020
  • So Hasegawa +3
  • Conference Article

Online Kernel Selection via Grouped Adversarial Bandit Model

  • Nov 01, 2019
  • Junfan Li +1
  • PDF
  • Research Article
  • Citations15

Mean‐ portfolio selection and ‐arbitrage for coherent risk measures

  • Aug 09, 2021
  • Mathematical Finance
  • Martin Herdegen +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.