Next
Livestream will start soon!
Livestream has already ended.
Presentation has not been recorded yet!
  • title: Quantile Constrained Reinforcement Learning: A Reinforcement Learning Framework Constraining Outage Probability
      0:00 / 0:00
      • Report Issue
      • Settings
      • Playlists
      • Bookmarks
      • Subtitles Off
      • Playback rate
      • Quality
      • Settings
      • Debug information
      • Server sl-yoda-v2-stream-007-alpha.b-cdn.net
      • Subtitles size Medium
      • Bookmarks
      • Server
      • sl-yoda-v2-stream-007-alpha.b-cdn.net
      • sl-yoda-v2-stream-007-beta.b-cdn.net
      • 1678031076.rsc.cdn77.org
      • 1932936657.rsc.cdn77.org
      • Subtitles
      • Off
      • English
      • Playback rate
      • Quality
      • Subtitles size
      • Large
      • Medium
      • Small
      • Mode
      • Video Slideshow
      • Audio Slideshow
      • Slideshow
      • Video
      My playlists
        Bookmarks
          00:00:00
            Quantile Constrained Reinforcement Learning: A Reinforcement Learning Framework Constraining Outage Probability
            • Settings
            • Sync diff
            • Quality
            • Settings
            • Server
            • Quality
            • Server

            Quantile Constrained Reinforcement Learning: A Reinforcement Learning Framework Constraining Outage Probability

            Nov 28, 2022

            Speakers

            WJ

            Whiyoung Jung

            Speaker · 0 followers

            MC

            Myungsik Cho

            Speaker · 0 followers

            JP

            Jongeui Park

            Speaker · 0 followers

            About

            Constrained reinforcement learning (RL) is an area of RL whose objective is to find an optimal policy that maximizes expected cumulative return while satisfying a given constraint. Most of the previous constrained RL works consider expected cumulative sum cost as the constraint. However, optimization with this constraint cannot guarantee a target probability of outage event that the cumulative sum cost exceeds a given threshold. This paper proposes a framework, named Quantile Constrained RL (QCR…

            Organizer

            N2
            N2

            NeurIPS 2022

            Account · 953 followers

            Like the format? Trust SlidesLive to capture your next event!

            Professional recording and live streaming, delivered globally.

            Sharing

            Recommended Videos

            Presentations on similar topic, category or speaker

            Transfer Learning with Deep Tabular Models
            16:07

            Transfer Learning with Deep Tabular Models

            Roman Levin, …

            N2
            N2
            NeurIPS 2022 2 years ago

            Total of 0 viewers voted for saving the presentation to eternal vault which is 0.0%

            Efficient Planning in a Compact Latent Action Space
            09:18

            Efficient Planning in a Compact Latent Action Space

            Zhengyao Jiang, …

            N2
            N2
            NeurIPS 2022 2 years ago

            Total of 0 viewers voted for saving the presentation to eternal vault which is 0.0%

            Lossless Compression of Deep Neural Networks: A High-dimensional Neural Tangent Kernel Approach
            01:00

            Lossless Compression of Deep Neural Networks: A High-dimensional Neural Tangent Kernel Approach

            Lingyu Gu, …

            N2
            N2
            NeurIPS 2022 2 years ago

            Total of 1 viewers voted for saving the presentation to eternal vault which is 0.1%

            Adjusting the Gender Wage Gap with a Low-Dimensional Representation of Job History
            02:48

            Adjusting the Gender Wage Gap with a Low-Dimensional Representation of Job History

            Keyon Vafa, …

            N2
            N2
            NeurIPS 2022 2 years ago

            Total of 0 viewers voted for saving the presentation to eternal vault which is 0.0%

            Quantization-Based Optimization : Alternative Stochastic Approximation of Global Optimization
            05:00

            Quantization-Based Optimization : Alternative Stochastic Approximation of Global Optimization

            Jinwuk Seok, …

            N2
            N2
            NeurIPS 2022 2 years ago

            Total of 0 viewers voted for saving the presentation to eternal vault which is 0.0%

            OccGen: Selection of Real-world Multilingual Parallel Data Balanced in Gender within Occupations
            04:43

            OccGen: Selection of Real-world Multilingual Parallel Data Balanced in Gender within Occupations

            Marta Costa-jussà, …

            N2
            N2
            NeurIPS 2022 2 years ago

            Total of 0 viewers voted for saving the presentation to eternal vault which is 0.0%

            Interested in talks like this? Follow NeurIPS 2022