1 research outputs found

    Balancing Exploitation and Exploration via Fully Probabilistic Design of Decision Policies

    No full text
    Adaptive decision making learns an environment model serving a design of a decision policy. The policy-generated actions influence both the acquired reward and the future knowledge. The optimal policy properly balances exploitation with exploration. The inherent dimensionality\ncurse of decision making under incomplete knowledge prevents the realisation of the optimal design
    corecore