# Bets way to handling policy change

**URL:** <https://discuss.ray.io/t/bets-way-to-handling-policy-change/6165>\
**Category:** RLlib\
**Created:** [May 17, 2022, 11:00am UTC](https://discuss.ray.io/t/bets-way-to-handling-policy-change/6165 "2022-05-17T11:00:06Z")\
**Posts on this page:** 1\
**Page:** 1

<div class="post-metadata">

**Author:** ![hossein836](https://sea2.discourse-cdn.com/flex020/user_avatar/discuss.ray.io/hossein836/32/2559_2.png) [@hossein836](https://discuss.ray.io/u/hossein836)\
**Post date:** [May 17, 2022, 11:00am UTC](https://discuss.ray.io/t/bets-way-to-handling-policy-change/6165/1 "2022-05-17T11:00:06Z")

</div>

I tested my env and I didn’t get the results that I expected, I repeated my RL with slightly modifications in model and etc but i got same results. the problem is results getting better until some big updates, I thought PPO clip param should address this problem but obviously it doesn’t (from what I get). I remember I saw an article (TL;DR) arguing PPO doesn’t solve this problem always and TRPO is more general but I don’t know what was the reason. any guide that how can I get through this?  
 ![2022-05-17 15_06_55-Bokeh Plot and 16 more pages - Personal - Microsoft​ Edge](https://us1.discourse-cdn.com/flex020/uploads/ray/original/2X/0/06465e1a2d6a11b14502d609cd8f619b1aaeb571.png)  
 ![2022-05-17 15_06_27-Settings](https://us1.discourse-cdn.com/flex020/uploads/ray/original/2X/c/c41a29bb61363a9c8ec814b3b2862b42d692d963.png)
