# Passing trained agents into Trainable

**URL:** <https://discuss.ray.io/t/passing-trained-agents-into-trainable/7504>\
**Category:** RLlib\
**Created:** [September 10, 2022, 5:12pm UTC](https://discuss.ray.io/t/passing-trained-agents-into-trainable/7504 "2022-09-10T17:12:05Z")\
**Posts on this page:** 1\
**Showing post:** 3

<div class="post-metadata">

**Author:** ![Kittiwin-Kumlungmak](https://sea2.discourse-cdn.com/flex020/user_avatar/discuss.ray.io/kittiwin-kumlungmak/32/3162_2.png) [@Kittiwin-Kumlungmak](https://discuss.ray.io/u/Kittiwin-Kumlungmak)\
**Post date:** [September 11, 2022, 4:47pm UTC](https://discuss.ray.io/t/passing-trained-agents-into-trainable/7504/3 "2022-09-11T16:47:01Z")

</div>

Hello @arturn

I think what you suggested is not exactly what I have in mind or I may not totally understand what you suggested. Actually, I want to train my “supervisor” in a reinforcement learning fashion rather than a supervised fashion. So, my idea is to have “supervisor” learns to select the right agent, “bull” or “bear”, for trading at the right time. Also, “bull” and “bear”, in this case, have already been trained separately.

I found this [discussion](https://discuss.ray.io/t/rllib-multiagent-with-one-pre-trained-policy-vs-another-adversarial-one/428/3) and I think I can do something similar for my project by passing “bull” and “bear” as policies into multi-agent env.

What do you think?

---

_[View the full topic](https://discuss.ray.io/t/passing-trained-agents-into-trainable/7504)._
