# How can I train multiple 'trainer' in same environment?(or embed trained trainer in environment?)

**URL:** <https://discuss.ray.io/t/how-can-i-train-multiple-trainer-in-same-environment-or-embed-trained-trainer-in-environment/8700>\
**Category:** RLlib\
**Created:** [December 19, 2022, 7:21am UTC](https://discuss.ray.io/t/how-can-i-train-multiple-trainer-in-same-environment-or-embed-trained-trainer-in-environment/8700 "2022-12-19T07:21:58Z")\
**Posts on this page:** 1\
**Showing post:** 3

<div class="post-metadata">

**Author:** ![coco](https://avatars.discourse-cdn.com/v4/letter/c/8797f3/32.png) [@coco](https://discuss.ray.io/u/coco)\
**Post date:** [January 9, 2023, 9:09am UTC](https://discuss.ray.io/t/how-can-i-train-multiple-trainer-in-same-environment-or-embed-trained-trainer-in-environment/8700/3 "2023-01-09T09:09:57Z")

</div>

![image](https://us1.discourse-cdn.com/flex020/uploads/ray/original/2X/d/d23f9ccefab737236af8e777ca84e67179e4c5b2.jpeg)  
 ![image](https://us1.discourse-cdn.com/flex020/uploads/ray/original/2X/4/4704edbe16838b6d50ebb138d82146df66cf552d.jpeg)

I am modifying my code by referring to this example.  
Since my source code is made to have multiple agents follow one policy,  
config[‘multiagent’] ['policy\_mapping\_fn '] was like  
" policy\_mapping\_fn = lambda x : original policy "  
By the way, I wanted to make one of these agents follow a policy other than the original policy.  
so,

![image](https://us1.discourse-cdn.com/flex020/uploads/ray/original/2X/4/4de61ae4065f5363ae4ec3dd07855d636948c308.png)

The code was written as shown in the example above.  
And then, it’s execution result seemed that ‘one policy’ was assigned to ‘every agent’ except ‘agent0’. (one policy per one agent)  
Can you tell me how to modify the code to have ‘multiple agents(exept ‘agent0’)’ refer to ‘only one policy’ ?  
thank you. Have a nice day.

---

_[View the full topic](https://discuss.ray.io/t/how-can-i-train-multiple-trainer-in-same-environment-or-embed-trained-trainer-in-environment/8700)._
