Discussion about this post

User's avatar
Avik De's avatar

I was looking for a survey of a similar topic, thanks! Are there any approaches where some RL is used in the loop to robustify the new task performance?

1 more comment...

No posts

Ready for more?