Unanswered
Hello, How Do You Manage To Unload A Model From Clearml-Serving Api?
I Am Trying To Unload A Model Through Grpc Via
Thank you for your answer, I added 100s models in the serving session, and when I send a post request it loads the willing model to perform an inference. I would like to be able to send a request to unload the model (because I cannot load all the models in gpu, only 7-8) or as @<1690896098534625280:profile|NarrowWoodpecker99> suggests add a timeout ? Or unload all the models if the gpu memory reach a limit ? Do you have a suggestion on how I could achieve that? Thanks!
56 Views
0
Answers
4 months ago
4 months ago