1. I tried with L40 for inference in MON1 with BentoML [@bentoml.service(resources={"gpu": 1}) class MyService: def __init__(self): import torch self.model = torch.load('model.pth').to('cuda:0')] and 2. RTX A4000 for…
Thanks. Looks good but do the GPUs like "1x H100 80GB NVLINK" support Infiniband?. Are these HGX Modules? (Since standard HGX comes in pairs of 4 or 8). Also, when will H200 be available? Great Stuff !
This is super! Best of luck ! BILLING PROFILE: bp7cc3b60e5ce167e99b1ce77
1. I tried with L40 for inference in MON1 with BentoML [@bentoml.service(resources={"gpu": 1}) class MyService: def __init__(self): import torch self.model = torch.load('model.pth').to('cuda:0')] and 2. RTX A4000 for…
Thanks. Looks good but do the GPUs like "1x H100 80GB NVLINK" support Infiniband?. Are these HGX Modules? (Since standard HGX comes in pairs of 4 or 8). Also, when will H200 be available? Great Stuff !
This is super! Best of luck ! BILLING PROFILE: bp7cc3b60e5ce167e99b1ce77