Facebook/bart-large-mnli inference when deployed on SageMaker

striki-ai · April 29, 2022, 5:12am

Hi,

Any idea/documentation on how to compose json payload for inferencing facebook/bart-large-mnli model, deployed on SageMaker, used the code as provided on the "Deploy → Amazon SageMaker → AWS:

from sagemaker.huggingface import HuggingFaceModel
import sagemaker

role = sagemaker.get_execution_role()

Hub Model configuration. Models - Hugging Face

hub = {
‘HF_MODEL_ID’:‘facebook/bart-large-mnli’,
‘HF_TASK’:‘zero-shot-classification’
}

create Hugging Face Model Class

huggingface_model = HuggingFaceModel(
transformers_version=‘4.17.0’,
pytorch_version=‘1.10.2’,
py_version=‘py38’,
env=hub,
role=role,
)

deploy model to SageMaker Inference

predictor = huggingface_model.deploy(
initial_instance_count=1, # number of instances
instance_type=‘ml.m5.xlarge’ # ec2 instance type
)

philschmid · April 29, 2022, 6:35am

Hello @striki-ai,

This documentation link should help you: Reference

example

predictor.predict({
  "inputs": "Hi, I recently bought a device from your company but it is not working as advertised and I would like to get reimbursed!",
  "parameters": {
    "candidate_labels": ["refund", "legal", "faq"]
  }
})

Topic		Replies	Views
Client_error_from_model Amazon SageMaker	1	428	March 15, 2023
Fine tuning facebook/bart-large-mnli zeroshot classifier Intermediate	2	905	June 30, 2023
Infer with SageMaker for a Private Model Amazon SageMaker	3	2423	June 30, 2022
Running batch transform in Sagemaker on a Huggingface model from the Hub with parameters Beginners	2	1708	February 2, 2023
Deploy big model to AWS Sagemaker fails Beginners	5	1079	July 31, 2023

Facebook/bart-large-mnli inference when deployed on SageMaker

Hub Model configuration. Models - Hugging Face

create Hugging Face Model Class

deploy model to SageMaker Inference

Related topics