In sagemaker, the docs talk about inference scripts requiring to have 4 specific functions. When we get a prediction, the python SDK sends a request to the endpoint.
Then the inference script runs. But I cannot find where in the SDK the inference script is run.
When I navigate through the sdk code the Predictor.predict() method calls the sagemaker session to post a request to the endpoint and get a response. That is the final step in the sdk. Sagemaker is obviously doing something when it receives that request.
What is the code that it runs?