Description
We need an AI/ML Engineer to help us deploy a small Hugging Face language model behind a FastAPI endpoint. The goal is simple. We want to send a text prompt to an API and receive a generated response back in a clean JSON format. This is a small proof of concept, so we do not need a complex production setup. The task includes setting up the FastAPI project, loading a suitable Hugging Face model, creating a text generation endpoint, adding basic error handling, testing the API with sample prompts, and providing short setup instructions. Experience with Python, FastAPI, Hugging Face, open source LLMs, model inference, text generation, Generative AI, or AI API development is required. Please share a brief example if you have worked with LLM deployment, Hugging Face models, or FastAPI based AI APIs before.