We are hiring an AI Engineer to deploy, fine-tune, and maintain production-grade machine learning models and large language models.
Deploy, fine-tune, and maintain production-grade ML models and LLMs (such as GPT-4, Claude, or open-source variants like Llama).
Build and optimize robust data pipelines to preprocess massive datasets for training and real-time inference.
Architect scalable microservices and APIs to serve AI-driven features with minimal latency.
Continuously evaluate model performance, cost, and bias while monitoring live production systems.
3+ years of professional software engineering experience, with at least 1-2 years focused on deploying AI/ML models in production systems.
Strong proficiency in Python and standard machine learning frameworks (e.g., PyTorch, TensorFlow, or JAX).
Hands-on experience with LLM orchestration tools like LangChain, LlamaIndex, and vector databases (e.g., Pinecone, Milvus, Chroma).
Experience with cloud infrastructure (AWS, GCP, or Azure) and containerized workflows using Docker and Kubernetes.
Bachelor’s or Master's degree in Computer Science, Data Science, Mathematics, or a related technical field (or equivalent practical experience).