As a Senior AI Engineer at OPSWAT, you will lead the deployment and optimization of Large Language Models (LLMs) for high-performance inference on GPU hardware. Your responsibilities include designing local LLM serving environments, optimizing model performance through quantization and compression techniques, and managing production serving infrastructure. You will collaborate with various teams to ensure efficient and cost-effective AI service delivery.
See how your résumé matches this role — and tailor it from what actually gets interviews.