AI Infrastructure / ML Engineer in Wyoming, MI

Published on CazVidat Capa Cloud

Join Capa Cloud as an AI Infrastructure / ML Engineer in Wyoming, MI. Optimize AI workloads, GPU computing, and model deployment. Full-time, remote eligible.

Verified by CazVid2 months agoOpen until: Sep 18, 2026

Salary

$5,000 per year

Location

United States

Employment type

Full time

Workplace

Remote

AI Infrastructure / ML Engineer in Wyoming, MI

$5,000 per year

Apply now

Job description

Capa Cloud is seeking a skilled AI Infrastructure / ML Engineer to join our team in Wyoming, Michigan. This full-time role focuses on optimizing AI workloads, model deployment, training, inference, and scalable GPU utilization on the CapaCloud platform. As an AI Infrastructure / ML Engineer, you will collaborate closely with infrastructure engineers to support modern AI workflows tailored for startups, researchers, and enterprise users. This position is perfect for professionals passionate about AI systems, MLOps, and large-scale GPU computing. Key Responsibilities Build and optimize AI deployment pipelines to enhance efficiency Improve GPU workload performance for AI applications Support infrastructure for AI training and inference tasks Optimize performance for PyTorch, TensorFlow, and large language model (LLM) workloads Develop scalable APIs and inference systems Create benchmarking and performance testing tools Collaborate with infrastructure teams on orchestration and containerization systems Support model deployment and containerized AI workloads Enhance developer experience for AI users on the platform Monitor and optimize AI compute resource utilization and performance Required Skills & Experience Proven experience with AI/ML infrastructure and MLOps practices Strong Python programming skills Hands-on experience with PyTorch, TensorFlow, or JAX frameworks Expertise in GPU computing and CUDA environments Familiarity with containerized deployment systems such as Docker Experience deploying AI models in production environments Understanding of inference optimization techniques and best practices Experience working with APIs and backend systems Strong debugging and analytical problem-solving skills Nice To Have Experience with large language model (LLM) infrastructure Familiarity with the Hugging Face ecosystem Experience with distributed training systems Knowledge of Kubernetes and orchestration platforms Experience with AI inference optimization tools Contributions to open-source AI projects What Success Looks Like Success in this role means delivering scalable, efficient AI infrastructure that empowers users to deploy and run AI workloads seamlessly, while continuously improving GPU utilization and developer experience. Location: Wyoming, Michigan, United States (Remote eligible) Salary: $5,000 per year Employment Type: Full-time

Similar jobs