AI Systems Engineer - DevOps& Observability Manager
Job Description
About this kind of role
What a AI Systems Engineer - DevOps& Observability Manager usually does. The employer's own details are in the listing.
Typically does: This role focuses on the design, implementation, and maintenance of infrastructure supporting artificial intelligence systems, with a strong emphasis on DevOps practices and observability. Responsibilities include automating deployment pipelines, monitoring system performance, and ensuring the reliability and scalability of AI infrastructure. The engineer will collaborate with data scientists and machine learning engineers to optimize the AI development lifecycle. Troubleshooting and resolving production issues related to AI systems and their underlying infrastructure are also key aspects.
Tools and skills: Proficiency in cloud platforms (e.g., AWS, Azure, GCP), containerization technologies (Docker, Kubernetes), scripting languages (Python, Bash), configuration management tools (Ansible, Terraform), and observability platforms (Prometheus, Grafana) is generally expected.
Good fit for: Individuals with a strong background in DevOps, systems engineering, and a passion for enabling AI innovation thrive in this role.
??? people applied to this job
View Similar Jobs
Similar jobs which you may be interested in. Typically using your existing skillset.
$100,000 - $150,000/Mo
2 years ago
Site Reliability Engineer
Team Remotely Inc.San Diego, USA
Site Reliability
Engineer
Infrastructure Maintenance
$100,000 - $150,000/Mo
2 years ago
Junior Site Reliability Engineer
Patterned Learning Career.San Diego, USA
Site Reliability
Engineering
Junior
$60,000 - $90,000/Mo
2 years ago
Site Reliability Engineer
Phoenix Recruitment.San Diego, USA
Site Reliability Engineer
DevOps
System Administrator
$95,000 - $160,000/Mo
2 years ago
Staff Site Reliability Engineer
2K.San Diego, USA
Site Reliability Engineering
Staff Engineer
Infrastructure Operations
$90,000 - $130,000/Mo
2 years ago
$90,000 - $150,000/Mo
2 years ago
Sr. Site Reliability Engineer
Veza.San Diego, USA
Site Reliability Engineer
Senior Engineer
Systems Architect
$120,000 - $160,000/Mo
2 years ago
Senior AWS Cloud & Site Reliability Engineer
INFY8.COM.San Diego, USA
AWS Expert
Cloud Infrastructure Specialist
Senior Site Reliability Engineer
$120,000 - $160,000/Mo
2 years ago
$110,000 - $150,000/Mo
2 years ago
Senior Systems Administrator
Ingenta.San Diego, USA
Systems Administration
IT Management
Network Infrastructure
$80,000 - $120,000/Mo
2 years ago