Site Reliability Engineer

Onboard

Role: Site Reliability Engineer (Remote)Location: Remote (Work from Anywhere)Payout: $40 - $70 /hour Role Overview: Join one of our clients, a global leader in the Artificial Intelligence industry, as a Site Reliability Engineer for a specialized project centered on training and optimizing AI models within cutting-edge containerized infrastructures. This role is a contractor position, requiring a systems-first approach, real-time troubleshooting, and dynamic process recovery. The ideal candidate will have a strong background in terminal-native problem-solving and containerized environment mastery. Key Responsibilities: • Lead the deployment, monitoring, and recovery of complex, containerized AI training environments using advanced terminal techniques. • Proactively identify, diagnose, and resolve infrastructure bottlenecks and failures in long-running processes. • Orchestrate resilient system builds and infrastructure management, ensuring stability and optimal resource utilization. • Collaborate with cross-functional teams to implement process improvements and ensure seamless integration with existing systems. • Continuously monitor and analyze system performance to identify areas for optimization. Required Skills & Qualifications: • Terminal-native problem-solving skills with expertise in Linux and containerized environments. • Experience with Python programming language and its use in AI model development. • Strong understanding of microservices architecture and container orchestration tools such as Kubernetes. • Excellent knowledge of network protocols and infrastructure management. • Ability to work independently and as part of a distributed team, with strong communication and collaboration skills. More About the Opportunity: This contract position offers a competitive hourly rate of $40-$70, with opportunities for future extension or transition into advanced phases for standout performers. Our client values expertise and innovation, and is committed to creating a collaborative and dynamic work environment. Equal Opportunity Employer: We hire based on skills and expertise. All qualified candidates are welcome regardless of background, experience, or prior employment history. Applications are reviewed solely on demonstrated technical ability and qualifications. Apply Now!

Last checked on May 24, 2026. We may earn a commission when you click through.

Advertisement
This listing may have expired or is no longer available. Information shown may be outdated.

Site Reliability Engineer

Onboard

Updated 3 months ago
Apply now

You'll be redirected to ae.linkedin.com

Dubai Remote متعاقد 📅 قبل يومين

About this role

Role: Site Reliability Engineer (Remote)Location: Remote (Work from Anywhere)Payout: $40 - $70 /hour

Role Overview:

Join one of our clients, a global leader in the Artificial Intelligence industry, as a Site Reliability Engineer for a specialized project centered on training and optimizing AI models within cutting-edge containerized infrastructures. This role is a contractor position, requiring a systems-first approach, real-time troubleshooting, and dynamic process recovery. The ideal candidate will have a strong background in terminal-native problem-solving and containerized environment mastery.

Key Responsibilities:

• Lead the deployment, monitoring, and recovery of complex, containerized AI training environments using advanced terminal techniques.

• Proactively identify, diagnose, and resolve infrastructure bottlenecks and failures in long-running processes.

• Orchestrate resilient system builds and infrastructure management, ensuring stability and optimal resource utilization.

• Collaborate with cross-functional teams to implement process improvements and ensure seamless integration with existing systems.

• Continuously monitor and analyze system performance to identify areas for optimization.

Required Skills & Qualifications:

• Terminal-native problem-solving skills with expertise in Linux and containerized environments.

• Experience with Python programming language and its use in AI model development.

• Strong understanding of microservices architecture and container orchestration tools such as Kubernetes.

• Excellent knowledge of network protocols and infrastructure management.

• Ability to work independently and as part of a distributed team, with strong communication and collaboration skills.

More About the Opportunity:

This contract position offers a competitive hourly rate of $40-$70, with opportunities for future extension or transition into advanced phases for standout performers. Our client values expertise and innovation, and is committed to creating a collaborative and dynamic work environment.

Equal Opportunity Employer:

We hire based on skills and expertise. All qualified candidates are welcome regardless of background, experience, or prior employment history. Applications are reviewed solely on demonstrated technical ability and qualifications.

Apply Now!

About the Company

Onboard is a prominent player in the Artificial Intelligence sector, focusing on innovative technology solutions.

Key Highlights

  • Remote position with flexible working hours.
  • Competitive pay rate of $40 - $70/hour.
  • Involves deploying and monitoring advanced AI training environments.
  • Requires expertise in terminal problem-solving and container orchestration.
  • Dynamic recovery processes for real-time troubleshooting.

💡 Honest Take: This role offers a unique opportunity to work with cutting-edge AI technologies, but candidates should be prepared for the complexities of remote systems management.

Apply for this position

You'll be redirected to ae.linkedin.com

You might also like

Related Articles