Verified today
DevOps Platform Engineer AI Evaluation
About the role
Gramian Consultancy is hiring a DevOps / Platform Engineer (AI Evaluation) to support the development of advanced AI systems through realistic engineering tasks based on modern cloud and infrastructure environments. The role combines hands-on engineering with evaluation and technical review, designing and reviewing production-style scenarios across CI/CD, containers, Kubernetes, cloud infrastructure, Infrastructure as Code, observability, and incident response. This is a remote contract role requiring a minimum of 20 hours per week, open to candidates in India, Romania, Brazil, France, Argentina, and Poland.
What you’ll do
- Design realistic DevOps and infrastructure tasks based on production-style scenarios.
- Create working reference solutions with clear setup, execution, and validation steps.
- Develop scenarios covering CI/CD, containers, orchestration, Infrastructure as Code, monitoring, and incident response.
- Review AI-generated technical outputs for correctness, completeness, and reproducibility.
- Identify configuration, infrastructure, or implementation failures and provide precise corrections.
- Validate that technical tasks can be reproduced consistently in the intended environment.
- Apply a high quality bar to infrastructure design, automation, and operational workflows.
- Document technical reasoning, expected outcomes, and troubleshooting steps clearly.
What you’ll bring
- At least 3 years of professional experience in DevOps, SRE, Platform Engineering, or Infrastructure Engineering.
- Hands-on coding ability in addition to infrastructure and automation expertise.
- Proficiency in Python, Go, Bash, or TypeScript.
- Experience with CI/CD, Docker, Kubernetes, and Linux.
- Ability to design realistic DevOps and infrastructure tasks based on production-style scenarios.
- Ability to create working reference solutions with clear setup, execution, and validation steps.
- Ability to review AI-generated technical outputs for correctness, completeness, and reproducibility.
- Ability to identify configuration, infrastructure, or implementation failures and provide precise corrections.