O que a vaga pede
JOB DESCRIPTION
Insight Global is looking for a Site Reliability Engineer remote in LATAM to join a AAA game company for a contract opportunity. You will be joining the Ops vertical of this engineering team and help keep AI-powered developer tools, Copilot, Claude Code, Cursor, GitLab Duo, and others, running reliably for the engineers building the company's highest profile games. You will productize third-party AI tooling, automate the operational work that currently consumes engineering capacity, and keep internal customers unblocked.
Reporting to the Technical Director, you will:
- Operate and support production AI developer tools across the company, including access management, seat triage, telemetry, observability, and chargeback reporting.
- Productize third-party tools through the legal, security, and procurement gates required for broad company rollout.
- Build and maintain the automated evaluation framework that benchmarks AI dev tools, and run periodic evaluation studies across vendors.
- Stand up and own dashboards (Grafana, PowerBI) that give leadership visibility into adoption, cost, and reliability.
- Run Office Hours and internal enablement, and drive the digital assistant as a frontline support channel.
- Maintain handed-off products from Core Development that no longer have a dedicated dev team, keep them stable, secure, and supported.
- Respond to incidents quickly, root-cause issues, and turn fixes into automation so the same toil doesn't recur.
REQUIRED SKILLS AND EXPERIENCE
- 3+ years in an SRE, DevOps, or Production Engineering role supporting enterprise developer tools or platforms.
- Hands-on with cloud platforms (Azure or AWS preferred), microservices, and distributed systems.
- Strong scripting and automation in Python (preferred), Bash, or PowerShell; comfortable reading C# or Node when needed.
- Familiarity with Docker, Kubernetes, Terraform, and CI/CD pipelines (GitHub Actions, GitLab CI, or Azure DevOps).
- Experience operating observability stacks (Grafana, Prometheus, Datadog, Splunk, or similar) and building actionable dashboards.
- Excellent written communication; comfortable being the customer-facing voice for an engineering team.
- Proven ability to drive operational projects across cross-functional teams (legal, security, finance) without dropping balls.
NICE TO HAVE SKILLS AND EXPERIENCE
- Cloud certifications (Azure, AWS) or recognized SRE certifications.
- Experience operating or integrating LLM APIs (OpenAI, Anthropic) and Generative AI dev tools (Copilot, Claude Code, Cursor, GitLab Duo) in an enterprise context.
- Experience building chargeback or FinOps reporting for cloud / SaaS spend.
- Familiarity with prompt evaluation frameworks, vector databases, or RAG pipelines.
- PowerBI report authoring and IT / finance stakeholder communication.
- Prior engagement with AAA-scale game studio environments.
Pay Transparency:
The expected Annual Salary for this position falls between $28-34/hr +any major perks
Artificial intelligence tools are used to assist in screening and assessing applicants for this position.