Agentic Coding Annotator - Online / Offline Tasks
About Turing
Turing is one of the world’s fastest-growing AI companies, accelerating the advancement and deployment of powerful AI systems. Turing helps customers in two ways: working with the world’s leading AI labs to advance frontier model capabilities in thinking, reasoning, coding, agentic behavior, multimodality, multilinguality, STEM, and frontier knowledge; and leveraging that work to build real-world AI systems that solve mission-critical priorities for companies.
Role Overview
We are looking for an experienced DevOps Engineer to build and operate GPU infrastructure and production LLM serving systems on Google Cloud Platform (GCP). You will own infrastructure across the lifecycle, from GPU provisioning and container orchestration to scalable model inference, observability, and cost optimization.
What does day-to-day look like
- Provision and manage GPU infrastructure on GCP using GKE, Compute Engine, Cloud Run, and related services.
- Deploy, tune, and operate production LLM serving stacks using vLLM, Triton Inference Server, and Hugging Face Transformers.
- Build containerized and serverless deployments with Docker, CI/CD, and Infrastructure as Code (IaC).
- Design and maintain reliable data pipelines supporting ML and inference workloads.
- Implement observability, autoscaling, reliability, and cost optimization for GPU fleets and production inference services.
Requirements
- Strong Python programming skills and hands-on experience building or operating production infrastructure.
- Strong practical experience with GCP, including GKE, Cloud Run, GCS, Pub/Sub, Cloud Functions, or equivalent services.
- Hands-on experience provisioning and operating GPU infrastructure, including CUDA/drivers, GPU node pools, quotas, and autoscaling.
- Production experience serving LLMs using vLLM and/or Triton Inference Server, with strong Docker and container orchestration fundamentals.
- Experience building data pipelines using Dataflow, Airflow/Cloud Composer, or similar technologies.
Perks of Freelancing With Turing
- Work on cutting-edge AI projects with leading foundation model companies
- Collaborate on high-impact work at the frontier of LLM evaluation and reasoning
- Remote, flexible opportunities with global teams
Offer Details
- Commitments Required: 8 hours per day with a 4-hour overlap with PST.
- Employment Type: Contractor position (Note: this role does not include medical/paid leave).
- Duration of Contract: 2 months [expected start date is immediate].
Evaluation process
- Profile review followed by an interview by delivery team (if required)
Details
- Platform: Turing (work.turing.com)
- Location: Remote
- Skills: Python, Google Cloud Platform, Docker
Apply directly through Turing using the link below. Turing matches vetted experts to frontier AI labs, with fast onboarding and remote flexibility.
Stay Updated on Roles Like This
Subscribe to receive fresh openings aligned with Engineering & Software and AI training roles across Turing, Mercor and JobHub by NeonLabs