We are looking for a skilled AI Engineer with deep expertise in Google Cloud Platform (GCP) to design, build, and deploy production-grade machine learning and generative AI solutions. You will be responsible for taking AI models from experimental prototypes to scalable, high-availability enterprise services using GCP’s native AI/ML ecosystem.
We are looking for a skilled AI Engineer with deep expertise in Google Cloud Platform (GCP) to design, build, and deploy production-grade machine learning and generative AI solutions.
Design, build, and maintain end-to-end ML and LLM pipelines using Vertex AI, Kubeflow, and Dataflow.
Fine-tune, evaluate, and integrate foundation models (e.g., Gemini) via Vertex AI Model Garden into application workflows using RAG architecture.
Automate continuous integration, deployment, and monitoring (CI/CD/CT) for machine learning models using GCP infrastructure and Terraform.
Work with data teams to optimize feature stores, data pipelines (BigQuery, Pub/Sub), and training datasets for scalable AI workflows.
Monitor model drift, latency, and throughput in production while optimizing GCP resource utilization and infrastructure costs.
Hands-on experience with Vertex AI (Pipelines, Feature Store, Model Registry, Endpoint deployment), BigQuery ML, and Cloud Run/GKE.
Proficiency in Python and standard ML frameworks (PyTorch, TensorFlow, JAX).
Experience with orchestration frameworks (LangChain, LlamaIndex), vector databases (Vertex AI Vector Search, Pinecone, or pgvector), and prompt engineering/tuning.
Familiarity with Docker, Kubernetes, Terraform, MLflow, or Kubeflow Pipelines.
Roles like this expire in about a week. Get new ML Engineer openings across the UK in your inbox, free, unsubscribe any time.