Security and Loss Prevention Cluster Manager, Worldwide Operations Security

Security and Loss Prevention Cluster Manager, Worldwide Operations Security

Schönefeld Vollzeit 63000 - 77000 € / Jahr (geschätzt) Homeoffice (teilweise)
A

Auf einen Blick

  • Aufgaben: Automatisiere CI/CD-Pipelines für KI und entwickle innovative Workflows.
  • Unternehmen: TechBiz Global, ein führendes Unternehmen im Bereich Rekrutierung.
  • Vorteile: Attraktives Gehalt, flexible Arbeitszeiten und Weiterbildungsmöglichkeiten.
  • Weitere Informationen: Spannende Karrierechancen in einem innovativen Team.
  • Warum dieser Job: Wachse in einem dynamischen Umfeld und forme die Zukunft der KI-Technologie.
  • Qualifikationen: Mindestens 10 Jahre Erfahrung in DevOps und 2 Jahre in MLOps.

Das prognostizierte Gehalt liegt zwischen 63000 - 77000 € pro Jahr.

Senior AI Dev Ops / LLMOps Specialist

At Tech Biz Global, we are providing recruitment service to our TOP clients from our portfolio. We are currently seeking an

Senior AI Dev Ops / LLMOps specialist to join one of our clients

' teams. If you're looking for an exciting opportunity to grow in an innovative environment, this could be the perfect fit for you.

Key Responsibilities

  • Automation of Build-to-Production
  • - Design and implement robust CI/CD pipelines tailored for AI, covering model weights, dataset versioning, and application code.
  • - Develop specialized workflows for Prompt Ops, ensuring that system prompts are version-controlled, tested for regressions, and deployed with the same rigor as traditional code.
  • - Automate the deployment of Agentic workflows, managing the complexities of stateful AI interactions and multi-agent handoffs.
  • 2. AI Infrastructure as Code (Ia C)
  • - Provision and manage high-performance compute environments (GPU clusters, TPU pods) using Terraform, Pulumi, or Ansible.
  • - Define and enforce Policy-as-Code for AI endpoints to ensure compliance with security, cost-usage limits, and data residency requirements.
  • - Maintain a consistent environment across Hybrid Infrastructure, ensuring seamless parity between On-Premises development and Cloud production.
  • 3. Safe Experimentation & Controlled Releases
  • - Architect Progressive Delivery strategies for AI, including Canary releases, Blue-Green deployments, and Shadowing (where new models run in parallel with production to compare outputs).
  • - Build “Evaluation-in-the-Loop” gates within the pipeline to automatically test for bias, hallucination, and performance degradation before a release.
  • - Implement A/B testing frameworks specifically designed for LLM outputs and agentic behavior.
  • 4. Monitoring & Observability
  • - Establish deep observability into Inference Endpoints, tracking metrics like tokens-per-second, latency, and drift in model accuracy.
  • - Integrate feedback loops that capture production “edge cases” to feed back into the training and fine-tuning pipelines.
  • Must-Have Technical Skills
  • - Orchestration: Advanced Kubernetes (K8s) skills, specifically with Kube Flow, Ray, or NVIDIA Triton.
  • - CI/CD & Ia C: Expertise in Git Hub Actions/Git Lab CI, and Terraform or Pulumi.
  • - AI Tooling: Experience with Weights & Biases, MLflow, Lang Smith, or Arize Phoenix.
  • - Hardware: Understanding of GPU virtualization, CUDA drivers, and on-premises hardware management.
  • - Security: Familiarity with Open Policy Agent (OPA) and secret management (Vault).

Experience

  • - 10+ years in Dev Ops, SRE, or Cloud Engineering.
  • - 2+ years of hands-on experience in MLOps or LLMOps, specifically moving LLMs from notebook to production.
  • - Proven experience managing Hybrid Cloud environments (e. g., AWS/Azure + Private
  • #J-18808-Ljbffr

Security and Loss Prevention Cluster Manager, Worldwide Operations Security Arbeitgeber: Amazon VZ Berlin-Brandenburg GmbH

Als Instandhaltungsplaner (m/w/d) in unserem engagierten Wartungs- und Instandhaltungsteam profitierst du von einer dynamischen Arbeitsumgebung, die Innovation und Teamarbeit fördert. Wir bieten dir nicht nur attraktive Weiterbildungsmöglichkeiten, sondern auch ein unterstützendes Arbeitsklima, das auf Vertrauen und Respekt basiert. Zudem liegt unser Standort in einer lebendigen Region, die eine hervorragende Lebensqualität und zahlreiche Freizeitmöglichkeiten bietet.

A

Kontaktdaten:

Amazon VZ Berlin-Brandenburg GmbH Recruiting-Team

Wir glauben, dass du diese Fähigkeiten brauchst, um Security and Loss Prevention Cluster Manager, Worldwide Operations Security mit Bravour zu bestehen

CI/CD Pipelines
Kubernetes (K8s)
KubeFlow
Ray
NVIDIA Triton
Terraform
Pulumi