Чтобы адаптировать резюме под вакансию или составить сопроводительное письмо, загрузите резюме
описание
Zencargo provides logistics and supply chain technology for businesses, supporting freight operations through software, cloud infrastructure, and AI-powered workloads.
задачи
Lead the design, implementation, and delivery of complex infrastructure projects;
Build and scale infrastructure for AI workloads, including cost, API reliability, and platform support for LLM-powered features;
Debug cluster behavior under load, service-to-service traffic, event-backbone consumer lag, and database performance;
Own the infrastructure-as-code estate, keep drift visible and reconciled, and set module standards for the wider team;
Advance the delivery pipeline through policy-as-code gates, progressive delivery, and tested, reviewed, auditable changes;
Contribute to internal AI tooling, including the MCP server and agentic workflows that reduce operational toil;
Engineer cost controls through right-sizing, reserved capacity, and waste removal against a measured baseline;
Advocate for targeted infrastructure spend;
Mentor peers through pairing, code review, and knowledge sharing;
Cover security and governance alongside platform work in collaboration with the infrastructure team;
Deliver operational outcomes that reduce manual effort and improve business speed, accuracy, cost, or service quality;
Improve platform reliability and incident outcomes through systemic fixes;
Deliver complex infrastructure projects end to end with recorded and followable decisions;
Maintain infrastructure-as-code and pipeline health, including drift reconciliation and wider adoption of standards.
требования
Write maintainable software in at least one programming language and move between languages when required;
Own production infrastructure on a major cloud, including during incidents;
Understand infrastructure-as-code state, module design, and reviewed, repeatable changes;
Debug misbehaving Kubernetes clusters in production;
Design CI/CD pipelines rather than only consume them;
Use AI tooling fluently and critically while understanding its limitations;
Work effectively in a fully remote, async-first environment;
Communicate a plan and work methodically under pressure during platform incidents;
Nice to have: Experience in freight, logistics, supply chain, B2B SaaS, or operational technology; workflow automation or orchestration experience, such as n8n; experience with service mesh, event streaming at scale, or managed Postgres-compatible databases under real load; familiarity with identity and access management, SSO, secret management, or SRE methodology; experience building with LLM APIs, MCP servers, or agentic tooling.
условия
London/Europe with GMT+0 to GMT+4 overlap with the UK team;
Fully remote role with offices in multiple locations;