Перейти к основному содержимому
LTG

Cloud Infrastructure Engineer

УдалённоUnited Kingdom только
Опубликовано
Роль
DevOps
Опыт
Синьор
Занятость
Полная занятость
Зарплата не указана
Проверьте доступность

Доступно для: GB only. Укажите, откуда вы работаете, чтобы проверить доступность.

Коротко по делу

Senior cloud infrastructure engineer for a UK-remote role building and operating an AWS-based multi-tenant SaaS hosting platform. Needs strong AWS, Terraform, Puppet/config management, Python, Linux systems, distributed systems, observability, and GitLab CI/CD experience.

Ключевые навыки

AWSTerraformPuppet

Обязательные навыки

EC2RDSS3SQSLambdaALBElastiCacheRoute 53IAMVPC networkingPythonLinuxUbuntuApacheNginxPHP-FPMVarnishsystemdPrometheusGrafanaLokiGitLab CI/CD

Желательные навыки

CephGlusterFSJuiceFSCubeFSAWS EFSetcdConsulZooKeeper

Чем предстоит заниматься

  • Design, build, and maintain AWS infrastructure using Terraform, including EC2, RDS, S3, SQS, Lambda, ALB, ElastiCache, Route 53, and VPC networking.
  • Write and maintain Puppet modules to configure and manage fleets of EC2 instances across multiple auto-scaling groups.
  • Maintain and extend Python-based automation and tooling that supports platform operations.
  • Operate and improve distributed service discovery and configuration management with etcd.
  • Manage and tune a multi-tier caching strategy using Varnish, Redis/Valkey, and PHP OPcache.
  • Run and scale the observability stack using Prometheus, Grafana, Loki, Fluentd, and PagerDuty.
  • Participate in on-call rotations.
  • Evaluate and implement distributed storage solutions as the platform evolves.
  • Improve deployment workflows and release processes.
  • Collaborate with internal teams on API contracts, integration patterns, and operational tooling.
  • Participate in incident response, root cause analysis, and platform reliability improvements.
  • Build, scale, and evolve a multi-tenant SaaS hosting platform on AWS.
  • Work across Terraform modules, Puppet manifests, Python automation, and observability pipelines.
  • Own and influence the platform's architecture and direction.

Что требуется

  • Senior Cloud Infrastructure Engineer experience.
  • Strong experience with AWS services in production, particularly EC2, RDS, S3, SQS, Lambda, ALB, ElastiCache, Route 53, IAM, and VPC networking.
  • Proficiency in authoring and maintaining Terraform modules for production infrastructure.
  • Proficiency in authoring and maintaining Puppet modules or equivalent agent-based configuration management for fleet management.
  • Solid Python skills for writing and maintaining production daemons, not just scripts.
  • Deep Linux systems knowledge on Ubuntu, including Apache/Nginx, PHP-FPM, Varnish, systemd, filesystem mounts, and networking fundamentals.
  • Understanding of distributed systems concepts including consensus, leader election, distributed locking, eventual consistency, and tradeoffs.
  • Proficiency in building and maintaining observability pipelines such as Prometheus, Grafana, Loki, or equivalent in production.
  • Comfortable working in a GitLab-based CI/CD workflow.
  • Clear communicator who can document architectural decisions and explain technical tradeoffs to technical and non-technical stakeholders.
  • Preferred: Hands-on experience with distributed storage systems such as Ceph, GlusterFS, JuiceFS, CubeFS, or AWS EFS, particularly in migration or evaluation contexts.
  • Preferred: Familiarity with etcd or similar distributed key-value stores like Consul or ZooKeeper, including watch APIs, TTL-based locking, and cluster operations.
  • Preferred: Experience with Varnish and VCL, especially dynamic backend routing or multi-tenant configurations.
  • Preferred: Working knowledge of PHP to understand and maintain integration scripts that bridge infrastructure and application layers.
  • Preferred: Background in multi-tenant SaaS platform design, particularly database-per-tenant models on shared infrastructure.
  • Preferred: Familiarity with Moodle LMS or education technology platforms.
  • Preferred: Experience with secrets management solutions such as AWS Secrets Manager, HashiCorp Vault, Parameter Store, and automated credential rotation.
  • Preferred: Experience designing zero-downtime deployment strategies for VM-based non-containerized environments.

Преимущества

  • Real ownership and influence over the platform's architecture and direction.

LTG

LTG is a global team of Engineers, Product Managers, Designers, and Program Managers across Hungary, the US, and many other countries. Bridge is software for learning, self-development, and career growth.

EdTech

Детали

Способ откликаDom
Зарплата не указана