Перейти к основному содержимому
Innodata

Application Reliability Engineer

УдалённоUnited States только
Опубликовано
Роль
DevOps
Опыт
Мидл
Занятость
Полная занятость
Размер компании
Крупная
$60k–$100k/yr
Проверьте доступность

Доступно для: US only. Укажите, откуда вы работаете, чтобы проверить доступность.

Коротко по делу

Mid-level Application Reliability/Support engineer (3–7 years) to maintain and run production microservices on Google Cloud (App Engine). Must be US-eligible/US-remote and experienced with GCP App Engine, deployments, incident response, and Python/Java/Node.js/Go.

Ключевые навыки

Google App EngineGCP

Обязательные навыки

PythonJavaNode.jsGoRESTgRPCSQLCloud SQLFirestoreDatastoreCloud SpannerBigQueryIAMservice accountsCloud MonitoringCloud LoggingError ReportingCloud TraceCI/CDCloud BuildJenkinsGitHub ActionsGitLab CI

Желательные навыки

Cloud RunGKECloud FunctionsApigeeAPI GatewayPub/SubCloud TasksCloud Scheduler

Чем предстоит заниматься

  • Provide production support and maintenance for enterprise applications hosted on Google App Engine, including standard and flexible environments.
  • Act as a first point of contact for user-impacting incidents: triage, diagnose, restore service, and drive issues to closure within agreed SLAs/SLOs.
  • Troubleshoot application errors, failed requests, latency and performance degradation, service-to-service failures, configuration issues, quota/scaling limits, and dependency or integration failures.
  • Design, develop, and deliver feature enhancements and functional improvements to existing applications based on user and business needs.
  • Support applications built on microservices architecture, including service boundaries, APIs/contracts, inter-service communication, authentication, and failure/retry behavior.
  • Own build, release, and deployment activities across development, test, pre-production, and production environments with appropriate validation, approvals, and rollback plans.
  • Manage App Engine deployments including versions, traffic splitting/migration, canary and staged rollouts, rollbacks, service configuration, and scaling settings.
  • Perform root-cause analysis for recurring production issues and implement sustainable fixes.
  • Build and maintain monitoring, alerting, logging, dashboards, and error reporting using Cloud Monitoring, Cloud Logging, Error Reporting, and Cloud Trace.
  • Support platform, framework, library, dependency, and runtime upgrades while maintaining stability, supportability, and compliance.
  • Support IAM, service accounts, access controls, secrets management, and operational governance.
  • Participate in change management, release-readiness reviews, and on-call/rotational support as required.
  • Collaborate with client and cross-functional Application Engineering, Product, QA, Data Engineering, Infrastructure, and Platform teams.
  • Create and maintain technical documentation, operational runbooks, troubleshooting guides, deployment procedures, and support playbooks.

Что требуется

  • 3-7 years of experience in Application Support, Application Engineering, Software Engineering, Cloud Engineering, or a related role.
  • Strong hands-on experience supporting, maintaining, and enhancing production applications on Google Cloud Platform (GCP).
  • Hands-on experience with Google App Engine, including deployment, configuration, scaling, versioning, and troubleshooting.
  • Solid understanding of microservices architecture, REST/gRPC APIs, service-to-service communication and authentication, distributed tracing, and cross-service debugging.
  • Practical experience with deployments and promotions across multiple environments, including release validation, rollback, and change control.
  • Strong programming skills in one or more of Python, Java, Node.js / JavaScript, or Go.
  • Working knowledge of SQL and application data stores such as Cloud SQL, Firestore/Datastore, Cloud Spanner, or BigQuery.
  • Good understanding of GCP IAM, service accounts, permissions, monitoring, logging, alerting, and production operations.
  • Experience with CI/CD pipelines and automated build/deployment tooling such as Cloud Build, Jenkins, GitHub Actions, or GitLab CI.
  • Experience troubleshooting complex production environments and performing root-cause analysis under time pressure.
  • Ability to quickly understand existing systems, codebases, services, configurations, and client-specific tools and workflows.
  • Strong communication and collaboration skills, including communication during user-impacting incidents.

Преимущества

  • None explicitly listed besides general company information and anti-scam guidance.

Innodata (Nasdaq: INOD) is a global data engineering company providing data, evaluation frameworks, human expertise, solutions, platforms, and services for Generative AI / AI builders and adopters.

🇺🇸 Соединенные ШтатыAI DataКрупнаяinnodata.com/

Что говорят о компании

4.1/ 5

$60k–$100k/yr