Перейти к основному содержимому
TensorWave

Staff Database Engineer

УдалённоUnited States только
Опубликовано
Роль
DevOps
Опыт
Стафф
Занятость
Полная занятость
Зарплата не указана
Проверьте доступность

Доступно для: US only. Укажите, откуда вы работаете, чтобы проверить доступность.

Коротко по делу

We’re looking for a Staff Database Engineer to join our team during an exciting phase of growth.

Ключевые навыки

PostgreSQLMySQL

Обязательные навыки

PerconaPrometheusGrafanaPMMQuery logsslow query logseBPF/BCCLinuxAnsible

Чем предстоит заниматься

  • Design and own database architecture for critical infrastructure and platform services, including PostgreSQL-backed internal platforms, Slurm accounting and operational databases, NetBox and infrastructure source-of-truth databases, custom internal applications and automation services, observability, inventory, and platform metadata systems, future database-backed control plane services.
  • Operate and improve production database environments across PostgreSQL, MySQL, Percona, and adjacent systems.
  • Serve as the senior database engineering owner for infrastructure-adjacent database platforms, including Slurm and NetBox.
  • Build deep database observability beyond basic dashboards.
  • Create database automation patterns that can be integrated with existing infrastructure tooling.

Что требуется

  • Define standard database patterns for high availability, replication, failover, backup and restore, point-in-time recovery, performance baselining, capacity planning, upgrade lifecycle management, access control and operational security.
  • Establish database design standards for new internal platforms, including schema review, indexing strategy, query design, service ownership boundaries, and production readiness requirements.
  • Own the lifecycle of database systems, including provisioning, configuration, version upgrades, replication topology design, performance tuning, backup validation, disaster recovery testing, decommissioning, documentation and runbook creation.
  • Troubleshoot and resolve production database issues involving query latency, lock contention, replication lag, storage I/O bottlenecks, connection exhaustion, poor indexing, schema design problems, database capacity constraints, backup or restore failures.
  • Drive root cause analysis for database-related incidents and convert findings into durable engineering improvements.

Our mission is simple: deliver seamless, secure, reliable, and resilient AI compute at scale. We've built a versatile cloud platform that eliminates infrastructure barriers, empowering builders to focus on innovation instead of fighting their stack. Because breakthrough AI should move at the speed of ideas, not infrastructure.

TechnologyСтартап
Зарплата не указана