Skip to main content
Meshy

Senior AI Data Infrastructure Engineer

RemoteChina only
Published
Role
Data Engineering
Experience
Senior
Employment
Full-time
Company size
Startup
Salary not disclosed
Check eligibility

Open to CN only. Set where you work from to check your eligibility.

No BS summary

Meshy is seeking a Senior AI Data Infrastructure Engineer in Shanghai to design, build, and operate distributed data systems for large-scale AI model training, focusing on processing pipelines for massive unstructured data and optimizing computing/storage efficiency.

Required skills

SparkRayFlinkTerraformKubernetes

Required languages

English basic

What you'll do

  • Design and maintain highly scalable ETL/ELT pipelines to support the ingestion and processing of structured and massive amounts of unstructured data (images, videos, 3D/2D assets).
  • Optimize data processing using frameworks such as Spark/Ray/Flink based on cloud object storage and data lakes, implementing data sharding, caching, and a robust monitoring and alerting mechanism.
  • Responsible for the preprocessing (format conversion, enhancement, feature extraction, etc.) and quality verification of model pre-training data, ensuring the lineage and reproducibility of datasets.
  • Manage environments using IaC tools such as Terraform/Kubernetes, promoting CI/CD best practices; collaborate closely across teams to quickly respond to R&D needs.

What they require

  • 5+ years of experience in data engineering or distributed system development
  • Basic English communication skills, capable of conducting all-English meetings and written communication.

Meshy is redefining 3D creation with generative AI, providing a pipeline for 3D content from text/image to 3D, texturing, texture editing, animation rigging, and related creator community features.

Generative AIStartup
Salary not disclosed