Перейти к основному содержимому
Antenna

Data Scientist

УдалённоВесь мир
Опубликовано
Роль
Data Science
Опыт
Мидл
Размер компании
Стартап
Зарплата не указана
Проверьте доступность

Доступно для: Worldwide. Укажите, откуда вы работаете, чтобы проверить доступность.

Коротко по делу

Data Scientist with 2+ years building ML models and data products in Python, strong SQL, and production engineering skills. Must work during US business hours, preferably from a US time zone, and have advanced English. LLM pipelines, evals, agentic coding tools, and large messy datasets are central to the role.

Ключевые навыки

PythonMachine LearningLLM pipelines

Обязательные навыки

SQLGitGCPBigQueryClaude Code/Cursor/Codex CLIApache Spark/PySpark/DaskDataprocCloud StorageCloud RunCloud BuildGKEPandasNumPy

Желательные навыки

PyTorchTensorFlowMLOpsmodel monitoringsynthetic datagenerative modelsGANsVAEs

Обязательные языки

English B2-C1

Чем предстоит заниматься

  • Build models and data products that make it all the way to production for our two flagship products: Subscriber Views, which connects what people watch to why they subscribe, and Subscriber Metrics, the market standard for sign-ups, churn, and retention
  • Solve the modeling problems that unlock new product capabilities - from viewership attribution and sports data to generative models on mixed data sources and new models for TV providers (MVPDs)
  • Dig into large, messy datasets to find the trends and patterns that turn into shipped features
  • Add the functions, classes, and tools to our core Python data science library that the rest of the team builds on
  • Make existing code faster and able to handle more data
  • Take on Antenna R&D work: explore new datasets and methods to answer real business questions and present what you find to senior stakeholders
  • Build LLM-powered pipelines and agents
  • Treat evals as a core deliverable: validated outputs, eval sets with clear pass/fail checks, and evidence they catch problems human review misses
  • Use agentic coding tools as part of your daily workflow - plan first, write tests and instructions up front, review every change
  • Design, debug, and defend clear, well-documented, testable code of your own
  • Collaborate with cross-functional stakeholders, including our co-founders
  • Clearly explain complex technical and system design decisions

Что требуется

  • You have 2+ years of experience building machine learning models and data products in Python, with the engineering skills to take them from prototype to production
  • Expert in Python with strong object-oriented design, software system design, and experience building high-quality, testable, production-grade code
  • Solid grasp of machine learning concepts and how models are built, used, and evaluated
  • Strong SQL skills working with large, complex datasets
  • Experience using Git, and comfortable working in cloud environments (GCP preferred) to pull and process data (e.g., BigQuery) and run your analytical work.
  • Excellent problem-solver, skilled at debugging complex distributed systems and optimizing them for performance and scale
  • Advanced English proficiency (B2-C1) with strong communication, teamwork, and consulting skills; able to clearly explain complex technical and system design decisions
  • You use agentic coding tools (Claude Code, Cursor, or Codex CLI) as part of your daily workflow: plan first, write tests and instructions up front, and review every change before accepting it.
  • You can still design, debug, and defend your own work without AI assistance, and our interview process tests this directly
  • You measure whether an LLM is actually right: you verify model outputs against ground truth by default, build simple eval sets or labeled samples to compare approaches, and can give a concrete example of a hallucination or silent error you caught.
  • Required proficiency: Python (expert)
  • Required proficiency: SQL (strong)
  • Required proficiency: Google Cloud (Dataproc, BigQuery, Cloud Storage, Cloud Run, Cloud Build, GKE strong experience expected)
  • Required proficiency: Git (expert)
  • Required proficiency: Pandas, NumPy (very good with these)
  • Preferred: Experience in or passion for the Subscription Economy, especially media and entertainment; experience working with media data or data clean rooms
  • Preferred: Hands-on experience with deep learning frameworks, MLOps practices, or synthetic data and generative models
  • Preferred: Experience with large-scale data processing tools
  • Preferred: Experience deploying and scaling services in the cloud
  • Preferred: Experience building and shipping LLM-powered agents or pipelines with an orchestration framework, including custom tool definitions, agent state and memory, and human review steps
  • Preferred: Advanced evaluation and observability practices
  • Preferred: Familiarity with RAG and context engineering for grounding model responses in proprietary data, or experience using LLMs for testing pipelines and QA workflows
  • Preferred: Experience with BI tools, building Python libraries that others use, or open-source contributions

Преимущества

  • Work from anywhere, during US business hours
  • Competitive compensation, including participation in Antenna equity program
  • Mentorship from experienced executives
  • Opportunity to grow and impact a rapidly growing start up
  • Travel to In-person team off-sites (visa-permitting)
  • And more!

Antenna is the leading provider of data and analytics for subscription media services in the U.S. Antenna provides standardized metrics, competitive benchmarks, syndicated insights, and market intelligence for media and entertainment brands.

Data AnalyticsСтартап

Детали

Способ откликаGreenhouse
Зарплата не указана