Harvey12 дней назад
Staff Software Engineer, Data Platform
Зарплата не указана
РЫНОК
14 882 ₽медиана по профессии
Data Engineer · 26 вакансий с указанной зарплатой
6 667половина предложений: 11 421–16 23421 667
Работодатель не указал зарплату — сравните с рынком сами.
Полная занятостьУдалёнка
Обязанности
- 01Own the data platform's architecture and technical direction, treating data infrastructure as a software product built from reusable frameworks and making deliberate build-vs-buy tradeoffs as the platform grows
- 02Build and operate the ingestion layer across streaming, batch, CDC, and third-party connectors, including schema evolution that safely absorbs upstream change so onboarding a new source is a paved path
- 03Land data into Snowflake with the freshness, completeness, and cost characteristics downstream consumers can plan around, and define a clean handoff for Analytics Engineering
- 04Own the orchestration platform — scheduling, retries, backfills, and dependency management across the full data graph
- 05Build the transformation and compute frameworks teams can use to process data at scale, and the self-serve tooling that lets product engineers and analysts stand up their own pipelines against primitives you've made safe
- 06Design and operate stream processing infrastructure for use cases that can't wait for batch — real-time product features, operational alerting, and near-live reporting
- 07Build the trust layer: quality and observability (freshness, validation, reconciliation, anomaly detection, alerting routed to the right owner) alongside lineage, cataloging, and discovery, so anyone can find data and know its origin and dependencies
- 08Build the patterns and tooling for PII and sensitive data — classification, masking, retention, access control — and for multi-region residency requirements
- 09Set the technical bar for data at Harvey through design reviews, standards, documentation, and mentorship as the team grows
Требования
- 0110+ years building and operating production data infrastructure, with ownership of systems other teams depend on
- 02Deep experience with cloud data warehouses — Snowflake strongly preferred (BigQuery, Databricks, or Redshift experience transfers well) — including performance tuning and cost management
- 03Hands-on experience building CDC and streaming pipelines with technologies like Kafka, Debezium, Flink, or Spark Streaming
- 04Experience with managed ingestion tooling (Fivetran, Airbyte, or similar) and clear judgment about when to buy the connector and when to build it
- 05Strong fluency with workflow orchestration — Temporal, Airflow, Dagster, or similar — operated at scale, not just configured
- 06Strong programming skills in Python and advanced SQL
- 07Experience building frameworks or internal tooling that other engineers use, and the product instinct to know when an abstraction is helping versus getting in the way
- 08Practical experience with data quality, observability, and lineage
Условия
- 01Based in San Francisco, CA or New York, NY