Data Engineer
๐ Location: EU-based, Remote
๐ผ Cooperation type: Full-time
Weโre looking for a skilled Data Engineer to join our Analytics team.
We are building the data platform that serves both our analytics and operational data needs. In this role you will own significant parts of the platform architecture โ from design through deployment โ and build the pipelines and services that make the companyโs data flow reliably between systems. You will work with large, complex datasets and high-throughput pipelines, and you will own a major greenfield project connecting our production systems with our CRM.
About us:
Stape is a global product-driven IT company and the #1 leader in the server-side tracking market. Weโre building a powerful, technically complex product that simplifies server-side tracking for marketers and website owners. Our platform processes over 5 billion requests daily, helping improve tracking accuracy and data privacy for more than 200, 000 clients worldwide. We work closely with top tech partners, including Meta, LinkedIn, Google, Snapchat, and TikTok to provide advanced tracking capabilities.
Skills and experience
- Minimum 4 years of experience as a data engineer (platform/pipeline building, not only report maintenance).
- At least an upper-intermediate level of English (written and spoken).
- Bachelorโs degree in Computer Science, Engineering, or another related technical field.
- Extensive knowledge of SQL and Python.
- Experience organizing, monitoring, and maintaining data infrastructure: managing ETL/ELT processes, writing tests for data pipelines, and monitoring data consistency.
- Practical experience with GCP: BigQuery, Cloud Logging, Cloud Storage, Cloud Functions, Cloud Run, Datastream, Pub/Sub.
- Practical experience with Dataform (or dbt).
- Experience with a workflow orchestrator (Apache Airflow or similar) โ scheduling, dependency management (DAGs), retries, and backfills.
- Experience with both batch and streaming data processing.
- Ability to build custom Python services and server-to-server API integrations (e.g. FastAPI).
- Understanding of data modeling for analytics โ including star schema and lakehouse/medallion (bronze/silver/gold) design patterns.
- Experience with Jira or similar project management tools.
Will be a plus
- Familiarity with containerization (Docker); Kubernetes.
- Experience integrating CRM or third-party SaaS platforms via their APIs (e.g. HubSpot, Salesforce) โ building bidirectional data syncs, handling webhooks, and reconciling records across systems.
- Change Data Capture (CDC) experience for near-real-time replication.
- Stream processing (Apache Flink, Dataflow) and open table formats (Apache Iceberg).
- Owning data platform architecture end-to-end โ technology selection and implementation standards.
- Experience with MySQL.
Duties and responsibilities
- Build and maintain infrastructure for reliable extraction, transformation, and loading of data from a variety of sources.
- Design and build batch and streaming pipelines that meet business requirements.
- Own the design and delivery of a bidirectional sync between our production systems and our CRM.
- Conduct SQL performance tuning for existing queries in cloud databases (primarily BigQuery).
- Create and maintain optimized data models and schemas (including star schema and medallion layers) for analytics.
- Evaluate data for integrity and accuracy; work with teams to backfill gaps where found.
- Design and implement data security and governance measures.
- Collaborate with analysts to improve data collection and processing.
- Work with stakeholders on data-related technical issues and support their infrastructure needs.
- Participate in strategic, data-related decision-making.