Senior Software Engineer โ€” Data Catalog / Ingestion

$$$$

ABOUT OUR CLIENT

Our client is a leading enterprise data platform company building an open, high-performance data lakehouse for AI and analytical workloads. The platform combines an intelligent SQL query engine, an AI-ready semantic layer, and an open catalog built on Apache Iceberg โ€” enabling Fortune 500 companies across finance, energy, manufacturing, and logistics to unify, query, and govern data at massive scale across cloud and on-premise sources.

ABOUT THE ROLE

We are looking for a Senior Software Engineer for the data catalog and enterprise data management layer of a large-scale lakehouse platform โ€” the services that automate data ingestion and autonomously optimize Apache Iceberg tables. You will own features through the full development cycle, from design through deployment, in a multithreaded, distributed environment where scalability, performance, and always-on availability are the baseline expectations.

This is a systems-level, backend engineering role focused on data infrastructure internals โ€” not application development or CRUD services.

RESPONSIBILITIES

  • Own the full software development cycle โ€” inception, design, development, testing, deployment โ€” for catalog and data-management services
  • Build and operate services for automated ingestion and autonomous optimization of Apache Iceberg tables
  • Reason about concurrency and parallelization to deliver scalability and performance in multithreaded, distributed systems
  • Drive performance tuning and system-efficiency improvements
  • Work with Apache Foundation open-source projects, contributing upstream where agreed
  • Participate in code and design reviews, upholding a high engineering bar
  • Collaborate with product managers, support teams, and solution architects on feature requests and design updates
  • Partner with US-based engineering leads on technical decisions

REQUIRED QUALIFICATIONS

  • B.S., M.S., or PhD in Computer Science or a related field
  • 5+ years of software engineering experience, preferably focused on database systems or related fields
  • Strong object-oriented programming skills in Java or C++
  • Solid grounding in data structures and algorithms
  • Hands-on experience with multithreaded and asynchronous programming patterns for scalable, performant systems
  • A passion for engineering quality โ€” zero-downtime upgrades, availability, resiliency, and uptime as first-class concerns
  • Comfort with a fast-moving environment; strong ownership; the confidence to defend a technical position and the openness to be mentored
  • Comfortable with AI-assisted development workflows โ€” using modern AI tools productively while critically validating their output
  • English Upper-Intermediate or higher (B2+) โ€” daily written and verbal communication with a US-based engineering team
  • Availability to work EU business hours shifted 2โ€“3 hours later for daily overlap with US West Coast mornings

DESIRED SKILLS

  • Database internals and query planning
  • Distributed systems: concurrency control, data replication, storage systems
  • Apache Iceberg or other open table formats (Delta Lake, Hudi)
  • Messaging systems: Kafka, NATS, or cloud pub/sub services
  • Contributions to open-source data infrastructure (Apache projects especially valued)

DETAILS

  • Engagement: Long-term contract
  • Location: Europe (EU / EEA / UK), remote
  • Working hours: EU business hours, shifted 2โ€“3 hours later for daily overlap with the US West Coast team

Required skills experience

Java 5 years
C/C++ 5 years
Algorithms 5 years
Multithreading 5 years

Required domain experience

SaaS 5 years

Required languages

English B2 - Upper Intermediate
Ukrainian A1 - Beginner
Kafka, NATS, Delta Lake, Apache Hudi, Data Replication
Published 28 August
11 views
ยท
0 applications
To apply for this and other jobs on Djinni login or signup.
Loading...