Senior Data Engineer
<div class="show-more-less-html__markup show-more-less-html__markup--clamp-after-5 relative overflow-hidden"> <br/><p>Our client is building the intelligence layer behind one of the world's largest rewarded advertising platforms, helping <strong>770+ million users</strong> discover and engage with new mobile apps every year. Their machine learning systems make <strong>200+ million decisions daily</strong>, processing <strong>100,000+ predictions per second</strong> from a <strong>1PB+ behavioural data platform</strong>. Backed by a <strong>$100 million strategic investment</strong>, they're expanding the engineering team responsible for the real-time data infrastructure powering next-generation AI systems.</p><p><br/></p><p>You'll join as a Senior Data Engineer, helping build one of the industry's highest-scale ML data platforms. You'll engineer real-time pipelines, optimise distributed data systems, and enable intelligent decision-making across infrastructure handling 100,000+ predictions every second.</p><p><br/></p><p><strong>Key Responsibilities</strong></p><ul><li>Build and scale real-time streaming data pipelines supporting production ML systems.</li><li>Develop and enhance feature engineering infrastructure used by Data Scientists.</li><li>Improve data quality through monitoring, validation and governance.</li><li>Optimise large-scale ETL and streaming workloads for performance and reliability.</li><li>Collaborate closely with Data Science and Backend Engineering teams.</li><li>Build ingestion frameworks for new and evolving data sources.</li><li>Help shape the future architecture of the company's cloud data platform.</li><li>Take ownership of technical decisions and platform scalability.</li></ul><p><br/></p><p><strong>Qualifications</strong></p><ul><li><strong>5+ years</strong> of commercial Data Engineering experience building production data platforms.</li><li>Strong experience with <strong>Apache Flink</strong>, <strong>Kafka</strong>, and real-time streaming architectures.</li><li>Professional Java development experience, with <strong>Go</strong> or <strong>Python</strong> considered a bonus.</li><li>Experience processing <strong>TB-scale datasets</strong> and high-throughput event streams.</li><li>Hands-on experience with <strong>AWS</strong>, <strong>Airflow</strong>, <strong>dbt</strong>, <strong>Terraform</strong>, and <strong>Kubernetes</strong>.</li><li>Experience working with Machine Learning or Data Science teams in production environments.</li><li>Understanding of <strong>Data Lakes</strong>, <strong>Feature Stores</strong>, <strong>Lakehouse architecture</strong>, and modern data modelling.</li><li>Comfortable owning architecture and working across multiple engineering teams.</li></ul><p><br/></p><p>If you're excited by the opportunity to build the data platform behind AI systems serving hundreds of millions of users worldwide, while owning architecture and solving genuinely complex engineering challenges at scale, we'd love to hear from you.</p><p><br/></p> </div>