Data Engineer
Ripjar · Cheltenham, United Kingdom
About The Role
Join our Data Engineering team as a Data Engineer, where you'll be responsible for the development and operation of the Data Collection Hub. You'll engineer distributed ingestion services, build high-throughput processing components, design and evolve data contracts, and improve platform reliability. You'll work with a technology stack that includes Python, Node.js, HDFS, HBase, Spark, MongoDB, OpenSearch, Airflow, and Kubernetes. Enjoy benefits such as a pension scheme, private family health insurance, 25 days of holidays, EMI share options, flexible working, a wellbeing allowance, and free lunch and snacks.
- Conception et développement de services d'ingestion distribués pour extraire des données de diverses sources.
- Création de composants de traitement à haut débit, en mettant l'accent sur la performance, l'évolutivité et le coût prévisible.
- Conception et évolution des contrats de données (schémas, règles de validation, versioning) pour garantir la confiance des équipes en aval.
- We’re looking for someone with 2+ years of industry experience building and operating production software who enjoys working across data pipelines, distributed systems, and operational reliability
- Fluency in at least one programming language (Python/Node.js a plus)
- 2+ years building and operating production software systems
- Strong fundamentals: data structures, testing, version control, Linux basics
- Experience debugging moderately complex systems and improving reliability/performance
- Spark/PySpark experience
- Hadoop ecosystem exposure (HDFS/HBase)
- Workflow orchestration (Airflow/Dagster/NiFi)
- Search/indexing (OpenSearch, MongoDB)
- Kubernetes and infrastructure-as-code
- Degree in Computer Science or numerical degree
Similar roles you might like
See all →This is an external listing. JobSpring does not represent or verify the employer. Report this listing
