Senior Data Engineer (Lisbon - Hybrid)
Descrição da Empresa
Serviços de Recrutamento e Selecção
Descrição da Função
We're fast learners, hard workers, natural collaborators... and we Make Modern Happen! Our ambition is to unlock the potential of our digital world so that organisations everywhere can innovate and thrive securely. We aim to achieve this goal by bringing together the world’s most talented people and the most powerful technologies, combining them to address our customers' challenges and to build something stronger together. If you share our vision, join us! Right now, we are looking for an Senior Data Engineer to integrate our talent team. Your responsibilities include: Re-architect, optimize, and scale big data processing components. Analyze workload patterns and drive improvements in performance, reliability, scalability, and cost efficiency. Ensure the stability and operational excellence of Apache Spark workloads running in cloud environments. Operate, maintain, and evolve Hadoop ecosystem components, including HDFS and YARN. Design, build, and improve data ingestion and processing pipelines. Enhance developer and power-user experiences across data science and analytics platforms. Collaborate with engineers, data scientists, product teams, and platform teams to deliver roadmap initiatives. Develop, deploy, monitor, and support services following a "you build it, you run it" DevOps mindset. Participate in an on-call rotation and take ownership of platform reliability and operational health. Contribute to the evolution of a scalable, cloud-native, microservices architecture. You must have: 5+ years of experience building and operating distributed big data systems. Strong hands-on experience with Apache Spark, including performance tuning, debugging, and orchestration. Solid software engineering background with strong programming skills in Java. Strong understanding of distributed systems and data processing architectures. Experience with the Hadoop ecosystem, including HDFS and YARN. Experience managing Linux-based environments in the cloud. Experience designing, building, and operating ETL/ELT pipelines. Familiarity with JupyterHub or JupyterLab environments. Experience with CI/CD, monitoring, observability, and production support. Ability to troubleshoot complex distributed systems and work autonomously on challenging technical problems. Strong ownership mindset and commitment to operational excellence. We value: Experience running Apache Spark on EMR or Kubernetes. Hands-on experience with Kubernetes, AWS S3, AWS Glue, Kafka, Apache Airflow, Apache Iceberg, Trino, Nessie. Experience building automation and data pipelines for large-scale data workflows. Experience developing or maintaining Data Science or Machine Learning platforms. Contributions to open-source projects, particularly within the Big Data ecosystem. Experience supporting high-volume, data-intensive platforms in production environments. We offer: Regular professional development; Health insurance, with family package; Regular teambuilding programs; Friendly workplace. Workplace: Lisbon - Hybrid Claranet, Make modern happen!
Localização
- Lisboa, Portugal