Data Engineer
Hace 8 meses
Madrid, España
Jobrapido
Jornada completa
Gratis con email o Google
Guarda esta oferta y sigue tu búsqueda
Crea una cuenta gratis para guardar empleos, crear alertas y volver a esta oferta desde tu panel.
Gratis con email o Google
Al continuar, aceptas nuestros Términos & Política de Privacidad.
We are an international technology services company founded in 1983 and currently have over 2,000 employees in 5 countries: France, Spain, Romania, Portugal, and LuxembourgWhat are we looking for?A Data Engineer to join a stable international project, based in Madrid.
Responsibilities:
- Contribute to the production support and maintenance of the DataHub, correct incidents and anomalies, resolve data-related issues, and implement functional and technical developments to ensure the stability of production processes and ensure the DataHub is available to users with minimal disruption during agreed uptime requirements.
- Modify existing code according to business requirements and continuously improve it for improved performance and maintainability.
- Ensure the performance and security of the data infrastructure and follow data engineering best practices.
- Data modeling and development of efficient data pipelines to enrich and transform large volumes of data with complex business rules, automate data pipelines, and optimize data ingestion. Design and implement scalable and secure data processing pipelines using Scala, Spark, cloud object storage, Hadoop, and COS.
- Integrate data from multiple sources and formats into the raw layer of the DataHub.
- Implement data transformation and quality to ensure data consistency and accuracy. Use programming languages such as Scala and SQL and tools such as Spark for data transformation and enrichment operations.
- Configure CI/CD pipelines to automate deployments, unit testing, and development management.
- Write and perform unit and validation tests to ensure the accuracy and integrity of the developed code.
- Implement various orchestrators and scheduling processes to automate data pipeline execution (Airflow as a Service).
- Migrate existing Hadoop infrastructure to cloud infrastructure on Kubernetes Engine, COS, Spark as a Service, and Airflow as a Service.
- Write technical documentation (specifications, operational documents) to ensure knowledge capitalization.
Requirements:
o Spark en Scalao CI/CD (Gitlab, Jenkins...)o HDFS and structured databases (SQL)o Apache Airflowo Streaming process (Kafka, event steam...)o Storage/COS S3o Shell Scripto Kuberneteso Elasticsearch and Kibanao HVaulto Dremio as tool to virtualize datao Dataikuo English and FrenchValuableo Elasticsearch and Kibanao Streaming process (Kafka, Event Steam...)o Vaulto DataikuWork Model:Hibrid.
Flexible hours, Monday to Friday.
We offer
Continuous trainingCareer plan tailored to employee preferencesProgression within the companyFlexible
working hours
Hybrid work modelLanguage training (English, French, Spanish).
Salary:
50.000€Would you like to join our team?If you have experience in data and are looking to grow technically and professionally, don't hesitate to apply for this position. Contact us
Responsibilities:
- Contribute to the production support and maintenance of the DataHub, correct incidents and anomalies, resolve data-related issues, and implement functional and technical developments to ensure the stability of production processes and ensure the DataHub is available to users with minimal disruption during agreed uptime requirements.
- Modify existing code according to business requirements and continuously improve it for improved performance and maintainability.
- Ensure the performance and security of the data infrastructure and follow data engineering best practices.
- Data modeling and development of efficient data pipelines to enrich and transform large volumes of data with complex business rules, automate data pipelines, and optimize data ingestion. Design and implement scalable and secure data processing pipelines using Scala, Spark, cloud object storage, Hadoop, and COS.
- Integrate data from multiple sources and formats into the raw layer of the DataHub.
- Implement data transformation and quality to ensure data consistency and accuracy. Use programming languages such as Scala and SQL and tools such as Spark for data transformation and enrichment operations.
- Configure CI/CD pipelines to automate deployments, unit testing, and development management.
- Write and perform unit and validation tests to ensure the accuracy and integrity of the developed code.
- Implement various orchestrators and scheduling processes to automate data pipeline execution (Airflow as a Service).
- Migrate existing Hadoop infrastructure to cloud infrastructure on Kubernetes Engine, COS, Spark as a Service, and Airflow as a Service.
- Write technical documentation (specifications, operational documents) to ensure knowledge capitalization.
Requirements:
o Spark en Scalao CI/CD (Gitlab, Jenkins...)o HDFS and structured databases (SQL)o Apache Airflowo Streaming process (Kafka, event steam...)o Storage/COS S3o Shell Scripto Kuberneteso Elasticsearch and Kibanao HVaulto Dremio as tool to virtualize datao Dataikuo English and FrenchValuableo Elasticsearch and Kibanao Streaming process (Kafka, Event Steam...)o Vaulto DataikuWork Model:Hibrid.
Flexible hours, Monday to Friday.
We offer
Continuous trainingCareer plan tailored to employee preferencesProgression within the companyFlexible
working hours
Hybrid work modelLanguage training (English, French, Spanish).
Salary:
50.000€Would you like to join our team?If you have experience in data and are looking to grow technically and professionally, don't hesitate to apply for this position. Contact us