Back to jobs

SENIOR DATA ENGINEER

Sigmasoftware2
Full-timesenior

Job description

<ul><li>Design and build scalable, cloud-native data platforms from greenfield to production</li><li>Implement near-real-time ingestion pipelines using event-driven patterns</li><li>Define and enforce platform standards, including Data Lake / Lakehouse principles, medallion architecture, and data contracts</li><li>Refactor and optimise existing Spark and PySpark scripts for performance and maintainability</li><li>Introduce best practices for code quality, testing, and CI/CD across data pipelines</li><li>Drive adoption of AI tooling and agentic workflows within the data engineering team</li><li>Ensure data quality, observability, and reliability across all pipelines and platforms</li><li>Develop self-service tooling and microservices to simplify platform usage for other teams</li></ul> <ul><li>5+ years of professional experience in Data Engineering</li><li>Strong Python and SQL development skills for pipeline development and optimisation</li><li>Proficiency in Apache Spark / PySpark, including query optimisation and performance tuning</li><li>Hands-on experience with Databricks (preferred) or Snowflake</li><li>Experience with at least one major cloud provider: Azure (preferred), AWS, or GCP</li><li>Experience with stream processing technologies (Kafka, Spark Structured Streaming)</li><li>Solid understanding of ETL/ELT patterns, data modelling (dimensional, Data Vault), and data warehousing</li><li>Experience with orchestration tools (Apache Airflow, Azure Data Factory, or equivalent)</li><li>Knowledge of Infrastructure as Code (Terraform or equivalent)</li><li>Understanding of production-grade system requirements: reliability, scalability, observability, and performance</li><li>Upper-Intermediate English level</li></ul><p><strong>WILL BE A PLUS</strong></p><ul><li>Familiarity with RAG pipeline design and LLM integration patterns</li><li>Knowledge of data governance frameworks and tools (Unity Catalog, Apache Atlas, or similar)</li><li>Experience with dbt for data transformation and modelling</li><li>Familiarity with MLflow, Feature Stores, or ML platform integration</li></ul> <p><strong>PERSONAL PROFILE</strong></p><ul><li>Self-driven and proactive in identifying improvements</li><li>Comfortable working in a fast-paced, innovative environment</li><li>Strong problem-solving mindset with attention to detail</li><li>Open to experimenting with emerging technologies and approaches</li></ul>

Skills

PythonSQLApache SparkPySparkDatabricksSnowflakeAzureAWSGCPKafkaSpark Structured StreamingApache AirflowAzure Data FactoryTerraformdbtMLflowFeature StoresRAGLLMUnity CatalogApache Atlas