wagey.ggwagey.gg
31,365  jobs31,365  jobs
Browse Tech JobsCompaniesFeaturesPricingFAQs
Log InGet Started Free
Jobs(31,365)/Data Engineer Role(444)/Simulmedia (1) - Senior Data Engineer
Pro members applied to this job 36 hours before you saw itGet Pro ›
Simulmedia

Simulmedia - Senior Data Engineer

Remote - Lviv, Kyiv2d ago
RemoteSeniorEMEACloud ComputingArtificial IntelligenceData EngineerSenior Data EngineerSQLPythonSnowflakeDatabricksRedshiftAirflowRESTFlaskFastAPIAWSDockerTemporalSentryJenkinsGrafana

Requirements

• Location: Ukraine is mandatory. Our offices are located in Kyiv and Lviv. Teams are located in Kyiv and Lviv and primarily work remotely with occasional offline meetings. • Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience • 7+ years of work experience as a data engineer • Proficiency in Python and using it as the primary development language in recent years • Expert-level SQL: comfortable writing, reading and tuning complex analytical queries against very large tables, and debugging why two result sets disagree • Hands-on experience with a distributed data processing platform (Spark/Databricks strongly preferred; EMR, Snowflake or BigQuery also relevant) and with a columnar data warehouse (Redshift, Snowflake, BigQuery, ClickHouse, etc) • Ability to design complex data models: normalized, dimensional and temporal (slowly changing dimensions, effective-dated records, point-in-time correctness) • Experience with workflow orchestration tools (Airflow or similar): building DAGs, managing dependencies and running backfills • Experience integrating third-party data feeds: handling schema drift, late or missing deliveries, vendor data-quality defects and versioned reference data • Experience building REST services in Python (FastAPI, Flask, etc) • Experience developing, maintaining, and debugging problems in large server-side code bases • Working knowledge of AWS (S3, IAM, ECS or similar compute) and Docker • Good knowledge of engineering best practices and testing (unit test, integration test, code review, CI/CD) • The desire to take a high level of ownership of the things you work on • Ability to learn new things quickly, maintain a high bar for quality, and be pragmatic • Must be able to communicate with U.S based teams • Experience with Delta Lake / medallion lakehouse architectures is a plus • Experience migrating legacy pipelines between platforms with strict parity requirements is a plus • Experience with advertising, media or measurement industry data is a plus • Ability to communicate effectively with the U.S.-based teams and work 11:00 AM — 8:00 PM EEST (11:00 - 20:00). • Almost everything we run is on AWS (S3, ECS, EMR, RDS and more) • Python is our primary language; SQL is everywhere • Databricks (Spark, Delta Lake) is our lakehouse platform; Redshift and Postgres are our warehouses and operational databases • Airflow orchestrates our pipelines • Docker for packaging; GitHub Actions and Jenkins for CI/CD • Grafana, Sentry and OpenSearch for observability • Datasets measured in billions of rows • Interview Process: • Pre-screening (30 mins) • Technical Interview (1h) • System Design (1.5h) • Product Interview (30 mins) • We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.

Responsibilities

• Design and build batch data pipelines that ingest, validate and transform multi-billion-row datasets from external data providers and internal systems • Model complex real-world data: dimensional models, reference data, and temporal data whose attributes change over time (e.g. slowly changing dimensions), and evolve those models safely as upstream sources change their schemas and semantics • Develop and operate workloads on our lakehouse platform (Databricks / Spark / Delta) and our data warehouse (Redshift), including migrating existing pipelines from the warehouse to the lakehouse • Orchestrate pipelines with Airflow: scheduling, dependencies, retries, backfills and alerting • Prove correctness, not just completion: design parity checks and reconciliation queries when replacing an existing pipeline, run large historical backfills, and investigate data discrepancies down to the row level • Build and maintain Python services and REST APIs that serve data to internal products • Optimize for performance and cost: query tuning, table design, workload management and right-sizing compute • Own what you ship: monitor production pipelines, participate in incident triage and root-cause analysis, and harden systems so the same failure does not happen twice • Collaborate cross-functionally with product managers, data scientists and stakeholders across the company to deliver on product roadmap • Work within an Agile team that releases cutting-edge new features regularly • Take a high degree of ownership and freedom to experiment with new technologies to improve our software

Apply in one click

Upload My Resume

Drop here or click to browse · Tap to choose · PDF, DOCX, DOC, RTF, TXT

Apply in One Click
Apply in One Click

Similar roles

DeepLDeepL - Senior Research Data Engineer5mo ago
·Berlin, Germany, Hybrid
In OfficeEMEASeniorCloud ComputingArtificial IntelligenceData EngineerSenior Data EngineerKubernetesAWSPythonGoData Analysis
The Quality Group GmbHThe Quality Group GmbH - (Senior) Data Engineer (gn)1mo ago
·Remote - Deutschland
RemoteEMEASeniorCloud ComputingData EngineerSenior Data EngineerGitdbtSQLPythonAWS
CapcoCapco - Senior Azure Data Engineer (Databricks)5mo ago
·UK - London - Hybrid
In OfficeEMEASeniorCloud ComputingData EngineerSenior Data EngineerAzureDatabricksLearning & DevelopmentPythonSQL
massive-rocketmassive-rocket - Senior Data Engineer (Databricks) - 8 months FTC2d ago
·Remote - UK
RemoteEMEASeniorCloud ComputingData EngineerSenior Data EngineerReportingRocketPythonSQLAWSAzureGCPGitPower BITableauB2CDatabricks
iFoodiFood - Engenheiro(a) de dados Sênior3mo ago
·Remote
RemoteEMEASeniorCloud ComputingSenior Data EngineerPythonSQLAWSAirflowDatabricks
RoomPriceGenieRoomPriceGenie - Remote Senior Data Engineer (m/f/d)5mo ago
·Germany
In OfficeEMEASeniorCloud ComputingData AnalyticsSenior Data EngineerPythonSnowflakeDatabricksRedshiftAirflow
The Quality GroupThe Quality Group - (Senior) Data Engineer (gn)1mo ago
·Remote - Deutschland
RemoteEMEASeniorCloud ComputingLogisticsData EngineerSenior Data EngineerSQLdbtGitPythonProduct Marketing
Obsidian SecurityObsidian Security - Senior Security Data Engineer1mo ago
·Manchester, UK·£91k/year/year + Equity
In OfficeEMEASeniorArtificial IntelligenceSoftwareData EngineerSenior Data EngineerPythondbtSQLClickHouseGoogle Workspace
2U2U - Senior Data Engineer5mo ago
·Remote - Cape Town, South Africa; Remote - South Africa
RemoteEMEASeniorCloud ComputingLogisticsData EngineerSenior Data EngineerMentoringGoal SettingGovernanceSQLDocker

Browse more by category

Show 444 moreData EngineerShow 286 moreSenior Data EngineerShow 2,750 moreSQLShow 4,958 morePythonShow 631 moreSnowflakeShow 347 moreDatabricksShow 115 moreRedshiftShow 321 moreAirflowShow 681 moreRESTShow 77 moreFlask
Privacy·Terms··Contact·FAQ·Wagey on X