wagey.ggwagey.gg
29,762  jobs29,762  jobs
Browse Tech JobsCompaniesFeaturesPricingFAQs
Log InGet Started Free
Jobs(29,762)/Data Engineer Role(401)/Dun & Bradstreet - Principal Data Engineer (R-19440)
Dun & Bradstreet

Dun & Bradstreet - Principal Data Engineer (R-19440)

Remote - Hyderabad - India3h ago
RemotePrincipalAPACCloud ComputingArtificial IntelligenceData EngineerPrincipalGitDockerKubernetesData QualitySQLAWSPythonRedshiftJiraConfluencePower BILookerTableauSnowflakeDatabricksData VisualizationGovernance

Requirements

• 8-12+ years of hands-on experience in data engineering or large-scale data processing. • Proven experience building and maintaining production-grade data pipelines and distributed systems. • Demonstrated experience architecting and delivering large-scale data platforms or mission-critical data systems. • Strong expertise in: SQL and relational databases (Postgres, BigQuery, Redshift, etc.), Python for data processing and analysis. • Experience with Google Cloud Platform (BigQuery, Dataflow, Pub/Sub, Cloud Storage, Cloud Functions) and/or AWS (S3, Redshift, EMR, RDS). • Experience working with large-scale datasets (hundreds of millions to billions of records). • Strong understanding of data modeling, partitioning, indexing, and query optimization. • Experience with distributed data processing and parallelization techniques. • Experience moving large volumes of data across systems and architectures. • Familiarity with CI/CD, containerization, and orchestration tools (Docker, Kubernetes, GitHub Actions, etc.). • Strong debugging and troubleshooting skills in complex data environments. • Experience with version control (Git) and Agile tools (Jira, Confluence, etc.). • Highly analytical with strong attention to detail and a data-driven mindset. • Ability to hit the ground running, quickly understand systems, and deliver independently. • Comfortable working in a remote, fast-paced, and collaborative environment. • Proven ability to drive system design and implementation. • Experience with identity graphs, entity resolution, or record linkage systems. • Background in AdTech, digital identity, cookies, or audience data platforms. • Experience with real-time or streaming data systems. • Familiarity with data quality, observability, and monitoring frameworks. • Experience with data visualization tools (Looker, Tableau, Power BI). • Knowledge of data privacy, compliance, and governance considerations. • Experience with modern data platforms such as Snowflake and Databricks. • Exposure to AI/ML technologies, including experience working with or integrating agentic frameworks. • We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please visit https://bit.ly/3LMn4CQ.

Responsibilities

• Design, build, and optimize scalable data pipelines and ETL/ELT workflows for large, complex datasets. • Design and implement foundational data architecture supporting identity resolution and ID graph systems. • Develop and enhance systems supporting identity resolution and ID graph construction (data ingestion, normalization, matching, and deduplication). • Process and unify multi-source datasets (cookies, device IDs, behavioral data, third-party and proprietary data). • Write efficient, testable, and maintainable code using Python and SQL for large-scale data processing. • Optimize data models, queries, and storage strategies for performance, scalability, and cost efficiency. • Build and maintain data validation, monitoring, and alerting systems to ensure data quality and reliability. • Troubleshoot, debug, and improve existing data pipelines and infrastructure. • Own and drive complex data problems end-to-end, from initial design through production deployment. • Make and influence key technical decisions related to data architecture, scalability, and system design. • Collaborate with data, platform, DevOps, and product teams to deliver scalable, production-ready solutions. • Translate business and product requirements into practical, performant data solutions. • Document data pipelines, systems, and workflows clearly. • Continuously improve system performance, data quality, and pipeline resilience. • Contribute to building new capabilities that improve how customers understand and leverage data insights

Apply in one click

Upload My Resume

Drop here or click to browse · Tap to choose · PDF, DOCX, DOC, RTF, TXT

Apply in One Click
Apply in One Click

Similar roles

Gather AIGather AI - Principal Data Engineer1w ago
·Remote - India
RemoteAPACPrincipalCloud ComputingInternet of ThingsData EngineerPrincipalData QualitydbtSQLAWSAzureGCPGovernanceDroneCustomer SuccessCSM
OKXOKX - Data Engineering Director3mo ago
·APAC
In OfficeAPACDirectorCloud ComputingArtificial IntelligenceData EngineerSQLAWSDatabricksData QualitySnowflake
Black Duck Software, Inc.Black Duck Software, Inc. - Principal Data Engineer3mo ago
·Belfast, UK - Hybrid
In OfficeEMEAPrincipalCloud ComputingData EngineerPrincipalMentoringSQLPythonAWSReporting
CodaMetrixCodaMetrix - Principal Data Engineer6d ago
·Boston, Massachusetts, United States - Hybrid
In OfficeNAPrincipalCloud ComputingData AnalyticsData EngineerPrincipalMedical RecordsSQLDatabricksScalaPythonTerraformAWSYAMLKafkaJenkinsGeminiClaudeData GovernanceTableauCustomer Success
Manus AIManus AI - Data Engineer6mo ago
·Singapore
In OfficeAPACMidCloud ComputingArtificial IntelligenceData AnalyticsData EngineerJavaScalaPythonSQLData Quality
AnaplanAnaplan - Principal Data Engineer2mo ago
·London, United Kingdom - Hybrid
In OfficeEMEAPrincipalArtificial IntelligenceCloud ComputingData AnalyticsData EngineerPrincipalPythonProspectingTraining DevelopmentVectorNoSQL
WIN Home InspectionWIN Home Inspection - Data Engineer1mo ago
·Greater Delhi Area
In OfficeAPACMidCloud ComputingData EngineerPythonPandasData VisualizationTableauPower BI
TemusTemus - Data Engineer1mo ago
·Singapore
In OfficeAPACMidCloud ComputingOil & GasData EngineerC#PythonSnowflakeDatabricksAWS
Tech HoldingTech Holding - BI Engineer (Contract)1w ago
·Remote - India
RemoteAPACSeniorCloud ComputingArtificial IntelligenceIntegration EngineerData EngineerBusiness IntelligencePower BISQLReportingAWSRedshiftSnowflakeGitStorytellingKnowledge TransferGovernance

Browse more by category

Show 401 moreData EngineerShow 728 morePrincipalShow 544 moreGitShow 757 moreDockerShow 1,454 moreKubernetesShow 555 moreData QualityShow 2,477 moreSQLShow 2,647 moreAWSShow 4,582 morePythonShow 86 moreRedshift
Privacy·Terms··Contact·FAQ·Wagey on X