wagey.ggwagey.gg
31,365  jobs31,365  jobs
Browse Tech JobsCompaniesFeaturesPricingFAQs
Log InGet Started Free
Jobs(31,365)/Data Engineer Role(453)/Infinity (5) - Data Engineer, Red Tape Index
Infinity

Infinity - Data Engineer, Red Tape Index

Remote - Argentina2d ago
RemoteLATAMCloud ComputingArtificial IntelligenceData EngineerPythonSQLPolarsAWSTerraformJupyterPrefectAirflowDagsterClaude

Requirements

• Our stack is deliberately modern (Python 3.14, uv, ruff, ty, polars, Prefect 3, marimo). We don't filter on those exact tools; we hire for Python and data depth and expect a short ramp. • Strong Python and SQL; you have designed Postgres schemas and owned migrations (SQLAlchemy and Alembic, or equivalents) in production • Data pipeline experience with a lakehouse/medallion mindset: idempotent ingestion, content hashing, and lineage are habits, not aspirations • Web scraping beyond requests: anti-bot evasion, browser automation, and resilience against messy or hostile sources • Statistics literacy for index methodology: winsorization, normalization, weighting, and sensitivity testing, and you can reason about whether an index's math supports its claims • Comfort with modern Python tooling and CI discipline: typing, linting, coverage gates, and conventional commits • Product discovery instincts: you talk with partners in plain language, assess data feasibility before committing, and flag what is proven versus assumed • End-to-end ownership: you are a pragmatic generalist who moves across data, backend, infrastructure, and basic product decisions in an uncertain environment • Prefect experience, or Airflow/Dagster with willingness to switch • AWS (ECS, S3) and Terraform • polars, pyarrow, and marimo or a Jupyter background • LLM-in-pipeline experience (pydantic-ai, AWS Bedrock, evals) • Actuarial, quantitative research, or data science background in ranking or index construction • Experience with government open data (permits, energy, environmental, or economic datasets) • Comfort working alongside AI tooling; our repos are agent-forward (Claude agent teams, spec-driven docs)

Responsibilities

• Ship scrapers and ingestion flows against messy, sometimes adversarial sources, using HTTP/2 clients, TLS-fingerprint evasion, and browser automation fallbacks, and keep them resilient as sources change • Own Postgres schema design and migrations end to end across per-country and per-domain schemas • Build and maintain medallion (bronze → silver → gold) transforms that are idempotent, content-hashed, and lineage-tracked • Implement and defend index methodology: normalization, weighting, and composite construction where the math verifiably says what it claims (our scoring core is held to 100% test coverage) • Assess data feasibility early, clarify requirements with partners, and convert ambiguous index ideas into executable plans • Take an index end to end: sourcing, validation, methodology, publication, and refresh planning • Operate pipelines on our orchestration stack (Prefect dispatching per-flow ECS Fargate tasks) with observability everywhere

Benefits

• High-impact work at the intersection of AI and critical infrastructure regulation • End-to-end ownership of indices, from raw source to published methodology • Small team with outsized influence; your feasibility calls shape what we build • Modern AI-native development environment (Claude Code, Cursor, multi-model orchestration) • Values We Hire For • Values We Hire For • Character: integrity and trustworthiness above all • Character: • Competency: evoking trust and reliably delivering • Competency: • Togetherness: family-level support and alignment • Togetherness: • Impact: meaningful outcomes over activity • Impact: • Commitment: ownership and follow-through • Commitment: • Equal Opportunity Statement

Apply in one click

Upload My Resume

Drop here or click to browse · Tap to choose · PDF, DOCX, DOC, RTF, TXT

Apply in One Click
Apply in One Click

Similar roles

Reality DefenderReality Defender - Data Engineer1w ago
·Remote - NA - USA *·$140k - $180k/year + Equity
RemoteNACloud ComputingArtificial IntelligenceData EngineerAWSKubernetesRayAirflowPrefectDagsterSQLPythonMLOps
Particle41Particle41 - Data Engineer (Tableau)1mo ago
·Remote - Argentina
RemoteLATAMMidCloud ComputingArtificial IntelligenceData EngineerPythonTableauPerformance ManagementAzureAWS
Goodway GroupGoodway Group - Global Data Engineer2mo ago
·Remote - Latin America *
RemoteLATAMSeniorCloud ComputingArtificial IntelligenceData EngineerData QualityScalaSQLPythonSnowflake
oscilaroscilar - Data Engineer1mo ago
·Remote - Brazil·$66k - $85k/year + Equity
RemoteLATAMSeniorArtificial IntelligenceData EngineerKafkaPythonSQLJavaAirflow
SunnyDataSunnyData - Mid-level Data Engineer5d ago
·Mexico
In OfficeLATAMMidCloud ComputingArtificial IntelligenceData EngineerReportingData QualityData GovernanceAWSGCPAzureKafkascikit-learnSQLPandasDatabricks
Vytalize HealthVytalize Health - Data Reliability Engineer1w ago
·Kansas, United States
In OfficeNAJuniorCloud ComputingArtificial IntelligenceData EngineerKPI TrackingAWSSQLPythonDatabricksData QualityAirflowDatadogClaudeNew RelicDocumentationData Governance
kueskikueski - Staff Data Engineer1mo ago
·Remote - LATAM (remote)·Equity
RemoteLATAMStaffCloud ComputingData AnalyticsData EngineerSQLScalaPythonTypeScriptApache Spark
ekumenlabsekumenlabs - Data Engineer2mo ago
·Remote - LATAM
RemoteLATAMMidCloud ComputingData EngineerJavaPythonDocumentationApache SparkAWS
ProfoundProfound - Data Engineer2mo ago
·New York, United States·$160k - $250k/year + Equity
In OfficeNAPrincipalCloud ComputingArtificial IntelligenceData EngineerSQLPythonLearning & DevelopmentdbtAWS

Browse more by category

Show 453 moreData EngineerShow 5,062 morePythonShow 2,815 moreSQLShow 13 morePolarsShow 3,036 moreAWSShow 920 moreTerraformShow 28 moreJupyterShow 43 morePrefectShow 331 moreAirflowShow 70 moreDagster
Privacy·Terms··Contact·FAQ·Wagey on X