• Degree in a quantitative field such as computer science, mathematics, or engineering
• Experience as a Red Teamer https://en.wikipedia.org/wiki/Red_team
• Experience working on large-scale data ingestion, crawling, or indexing systems
• Experience with Apache Spark, Databricks, or other distributed data platforms
• Experience with streaming data systems (Kafka, Pub/Sub, Spark Streaming, etc.)
• Proficiency with SQL and data warehousing (Snowflake, Redshift, BigQuery, or similar)
• Experience with cloud platforms (AWS preferred, GCP or Azure also great)
• Understanding of modern data storage and design patterns (parquet, Delta Lake, partitioning, incremental updates)
• Knowledge of modern data design and storage patterns (e.g., incremental updating, partitioning and segmentation, rebuilds and backfills)
• Experience building and maintaining data pipelines on modern big-data or cloud platforms (Databricks, Spark, or equivalent)