Skip to content
Back to talent search

Candidate profile

Principal Fractional Distributed Systems Data Engineer

BigQueryKafkaTokioPrestoSQLKinesisDBTClickhouseScalaAerospikeSpark StreamingApache IcebergFlinkCUDAHiveSQLGolangAirbyteZookeeperDelta LakeVector DBsRedshiftTemporalgRPC/protobufRustTrinoSnowflakeDatabricksPyO3SparkCassandraStormPythonAirflowTerraform

Description

Pursuing a role in the high-scalability distributed systems space. A well-rounded Scala/Rust/Python developer, well-versed in data engineering, with deep knowledge of the internals of distributed datastores. Experienced with data modeling for high throughput database activity and a strong understanding of which workloads and data access patterns are scalable and what datastore the data should reside in. Has a highly confident ability to lead a data / MLops project from start to finish. Core Focus Areas include: Cassandra (Data Modeling, Troubleshooting Performance And Operational Issues), Apache Iceberg (Scaling, Tuning, Self-Hosted Setup), Stream Processing At Scale (Kafka, Flink, Spark Streaming, Storm), Custom-Crafted Contextualized Embeddings, Vector-Based Semantic Search, Deep Intent Recognition In Search Engine Queries. Languages: Scala, Rust, Python, SQL (proficient), Golang (ramping up). Educational Background: Computer Science. Solid experience working remotely and working with teams that are distributed geographically. Typically works Pacific Time hours. Asking for $385K base, pro-rated.