Job Title: Full Stack Software Engineer
Location: Bloomington, IL (90% remote and 10% in-office)
Duration: 6 Months
Must-Have Technical Skills:
- AWS Data Services: S3, Redshift, Glue, EMR, Kinesis, Athena, Lake Formation, Step Functions, Lambda
- Programming: Python (PySpark, boto3) advanced proficiency; SQL (window functions, query optimization)
- Big Data / Spark: Apache Spark - batch and streaming; performance tuning
- Pipeline Orchestration: Apache Airflow (MWAA) or equivalent (Prefect / Dagster)
- Data Modeling: Dimensional modeling (star/snowflake schema); data lakehouse design (Iceberg / Delta Lake / Hudi)
- Infrastructure as Code: Terraform or AWS CDK - mandatory
- CI/CD: GitHub Actions, AWS CodePipeline, or equivalent; automated testing for pipelines
- Data Transformation: dbt (preferred)
Good to Have:
- Snowflake or Databricks experience alongside AWS
- Apache Kafka / Amazon MSK for streaming pipelines
- AWS SageMaker or ML pipeline support experience
- Multi-cloud or on-prem to AWS migration experience
Soft Skills & Expectations:
- Ability to independently design and deliver end-to-end data pipelines
- Strong communication skills to engage with analysts, scientists, and business stakeholders
- Experience mentoring junior engineers and conducting code reviews
- Proven track record of debugging and resolving production pipeline issues
Key Skills: Data Engineer, AWS SageMaker, Apache Kafka, Databricks