Data Engineer
ATC · Texas, United States
Apply directly on ATC’s careers site — no account needed.
About the role
Job Summary
We are seeking a Data Engineer with 5+ years of experience in designing, developing, and optimizing scalable data pipelines and cloud-based data platforms. The ideal candidate will have strong expertise in Python, SQL, Databricks, Apache Spark, and ETL/ELT development, along with hands-on experience in at least one cloud platform (AWS, Azure, or Google Cloud Platform).
Key Responsibilities
- Design, develop, and maintain scalable ETL/ELT data pipelines.
- Build and optimize data ingestion, transformation, and integration workflows using Databricks and Apache Spark.
- Develop and maintain data lakes and cloud-based data warehouses.
- Create scalable batch and real-time data processing solutions.
- Optimize data pipelines for performance, reliability, and scalability.
- Develop reusable data engineering frameworks and automation solutions.
- Collaborate with data analysts, data scientists, and business stakeholders to deliver high-quality data solutions.
- Implement data quality, governance, monitoring, and security best practices.
- Troubleshoot production issues and continuously improve data platform performance.
- Follow DevOps and CI/CD best practices for data engineering projects.
Required Qualifications
- Master’s degree in Computer Science, Information Technology, Engineering, or a related field.
- 5+ years of experience as a Data Engineer.
- Strong programming skills in Python.
- Advanced SQL proficiency.
- Hands-on experience with Databricks.
- Strong experience with Apache Spark (PySpark preferred).
- Experience building and maintaining ETL/ELT pipelines.
- Experience with Apache Airflow or similar workflow orchestration tools.
- Experience with data warehousing technologies such as Snowflake, Amazon Redshift, Google BigQuery, or Azure Synapse Analytics.
- Hands-on experience with at least one cloud platform (AWS, Azure, or Google Cloud Platform).
- Experience with Git, Docker, and CI/CD pipelines.
- Knowledge of relational and NoSQL databases.
Preferred Qualifications
- Experience with Kafka or other streaming platforms.
- Experience with Delta Lake.
- Familiarity with Apache Iceberg or Apache Hudi.
- Experience with Infrastructure as Code (Terraform or CloudFormation).
- Knowledge of data governance and data quality frameworks.
- Exposure to DevOps and MLOps practices.
Description sourced from the public LinkedIn listing — this role isn't indexed from the company's career page yet.
Skills
- Python
- SQL
- Databricks
- Airflow
- Snowflake
- Redshift
- BigQuery
- AWS
- GCP
- Docker
- GitHub Actions
Never be applicant #200 again
Every job here is indexed straight from company career pages — often hours after it opens, before it reaches the big boards. Create a free account and get your best matches in a twice-daily digest.
- Your best matches, twice a day
- No duplicates, no ghost jobs, no recruiter spam
- Every job free to browse — pay only when you apply
Free account — no card required
93 115 live jobs · 17 641 companies tracked · 159 added today