Middle Data Engineer ID86295
Compensation
$95k - $140k / yr
💡 Compensation Note
Estimated market benchmark or extracted range. Exact salary, benefits, and bonuses are confirmed directly with the employer upon applying.
Job & Contract
Full-Time • Permanent
Work Policy
100% Remote Role
Direct Employer
Verified ATS Link
Job Overview & Responsibilities
AgileEngine is an Inc. 5000 company that creates award-winning software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has earned us multiple Best Place to Work awards.
WHY JOIN US
If you're looking for a place to grow, make an impact, and work with people who care, we'd love to meet you!
ABOUT THE ROLE
We are looking for a Middle Data Engineer to help modernize a 15-year-old data warehouse into a governed Databricks Lakehouse. You will build batch and streaming pipelines with PySpark and Delta Lake, following a medallion architecture across bronze, silver, and gold layers. This role also uses AI tools like Claude and GitHub Copilot to speed up development.
WHAT YOU WILL DO
- Design, build, and operate batch and streaming data pipelines on Databricks using PySpark, Delta Lake, and Databricks Workflows.
- Model and maintain a medallion (bronze/silver/gold) architecture serving analytics, reporting, and machine learning consumers.
- Migrate legacy ETL and data warehouse workloads onto the Lakehouse with validated data parity and minimal business disruption.
- Use Claude or Github Copilot as a development accelerator, generating code scaffolding, writing and reviewing tests, creating documentation and prototyping solutions.
- Write clean, well-tested Python and SQL; maintain high standards through code review and documentation.
- Optimize Spark jobs and Delta tables for performance and cost, including partitioning, clustering, caching, and cluster sizing.
- Implement data quality, lineage, and governance controls using Unity Catalog and automated validation checks.
- Debug, troubleshoot, and resolve pipeline failures, data defects, and production incidents.
- Collaborate with DevOps, platform, and analytics engineers on observability, security, and compliance best practices.
MUST HAVES
- 3+ years of professional experience in data engineering, featuring direct expertise with Apache Spark and cloud-based data architectures.
- Strong hands-on experience building data pipelines with Databricks, Apache Spark (PySpark), and Delta Lake.
- Advanced SQL and Python, with strong data modeling skills across dimensional and Lakehouse patterns.
- Experience with streaming ingestion using Structured Streaming, Auto Loader, Kafka, or Event Hubs.
- Experience with workflow orchestration (Databricks Workflows, Airflow, or Azure Data Factory).
- Experience with legacy platform migrations, ETL modernization, or managing data hygiene when porting old systems.
- Strong problem-solving, collaboration, and communication skills.
- Familiarity with Unity Catalog, data governance, access control, and PII handling.
- Experience with dbt or an equivalent transformation framework.
- Familiarity with secure coding standards and industry security best practices.
- Experience delivering production data platforms at scale.
- Upper-intermediate English level.
NICE TO HAVES
- Experience with Infrastructure as Code (IaC) using Terraform and CI/CD using Azure DevOps.
- Experience working with relational databases (specifically PostgreSQL) and data persistence concepts.
- Familiarity with logging and monitoring tools (e.g., Dynatrace, CloudWatch, Databricks system tables).
- Experience working in Agile or team-based development environments preferred.
PERKS AND BENEFITS
- Growth without limits: build your skills through mentorship, internal TechTalks, challenging projects, and a dedicated annual learning budget
- Competitive compensation: get recognition that reflects your skills and impact, with regular performance and compensation reviews
- Flexibility: work 100% remotely with flexible hours that support focus, autonomy, and a healthy work rhythm
- Meaningful, modern projects: build impactful products using modern technologies alongside global teams and leading brands
- Collaborative culture: join a supportive environment with zero micromanagement where ideas are welcomed and contributions are recognized
- Well-being & support: access local well-being programs and people-focused support tailored to your location
Meet Our Recruitment Process
Application → Coding Challenge → Video Interview → Technical Interview or Hiring Manager Interview
Each step helps us understand your skills and overall fit.
If it’s a match, you’ll receive an offer.
Direct Employer ATS Application
Ready to join AgileEngine?
Apply directly on the official company career portal with verified fast-track review.
Frequently Asked Questions
Essential verified facts about this Middle Data Engineer ID86295 opportunity.
Is this role 100% remote? ▾
Yes, this vacancy has been verified for remote and flexible working arrangements. Candidates can work from home or anywhere in supported timezones.
What is the expected compensation? ▾
The compensation for this role is estimated at $95k - $140k / yr. Exact salary, equity packages, and bonuses are confirmed directly with AgileEngine during interview stages.
How do I submit my application securely? ▾
Click on the "Apply for this Role" button above. You will be routed directly to the official AgileEngine career portal with zero third-party recruiter interception.
Related Opportunities
Similar Remote Jobs You May Like
Explore more verified 100% remote roles matching this specialization.