Senior Data Engineer
- Design, build, and maintain production-grade data pipelines using PySpark and Spark Declarative Pipelines (Delta Live Tables) on Databricks.
- Develop modular, tested, and documented dbt transformation models as part of the team's analytics engineering layer.
- Implement and maintain multi-layer data architectures (Bronze/Silver/Gold) on Delta Lake following medallion architecture principles.
- Contribute to the migration of data pipelines, stored procedures, and SSIS packages from SQL Server to Databricks, ensuring logic fidelity and data integrity.
- Ensure pipeline reliability through robust error handling, data quality checks, and SLA-aligned alerting.
- Manage Databricks compute resources, cluster configurations, and job scheduling via Databricks Workflows.
- Enforce data security and access controls within pipeline and storage design via Unity Catalog.
- Participate in code reviews, contribute to engineering discussions, and help uphold team-wide coding standards.
- Write clean, maintainable Python and SQL code following version control and CI/CD best practices.
- 5+ years of experience in data engineering with a strong focus on pipeline development.
- Hands-on expertise with Databricks, including Delta Lake, Databricks Workflows, and Unity Catalog.
- Proficiency in PySpark and Apache Spark for large-scale distributed data processing.
- Experience building pipelines with Spark Declarative Pipelines (Delta Live Tables).
- Working knowledge of dbt for SQL-based transformation development and testing.
- Experience processing healthcare data including claims, clinical, EMR/EHR, eligibility, or provider data.
- Proficiency in Python and SQL with a disciplined approach to code quality and testing.
- Experience with cloud infrastructure (Azure, AWS, or GCP) and CI/CD pipeline tooling.
- Understanding of data modeling methodologies: Data Vault 2.0, 3NF, and dimensional modeling (Kimball).
- Familiarity with healthcare data compliance considerations (HIPAA) and data standards (HL7, FHIR, ICD-10, CPT, NPI).
- Experience with data orchestration tools such as Apache Airflow or Azure Data Factory.
- Exposure to streaming data pipelines using Kafka or Databricks Structured Streaming.
- Databricks Certified Data Engineer Associate or Professional certification.
Healthcare data experience is essential.
- Major Medical Expense Insurance
- Life Insurance
- Dental and Vision Insurance
- Mental Health Support
- IMSS (Mexican Social Security)
- Seniority Bonus
- Savings Fund Program
- Career Development Plan
- Christmas Bonus (Aguinaldo)
- Vacation Bonus
- Corporate Retirement Plan
- Certifications and Training Programs
- Internal Events
- TotalPass Wellness Program
- Additional Paid Time Off
Additional Protection and Discounts
- Auto and Motorcycle Insurance
- Pet Insurance
- Personal Belongings Insurance