Data Engineer
Analytics
Business Analytics
Business Intelligence
Data
Data Analysis
Data Analytics
Data Architecture
Data Engineer
Data Integration
Data Pipeline
Data Platform
Data Processing
Data Visualization
Database
Databases
Design
ETL
Hr Technology
Informatica
Information Technology (IT)
Integration
Lakehouse
Microsoft
Microsoft Office
Power BI
Power Platform
Reporting and Analytics
SQL
Visual Design
Job Description
TWFG Insurance is hiring a Data Engineer to help build and maintain the data foundation that powers enterprise analytics across insurance operations. Based in Spring, TX (onsite), this role focuses on scalable data pipelines, lakehouse architectures, and curated data products that support reporting and operational insights.
You will design and deliver ELT/ETL solutions using Microsoft Fabric and PySpark, collaborating with cross-functional teams to integrate data from multiple internal and external systems. The work also includes data quality, governance, performance tuning, and supporting analytics capabilities for Power BI Embedded.
What you’ll do
- Design, build, and maintain scalable ELT/ETL pipelines using Microsoft Fabric and PySpark notebooks.
- Develop and optimize lakehouse layers using the Medallion architecture.
- Create curated dimensional models, fact tables, and semantic-ready datasets for reporting and analytics.
- Integrate data from internal and external platforms, including policy administration systems, carrier systems, vendor platforms, and APIs.
- Implement data quality controls, validation frameworks, and monitoring processes.
- Develop and maintain Fabric Lakehouses, Mirrored Databases, Dataflows, Notebooks, and Pipelines.
- Optimize Direct Lake architectures and semantic model performance.
- Support Power BI Embedded analytics and enterprise reporting solutions.
- Manage large-scale datasets and ensure optimal query performance across analytical workloads.
- Contribute to data governance, metadata management, lineage tracking, and cataloging efforts.
- Contribute to CI/CD, testing, release management, and operational monitoring for data assets.
- Maintain documentation, data dictionaries, and architecture diagrams.
- Promote engineering best practices and reusable development patterns.
- Work on certification for the organization’s Cloud Provider’s Data platform if it is not already obtained.
Requirements
- Bachelor’s degree in Computer Science, Information Systems, Engineering, Mathematics, or a related discipline.
- Equivalent professional experience will be considered.
- 3+ years of experience in Data Engineering, Analytics Engineering, or Data Warehousing.
- Experience designing and supporting enterprise-scale data pipelines.
- Strong SQL development skills.
- Experience with Python and/or PySpark.
- Experience working with cloud-based data platforms: Azure, AWS, or GCP.
Technologies you’ll work with
- Microsoft Fabric, PySpark, ELT/ETL
- Lakehouse, Medallion architecture
- Power BI Embedded
- Direct Lake
- Azure, AWS, GCP
- SQL, Python, CI/CD
- REST APIs
- Delta Lake, OneLake
- Row-Level Security
- Microsoft Purview
Preferred qualifications
- Hands-on experience with Microsoft Fabric.
- Experience with Power BI semantic models and analytics workloads.
- Experience with Azure Data Engineering technologies.
- Knowledge of Delta Lake, OneLake, and Direct Lake architectures.
- Experience implementing Row-Level Security and enterprise data governance practices.
- Insurance, financial services, or highly regulated industry experience.
- Familiarity with data cataloging and lineage tools such as Microsoft Purview.
Compensation and benefits
Salary: USD 85,000 - 110,000 per year.
- Medical/Dental/Vision Insurance
- Life Insurance
- 401K with a matching component
- LTD (employer paid)
- Education reimbursement based on tenure
- Paid Time Off
- 9 Paid Holidays a year
- 401(k) matching
- Dental insurance
- Health insurance
- Vision insurance
- Work Location: In person