Data Engineer
Job Description
Data Engineer role at Todata Analytics in Omaha, NE (onsite), focused on governed ingestion and transformation pipelines for products and AI in a regulated environment.
Responsibilities
- Build and maintain data ingestion and transformation pipelines across the platform
- Design and evolve data models that support products and analytics
- Implement data governance and access controls for a regulated environment, ensuring correct client data isolation
- Apply engineering fundamentals to the codebase, including CI/CD, consistent naming conventions, and environment separation
- Collaborate with product and software development to translate client needs into well-modeled, usable data
- Support infrastructure that makes governed data available to downstream consumers, including AI products
- Monitor data quality and lineage, and address issues before they reach clients
Requirements
- 3+ years in data engineering with hands-on Databricks experience (Spark, Delta Lake, Unity Catalog)
- Strong SQL skills: complex queries, joins, window functions, performance optimization
- Ability to design dimensional/star-schema data models from ambiguous requirements
- Solid Python experience for data processing (for example, PySpark, pandas) and scripting
- Hands-on Databricks usage (notebooks, Delta Lake, jobs, clusters)
- Excellent problem-solving and debugging: trace issues through logs, code, and data to find root causes
- Demonstrated data governance practices and multi-tenant isolation, including dataset separation and permissioning
- Experience with CI/CD, version control, and disciplined naming/environment conventions for data work
- Proven ability to independently deliver foundational infrastructure work without close supervision
- Clear written communication, including documentation of what you build
Technologies
- SQL, Python, PySpark, pandas
- Databricks, Spark, Delta Lake, Unity Catalog
- CI/CD
Compensation & Location
- Location: Omaha, NE (onsite)
- Salary: USD 80,000 - 140,000 per year
- Experience: 3+ years
Benefits
- 401(k) matching
- Dental insurance
- Health insurance
- Life insurance
- Paid time off
- Professional development assistance
- Retirement plan
- Vision insurance
Nice to Have
- Experience in a HIPAA, SOC 2, or otherwise regulated data environment
- Familiarity with healthcare/clinical research or financial/accounting data domains
- Exposure to enabling AI/LLM consumers of a governed semantic layer
- Databricks certification