DeveloperJobs.io
← Back to all jobs

Job Description

Data Engineer role at Caterpillar’s Physical AI Platform team, focused on scalable data pipelines, microservices, and AWS-based data systems for reliable real-time and batch processing.

Responsibilities

  • Design, build, and optimize data pipelines and microservices using Python for real-time and batch processing
  • Develop cloud-based data ingestion solutions using AWS services such as Kinesis, S3, DynamoDB, and EventBridge
  • Build and maintain data integration workflows for source data pipelines aligned to CI Autonomy
  • Translate business requirements into reliable data workflows, mappings, and system designs
  • Implement automated testing and validation to support data integrity across distributed systems
  • Monitor and troubleshoot pipelines using CloudWatch to maintain reliability

Requirements

  • Ability to analyze issues in distributed data systems and deliver scalable solutions
  • Ability to clearly document data flows, mappings, and system behavior for cross-team use
  • Experience building backend systems and pipelines using Python, Java, and modern frameworks
  • Experience delivering solutions in an agile environment
  • Experience integrating APIs, streaming platforms, and databases
  • Ability to design scalable, event-driven data systems
  • Strong understanding of AWS services and data engineering tools
  • Experience implementing testing strategies to ensure data quality and system reliability

Technologies

  • Python, Java, AWS
  • Kinesis, S3, DynamoDB, EventBridge
  • Event-driven data systems, CloudWatch, SQL
  • CI/CD tools: Azure DevOps, Jira, Jenkins
  • CI Autonomy, OAuth 2.0, APIs, microservices
  • Relational databases, noSQL databases, streaming platforms, distributed systems

Top Candidates Will Have

  • Bachelor’s degree in Computer Science, Computer Engineering, or related field
  • 2+ years development experience on modern, large scale, complex data platforms
  • 3+ years developing and deploying Java or Python solutions to a production environment
  • Experience building high-throughput, scalable data pipelines
  • Hands-on AWS data services experience (Kinesis, S3, DynamoDB, EventBridge, etc.)
  • SQL proficiency, including data quality and validation practices
  • CI/CD deployment experience using tools such as Azure DevOps, Jira, Jenkins
  • Experience developing microservices for real-time data ingestion
  • Experience with relational and noSQL databases
  • Ability to ensure data integrity across distributed and streaming systems
  • API development and secure integration experience including OAuth 2.0
  • Proven experience with monitoring, testing, and automation in large-scale data environments

Benefits

  • Medical, dental, and vision benefits
  • Paid time off plan (Vacation, Holidays, Volunteer, etc.)
  • 401(k) savings plans
  • Health Savings Account (HSA)
  • Flexible Spending Accounts (FSAs)
  • Health Lifestyle Programs
  • Employee Assistance Program
  • Voluntary Benefits and Employee Discounts
  • Career Development
  • Incentive bonus
  • Disability benefits
  • Life Insurance
  • Parental leave
  • Adoption benefits
  • Tuition Reimbursement

Location and Work Details

  • Irving, TX (onsite)
  • Full-time role at the Irving, TX office (Dallas)
  • Domestic relocation assistance is available
  • Visa sponsorship is available

Compensation

$97,530.00 - $158,480.00 per year

Posting Notes

  • Any offer of employment is conditioned upon successful completion of a drug screen
  • Caterpillar is an Equal Opportunity Employer, including Veterans and Individuals with Disabilities
  • Qualified applicants of any age are encouraged to apply

Similar Jobs