Caterpillar Logo

Caterpillar

Data Engineer

Posted Yesterday
Hybrid
Irving, TX
98K-158K Annually
Entry level
Hybrid
Irving, TX
98K-158K Annually
Entry level
Build and maintain scalable data platforms, batch and real-time pipelines, cloud data solutions, databases, and data integration processes. Improve automation, scalability, monitoring, testing, and data quality while operationalizing data jobs and models. Partner with data scientists, analysts, and business stakeholders to deliver trusted data products supporting pricing decisions, analytics, automation, and AI-enabled outcomes.
The summary above was generated by AI
Career Area:
Technology, Digital and Data
Job Description:
Your Work Shapes the World at Caterpillar Inc.
When you join Caterpillar, you're joining a global team who cares not just about the work we do - but also about each other. We are the makers, problem solvers, and future world builders who are creating stronger, more sustainable communities. We don't just talk about progress and innovation here - we make it happen, with our customers, where we work and live. Together, we are building a better world, so we can all enjoy living in it.
As a Data Engineer on the Global Parts Pricing Data Engineering team, you will design and build scalable data platforms, pipelines, and cloud solutions that power pricing decisions across a global business. You will solve complex data challenges, transform diverse data sources into trusted and actionable insights, and deliver solutions that drive significant business outcomes. Working alongside data scientists, analysts, and business stakeholders, you will develop high-performance data products, automate critical processes, and help shape the future of data and AI-driven pricing determination.
What You Will Do:
  • Build infrastructure required for optimal extraction, transformation, and loading of data from a wide variety of data sources
  • Identify, design, and implement internal process improvements: automating manual processes, optimizing data delivery, re-designing infrastructure for greater scalability
  • Perform debugging, troubleshooting, modifications and testing of integration solutions
  • Operationalize the developed jobs, processes and models.
  • Create databases and infrastructure to process data at scale
  • Create solutions and methods to monitor systems and solutions
  • Automate code testing and pipelines
  • Engage with business partners to participate in design and development of data integration/transformation solutions per functional requirements.
  • Engage and actively seek industry perspectives through, continuous learning, peer groups, etc.

What You Will Have:
  • Data Pipeline Development: Ability to design, build, and optimize scalable batch and real-time data pipelines that deliver trusted, high-quality data for analytics, reporting, and AI applications.
  • Cloud Data Engineering: Experience with cloud-based data platforms, storage, and processing technologies; ability to develop resilient, scalable, and secure data solutions.
  • Data Modeling & Integration: Knowledge of data modeling, transformation, and integration techniques; ability to consolidate diverse data sources into well-governed, consumable data products.
  • Data Observability & Quality: Experience implementing data quality frameworks, monitoring, testing, and automated validation to ensure reliable and trusted data assets.
  • Collaboration & Business Impact: Ability to partner with cross-functional teams to understand business needs and deliver data solutions that support pricing decisions, analytics, automation, and AI-enabled outcomes.

Top Candidates Will Have:
  • Extensive experience in designing and developing software applications in Python.
  • Experience with AI assisted software development (ex. Github Copilot, Cursor, Codex, etc)
  • Extensive experience working with Git version control
  • Extensive experience deploying software using CI/CD tools such as Github Actions, Azure Devops etc.
  • Experience with AWS components such as Bedrock, Lambda, Fargate, S3, Sagemaker, IAM and RDS
  • Experience with databases such as PostgreSQL, Snowflake, Oracle, etc.
  • Demonstrated strong learning ability and a proactive approach to staying current with the latest technologies and industry trends.

Summary Pay Range:
$97,530.00 - $158,480.00
Compensation and benefits offered may vary depending on multiple individualized factors, job level, market location, job-related knowledge, skills, individual performance and experience. Please note that salary is only one component of total compensation at Caterpillar.
Benefits:
Subject to plan eligibility, terms, and guidelines. This is a summary list of benefits.
  • Medical, dental, and vision benefits*
  • Paid time off plan (Vacation, Holidays, Volunteer, etc.)*
  • 401(k) savings plans*
  • Health Savings Account (HSA)*
  • Flexible Spending Accounts (FSAs)*
  • Health Lifestyle Programs*
  • Employee Assistance Program*
  • Voluntary Benefits and Employee Discounts*
  • Career Development*
  • Incentive bonus*
  • Disability benefits
  • Life Insurance
  • Parental leave
  • Adoption benefits
  • Tuition Reimbursement

* These benefits also apply to part-time employees
This position requires working onsite five days a week.
Relocation is available for this position.
Visa Sponsorship is not available for this position.
Posting Dates:
October 7, 2026 - October 14, 2026
Any offer of employment is conditioned upon the successful completion of a drug screen.
Caterpillar is an Equal Opportunity Employer, Including Veterans and Individuals with Disabilities. Qualified applicants of any age are encouraged to apply.
Not ready to apply? Join our Talent Community.

Similar Jobs at Caterpillar

6 Days Ago
Hybrid
98K-158K Annually
Mid level
98K-158K Annually
Mid level
Artificial Intelligence • Cloud • Internet of Things • Software • Cybersecurity • Industrial • Industrial Equipment
Designs, builds, and maintains scalable data pipelines, microservices, and cloud-based data systems. Develops real-time and batch processing solutions using Python, Java, and AWS services; creates data integrations and workflows; translates requirements into system designs; implements testing and validation; and monitors distributed pipelines for reliability, performance, and data quality.
Top Skills: Amazon CloudwatchAmazon DynamodbAmazon EventbridgeAmazon KinesisAmazon S3APIsAWSAzure DevopsCi/CdJavaJenkinsJIRAMicroservicesNosql DatabasesOauth 2.0PythonRelational DatabasesSQL
14 Days Ago
Hybrid
128K-209K Annually
Senior level
128K-209K Annually
Senior level
Artificial Intelligence • Cloud • Internet of Things • Software • Cybersecurity • Industrial • Industrial Equipment
Lead the design and development of scalable data pipelines, microservices, cloud-native ingestion systems, and streaming solutions. Collaborate with architects, engineers, product teams, and business stakeholders to build reliable data platforms using Python, Java, AWS services, SQL, and relational and NoSQL databases. Establish data quality, testing, monitoring, automation, and validation frameworks while optimizing performance and resolving production issues across distributed systems.
Top Skills: Amazon CloudwatchAmazon DynamodbAmazon EventbridgeAmazon KinesisAmazon S3APIsAWSAzure DevopsCi/CdJavaJenkinsJIRAMicroservicesNosql DatabasesPythonRelational DatabasesSQL
19 Days Ago
Hybrid
113K-183K Annually
Senior level
113K-183K Annually
Senior level
Artificial Intelligence • Cloud • Internet of Things • Software • Cybersecurity • Industrial
Develop and deploy AI-driven annotation, labeling, and captioning systems for robotics, computer vision, autonomy, and Vision-Language-Action datasets. Train and evaluate machine learning models, build scalable data pipelines, integrate human-in-the-loop workflows, and improve annotation quality. The role requires production software engineering, distributed systems, sensor-data expertise, and experience with computer vision or autonomous systems. Work is onsite in Irving, Texas, with relocation assistance and visa sponsorship available.
Top Skills: Computer VisionDockerKubernetesLabelboxLidarMlflowMlopsNvidia Cosmos CuratorNvidia Cosmos ReasonOnnxPythonPyTorchRadarRoboflowSuperviselyTensorFlowWeights & Biases

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account