Data Engineer

Posted Yesterday
Ankeny, IA, USA
In-Office
Mid level
Design
The Role
Designs and maintains Microsoft Fabric data pipelines, medallion Lakehouse architecture, data models, governance, orchestration, and quality processes. Integrates ERP, HRIS, MRP, finance, project, employee, service, and manufacturing data into reliable data products. Supports reporting, business intelligence, AI and machine learning use cases, including RAG and vector embeddings. Troubleshoots pipeline issues, manages Fabric capacity and documentation, and collaborates with analysts, data scientists, developers, DevOps teams, and business system owners.
Summary Generated by Built In

PURPOSE

The Data Engineer is responsible for designing, building, and maintaining the data pipelines and infrastructure that power Baker Group's Microsoft Fabric data warehouse, serving as the organization's single source of truth. This role owns the ingestion, transformation, and orchestration of data from disparate internal systems (ERP, HRIS, MRP and other structured data sources) into governed, reliable data products used by Data Analysts and developers to deliver insights to executive and operational teams and ensures that data is structured to support both traditional reporting and emerging AI and machine learning use cases. The Data Engineer curates and maintains core datasets spanning employees, finance, construction and manufacturing projects, and service, and partners with the Data Scientist, Data Analyst, and Software Development roles to ensure data is trustworthy, well-structured, and fit for downstream use.

 

ESSENTIAL FUNCTIONS AND RESPONSIBILITIES

The following duties are typical for this job. These are not to be constructed as exclusive or all inclusive. Other duties may be required and assigned.

  • Designs, builds, and maintains ETL/ELT pipelines that ingest data from enterprise systems into Microsoft Fabric.
  • Architects and maintains the Fabric medallion Lakehouse structure (bronze, silver, gold layers) as Baker Group's single source of truth.
  • Develop and implement best practices for the data infrastructure and environment (e.g. Development/Test/Production environments, Git for version control).
  • Owns pipeline orchestration, scheduling, and monitoring to ensure reliable, timely, and accurate data availability.
  • Curates and maintains core datasets across employee, finance, project, service, and manufacturing domains.
  • Establishes and enforces data quality, validation, and reconciliation processes across all pipelines.
  • Designs and manages data models, schemas, and semantic layers that support Data Analyst reporting and Data Scientist modeling work.
  • Defines and maintains data ontologies and canonical business definitions (for example, what constitutes a "project," "employee," or "cost code") to ensure consistent meaning across systems and consumers.
  • Prepares and structures data to support AI and machine learning use cases, including feature-ready datasets, retrieval-augmented generation (RAG) pipelines, and vector embedding storage.
  • Manages Fabric capacity planning, workspace organization, and performance optimization.
  • Implements data governance practices, including access controls, lineage tracking, and metadata management, consistent with Baker Group's data classification standards.
  • Partners with business system owners (ERP, HRIS, MRP, etc.) to understand upstream data structures and manage change impacts.
  • Collaborates with the Data Scientist to ensure pipeline outputs support analytical and machine learning use cases.
  • Collaborates with Data Analysts to ensure data products support paginated reporting, dashboards, and self-service BI needs.
  • Collaborates with Software Development and DevOps Teams to ensure data products support application development needs.
  • Coordinates with 3rd party consultants when necessary to deliver data engineering projects and augment capacity for demanding business needs.
  • Develops and maintains documentation for pipelines, schemas, and integration logic.
  • Troubleshoots and resolves pipeline failures, latency issues, and data quality incidents.
  • Monitors and maintains data-specific infrastructure, including Fabric capacity, pipeline orchestration tools, and monitoring/alerting systems.
  • Evaluates and recommends new data engineering tools, patterns, and best practices.
  • Stays current on emerging trends in data engineering, cloud data platforms, and integration techniques. 

 

MINIMUM EDUCATION and EXPERIENCE REQUIRED TO PERFORM ESSENTIAL FUNCTIONS

  • Bachelor's degree in Computer Science, Data Engineering, Information Systems, or other relevant quantitative field
  • Three to five years of experience in data engineering, ETL/ELT development, or a related field
  • Proficiency with SQL and database technologies for data extraction, transformation, and loading
  • Experience with Microsoft Fabric, Azure Data Factory, or similar cloud ETL/orchestration tools
  • Experience with medallion architecture and modern data warehousing patterns
  • Experience with a programming language such as Python, PySpark, or T-SQL for data transformation
  • Familiarity with data modeling techniques (dimensional modeling, star schema)
  • Understanding of data governance, data quality, and metadata management practices
  • Experience preparing data for AI/ML consumption (e.g., vector embeddings, RAG architectures) is a plus
  • Business acumen and understanding of construction or related industries is a plus


CERTIFICATES, LICENSES, REGISTRATIONS

  • No specific requirements; however, relevant certifications such as Microsoft Certified: Fabric Data Engineer Associate, Azure Data Engineer Associate, or similar cloud platform certifications are a plus

 

MENTAL AND PHYSICAL COMPETENCIES REQUIRED TO PERFORM ESSENTIAL FUNCTIONS

  • Strong analytical and troubleshooting skills with the ability to diagnose and resolve complex pipeline and data quality issues 
  • Excellent time and project management skills with the ability to prioritize across multiple pipeline and infrastructure projects 
  • Current with industry trends in data engineering, cloud platforms, and integration best practices 
  • Strong communication skills with the ability to translate technical data structures for non-technical stakeholders 
  • Team player with strong collaboration skills, particularly with the Data Scientist, Data Analysts, and business system owners 
  • Must be able to focus on complex technical problems and work independently with minimal supervision 
  • Ability to work in a fast-paced environment and adapt to changing business priorities 
  • Meticulous attention to detail and commitment to producing reliable, well-documented data infrastructure


ENVIRONMENTAL ADAPTABILITY

  • Prolonged periods of sitting at a desk and working on a computer
  • Must be able to lift 10 pounds occasionally
  • May have occasional visits to a job site which would require periods of standing, walking and/or climbing stairs


EQUIPMENT/TOOLS

  • Use a computer for 8 hours a day


Baker Group is an Equal Opportunity Employer. In compliance with the Americans with Disabilities Act, Baker Group will consider reasonable accommodations for qualified individuals with disabilities and encourage prospective employees and incumbents to discuss potential accommodations with the Employer.

Skills Required

  • Bachelor's degree in Computer Science, Data Engineering, Information Systems, or another relevant quantitative field
  • Three to five years of experience in data engineering, ETL/ELT development, or a related field
  • Proficiency with SQL and database technologies for data extraction, transformation, and loading
  • Experience with Microsoft Fabric, Azure Data Factory, or similar cloud ETL/orchestration tools
  • Experience with medallion architecture and modern data warehousing patterns
  • Experience with Python, PySpark, T-SQL, or another programming language for data transformation
  • Familiarity with dimensional modeling and star schema techniques
  • Understanding of data governance, data quality, and metadata management practices
  • Experience preparing data for AI and machine learning consumption, including vector embeddings or RAG architectures
  • Business acumen and understanding of construction or related industries
  • Relevant Microsoft Fabric, Azure Data Engineer, or similar cloud platform certification
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Ankeny, Iowa
847 Employees
Year Founded: 1964

What We Do

Established in 1963, Baker Group is Iowa’s premier full-service specialty contractor—serving commercial, industrial, institutional, and mission-critical markets across the country. We offer integrated solutions in mechanical, electrical, sheet metal, plumbing, piping, building automation, fire alarm, access control, CCTV, and 24/7 service and maintenance. Our offsite manufacturing capabilities and modular construction approach allow us to deliver faster, safer, and more cost-effective projects with consistent quality. What sets us apart is our collaborative, data-driven process. By aligning design, field, and BIM teams from day one, we reduce risk, increase efficiency, and deliver long-term value. We don’t just build systems—we build trusted partnerships. As an employee-owned company, we’re deeply invested in exceptional outcomes for our clients and communities. At Baker Group, you can always Expect the Best®.

Similar Jobs

Octus Logo Octus

Data Engineer

Fintech • News + Entertainment • Software • Database • Financial Services
Easy Apply
Remote or Hybrid
United States
808 Employees

PwC Logo PwC

Data Engineer

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Hybrid
42 Locations
370000 Employees
77K-202K Annually

Samsara Logo Samsara

Data Engineer

Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Easy Apply
Remote or Hybrid
United States
4000 Employees
118K-179K Annually

SoFi Logo SoFi

Senior Data Engineer

Fintech • Mobile • Software • Financial Services
Easy Apply
Remote or Hybrid
United States
4500 Employees
125K-234K Annually

Similar Companies Hiring

Million Dollar Baby Co. Thumbnail
eCommerce • Kids + Family • Retail • Sales • Design • Manufacturing
Pico Rivera, CA
200 Employees
Tapestry - Coach and Kate Spade Thumbnail
eCommerce • Fashion • Retail • Sales • Wearables • Design
New York, NY
16000 Employees
Munchkin, Inc. Thumbnail
Consumer Web • eCommerce • Food • Kids + Family • Design • Manufacturing
Milton, Ontario
325 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account