We are hiring a Senior Cloud Infrastructure Engineer to build the core infrastructure that powers the next generation of AI agents. Our agents must ingest, enrich, and vectorize massive, multi-modal datasets from over 100 customer sources. The core challenge is twofold: how do we do this in a way that is radically cost-efficient, while still allowing agents to deliver thorough responses in seconds?
This position reports to a Senior Manager, Software Engineering and will be 2 days per week hybrid role in our Austin office. We’re looking for someone to join us immediately.
What You’ll DoBuild & Optimize for Scale & Cost: Implement and optimize our highly scalable, multi-tenant data ingestion and vectorization pipeline, with a relentless focus on improving cost and performance.
Implement Robust Monitoring: Create and maintain dashboards, alerts, and logging to ensure system health, identify performance bottlenecks, and provide immediate visibility into production issues.
Contribute to Millisecond Latency: Be a key contributor in performance tuning across the stack. You will help identify and eliminate bottlenecks in data retrieval, model inference, and agent response times to ensure a snappy, real-time user experience.
Engineer Multi-Tenant Architectures: Design and implement scalable, secure, and cost-effective multi-tenant infrastructures for our SaaS-based AI products, ensuring strict tenant isolation and fair resource allocation.
5+ years of hands-on experience with cloud platforms
Hands-on experience with cloud platforms, with strong expertise in Google Cloud Platform (GCP) and Infrastructure as Code (e.g., Terraform)
A demonstrated history of performance analysis, tuning, and infrastructure cost optimization, with the ability to speak about trade-offs and quantified impact
Experience building or working on multi-tenant SaaS platforms
Experience setting up end-to-end observability, including logging, metrics, and alerting using tools like Prometheus, Grafana, Datadog, or GCP Operations suite
A fundamental understanding of the challenges in training and serving large machine learning models (e.g., memory constraints, computational complexity)
Strong understanding of VectorDBs, LLMs, and Agentic Observability tools (Datagrid uses Milvus for our VectorDB and Arize for agent tracing)
Experience with the Gemini API, specifically managing LLM quotas and load balancing
Hands-on experience with LLM serving frameworks and optimization techniques (quantization, tensor parallelism, FlashAttention)
Experience designing multi-tenant SaaS architectures and implementing resource quotas and cost allocation
High-level knowledge of agentic systems and best practices
Base Pay Range:
140,960.00 - 193,820.00 USD AnnualThis role may also be eligible for Equity Compensation and/or Bonus Incentive Compensation. Procore is committed to offering competitive, fair, and commensurate compensation. Actual compensation will be based on a candidate’s job-related skills, experience, education or training, and location.
For Los Angeles County (unincorporated) Candidates:Procore will consider for employment all qualified applicants, including those with arrest or conviction records, in accordance with the requirements of applicable federal, state, and local laws, including the City of Los Angeles’ Fair Chance Initiative for Hiring Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act.
A criminal history may have a direct, adverse, and negative relationship on the following job duties, potentially resulting in the withdrawal of the conditional offer of employment: 1. appropriately managing, accessing, and handling confidential information including proprietary and trade secret information, as well as accessing Procore's information technology systems and platforms; 2. interacting with and occasionally having unsupervised contact with internal/external customers, stakeholders, and/or colleagues; and 3. exercising sound judgment.
Skills Required
- 5+ years of hands-on experience with cloud platforms
- Hands-on experience with Google Cloud Platform (GCP) and Infrastructure as Code (Terraform)
- Demonstrated history of performance analysis, tuning, and infrastructure cost optimization with quantified impact
- Experience building or working on multi-tenant SaaS platforms
- Experience setting up end-to-end observability (logging, metrics, alerting) using Prometheus, Grafana, Datadog, or GCP Operations suite
- Fundamental understanding of challenges in training and serving large ML models (memory, compute)
- Understanding of VectorDBs, LLMs, and agentic observability tools (examples: Milvus, Arize)
- Experience with the Gemini API, including managing LLM quotas and load balancing
- Hands-on experience with LLM serving frameworks and optimization techniques (quantization, tensor parallelism, FlashAttention)
- Experience designing multi-tenant SaaS architectures with resource quotas and cost allocation
- High-level knowledge of agentic systems and best practices
Procore Technologies Compensation & Benefits Highlights
The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Procore Technologies and has not been reviewed or approved by Procore Technologies.
-
Parental & Family Support — Family-building benefits such as fertility assistance on eligible plans, cash support for adoption and surrogacy, and substantial paid parental leave with a supported return-to-work indicate strong support for parents. Feedback suggests these offerings are a standout component of the total rewards package.
-
Leave & Time Off Breadth — Open PTO with no accruals, a company-wide Wellness Week, and separate sick time reflect broad time-off flexibility. Feedback suggests employees value the ability to take time away in addition to standard holidays.
-
Wellbeing & Lifestyle Benefits — A quarterly Procore Perks stipend, mental-health resources through an EAP/Modern Health, and free meals/snacks with WFH reimbursements demonstrate ongoing investment in wellbeing and daily convenience. Feedback suggests these benefits add meaningful everyday value beyond base pay.
Procore Technologies Insights
What We Do
At Procore Technologies, we’re collectively building towards what’s next for our employees, industry, customers, and global communities. Our cloud-based construction management software streamlines the entire lifecycle of a construction project, connecting field and office teams, centralizing data to mitigate risks, providing real-time financials, and more to help clients efficiently build everything from skyscrapers to hospitals to airports. Procore was founded in 2002, and we’ve since grown into a global company of groundbreakers working throughout North America, EMEA, and APAC. Coming together from across diverse backgrounds to be our best, we embrace a culture of ownership and excellence that gives our teams the tools to grow and thrive as they shape their careers – and the Procore of tomorrow. To learn more about Procore and how you can build what comes next for your career, visit us at https://careers.procore.com/.
Why Work With Us
We make each other better at Procore. Here, your career is not pre-defined and it can take many paths. While you own your career, we provide you with the support and opportunities to help you succeed. You can help us transform an industry while you are transforming your career.
Gallery
.jpg)








