Platform Engineer

Posted Yesterday
Be an Early Applicant
Hyderabad, Telangana, IND
In-Office
Mid level
Database • Analytics
The Role
Build and own the platform observability layer: design Grafana dashboards, instrument metrics (including LLM signals), connect Azure Monitor/OpenTelemetry, set up alerting, and coach the team. Contribute to CI/CD, deployment tooling, and cloud infrastructure as needed.
Summary Generated by Built In
Company Description

Blend is a premier AI services provider, committed to creating meaningful impact for its clients through the power of data science, AI, technology, and people. We help organisations solve complex business challenges by combining deep domain understanding with modern data and AI capabilities. Our teams work across strategy, analytics, engineering, and product delivery to create scalable, high-value solutions that improve decision-making, efficiency, and growth

Job Description

You’d be joining the software engineering team behind a production-grade, event-driven platform at the intersection of cloud infrastructure and AI. The system processes complex, multi-step workflows in near real-time with AI capabilities woven throughout — running on Azure, managed as a Python Nx monorepo.
This isn’t a role that values depth in one discipline above all else — it’s a role that values breadth. We’re looking for a mid-level engineer who has worked meaningfully across software engineering, data engineering, and DevOps, and who can use that perspective to understand how a complex distributed system behaves end-to-end. Your primary focus will be building and owning the observability layer for the platform: Grafana dashboards that give the engineering team clear visibility into system health, performance, and behaviour. You’ll work collaboratively with the team to discover what needs to be measured, build the tooling to surface it, and grow into the team’s subject matter expert on observability.

Responsibilities
• Design and build Grafana dashboards that monitor the platform across multiple dimensions: service health, throughput, latency, queue depths, error rates, and alerting for impending or ongoing issues
• Instrument and surface metrics from across the system — including AI-specific signals such as LLM token consumption and inference timing
• Work closely with the engineering team to identify the data points that matter and translate them into useful, actionable views
• Connect observability tooling to Azure Monitor and OpenTelemetry data sources
• Set up and maintain alerting to give the support team early warning of degradation or 
failure
• Act as the team’s SME on observability — helping engineers understand what the dashboards reveal, coaching them to build their own dashboard components, and growing collective capability across the team
• Contribute to the broader platform engineering effort across CI/CD, deployment 
tooling, and cloud infrastructure as needed

Qualifications

  • Meaningful experience across at least two of: software engineering, data 
  • engineering, and DevOps/platform engineering — breadth here is genuinely the 
  • point, not a compromise
  • • Ability to read and understand a complex distributed system and reason about what 
  • to measure and why
  • • Hands-on experience with Grafana, including building dashboards from real data 
  • sources
  • • Familiarity with Azure monitoring tooling — Azure Monitor, Application Insights, or 
  • equivalent
  • • Understanding of observability concepts: metrics, traces, logs, alerting, SLIs/SLOs
  • • Experience working with event-driven or microservices architectures — you need to 
  • understand how the system works to know what to watch
  • • Python skills sufficient to navigate and contribute to the codebase

Skills Required

  • Meaningful experience across at least two of software engineering, data engineering, and DevOps/platform engineering
  • Ability to read and understand a complex distributed system and reason about what to measure
  • Hands-on experience with Grafana, including building dashboards from real data sources
  • Familiarity with Azure monitoring tooling (Azure Monitor, Application Insights) or equivalent
  • Understanding of observability concepts: metrics, traces, logs, alerting, SLIs/SLOs
  • Experience with event-driven or microservices architectures
  • Python skills sufficient to navigate and contribute to the codebase

Blend360 Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Blend360 and has not been reviewed or approved by Blend360.

  • Fair & Transparent Compensation Pay is considered fair-to-good by many, and public salary postings for common data roles indicate competitive packages in numerous markets. Feedback suggests overall company sentiment aligns with acceptable compensation relative to peers in consulting and analytics.
  • Flexible Benefits Flexible and remote/hybrid work arrangements are consistently highlighted in official materials and role descriptions. Feedback suggests flexibility is a meaningful part of the total rewards experience.
  • Retirement Support A 401(k) with company match is part of the core package. Feedback suggests retirement offerings are standard and contribute to a complete benefits set.

Blend360 Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Columbia, MD
390 Employees
Year Founded: 2016

What We Do

Our Vision is to build a company of world-class people that helps our clients optimize business performance through data, technology and analytics. Blend360 has two divisions: Data Science Solutions: We work at the intersection of data, technology and analytics. Talent Solutions: We live and breathe the digital and talent marketplace.

Similar Jobs

MetLife Logo MetLife

Platform Engineer

Fintech • Information Technology • Insurance • Financial Services • Big Data Analytics
Hybrid
Hyderabad, Telangana, IND
43000 Employees

MetLife Logo MetLife

Platform Engineer

Fintech • Information Technology • Insurance • Financial Services • Big Data Analytics
Hybrid
Hyderabad, Telangana, IND
43000 Employees

Accenture Logo Accenture

Platform Engineer

Information Technology
In-Office
5 Locations
456553 Employees

Accenture Logo Accenture

Platform Engineer

Information Technology
In-Office
5 Locations
456553 Employees

Similar Companies Hiring

Northslope Thumbnail
Artificial Intelligence • Information Technology • Software • Analytics • Consulting • Generative AI
London, GB
100 Employees
Scotch Thumbnail
Artificial Intelligence • eCommerce • Fintech • Payments • Retail • Software • Analytics
US
35 Employees
Milestone Systems Thumbnail
Artificial Intelligence • Security • Software • Analytics • Big Data Analytics
Lake Oswego, OR
1500 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account