What You’ll Be Doing
- Platform Design, Development & Evolution: Architect, build, and continuously evolve the core M&O platform and services, leveraging modern technologies and best practices to provide comprehensive observability and data quality functions.
- Ensuring System Reliability & Performance: Actively maintain and enhance the stability, availability, and performance of critical applications and data infrastructure by integrating Site Reliability Engineering (SRE) principles directly into the platform's design and operation.
- Proactive Issue Detection & Resolution: Develop and integrate intelligent systems within the platform to proactively identify, diagnose, and trigger automated or semi-automated resolution for technical issues, performance bottlenecks, and anomalies across all operational and data systems.
- Advanced Data Quality Platform Implementation: Build and integrate capabilities within the M&O platform for continuously measuring, monitoring, and reporting on all critical data quality dimensions (Timeliness, Consistency, Completeness, Accuracy, Validity, Uniqueness) across diverse supply chain pipelines, including Warehousing, Transformation, SAP, Manufacturing, and other critical data sources.
- Centralized Insights & Dashboarding: Develop and manage a unified dashboarding interface (e.g., Grafana, Power BI, Databricks) within the platform to visualize key performance indicators, system health, operational metrics, financial insights (FinOps), and granular data quality metrics for various stakeholders.
- Automating Infrastructure & Operations: Drive the platform's automation capabilities through Infrastructure as Code (IaC), Continuous Integration/Continuous Delivery (CI/CD) pipelines, and extensive scripting (Python, Shell) for provisioning, deployment, and operational workflows.
- Managing Core Data & Cloud Technologies: Integrate and optimize essential technologies like Fivetran, AKS, Kafka, Azure SQL Server, Databricks, and Flink into the M&O platform, ensuring seamless operation and data flow within the Azure cloud environment.
- Optimizing Cloud Resources & Costs: Embed FinOps practices and reporting into the platform to monitor cloud resource utilization and spending, identifying opportunities for cost reduction and ensuring efficient allocation of infrastructure investments.
- Fostering a Culture of Continuous Improvement: Leverage platform data and SRE practices to continuously analyze operational incidents, performance trends, and data quality issues, driving ongoing enhancements to both the platform and the systems it monitors.
- Enabling Data-Driven Decision Making & Compliance: Ensure the M&O platform provides accurate, timely data and insights, derived from both monitoring and high-quality data across all domains, to support informed business decisions, while also guaranteeing compliance with regulatory requirements for data integrity and auditability.
What We’re Looking For
- At least 4 years of experience in a similar position.
- Fundamental understanding of SRE methodologies for reliability, scalability, and performance.
- Experience with Cloud & Kubernetes (MS Azure, AKS, Kafka, HVR, Databricks).
- Expertise in coding & scripting (Python, Java, Shell Scripting). Strong automation skills using scripting languages for infrastructure management and tool development.
- Expertise in monitoring, alerting, and logging systems to ensure visibility into system health. Prior experience in working on Grafana/ Datadog/Dynatrace/ or someone who has worked in teams which built inhouse platform using open-source tools (Grafana, Prometheus, etc.).
- Experience with CI/CD (Terraform, GitHub Actions). Ability to manage infrastructure as code and implement continuous integration/continuous delivery pipelines for reliable deployments.
- Previous working experience in working on building Enhanced data resiliency by defining and tracking key Data Quality (DQ) metrics, including Timeliness, Consistency, Completeness, and Accuracy.
- Strong problem-solving skills, attention to detail, and a willingness to work collaboratively with development and infrastructure teams.
- English at least B2.
Offer
- Stable employment. On the market since 2008, 1500+ talents currently on board in 7 global sites.
- “Office as an option” model. You can choose to work remotely or in the office.
- Workation. Enjoy working from inspiring locations in line with our workation policy.
- Great Place to Work® certified employer.
- Flexibility regarding working hours and your preferred form of contract.
- Comprehensive online onboarding program with a “Buddy” from day 1.
- Cooperation with top-tier engineers and experts.
- Unlimited access to the Udemy learning platform from day 1.
- Certificate training programs. Lingarians earn 500+ technology certificates yearly.
- Upskilling support. Capability development programs, Competency Centers, knowledge sharing sessions, community webinars, 110+ training opportunities yearly.
- Grow as we grow as a company. 76% of our managers are internal promotions.
- A diverse, inclusive, and values-driven community.
- Autonomy to choose the way you work. We trust your ideas.
- Create our community together. Refer your friends to receive bonuses.
- Activities to support your well-being and health.
- Plenty of opportunities to donate to charities and support the environment.
- Modern office equipment. Purchased for you or available to borrow, depending on your location.
Skills Required
- At least 4 years of experience in a similar position.
- Hands-on experience with Microsoft Azure and core Azure services.
- Ability to design cloud environment architecture.
- Experience with monitoring and observability tools, particularly Grafana and Prometheus.
- Experience creating Grafana dashboards, alerts and collecting Prometheus metrics.
- Practical experience with Docker and managing containerised applications.
- Experience working with Kubernetes or other container orchestration platforms.
- Knowledge of Infrastructure as Code (IaC) principles and hands-on experience with Terraform.
- Familiarity with version control systems, especially GitHub and/or Azure Repos.
- Understanding of CI/CD processes and experience with GitHub Actions and Azure Pipelines.
- Basic knowledge of relational databases and SQL, including experience with MySQL and/or PostgreSQL.
- Strong problem-solving skills, attention to detail, and willingness to work collaboratively with development and infrastructure teams.
- English language proficiency at least B2.
What We Do
You’ve got the data — now what’s next? Many businesses are overwhelmed by data and struggle to turn it into real business impact. At Lingaro, we empower global brands and companies to achieve more with data. From strategy development to scalable solutions, we guide you every step of the way. We transform raw data into actionable insights with deep tech expertise, sharp business sense, and a user-centric mindset. Rooted in Europe and operating worldwide, we deliver with agility, a fresh perspective, and a proven track record. Join 75+ leading CPG companies across 30+ countries that have already transformed their business with Lingaro. Be data ready. Be business ready. Be future ready. Contact us at https://lingarogroup.com/contact_us to get started. Want to know more? Check our recognitions & awards below. 🏆 Positioned as a Leader in ISG Provider Lens™ 2025 – Generative AI Services in Strategy & Consulting and Development & Deployment 🏆 Identified as a Leader in ISG Provider Lens™ 2025 – Specialty Analytics Services in Supply Chain and Retail & CPG 🏆 Strong Performer in the Gartner® Peer Insights™ 2022 and 2024 "Voice of the Customer" reports 🏆 Major Contender in the Everest Group® 2024 Analytics and AI Services Specialists PEAK Matrix® Assessment 🏆 Listee in the Gartner® 2024 Guide to Service Providers for GenAI Initiatives 🏆 Great Place to Work® in Poland, the Philippines, and India in 2024









