L1 - Linux System Administrator

Posted 3 Days Ago
Be an Early Applicant
Mumbai, Maharashtra, IND
In-Office
Junior
Artificial Intelligence • Cloud • Infrastructure as a Service (IaaS)
The Role
Monitor AI platforms and infrastructure, respond to incidents, troubleshoot Linux, application, and network issues, and escalate problems according to ITIL and SLA processes. Maintain proactive monitoring using Nagios, Prometheus, or Grafana; document incidents and resolutions; support patching and system maintenance; participate in root cause analysis; and collaborate with administrators, engineers, and developers to improve service reliability.
Summary Generated by Built In

Job Title: L1 Linux System Administrator – Monitoring Desk
Location: Mumbai
Type: Onsite – Work from office

About Neysa:
Neysa is an AI Acceleration Cloud System provider, dedicated to democratizing AI adoption with purpose-built platforms and services for AI-native applications and workloads. Co-founded by industry leaders, we empower businesses to discover, deploy, and scale Generative AI (Gen AI) and AI use cases securely and cost-effectively. Our flagship platforms—Neysa Velocis, Neysa Overwatch, and Neysa Aegis—accelerate AI deployment, optimize network performance, and safeguard AI/ML landscapes. We are committed to enabling AI-led innovation across industries and geographies.

Position Overview
We are looking for a Monitoring Desk Associate to join our Service Assurance Team. This position will play a key role in ensuring the optimal performance of Neysa’s AI platforms by monitoring system health, responding to incidents, and performing troubleshooting and resolution in real time. The ideal candidate will have hands-on experience with Linux systems, a passion for operational excellence, and the ability to quickly resolve issues impacting service availability.

Key Responsibilities

  • Incident Monitoring & Response: Monitor Neysa’s AI platforms and infrastructure for any system alerts, performance issues, or service disruptions. Respond to incidents promptly and escalate issues as needed to ensure timely resolution.

  • Incident Management: Follow defined processes for incident identification, classification, and escalation. Ensure incidents are managed effectively, with minimal disruption to service and in alignment with service level agreements (SLAs).

  • Troubleshooting & Resolution: Use your Linux expertise to investigate, diagnose, and resolve incidents affecting system performance. Troubleshoot system-level issues, application failures, and network-related problems.

  • Proactive Monitoring: Continuously monitor the operational status of servers, applications, and networks, proactively identifying potential issues before they impact customers. Utilize monitoring tools such as Nagios, Prometheus, or Grafana to track system health.

  • Documentation & Reporting: Accurately document incidents, actions taken, and resolutions in incident management systems. Provide detailed reports on recurring issues, root causes, and preventive measures for the Service Assurance team.

  • Collaboration with Technical Teams: Work closely with system administrators, engineers, and developers to identify areas of improvement, share insights, and ensure issues are resolved with minimal business impact.

  • Root Cause Analysis: Participate in post-incident reviews to analyze root causes, provide feedback, and suggest improvements to incident management processes.

  • System Maintenance: Support periodic system checks, patch management, and routine maintenance to ensure systems are secure, optimized, and operating at peak efficiency.

Qualifications

  • Experience: 1-5 years of experience in an operations or service assurance role, with a focus on incident management and system monitoring in a Linux environment.

  • Linux Skills: Solid experience with Linux operating systems (e.g., CentOS, Ubuntu, RHEL), including system administration, basic troubleshooting, and performance tuning.

  • Incident Management: Knowledge of ITIL processes, specifically incident management, with the ability to handle incidents efficiently while maintaining communication with stakeholders.

  • Troubleshooting Skills: Strong ability to troubleshoot technical issues in a timely manner, including server failures, network connectivity issues, and application problems.

  • Monitoring Tools: Experience with monitoring and alerting tools such as Nagios, Prometheus, Grafana, or similar, and a strong understanding of how to use these tools to monitor system health and performance.

  • Communication: Excellent communication skills, both verbal and written, with the ability to provide clear and concise updates to both technical and non-technical stakeholders.

  • Team Player: Ability to work effectively in a collaborative team environment, with a proactive approach to problem-solving and incident resolution.

  • Technical Aptitude: Basic understanding of cloud platforms (AWS, Azure, or Google Cloud) and networking fundamentals is a plus.

Preferred Qualifications

  • Experience with containerized environments (e.g., Docker, Kubernetes) is a plus.

  • Familiarity with automated scripting for incident resolution and process improvement (e.g., Bash, Python).

  • ITIL certification or similar incident management qualifications.

Why This Role is a Unique Opportunity at Neysa

  • Work closely with leadership on key business decisions

  • High visibility and impact across teams

  • Opportunity to build legal processes from the ground up

  • Fast learning environment in a scaling startup

Team Culture and Inclusion

  • Open and collaborative work environment

  • Strong focus on ownership and accountability

  • A culture where ideas and initiative are valued

Love what this role has to offer? Discover the world of Neysa:
Website: https://neysa.ai/
Socials
LinkedIn | YouTube | Reddit | Instagram

Skills Required

  • 1-5 years of experience in operations or service assurance, focused on incident management and Linux system monitoring
  • Hands-on experience administering and troubleshooting Linux systems, including CentOS, Ubuntu, or RHEL
  • Knowledge of ITIL incident management processes
  • Ability to troubleshoot server failures, network connectivity issues, application problems, and system performance issues
  • Experience with monitoring and alerting tools such as Nagios, Prometheus, Grafana, or similar
  • Strong verbal and written communication skills
  • Ability to collaborate effectively and resolve incidents proactively
  • Basic understanding of AWS, Azure, or Google Cloud and networking fundamentals
  • Experience with Docker or Kubernetes
  • Familiarity with Bash or Python scripting for incident resolution and process improvement
  • ITIL certification or similar incident management qualification
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
141 Employees
Year Founded: 2023

What We Do

Neysa is an India-based AI acceleration cloud provider focused on democratizing enterprise AI adoption. Co-founded by Sharad Sanghi and Anindya Das, it combines cloud infrastructure, AI systems, and cybersecurity expertise. Its flagship Neysa Velocis platform supports AI training, fine-tuning, inference, compute orchestration, security, and observability, helping organizations across high-growth markets such as India and beyond deploy and scale AI more quickly, safely, and cost-effectively.

Similar Jobs

Neysa Logo Neysa

L1 - Linux System Administrator

Artificial Intelligence • Cloud • Infrastructure as a Service (IaaS)
In-Office
Mumbai, Maharashtra, IND
141 Employees

Coursera + Udemy  Logo Coursera + Udemy

Content Marketing Manager

Artificial Intelligence • Consumer Web • Edtech • Enterprise Web • HR Tech • Social Impact • Generative AI
Remote or Hybrid
India
1500 Employees
106K-143K Annually

Mastercard Logo Mastercard

Software Engineer

Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Hybrid
Pune, Maharashtra, IND
38800 Employees

Mastercard Logo Mastercard

Senior Software Engineer

Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Hybrid
Pune, Maharashtra, IND
38800 Employees

Similar Companies Hiring

Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software • Productivity
US
15 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account