Top Linux Jobs in Miami, FL

19 Days AgoSaved
Remote
Miami, FL
Senior level
Senior level
Software
Serves as an escalation resource for complex platform support issues, conducting root-cause analysis through database, log, and system diagnostics. Manages technical bug backlogs, monitoring and alerting platforms, access administration, runbooks, and support tooling. Partners with Engineering and customer experience teams on critical cases, escalation criteria, process improvements, automation, and platform reliability. Provides technical mentorship and participates in on-call support for high-priority incidents.
Top Skills: ConfluenceDebugviewElkFiddlerGrafanaHrisKeycloakLinuxMicrosoft Sql ServerProcess MonitorSaaSSalesforceService-Oriented ArchitectureSite24X7SQLWindowsWireshark
19 Days AgoSaved
Remote
Miami, FL
116K-194K Annually
Senior level
116K-194K Annually
Senior level
Fintech
Leads the design, development, configuration, integration, modernization, and production support of Nasdaq Calypso solutions across capital markets functions. Translates business requirements into technical designs, oversees testing and deployments, optimizes performance and resiliency, manages incidents, and ensures compliance with security, regulatory, architecture, and risk standards. Mentors technical teams, coordinates lifecycle initiatives, and partners with business stakeholders, vendors, infrastructure, cybersecurity, and enterprise architecture teams.
Top Skills: Ci/CdEnterprise Application Design PatternsFile-Based IntegrationsGitJavaJSONLinuxMessagingNasdaq CalypsoOraclePostgresRestShell ScriptingSoapSQLUnixXML
19 Days AgoSaved
Remote
Miami, FL
105K-140K Annually
Senior level
105K-140K Annually
Senior level
Fintech
Supports and improves enterprise hybrid infrastructure across Linux, Windows, Azure, and on-premises environments. Responsibilities include VoIP and telecommunications support, application and systems troubleshooting, Microsoft 365 administration, vulnerability remediation, monitoring, disaster recovery, documentation, incident response, and on-call support. The role collaborates with technical teams, stakeholders, vendors, and leadership to deliver reliable infrastructure solutions.
Top Skills: Active DirectoryAnsibleAsteriskAzure DevopsAzure Virtual DesktopDatadogDnsDockerEthernetExchangeGitGroup PolicyLinuxMicrosoft 365AzureNetwrix AuditorOnedriveOsiPodmanPowershellRhel 9+SharepointTcp/IpTeamsWindows Server
19 Days AgoSaved
Remote
Miami, FL
Mid level
Mid level
Security • Software
Design, build, and maintain highly available AWS GovCloud infrastructure supporting FedRAMP environments. Manage ECS scaling, RDS PostgreSQL performance and failover, infrastructure as code, disaster recovery, monitoring, alerting, CI/CD automation, incident response, troubleshooting, and post-mortems. Collaborate with engineering teams to improve system reliability, security, and performance.
Top Skills: Amazon EcsAmazon RdsAmazon S3Amazon SnsAmazon SqsAmazon VpcAws CdkAws CloudformationAws CloudwatchAws Ec2Aws GovcloudAws IamAws LambdaBashCi/CdDatadogDockerFedrampGrafanaKubernetesLinuxPostgresPrometheusPythonTerraform
19 Days AgoSaved
Remote
Miami, FL
185K-220K Annually
Expert/Leader
185K-220K Annually
Expert/Leader
Big Data • Cloud • Hardware • Software • App development
Architect and deliver high-scale NVIDIA GPU compute environments for enterprise AI workloads. Lead bare-metal provisioning, cluster management, operating-system hardening, scheduler and workload configuration, monitoring, performance validation, and day-two operations. Serve as a customer-facing technical authority, supporting presales discovery, architecture, bills of materials, estimation, and client workshops. Design repeatable AI infrastructure across DGX, HGX, MGX, NVL72, and Cisco AI POD platforms.
Top Skills: AnsibleAWSBlackwellCisco Ai PodCudaCudnnGb200Gb300 Nvl72GCPGrace BlackwellGrace HopperHgxHopperHplInfinibandKubeflowKubernetesKueueLinuxMgxAzureMulti-Instance GpuNcclNumaNvidia Base Command ManagerNvidia Data Center Gpu ManagerNvidia Dgx BasepodNvidia Dgx SuperpodNvidia GpusNvidia Mission ControlNvidia OmniverseNvidia Run:AiNvlinkOciPciePythonRafayRed Hat OpenshiftRed Hat Openshift AiRhelRocev2SlurmTerraformUbuntuVolcano
19 Days AgoSaved
Remote
Miami, FL
185K-220K Annually
Expert/Leader
185K-220K Annually
Expert/Leader
Cloud • Information Technology • Productivity • Security • Software
Leads architecture, provisioning, automation, performance validation, and day-two operations for large-scale NVIDIA GPU clusters and AI compute environments. Designs repeatable bare-metal and Kubernetes-based infrastructure using schedulers, monitoring, networking, and infrastructure-as-code. Serves as the technical authority for DGX, HGX, MGX, NVL72, and Cisco AI POD deployments, while supporting pre-sales discovery, bills of materials, solution estimates, client workshops, and executive-level communication.
Top Skills: AnsibleAWSCisco Ai PodCudaCudnnGCPGrace BlackwellGrace HopperHgxHplInfiniband HdrInfiniband NdrKubeflowKubernetesKueueLinuxMgxAzureMulti-Instance GpuNcclNumaNvidia Base Command ManagerNvidia BlackwellNvidia Data Center Gpu ManagerNvidia Dgx BasepodNvidia Dgx SuperpodNvidia Gb200Nvidia Gb300Nvidia HopperNvidia Mission ControlNvidia OmniverseNvidia Run:AiNvidia-Certified SystemsNvl72NvlinkNvmeOracle Cloud InfrastructurePciePythonRafayRed Hat Enterprise LinuxRed Hat OpenshiftRed Hat Openshift AiRocev2SlurmTerraformUbuntuVolcano
20 Days AgoSaved
Remote
Miami, FL
Entry level
Entry level
Other
Design and operate cloud platforms supporting backend telecom services. Automate deployments, scaling, recovery, and infrastructure provisioning; monitor production systems; maintain observability, alerting, and dashboards; support incident response and on-call operations; manage CI/CD pipelines; and enable engineering, telecom, and data teams through reliable tools and infrastructure.
Top Skills: AnsibleAWSAzureBashCassandraCircleCICloudFormationDatadogDnsDockerElasticsearchElk StackGitlab CiGoGCPGrafanaHttp/HttpsIamJaegerJenkinsKafkaKubernetesKvmLinuxNoSQLOpentelemetryPerlPrometheusPythonRubySaltstackSplunkSQLTcp/IpTerraformUnixVMware
20 Days AgoSaved
Remote
Miami, FL
75K-165K Annually
Junior
75K-165K Annually
Junior
Software
Develop data conversion code to migrate clients to TaxSys using Perl, SQL, Linux, MySQL, and a proprietary parallel-computing framework. Debug and resolve client data issues, collaborate with data analysts and clients, validate converted data, and support agile teams through successful go-live. Occasional on-site travel is required.
Top Skills: BashDockerLinuxMakefilesMySQLParallel Computing FrameworkPerlRelational DatabasesSQL
20 Days AgoSaved
Remote
Miami, FL
119K-161K Annually
Senior level
119K-161K Annually
Senior level
Aerospace • Information Technology • Professional Services • Security • Software
Designs and manages AWS cloud infrastructure and Amazon EKS platforms using GitLab CI/CD and Infrastructure as Code. Builds ETL workflows and scalable data pipelines for analytics and AI workloads, supports MLOps/AIOps and FinOps capabilities, troubleshoots Linux systems, and improves platform performance, security, scalability, and availability.
Top Skills: Ai/Ml FrameworksAiopsAmazon EksAWSBashDatadogETLFinopsGitlab Ci/CdGoInfrastructure As Code (Iac)KubernetesLinuxMlopsPythonSQLVector Databases
20 Days AgoSaved
Remote
Miami, FL
Senior level
Senior level
Agency • Cloud • Professional Services • Software
Improve AWS production infrastructure reliability, observability, performance, and operational maturity. Build Terraform infrastructure, enhance CI/CD, automate operational work, manage incident response and on-call operations, lead postmortems, improve application resilience, support capacity planning and database reliability, and collaborate on security hardening and compliance. Mentor engineers and promote reliability practices across the organization.
Top Skills: AWSCi/CdCircleCIDatadogGithub ActionsGitlab CiLinuxNew RelicPostgresRubyRuby On RailsSlisSlosTerraform
20 Days AgoSaved
Remote
Miami, FL
Mid level
Mid level
Software • Automation
Supports and maintains production cloud infrastructure across Verint’s product portfolio. Responsibilities include troubleshooting incidents, provisioning environments, managing Windows and Linux systems across AWS, Azure, and GCP, supporting deployment pipelines, collaborating on client projects, documenting procedures, and contributing to infrastructure upgrades, security initiatives, monitoring, resiliency, and continuous improvement. The role also provides distributed 24/7 operational coverage.
Top Skills: Amazon Ec2AWSConfluenceDatadogDeployment PipelinesGoogle Cloud PlatformItilJIRALinuxAzureServicenowSQL ServerTomcatWeblogicWindows Server
Reposted 25 Days AgoSaved
Remote or Hybrid
Miami, FL
285K-340K Annually
Senior level
285K-340K Annually
Senior level
Artificial Intelligence • Machine Learning • Natural Language Processing • Software • Generative AI
Build and scale ML-optimized HPC infrastructure, manage Kubernetes-based GPU/TPU superclusters, optimize for AI/ML training, and mentor teams while innovating in ML infrastructure.
Top Skills: GoGpuJaxKubernetesLinuxNcclPythonPyTorchRdmaTensorFlowTpu
New

Cut your apply time in half.

Use ourAI Assistantto automatically fill your job applications.

Use For Free
Application Tracker Preview
25 Days AgoSaved
Remote
Miami, FL
110K-135K Annually
Mid level
110K-135K Annually
Mid level
Big Data • Cloud • Hardware • Software • App development
Evaluate and develop compute technologies for an advanced technology lab. Design and build compute environments for sandboxes, proofs of concept, and lab-as-a-service engagements. Create architectural diagrams, access guides, test plans, use cases, and technical designs. Collaborate with customers, sales, partners, and internal teams; support business development and marketing initiatives; maintain OEM partnerships; and achieve certifications required for partner status.
Top Skills: CloudDell ComputeHpe ComputeLinuxNetworkingS3StorageVirtualizationVMwareWindows Server
25 Days AgoSaved
Remote
Miami, FL
110K-135K Annually
Mid level
110K-135K Annually
Mid level
Cloud • Information Technology • Productivity • Security • Software
Evaluate and develop compute technologies for the Advanced Technology Center. Design and build compute environments for sandboxes, proofs of concept, and lab-as-a-service engagements. Create architectural diagrams, access guides, test plans, use cases, and high- and low-level designs. Collaborate with customers, partners, sales, and internal teams; support business development and marketing; foster OEM partnerships; and maintain required industry certifications.
Top Skills: Amazon S3CloudDell ComputeHpe ComputeLinuxNetworkingStorageVirtualizationVMwareWindows Server
25 Days AgoSaved
Remote
Miami, FL
Senior level
Senior level
Cloud • Information Technology • Cybersecurity • Infrastructure as a Service (IaaS)
Owns the hardware lifecycle for GPU and infrastructure assets, including fleet health monitoring, vendor RMA workflows, firmware and BIOS upgrades, burn-in testing, failure investigation, inventory accuracy, capacity planning, spare-parts strategy, runbook creation, and new platform qualification. The role serves as the technical owner of physical compute platforms and supports high-density GPU infrastructure across sites.
Top Skills: Base Command ManagerBashBiosBmcBright Cluster ManagerCmdbGpu SystemsIpmiLinuxLiquid CoolingNcclNvidia DgxNvidia HgxNvidia Mission ControlPythonRedfishX86 Server Architecture
25 Days AgoSaved
Remote
Miami, FL
200K-275K Annually
Senior level
200K-275K Annually
Senior level
Cloud • Information Technology • Cybersecurity • Infrastructure as a Service (IaaS)
Architect, deploy, optimize, and operate large-scale GPU clusters for AI training and inference. Responsibilities include distributed PyTorch and NCCL tuning, GPU networking and storage optimization, scheduling, benchmarking, troubleshooting, automation, monitoring, and performance engineering across compute, networking, storage, and software layers. The role supports production environments with hundreds to thousands of GPUs and partners with ML engineers to improve training scalability and inference efficiency.
Top Skills: AnsibleAWSAzureBashCudaDcgm ExporterEnrootGCPGpudirect RdmaGpudirect StorageGrafanaInfinibandKubernetesKueueLinuxMlperfMpiNcclNccl-TestsNsight SystemsNumaNvidia DcgmNvidia Gpu OperatorNvlinkNvmeNvswitchPciePrometheusPythonPyTorchPyxisRdmaRoce V2Run:AiSglangSlurmTensorrt-LlmTerraformTritonUcxVllmVolcano
25 Days AgoSaved
Remote
Miami, FL
Mid level
Mid level
Cloud • Information Technology • Cybersecurity • Infrastructure as a Service (IaaS)
Monitor GPUaaS infrastructure, triage alerts, execute runbooks, coordinate incident response, communicate with customers, manage maintenance windows, resolve Tier 1 tickets, and improve monitoring and operational procedures. The role provides 24/7 rotating coverage, including nights, weekends, and holidays, while maintaining shift handoffs and active-incident documentation.
Top Skills: BashDatadogGrafanaLinuxPagerdutyPrometheusPython
25 Days AgoSaved
Remote
Miami, FL
Entry level
Entry level
Artificial Intelligence • Information Technology
Operate and improve large-scale NVIDIA GPU clusters across bare-metal and rented capacity. Build automation for provisioning, health checks, remediation, monitoring, scheduling, and capacity management. Own Linux images, drivers, CUDA, containers, networking, storage, and Slurm. Diagnose performance and hardware issues, evaluate GPU providers, support ML training workloads, and maintain secure infrastructure. The role also includes hardware installation, datacenter coordination, and vendor management.
Top Skills: AnsibleBashBmcContainersCudaDcgmInfinibandIpmiLinuxNcclNvidia GpusPromqlPxePythonRedfishRoceSlurmTerraformVastWeka
25 Days AgoSaved
Remote
Miami, FL
128K-267K Annually
Senior level
128K-267K Annually
Senior level
AdTech • Digital Media • Information Technology • Other
Own and advance Yahoo’s enterprise Node.js development environment, including runtimes, packages, container images, internal integrations, and CI/CD workflows. Manage runtime lifecycles, security updates, dependencies, releases, testing, automation, and troubleshooting. Support GitHub Enterprise and GitLab Enterprise while integrating AI programming tools to improve developer productivity. Partner across engineering and security teams, and document platform changes, migrations, and deprecations.
Top Skills: Ci/CdClaude AiClaude CodeContainersCursorGitGithub ActionsGithub CopilotGithub EnterpriseGitlab EnterpriseGoJavaScriptLinuxNode.jsPythonRpm PackagingTypescript
26 Days AgoSaved
Remote
Miami, FL
Senior level
Senior level
Blockchain • Payments • Financial Services
Build and manage Tempo’s infrastructure stack, improve developer velocity and experience, maintain blockchain, validator, and explorer reliability, and support enterprise validator onboarding. The role involves production bare-metal and cloud environments, Kubernetes, infrastructure as code, observability, Linux, networking, security hardening, troubleshooting, and tooling development using Rust, Go, or Python.
Top Skills: ArgocdBare Metal ServersBlockchain ValidatorsCloud InfrastructureConfiguration ManagementGoGrafanaHelmInfrastructure As CodeKubernetesLinuxNetworkingPrometheusPythonRustTerraform
26 Days AgoSaved
Remote
Miami, FL
Senior level
Senior level
Healthtech • Software
Manage and optimize a multi-account AWS environment supporting production healthcare applications and analytics platforms. Responsibilities include AWS infrastructure administration, CI/CD and infrastructure automation, observability, incident response, on-call support, disaster recovery validation, HIPAA/HiTrust compliance, IAM and security controls, and support for containerized Java and Python applications. The role also contributes to Kubernetes and EKS modernization initiatives and partners with developers to resolve complex production issues.
Top Skills: Amazon EksApi GatewayArgocdAuroraAWSAws CloudformationAws CodepipelineAws Security HubBashCloudfrontCloudwatchCortex CloudDatadogDockerEc2EcsFargateGitGuarddutyHelmIamJavaJenkinsKubernetesLambdaLinuxPrismaPythonRdsS3Spring BootUbuntuZabbix
Reposted One Month AgoSaved
In-Office
Miami, FL
150K-250K Annually
Mid level
150K-250K Annually
Mid level
Fintech • Financial Services
The C++ Developer will implement, test and deploy features for a high-frequency trading system, optimize performance, and debug issues.
Top Skills: BoostC++GitLinuxStl
26 Days AgoSaved
Remote
Miami, FL
125K-140K Annually
Senior level
125K-140K Annually
Senior level
Information Technology • Professional Services • Consulting
Owns the technical delivery and ongoing support of key customer accounts across AV, unified communications, video conferencing, and cloud collaboration platforms. Leads deployments, upgrades, troubleshooting, escalations, root-cause analysis, documentation, operational readiness, and service transitions. Partners with Sales, Customer Success, Program Management, Service Delivery, vendors, and customer engineering teams. Serves as a senior technical advisor, subject matter expert, and cross-functional leader while driving process improvements and platform administration.
Top Skills: Active DirectoryAudio VisualAzureBiamp TesiraCisco Collaboration SolutionsCloud Collaboration PlatformsCrestronExtronIp NetworkingItilLinuxMicrosoft Teams RoomsUnified CommunicationsUtelogyVideo ConferencingVMwareVnocWebex Control Hub
Reposted One Month AgoSaved
Remote
Miami, FL
150K-200K Annually
Senior level
150K-200K Annually
Senior level
Artificial Intelligence • Cloud • Software • Infrastructure as a Service (IaaS)
Ensure stability and resilience of Runpod's distributed AI platform by defining SLIs/SLOs, leading incident response, building observability and reliability tooling, automating operational workflows, and partnering with engineering teams to reduce toil and improve production readiness.
Top Skills: BashCi/CdContainerized Production SystemsGoGpu Observability ToolingGrafanaInfrastructure As CodeLinuxPrometheusPython
26 Days AgoSaved
Remote
Miami, FL
Senior level
Senior level
Edtech
Lead infrastructure modernization and platform reliability across multiple cloud providers. Design infrastructure as code, operate Kubernetes and Linux environments, improve CI/CD and deployment tooling, establish SLI/SLO practices, strengthen observability, lead incident response, manage cloud costs, and partner on security and compliance. Provide technical leadership through architecture guidance, mentorship, engineering standards, and roadmap development while participating in on-call support.
Top Skills: AWSCi/CdGCPJenkinsKubernetesLinuxPythonRubyRuby On RailsSoc 2SpinnakerTerraform
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account