Top Tech Jobs & Startup Jobs

Reposted 16 Days AgoSaved
In-Office or Remote
Barcelona, Cataluña, ESP
Junior
Junior
Software
Operate, troubleshoot, and maintain InfiniBand and Ethernet networks for HPC/AI infrastructures. Support Fortinet firewall/VPN management, monitor network performance, collaborate with compute and storage teams, participate in incident response, document configurations, and learn automation and scripting (Ansible, Bash/Python).
Top Skills: AnsibleBashEthernetFortigateFortinetHcaInfinibandKubernetesLinuxMpiNvidia GpusPythonRdmaRoutingSubnet ManagerSwitchingTcp/IpVlansVpn
Reposted 16 Days AgoSaved
Remote
Sofia-grad, BGR
Mid level
Mid level
Software
Operate, troubleshoot, and maintain InfiniBand and Ethernet networks for HPC/AI clusters. Support Fortinet firewall/VPNs, monitor performance, assist with routing and segmentation, collaborate with compute/storage teams, participate in incident response, document configurations, and learn automation (Ansible/scripting).
Top Skills: AnsibleBashEthernetFirewallFortigateFortinetInfinibandKubernetesLinuxMpiNvidia GpuPythonRdmaRoutingSwitchingTcp/IpVlansVpn
Reposted 16 Days AgoSaved
In-Office or Remote
Praha, Hlavní město Praha, CZE
Mid level
Mid level
Software
Operate and support large-scale AI infrastructure: monitor Kubernetes clusters, NVIDIA GPU platforms, and high-performance networking; investigate incidents, troubleshoot hardware/network/platform issues, run incident response and RCA, improve observability and automation, and maintain runbooks and operational documentation.
Top Skills: ElkGrafanaHigh-Performance NetworkingInfinibandInfrastructure-As-CodeK0Rdent AiKubernetesLinuxNvidia GpusNvidia UfmOpentelemetryPrometheus
Reposted 16 Days AgoSaved
Remote
Sofia-grad, BGR
Mid level
Mid level
Software
Operate, monitor, and support large-scale AI infrastructure (NVIDIA GPUs, HPC networking, Kubernetes). Troubleshoot incidents, collaborate with vendors and datacenter teams, improve observability, automation, runbooks, and participate in incident response and root cause analysis.
Top Skills: ElkGrafanaInfinibandInfrastructure-As-CodeK0Rdent AiKubernetesLinuxNvidia GpusNvidia UfmOpentelemetryPrometheus
17 Days AgoSaved
Remote
Sofia-grad, BGR
Senior level
Senior level
Software
Design, deploy, and operate cloud-native AI infrastructure (Kubernetes/OpenStack) on NVIDIA-certified hardware. Improve reliability, performance, and security; troubleshoot networking, storage, and Linux issues; implement CI/CD and automation; conduct code reviews; mentor customers and teammates; collaborate with international teams; and support architecture and integration for production AI workloads.
Top Skills: Ci/CdDcgm)GoGpu SchedulingInfinibandJavaScriptKubernetesLinuxNvidia Gpu (MigNvlinkOpenshiftOpenstackPythonRancherRdmaRoceVgpuVMware
New

Track Smarter, Apply Better.

Ditch the spreadsheets. Organize your job search with our freeApplication Tracker.

Use For Free
Application Tracker Preview
17 Days AgoSaved
In-Office or Remote
Brussels, BEL
Senior level
Senior level
Software
Lead pre-sales architecture and PoCs for enterprise AI infrastructure across regulated French-speaking European accounts. Own technical relationships, design Kubernetes/OpenStack-based solutions, deploy and debug PoCs, and guide platform and security stakeholders from discovery through expansion.
Top Skills: CloudContainersGpu OrchestrationHybridInferencingKubernetesModel ServingMulti-CloudOpenstack
17 Days AgoSaved
In-Office or Remote
Rīga, LVA
Senior level
Senior level
Software
Design, deploy, and operate cloud and AI infrastructure (Kubernetes/OpenStack) on NVIDIA-certified hardware. Improve reliability, performance, and security of container platforms; troubleshoot networking, storage, and Linux issues; implement CI/CD and automation; mentor teammates and customers; and collaborate across distributed international teams.
Top Skills: Ci/CdDcgmGoGpu SchedulingInfinibandJavaScriptKubernetesLinuxMigNvidia Ai EnterpriseNvlinkOpenshiftOpenstackPythonRancherRdmaRoceVgpuVMware
17 Days AgoSaved
In-Office or Remote
Barcelona, Cataluña, ESP
Senior level
Senior level
Software
Design, deploy, and operate Kubernetes-based AI infrastructure on NVIDIA-certified hardware. Improve reliability, performance, and security of container platforms, mentor teams and customers, troubleshoot complex networking/storage issues, implement CI/CD and automation, and collaborate with distributed stakeholders.
Top Skills: Ci/CdCncfDcgmGoGpu SchedulingInfinibandJavaScriptKubernetesLinuxMigNvidia Ai EnterpriseNvidia-Certified HardwareNvlinkOpenshiftOpenstackPythonRancherRdmaRoceVgpuVMware
17 Days AgoSaved
In-Office or Remote
Praha, Hlavní město Praha, CZE
Senior level
Senior level
Software
Design, deploy, operate, and optimize cloud and AI infrastructure (Kubernetes/OpenStack) on NVIDIA-certified hardware. Troubleshoot networking, storage, and Linux issues, implement CI/CD and automation, mentor customers and teams, and drive reliability, security, and performance improvements across globally distributed environments.
Top Skills: Ci/CdDcgmGoGpu SchedulingInfinibandJavaScriptKubernetesLinuxMig/VgpuNvidiaNvlinkOpenshiftOpenstackPythonRancherRdmaRoceVMware
Reposted 17 Days AgoSaved
Remote
USA
Senior level
Senior level
Software
Own the networking vision and roadmap for k0rdent AI, defining underlay/overlay fabrics, RDMA, DNS/IPAM, network automation, and DPU/SmartNIC integration. Translate customer and engineering requirements into priorities, track emerging interconnect standards, and support positioning, field engagements, and reference architectures.
Top Skills: Amd PensandoBgpDcqcnDnsDpuEcnEvpnFabric TelemetryGpudirect RdmaInfinibandIpamKubernetes NetworkingLinux NetworkingNcclNetwork AutomationNvidiaOverlay NetworkingOvnOvsPfcRcclRdmaRocev2Scale-Up EthernetSdnSmartnicSr-IovUalinkUltra Ethernet (Uec)VrfVxlan
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account