Top Tech Jobs & Startup Jobs

Reposted 22 Days AgoSaved
In-Office
San Francisco, CA, USA
240K-280K Annually
Senior level
240K-280K Annually
Senior level
Artificial Intelligence • Information Technology
Design and implement manifest-driven provisioning and reconciliation systems for GPU hosts and clusters. Build durable workflow engines, declarative self-service APIs, event-driven automation, self-healing pipelines, and production-grade control planes. Own coding, testing, CI/CD, and operate the platform in production to eliminate manual provisioning.
Top Skills: BgpBmcCadenceCi/CdCudaGoInfinibandIpmiIpxeKafkaKubernetesKubernetes ControllersNatsNcclPxePythonRedfishRoceRustSqsTemporalVlan
Reposted 23 Days AgoSaved
In-Office
Francis Corners, Town of Malta, NY, USA
300K-370K Annually
Senior level
300K-370K Annually
Senior level
Artificial Intelligence • Information Technology
Own full sales cycle for strategic AI-native and enterprise accounts: generate pipeline, close new business, expand customers, and partner with Solutions, Engineering, and Research to run POCs and scale deployments. Provide technical and business insights to inform product and sales strategy.
Top Skills: Ai InfrastructureFine-TuningFlashattentionFlexgenGpu ClustersHyenaInferenceRedpajama
Reposted 23 Days AgoSaved
In-Office or Remote
2 Locations
Senior level
Senior level
Artificial Intelligence • Information Technology
Provide technical support to customers on AI solutions, resolve complex issues, collaborate with teams, and improve product offerings.
Top Skills: AIAnsibleGpuJavaScriptKubernetesMlPythonSlurmTypescript
Reposted 25 Days AgoSaved
In-Office or Remote
Bangalore, Bengaluru Urban, Karnataka, IND
Senior level
Senior level
Artificial Intelligence • Information Technology
Operate and scale Together AI's production infrastructure: run GPU-enabled Kubernetes clusters, automate with Ansible and Terraform, build monitoring/observability, respond to incidents (on-call), debug production issues, design deployment/upgrade processes, and plan infrastructure growth for high availability and scalability.
Top Skills: AnsibleCloud ServicesGpu-Enabled KubernetesKubernetesPagerdutyTerraform
Reposted 25 Days AgoSaved
In-Office
San Francisco, CA, USA
150K-230K Annually
Senior level
150K-230K Annually
Senior level
Artificial Intelligence • Information Technology
As a Senior Developer Productivity Engineer, you will enhance CI/CD processes, optimize workflows, and build tools for efficient software delivery.
Top Skills: AnsibleArgocdGithub ActionsGoGrafanaHoneycombJavaScriptJestNext.JsPrometheusPulumiPythonReactSkaffoldTerraformTypescript
New

Cut your apply time in half.

Use ourAI Assistantto automatically fill your job applications.

Use For Free
Application Tracker Preview
Reposted 25 Days AgoSaved
In-Office
Amsterdam, NLD
Senior level
Senior level
Artificial Intelligence • Information Technology
The Site Reliability Engineer ensures user-facing services run smoothly, focuses on systems reliability and scalability, and enhances operational practices through automation and engineering principles.
Top Skills: AnsibleCloud ServicesKubernetesProgramming/Scripting LanguagesTerraform
28 Days AgoSaved
Remote
AUS
160K-230K Annually
Mid level
160K-230K Annually
Mid level
Artificial Intelligence • Information Technology
Provide customer-facing SRE/technical support for Kubernetes GPU clusters and HPC environments. Troubleshoot GPU hardware, networking (InfiniBand/RDMA/NVLink), distributed storage (Weka/NFS), Slurm workflows, and container workloads. Maintain cluster health, perform node maintenance and migrations, document procedures, and collaborate with Engineering, Product, and Sales to drive customer success and product improvements. Provide weekend and on-call coverage.
Top Skills: AnsibleBmcContainer InfrastructureGpuHpcInfinibandKubernetesNfsNvlinkRdmaSlurmWeka
28 Days AgoSaved
Remote
AUS
160K-230K Annually
Senior level
160K-230K Annually
Senior level
Artificial Intelligence • Information Technology
Customer-facing technical role providing SRE-style support for GPU clusters, LLM inference, and fine-tuning. Troubleshoot production incidents, run infra-as-code changes, validate migrations, monitor observability dashboards, and translate technical findings to customers and engineering. Drive improvements by surfacing patterns and maintaining documentation.
Top Skills: AnsibleAWSAzureContainer InfrastructureCurlGCPGitGpu ClustersGrafanaHpcInfrastructure As CodeJavaScriptKubernetesLlm Inference FrameworksLoraNfsPostmanPrometheusPythonRest ApiSlurmTypescriptVastWeka
28 Days AgoSaved
In-Office
Singapore, SGP
Senior level
Senior level
Artificial Intelligence • Information Technology
Partner with strategic customers to optimize LLM inference and post-training pipelines, tune inference engines for latency/throughput, run and productionize fine-tuning (LoRA, SFT, DPO, RLHF), and feed field insights into product and model roadmaps to ensure successful POCs and deployments.
Top Skills: DpoFlashattentionFlexgenGrpoHyenaKv CacheLoraPipeline ParallelismPythonQuantizationRedpajamaRlhfSftSglangSpeculative DecodingTensor ParallelismTensorrt-LlmVllm
Reposted 28 Days AgoSaved
In-Office
San Francisco, CA, USA
200K-230K Annually
Senior level
200K-230K Annually
Senior level
Artificial Intelligence • Information Technology
Provide legal support for infrastructure procurement and enterprise commercial deals. Lead negotiations for GPU/cloud capacity, data center, networking, and vendor agreements; align vendor terms with customer contracts; manage export controls and data/privacy compliance; build templates, playbooks, and scalable procurement and sales workflows; support partnerships, reseller and go-to-market deals; and run procurement processes and vendor template libraries.
Top Skills: Ai ToolsCloudColocationData CenterGpuNetworking
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account