Top Tech Jobs & Startup Jobs

Reposted 19 Days AgoSaved
In-Office
Seoul, KOR
Junior
Junior
Artificial Intelligence • Cloud • Generative AI • Infrastructure as a Service (IaaS)
Act as technical liaison for customers using FriendliAI's inference platform: onboard developers, provide technical support, create documentation and tutorials, debug production issues with engineering, and run demos and Q&A sessions.
Top Skills: APIsCli ToolsDeployment WorkflowsHugging FaceLangchainPython
Reposted 19 Days AgoSaved
Hybrid
San Francisco, CA, USA
Senior level
Senior level
Artificial Intelligence • Cloud • Generative AI • Infrastructure as a Service (IaaS)
Design, build, and maintain a scalable web platform and APIs for deploying and monitoring multimodal AI models and agent workflows. Collaborate with product, infrastructure, and design teams to optimize performance, ensure reliability, drive CI/CD and testing, and contribute to long-term architecture decisions for a cloud-native, multi-tenant SaaS system.
Top Skills: Ci/CdCloud-NativeFastapiGraphQLGrpcHugging FaceKubernetesLlmNext.JsOpentelemetryPostgresPythonRbacReactRestSQLTypescript
Reposted 19 Days AgoSaved
In-Office
Seoul, KOR
Senior level
Senior level
Artificial Intelligence • Cloud • Generative AI • Infrastructure as a Service (IaaS)
Lead strategy and roadmap for FriendliAI's inference platform, owning initiatives end-to-end. Mentor junior PMs/designers, drive customer discovery, define product requirements and KPIs, partner with engineering/research, and align with GTM and sales to deliver scalable model APIs, deployment workflows, and developer features.
Top Skills: Ai/Ml SystemsAPIsCloud-NativeDeveloper PlatformsGpu-Based InfrastructureHugging FaceInference PlatformsLlm DeploymentModel ApisMulti-Tenant Saas
Reposted 19 Days AgoSaved
Hybrid
San Francisco, CA, USA
Senior level
Senior level
Artificial Intelligence • Cloud • Generative AI • Infrastructure as a Service (IaaS)
Lead end-to-end enterprise sales for FriendliAI's AI inference platform: generate pipeline, close high-value deals, run technical POCs, engage AI/ML communities, collaborate with engineering, and inform product roadmap.
Top Skills: Ai/MlCloud PlatformsDeveloper ToolsHugging FaceInference ServingLlmMlops
Reposted 19 Days AgoSaved
Hybrid
San Francisco, CA, USA
Senior level
Senior level
Artificial Intelligence • Cloud • Generative AI • Infrastructure as a Service (IaaS)
Design, implement, and optimize GPU kernels, kernel compiler, memory planner, and runtime for low-latency generative AI inference. Analyze performance bottlenecks across hardware and software, collaborate with infrastructure teams, and maintain production profiling, benchmarking, and validation tooling while supporting new model architectures and multi-GPU strategies.
Top Skills: BenchmarkingC++Compiler InfrastructureDiffusion ModelsDistributed InferenceGpu KernelsKernel CompilerMulti-GpuProfilingPythonRuntime SystemsTransformer Models
Reposted 19 Days AgoSaved
Hybrid
San Francisco, CA, USA
Mid level
Mid level
Artificial Intelligence • Cloud • Generative AI • Infrastructure as a Service (IaaS)
Design, deploy, and operate large-scale LLM and multimodal inference architectures. Work hands-on with customer engineering teams to containerize, scale, monitor, and troubleshoot GPU-based inference workloads across Kubernetes, CI/CD, and hybrid/on-prem environments. Create Helm charts, Terraform modules, and observability tooling while delivering workshops and platform reliability insights.
Top Skills: AWSCi/CdDeepspeed-InferenceDockerDocker ImagesEksElkGCPGpu ComputingGrafanaHelmHugging FaceKubernetesLokiOciOtelPrometheusTensorrtTerraformTritonVllm
Reposted 19 Days AgoSaved
In-Office
Seoul, KOR
Mid level
Mid level
Artificial Intelligence • Cloud • Generative AI • Infrastructure as a Service (IaaS)
Design, build, and maintain agent APIs and production agent applications (document understanding, RAG, automation). Integrate open-source LLMs and multimodal models, collaborate with backend and infra teams for deployment, and ensure APIs are reliable, scalable, and developer-friendly with strong documentation and monitoring.
Top Skills: HuggingfaceKubernetesLangchainLlamaindexLlmsMultimodal ModelsOcrPythonRag
Reposted 19 Days AgoSaved
In-Office
Seoul, KOR
Senior level
Senior level
Artificial Intelligence • Cloud • Generative AI • Infrastructure as a Service (IaaS)
Build and optimize GPU kernels and core inference engine components (compiler, memory planner, runtime) for latency-critical generative AI workloads. Profile and benchmark performance, collaborate with cloud/infrastructure teams, support new model architectures and multi‑GPU/distributed inference, and maintain production-grade validation tools.
Top Skills: Benchmarking ToolsC++Diffusion ModelsDistributed InferenceGpu KernelsHugging FaceKernel CompilerMemory PlannerMulti-GpuProfiling ToolsPythonRuntime SystemsTransformer Models
New

Track Smarter, Apply Better.

Ditch the spreadsheets. Organize your job search with our freeApplication Tracker.

Use For Free
Application Tracker Preview
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account