We’re Trezor, a leading company in crypto security that has pioneered the hardware wallet industry as the inventor of the world’s first hardware wallet.
We’re building our own AI inference infrastructure, a dedicated multi-GPU server hosted in our datacenter, so that Trezor employees can use large language models on hardware we control, with no data ever leaving our premises. We are looking for an AI Platform Engineer to own that machine end-to-end: the hardware and OS underneath it with IT’s help, the model serving stack on top of it, and the people who use it every day.
This is a hands-on, broad role. Some days you will be benchmarking a newly released open-weight model, another one sitting with a developer helping them wire an agent into their workflow.
If you like owning a system completely rather than a narrow slice of one, this is for you.
👉 What You’ll DoRun the machine
Collaborate with IT on owning the full lifecycle of our GPU server: OS, storage, networking…
Set up observability — GPU utilization, thermals, memory, request latency, throughput — and alerting that actually catches problems before users do
Run the models
Bring to life an inference stack (vLLM / SGLang, LiteLLM as a gateway, Open WebUI as the front end as an example)
Deploy, upgrade and tune open-weight models
Evaluate new model releases as they land, benchmark them on our hardware for quality, throughput and latency, and recommend what we should be running
Own quotas, routing and cost/usage reporting across teams
Collaborate in optimizing cloud usage as well, if needed. Sometimes we do have to use closed models
Support the engineers
Be the go-to person for developers integrating the internal models into their tooling — IDE assistants, agents, CI pipelines, internal apps
Maintain API keys, endpoints and documentation; write the internal guides that make onboarding self-service
Run internal enablement: short workshops, office hours, examples of what good usage looks like
Consider how queueing and prioritization will work. Who has priority and what long-term tasks are running over night?
Coding experience and Infrastructure as Code skills (Ansible, Terraform, Docker/Kubernetes) to automate your own work
Some Linux systems administration: networking, storage, containers, systemd, troubleshooting from the kernel up. IT will collaborate here, though
Genuine interest in the open-weight model ecosystem — you already know which models matter this month
Service mindset: you enjoy unblocking other engineers and writing things down
English for daily work; Czech is a plus
Nice to have
Experience serving LLMs in production or a serious homelab: vLLM, SGLang, llama.cpp, Ollama or similar
Working knowledge of NVIDIA GPU operations: drivers, CUDA, NVLink, MIG, nvidia-smi, DCGM
Datacenter experience: rack power budgets, liquid cooling, hardware vendor support processes
Fine-tuning / LoRA, quantization, model evaluation methodology
A unique opportunity to be part of a pioneering, security-first company in the crypto industry
A role where you can build, implement, and see the real impact of your work
A high level of ownership and freedom
The chance to work in an open-source company where transparency, trust, and security are part of how we think
Option to get paid in bitcoin
Flexible working hours and a supportive team
Budget for professional development, including training programs, courses, and workshops of your choice
Friendly, open culture with regular company events and fun get-togethers
Renovated offices with a gym, massages, football table, billiards, PlayStation, 3D printer and free on-site parking
Additional benefits such as a MultiSport card, company mobile phone tariff, yoga, fitness classes, and more
👋 Interested? We’d love to hear from you. Send us your CV and a few words about yourself, and we’ll get back to you as soon as we review your application.
Skills Required
- Coding experience
- Infrastructure as Code experience with Ansible, Terraform, Docker, or Kubernetes
- Linux systems administration experience, including networking, storage, containers, systemd, and troubleshooting
- English proficiency for daily work
- Czech language skills
- Experience serving LLMs in production or in a serious homelab using vLLM, SGLang, llama.cpp, Ollama, or similar
- Working knowledge of NVIDIA GPU operations, including drivers, CUDA, NVLink, MIG, nvidia-smi, and DCGM
- Datacenter experience with rack power budgets, liquid cooling, and hardware vendor support processes
- Experience with fine-tuning, LoRA, quantization, or model evaluation methodology
What We Do
We are technology pioneers developing products that secure individual autonomy and privacy. Innovation drives our brand success, putting us in a unique position where we get to focus on projects that matter, without having to compromise. SatoshiLabs established the cryptocurrency hardware wallet industry in 2013 with Trezor and hasn’t stopped innovating since. Our pioneering team delivers secure solutions to real-world problems, with Tropic Square creating the first auditable secure chip, Invity bringing new ways to access crypto, and Trezor developing new secure hardware and interfaces. Trezor, the original bitcoin hardware wallet, has developed into a multi-functional device that meets the needs of advanced and beginner users alike. Its success has enabled SatoshiLabs’ rapid expansion across its three constituent companies, as part of a unified strategy to meet new global challenges with cutting-edge solutions. We champion independence, innovation and secure access to finance for everyone, everywhere. Learn more about our companies at the links below: • SatoshiLabs - https://satoshilabs.com • Trezor - https://trezor.io • Invity - https://invity.io • Tropic Square - https://tropicsquare.com • Vexl - https://vexl.it







