The Role
Build and ship runtime instrumentation and observability components that operate inside production applications. Investigate runtime internals such as garbage collection, JITs, event loops, tracing hooks, and bytecode; optimize CPU and memory overhead; and ensure safe, backward-compatible operation. Debug distributed-system issues, design resilient fail-open systems, and work with product teams on technical trade-offs. The role also involves mentoring, technical leadership, and potentially developing security agents, APM tooling, telemetry systems, eBPF integrations, and low-overhead tracing.
Summary Generated by Built In
About the Role
Job requirements
Experience
Job responsibilities
Engineering Excellence / Mindset
We’re building a runtime code sensor that operates where most tools don’t: inside running applications. Our goal is to give engineers and AI agents real-time, high-fidelity visibility into how code actually behaves in production – under real load, real traffic, and real failures.
This role blends deep systems engineering with applied research. You’ll dive into runtime internals, explore undocumented behavior, design low-overhead instrumentation, and turn research insights into production-grade components that safely run in customer environments. One day you might analyze GC or JIT behavior; the next, design tracing mechanisms that survive real-world distributed systems.
If you get excited about runtime internals, enjoy breaking (and fixing) complex systems, think like both a researcher and a production engineer, and care deeply about performance, safety, and correctness – you’ll feel right at home here. This is a hands-on, high-impact role for engineers who want to ship technology that engineers actually trust to run in their most critical services.
This role blends deep systems engineering with applied research. You’ll dive into runtime internals, explore undocumented behavior, design low-overhead instrumentation, and turn research insights into production-grade components that safely run in customer environments. One day you might analyze GC or JIT behavior; the next, design tracing mechanisms that survive real-world distributed systems.
If you get excited about runtime internals, enjoy breaking (and fixing) complex systems, think like both a researcher and a production engineer, and care deeply about performance, safety, and correctness – you’ll feel right at home here. This is a hands-on, high-impact role for engineers who want to ship technology that engineers actually trust to run in their most critical services.
Job requirements
Experience
- 5+ years of hands-on research or development roles.
- Deep expertise in at least one runtime (Node.js, Python, or Java/JVM), including understanding of internals (event loop, GC, tracing hooks, bytecode/JIT, etc.).
- Hands-on experience building in-process production components (SDKs, agents, profilers, monitoring/security tools) that must be safe, stable, and backward-compatible.
- Strong performance engineering skills – profiling CPU/memory, avoiding overhead, understanding how instrumentation affects runtime behavior.
- Defensive engineering mindset – experience designing systems that fail-open, degrade gracefully, protect the host application, and never introduce instability.
- Track record debugging production issues (latency, memory leaks, regressions, deadlocks) in real-world distributed systems.
- Solid understanding of modern backend architectures – experience with microservices, distributed systems, async and event-driven patterns, containers/orchestration (Docker/K8s), cloud runtimes, and the performance or reliability challenges they introduce.
- Proven ability to ship stable, resilient, maintainable systems in production.
Job responsibilities
Engineering Excellence / Mindset
- Ability to anticipate technical risks, identify bottlenecks, and drive long-term engineering improvements.
- Takes ownership of code quality, documentation, reliability, and observability.
- Comfortable working with product teams to balance technical trade-offs with user and business needs.
- Autonomous and proactive; capable of mentoring others or leading technical initiatives.
- Background in security agents, observability tools, or other components deployed directly into customer environments.
- Experience with APM agents, JVM agents, Python tracing, V8 internals, or other instrumentation/profiling frameworks.
- Experience with telemetry systems (metrics, tracing, logging) including batching, rate-limiting, and safe data collection.
- Familiarity with sampling techniques, bytecode manipulation, eBPF, or low-overhead tracing.
- Exposure to safety-critical or high-throughput environments where reliability and minimal overhead are mandatory.
- Contributions to open-source instrumentation, tracing, or internals-related projects.
- This is a full-time on-site position located in Tel Aviv.
- Ability to thrive in a dynamic, fast-paced startup environment is essential.
Skills Required
- 5+ years of hands-on experience in research or development roles
- Deep expertise in at least one runtime: Node.js, Python, or Java/JVM, including runtime internals
- Hands-on experience building in-process production components such as SDKs, agents, profilers, monitoring tools, or security tools
- Strong performance engineering skills, including CPU and memory profiling and overhead reduction
- Experience designing defensive systems that fail open, degrade gracefully, protect host applications, and avoid instability
- Experience debugging production issues including latency, memory leaks, regressions, and deadlocks in distributed systems
- Understanding of microservices, distributed systems, asynchronous and event-driven patterns, containers, orchestration, cloud runtimes, and performance or reliability challenges
- Proven ability to ship stable, resilient, maintainable production systems
- Ability to anticipate technical risks, identify bottlenecks, and drive long-term engineering improvements
- Ability to take ownership of code quality, documentation, reliability, and observability
- Ability to collaborate with product teams to balance technical trade-offs with user and business needs
- Autonomous and proactive working style, with ability to mentor others or lead technical initiatives
- Full-time onsite availability in Tel Aviv
- Security agents, observability tools, or components deployed in customer environments
- Experience with APM agents, JVM agents, Python tracing, V8 internals, or instrumentation and profiling frameworks
- Experience with telemetry systems including batching, rate limiting, and safe data collection
- Familiarity with sampling techniques, bytecode manipulation, eBPF, or low-overhead tracing
- Experience in safety-critical or high-throughput environments
- Open-source contributions to instrumentation, tracing, or runtime-internals projects
Am I A Good Fit?
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.
Success! Refresh the page to see how your skills align with this role.
The Company
What We Do
Hud Technology Inc. develops runtime intelligence software for coding agents. Its platform runs with production code, detects errors, performance degradation, and CPU spikes, and captures forensic execution context. Hud uses that data to provide AI coding tools with root-cause evidence and support automatically generated, risk-analyzed fixes. The company aims to help engineering teams understand production behavior, ship software more confidently, and make AI-generated code safer.


.jpg)
.jpeg)





