Senior Software Engineer Runtime Internals

Posted 3 Hours Ago
Be an Early Applicant
Hiring Remotely in Israel
Remote
Senior level
Artificial Intelligence • Software • Automation
The Role
Build and ship runtime instrumentation and observability components that operate inside production applications. Investigate runtime internals such as garbage collection, JITs, event loops, tracing hooks, and bytecode; optimize CPU and memory overhead; and ensure safe, backward-compatible operation. Debug distributed-system issues, design resilient fail-open systems, and work with product teams on technical trade-offs. The role also involves mentoring, technical leadership, and potentially developing security agents, APM tooling, telemetry systems, eBPF integrations, and low-overhead tracing.
Summary Generated by Built In
About the Role
We’re building a runtime code sensor that operates where most tools don’t: inside running applications. Our goal is to give engineers and AI agents real-time, high-fidelity visibility into how code actually behaves in production – under real load, real traffic, and real failures.
This role blends deep systems engineering with applied research. You’ll dive into runtime internals, explore undocumented behavior, design low-overhead instrumentation, and turn research insights into production-grade components that safely run in customer environments. One day you might analyze GC or JIT behavior; the next, design tracing mechanisms that survive real-world distributed systems.
If you get excited about runtime internals, enjoy breaking (and fixing) complex systems, think like both a researcher and a production engineer, and care deeply about performance, safety, and correctness – you’ll feel right at home here. This is a hands-on, high-impact role for engineers who want to ship technology that engineers actually trust to run in their most critical services.

Job requirements
Experience
  • 5+ years of hands-on research or development roles.
  • Deep expertise in at least one runtime (Node.js, Python, or Java/JVM), including understanding of internals (event loop, GC, tracing hooks, bytecode/JIT, etc.).
  • Hands-on experience building in-process production components (SDKs, agents, profilers, monitoring/security tools) that must be safe, stable, and backward-compatible.
  • Strong performance engineering skills – profiling CPU/memory, avoiding overhead, understanding how instrumentation affects runtime behavior.
  • Defensive engineering mindset – experience designing systems that fail-open, degrade gracefully, protect the host application, and never introduce instability.
  • Track record debugging production issues (latency, memory leaks, regressions, deadlocks) in real-world distributed systems.
  • Solid understanding of modern backend architectures – experience with microservices, distributed systems, async and event-driven patterns, containers/orchestration (Docker/K8s), cloud runtimes, and the performance or reliability challenges they introduce.
  • Proven ability to ship stable, resilient, maintainable systems in production.

Job responsibilities
Engineering Excellence / Mindset
  • Ability to anticipate technical risks, identify bottlenecks, and drive long-term engineering improvements.
  • Takes ownership of code quality, documentation, reliability, and observability.
  • Comfortable working with product teams to balance technical trade-offs with user and business needs.
  • Autonomous and proactive; capable of mentoring others or leading technical initiatives.
Bonus Points
  • Background in security agents, observability tools, or other components deployed directly into customer environments.
  • Experience with APM agents, JVM agents, Python tracing, V8 internals, or other instrumentation/profiling frameworks.
  • Experience with telemetry systems (metrics, tracing, logging) including batching, rate-limiting, and safe data collection.
  • Familiarity with sampling techniques, bytecode manipulation, eBPF, or low-overhead tracing.
  • Exposure to safety-critical or high-throughput environments where reliability and minimal overhead are mandatory.
  • Contributions to open-source instrumentation, tracing, or internals-related projects.
Requirements
  • This is a full-time on-site position located in Tel Aviv.
  • Ability to thrive in a dynamic, fast-paced startup environment is essential.

Skills Required

  • 5+ years of hands-on experience in research or development roles
  • Deep expertise in at least one runtime: Node.js, Python, or Java/JVM, including runtime internals
  • Hands-on experience building in-process production components such as SDKs, agents, profilers, monitoring tools, or security tools
  • Strong performance engineering skills, including CPU and memory profiling and overhead reduction
  • Experience designing defensive systems that fail open, degrade gracefully, protect host applications, and avoid instability
  • Experience debugging production issues including latency, memory leaks, regressions, and deadlocks in distributed systems
  • Understanding of microservices, distributed systems, asynchronous and event-driven patterns, containers, orchestration, cloud runtimes, and performance or reliability challenges
  • Proven ability to ship stable, resilient, maintainable production systems
  • Ability to anticipate technical risks, identify bottlenecks, and drive long-term engineering improvements
  • Ability to take ownership of code quality, documentation, reliability, and observability
  • Ability to collaborate with product teams to balance technical trade-offs with user and business needs
  • Autonomous and proactive working style, with ability to mentor others or lead technical initiatives
  • Full-time onsite availability in Tel Aviv
  • Security agents, observability tools, or components deployed in customer environments
  • Experience with APM agents, JVM agents, Python tracing, V8 internals, or instrumentation and profiling frameworks
  • Experience with telemetry systems including batching, rate limiting, and safe data collection
  • Familiarity with sampling techniques, bytecode manipulation, eBPF, or low-overhead tracing
  • Experience in safety-critical or high-throughput environments
  • Open-source contributions to instrumentation, tracing, or runtime-internals projects
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
32 Employees
Year Founded: 2020

What We Do

Hud Technology Inc. develops runtime intelligence software for coding agents. Its platform runs with production code, detects errors, performance degradation, and CPU spikes, and captures forensic execution context. Hud uses that data to provide AI coding tools with root-cause evidence and support automatically generated, risk-analyzed fixes. The company aims to help engineering teams understand production behavior, ship software more confidently, and make AI-generated code safer.

Similar Jobs

HiBob Logo HiBob

Data Analyst

HR Tech • Information Technology • Professional Services • Sales • Software
Remote or Hybrid
IL
1350 Employees

Tufin Logo Tufin

Technical Support

Security • Cybersecurity
Remote or Hybrid
Tel Aviv, ISR
500 Employees

Cloudera Logo Cloudera

Account Manager

Artificial Intelligence • Cloud • Software • Big Data Analytics
Remote
Israel
3092 Employees

Micron Technology Logo Micron Technology

Design Engineer

Artificial Intelligence • Hardware • Information Technology • Machine Learning
Remote
Israel
45000 Employees

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account