Senior GenAI Engineer

Posted 13 Days Ago
Be an Early Applicant
Hiring Remotely in POL
Remote
Senior level
Agency • Information Technology • Professional Services • Consulting
The Role
Build and operate production-grade LLM features using agentic flows, RAG, tool calling, and cloud APIs. Develop FastAPI services, evaluation datasets, regression suites, and LLM-as-judge checks. Improve retrieval, prompts, answer quality, latency, observability, and cost efficiency. Debug production traces with Langfuse and manage model routing across providers. Deploy and maintain services on Azure while collaborating with teammates and potentially mentoring engineers.
Summary Generated by Built In

This is a remote position.

We are looking for a Senior AI Engineer to join a team building and operating a production-grade LLM system used by real users.  This is not a proof-of-concept project. You will be responsible for taking LLM-powered features from idea through development and deployment to continuous improvement, with a strong focus on answer quality, performance, observability and cost efficiency. You will work with modern LLM technologies, including LangGraph, RAG, tool calling, Azure OpenAI, Gemini and Claude, and have a real impact on how AI-powered products are built and operated in production.
 
Responsibilities: 
  • Build LLM-powered features end to end. Design and implement agentic flows, retrieval, and tool calling using LangGraph - then ship them as FastAPI services with streaming, persistence and proper tests.
  • Own answer quality. Build evaluation datasets, regression suites and LLM-as-judge checks so we know whether a prompt or model change made things better before it reaches users.
  • Get the right context to the model. Turn user questions into effective queries against our search platform, orchestrate multi-step research loops, and shape the context the model reasons over. When an answer is wrong, work out whether retrieval, the query or the prompt is at fault - and fix the right one.
  • Debug production. Instrument flows with tracing (Langfuse), investigate bad answers from real traces, and manage latency, token and cost budgets - including routing across model sizes and families behind an AI gateway.


Requirements
  • Experience in building and operating backend services - APIs, async, testing
  • Hands-on experience taking LLM features to production and keeping them running - not only prototypes
  • Agent / orchestration frameworks - LangGraph ideally
  • Practical RAG experience
  • Experience debugging LLM systems in production - tracing, evaluation, cost and latency
  • Experience running services in the cloud (we're on Azure)
  • Strong problem-solving skills, analytical thinking, and technical decision-making
  • Fluent in English, proactive communicator, and a collaborative team player
  • Open-minded, creative, and motivated to push boundaries in AI and automation
Nice to have:
  • Azure OpenAI, AI Search, App Service
  • Infrastructure-as-code (Bicep)
  • Mentoring or tech-lead experience
Tech stack:
  • Python
  • FastAPI
  • LangGraph
  • Azure OpenAI, Google Gemini, Anthropic Claude
  • FAISS
  • PostgreSQL
  • Langfuse
  • Azure App Insights
  • Bicep
  • Docker/Podman




Benefits
  • B2B contract
  • 100% remote work
  • Long-term engagement
  • Work on a live LLM product
  • Modern AI/LLM technology stack
  • Flexible working environment


Skills Required

  • Experience building and operating backend services, including APIs, asynchronous systems, and testing
  • Hands-on experience taking LLM features into production and operating them beyond prototypes
  • Experience with agent or orchestration frameworks, ideally LangGraph
  • Practical retrieval-augmented generation experience
  • Experience debugging production LLM systems using tracing, evaluation, cost, and latency analysis
  • Experience running cloud services, particularly on Azure
  • Strong problem-solving, analytical thinking, and technical decision-making skills
  • Fluent English communication and collaborative teamwork
  • Experience with Azure OpenAI, Azure AI Search, or Azure App Service
  • Experience with infrastructure as code using Bicep
  • Mentoring or technical leadership experience
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
Year Founded: 2012

What We Do

Look4IT is a recruitment agency specializing in providing top-quality recruitment and outsourcing services within the IT sector, focusing on staffing solutions for permanent and contract positions.

Similar Jobs

Provectus Logo Provectus

Machine Learning Engineer

Artificial Intelligence • Information Technology • Consulting
In-Office or Remote
8 Locations
572 Employees

Provectus Logo Provectus

Machine Learning Engineer

Artificial Intelligence • Information Technology • Consulting
In-Office or Remote
7 Locations
572 Employees

Provectus Logo Provectus

Machine Learning Engineer

Artificial Intelligence • Information Technology • Consulting
In-Office or Remote
9 Locations
572 Employees

Provectus Logo Provectus

Machine Learning Engineer

Artificial Intelligence • Information Technology • Consulting
In-Office or Remote
8 Locations
572 Employees

Similar Companies Hiring

Standard Template Labs Thumbnail
Artificial Intelligence • Information Technology • Software
New York, NY
25 Employees
NODA AI Thumbnail
Artificial Intelligence • Information Technology • Software • Cybersecurity
Sydney, AU
54 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account