Prompt Engineering Specialist

Posted 4 Days Ago
Be an Early Applicant
Pune, Maharashtra, IND
In-Office
Mid level
Artificial Intelligence • Software • Analytics • Business Intelligence
The Role
Design, test, and refine prompt systems for production LLM agents supporting code generation, specification parsing, code review, documentation, and enterprise reasoning. Build versioned prompt libraries and evaluation frameworks, measure performance, investigate hallucinations and other failure modes, and collaborate with engineering teams to integrate reliable prompts into AI workflows. The role also requires documenting prompt decisions and staying current with techniques such as RAG, structured outputs, chain-of-thought, and tool use.
Summary Generated by Built In
About the Role

Rubiscape is building the Decision Intelligence Platform that India’s enterprises will run on for the next decade — and prompt engineering is the craft layer that makes our AI-first development workflows reliable at scale. As a Prompt Engineering Specialist, you will design, test, and iterate on the prompt systems that power Rubiscape’s internal LLM agents: from code generation and spec parsing to enterprise domain reasoning across BFSI, manufacturing, and healthcare verticals. You will work with real production LLMs, real enterprise complexity, and real performance constraints — building prompt libraries and evaluation frameworks that directly accelerate how Rubiscape ships software and how our platform surfaces AI insights to customers.

 

Key Responsibilities

·         Design, write, and iteratively refine prompt templates for internal LLM agents covering code generation, spec interpretation, code review, documentation, and enterprise domain reasoning tasks.

·         Build and maintain a structured prompt library: versioned, tested, and annotated prompts organised by task type, model target, and domain context (BFSI, manufacturing, healthcare, etc.).

·         Develop prompt evaluation frameworks: define success metrics, construct test sets, run A/B comparisons, and document prompt performance regressions and improvements.

·         Collaborate with Spec Engineering and AI Toolchain teams to ensure prompts are correctly parameterised by upstream specs and that outputs meet downstream code quality standards.

·         Investigate and mitigate failure modes in LLM outputs: hallucination, instruction drift, context loss, and domain-specific reasoning errors relevant to enterprise data and analytics workflows.

·         Stay current with advances in prompt engineering techniques (chain-of-thought, retrieval-augmented generation, structured output forcing, tool use) and adapt internal practices accordingly.

·         Document prompt design decisions, rationale, and empirical results in a shared knowledge base accessible to the entire engineering organisation.

Nice to Have

·         Experience with retrieval-augmented generation (RAG) architectures and vector databases (Pinecone, Weaviate, pgvector) in enterprise contexts.

·         Familiarity with the data engineering, BI, or analytics domain — ability to write prompts that reason correctly about SQL, data pipelines, ML model outputs, or business KPIs.

·         Prior work on multi-agent LLM systems where prompt design affects agent-to-agent communication and task decomposition.

·         Contributions to open-source prompt engineering toolkits, evaluation frameworks, or published benchmarks.

 

 

 

About Rubiscape

Rubiscape is India’s leading Decision Intelligence Platform, unifying data engineering, BI, machine learning, and agentic AI in a single governed platform. Built in Pune and trusted by Fortune 500 enterprises across BFSI, manufacturing, healthcare, and government. 8 international innovation patents. 10 Industry-Academia Labs & COEs. From BI to AI — One Platform. Every Decision.



RequirementsRequirements

·         3+ years of hands-on experience in prompt engineering, NLP engineering, or applied LLM development, with a demonstrable portfolio of production prompt systems.

·         Deep practical knowledge of major LLM APIs (OpenAI GPT-4/o series, Anthropic Claude, Google Gemini) including token economics, context window management, and structured output techniques.

·         Experience building prompt evaluation pipelines: automated test harnesses, LLM-as-judge patterns, or human evaluation workflows.

·         Strong Python skills for scripting prompt experiments, parsing LLM outputs, and integrating with LangChain, LlamaIndex, or equivalent orchestration frameworks.

·         Ability to reason clearly about enterprise domain complexity and encode that domain knowledge into reliable, reusable prompt structures.

·         B.E. / B.Tech in Computer Science, AI/ML, or Linguistics; or equivalent practical experience with a strong portfolio.



Skills Required

  • 3+ years of hands-on experience in prompt engineering, NLP engineering, or applied LLM development
  • Demonstrable portfolio of production prompt systems
  • Practical knowledge of major LLM APIs, including OpenAI GPT-4/o, Anthropic Claude, and Google Gemini
  • Knowledge of token economics, context window management, and structured output techniques
  • Experience building prompt evaluation pipelines, automated test harnesses, LLM-as-judge patterns, or human evaluation workflows
  • Strong Python skills for scripting experiments, parsing LLM outputs, and integrating orchestration frameworks
  • Ability to reason about enterprise domain complexity and encode domain knowledge into reusable prompt structures
  • B.E. or B.Tech in Computer Science, AI/ML, or Linguistics, or equivalent practical experience with a strong portfolio
  • Experience with retrieval-augmented generation architectures and vector databases such as Pinecone, Weaviate, or pgvector
  • Familiarity with data engineering, business intelligence, or analytics domains
  • Prior work on multi-agent LLM systems
  • Contributions to open-source prompt engineering toolkits, evaluation frameworks, or published benchmarks
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
30 Employees
Year Founded: 2018

What We Do

Rubiscape is a decision intelligence platform that unifies business intelligence, analytics, data science, and artificial intelligence in one place, helping organizations move from BI to AI. Its technology and product-development work spans AI-focused development, software engineering, product management, agile delivery, security operations, and DevOps, reflecting a software platform mission centered on enabling data-driven decisions across modern organizations.

Similar Jobs

The Aerospace Corporation Logo The Aerospace Corporation

Systems Engineer

Aerospace • Artificial Intelligence • Cloud • Machine Learning • Software • Cybersecurity • Defense
Remote or Hybrid
India
4600 Employees

Mastercard Logo Mastercard

Principal Software Engineer

Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Hybrid
Pune, Maharashtra, IND
38800 Employees

Mastercard Logo Mastercard

Senior Software Engineer

Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Hybrid
Pune, Maharashtra, IND
38800 Employees

Mastercard Logo Mastercard

Software Engineer

Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Hybrid
Pune, Maharashtra, IND
38800 Employees

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account