AI/ML Data Engineer | Unstructured Data, PySpark, Vector Search, Retrieval-Augmented Generation (RAG), Cloud (AWS/Azure/GCP)

Sorry, this job was removed at 02:04 p.m. (UTC) on Thursday, Apr 23, 2026
Be an Early Applicant
Hyderabad, Telangana, IND
In-Office
Fintech • Financial Services
The Role

Job Summary
Synechron is seeking a highly experienced Python Data & AI Engineer to lead the design, development, and deployment of intelligent data pipelines supporting AI/ML applications. This role involves integrating unstructured data processing, working with large-scale data storage, and supporting AI model deployment in production environments. The successful candidate will collaborate with cross-functional teams to innovate data solutions, optimize performance, and enable AI-driven insights aligned with enterprise goals.

Software Requirements

  • Required:

    • Expertise in Python (latest stable version) for data processing, automation, and ML workflows

    • Hands-on experience with PySpark for distributed data processing at scale

    • Experience with data ingestion, cleansing, and transformation for unstructured data (PDFs, emails, forms)

    • Familiarity with NLP and AI frameworks such as Hugging Face, Transformers, LangChain, and FAISS for semantic search and retrieval-augmented generation (RAG) systems

    • Knowledge of data storage solutions including NoSQL (MongoDB, DynamoDB) and relational databases (PostgreSQL, MySQL)

    • Experience working with cloud platforms such as AWS, Azure, or GCP for deployment and scalability

    • Understanding of data quality metrics, data governance, and data security best practices

  • Preferred:

    • Experience with AI/ML lifecycle management including model training, fine-tuning, prompt engineering, and inference deployment

    • Familiarity with containerization (Docker) and orchestration (Kubernetes) for scalable AI/ML systems

    • Knowledge of data orchestration tools like Apache Airflow or Prefect

Overall Responsibilities

  • Design, develop, and optimize scalable data pipelines for unstructured data, supporting AI/ML applications and retrieval systems

  • Build and integrate document classification, enrichment, and metadata tagging workflows to prepare high-quality data for model consumption

  • Implement and support RAG architectures supporting semantic search, question-answering, and document retrieval workflows

  • Engineer and manage embeddings, document chunking, and vector database integrations supporting enterprise search capabilities

  • Collaborate with AI architects, data scientists, and platform teams to deliver end-to-end data solutions supporting intelligent applications

  • Automate data ingestion, processing, and deployment workflows for efficiency and reproducibility

  • Conduct system performance monitoring, troubleshoot issues, and implement continuous improvements

  • Ensure systems adhere to security, compliance, and data governance standards

Technical Skills (By Category)

  • Programming Languages:
    Required: Python, PySpark for distributed data processing
    Preferred: SQL, Java, or Scala for integration and performance optimization

  • Data Management & Storage:
    NoSQL (MongoDB, DynamoDB), relational databases (PostgreSQL, MySQL), data modeling, and query optimization

  • Cloud Technologies:
    AWS, Azure, or GCP services supporting scalable data pipelines and AI/ML deployment (e.g., BigQuery, S3, Dataflow, GKE)

  • Frameworks & Libraries:
    Hugging Face Transformers, LangChain, FAISS, Spark MLlib, NLP libraries for document understanding and retrieval

  • Data Orchestration & Automation:
    Apache Airflow, Prefect, Terraform, Docker, Kubernetes for deployment and pipeline management

  • Security & Governance:
    Data encryption, access controls, and compliance with enterprise security standards

Experience Requirements

  • Minimum of 6 years supporting data engineering, AI/ML deployment, or unstructured data processing in enterprise environments

  • Proven expertise in building large-scale, scalable data pipelines supporting AI workflows

  • Hands-on experience with vector databases, embedding models, and retrieval-augmented generation systems

  • Demonstrated success integrating cloud-based data systems with AI/ML models in production environments

  • Industry experience in financial services, healthcare, or enterprise analytics preferred but not mandatory

Day-to-Day Activities

  • Develop, test, and optimize data pipelines for unstructured data ingestion, cleansing, and transformation workflows

  • Support the development and deployment of AI models, including prompt engineering, fine-tuning, and inference workflows

  • Collaborate with data scientists and platform teams to design scalable, cloud-enabled AI and data solutions

  • Troubleshoot and resolve technical bottlenecks related to data pipelines and model deployment

  • Automate data workflows, manage infrastructure as code, and supports cloud migration strategies

  • Monitor system health, optimize data storage and processing performance, and ensure security and compliance standards are maintained

  • Document system architecture, data workflows, and operational procedures

Qualifications

  • Bachelor’s or Master’s degree in Computer Science, Data Science, or related field

  • 6+ years supporting or developing data engineering or AI pipelines supporting enterprise applications

  • Certifications such as GCP Professional Data Engineer, AWS Certified Machine Learning, or equivalent are advantageous

  • Proven experience with unstructured data processing, vector search, and retrieval-augmented generation workflows

Professional Competencies

  • Strong analytical and troubleshooting skills for complex data and AI systems

  • Excellent communication skills to effectively engage with technical teams and stakeholders

  • Leadership qualities to mentor junior engineers and promote best practices in data and AI engineering

  • Strategic thinking to design scalable, secure, and compliant data platforms

  • Adaptability to new tools, frameworks, and emerging AI/ML trends

  • Time management and organizational skills to handle multiple projects effectively

S​YNECHRON’S DIVERSITY & INCLUSION STATEMENT
 

Diversity & Inclusion are fundamental to our culture, and Synechron is proud to be an equal opportunity workplace and is an affirmative action employer. Our Diversity, Equity, and Inclusion (DEI) initiative ‘Same Difference’ is committed to fostering an inclusive culture – promoting equality, diversity and an environment that is respectful to all. We strongly believe that a diverse workforce helps build stronger, successful businesses as a global company. We encourage applicants from across diverse backgrounds, race, ethnicities, religion, age, marital status, gender, sexual orientations, or disabilities to apply. We empower our global workforce by offering flexible workplace arrangements, mentoring, internal mobility, learning and development programs, and more.

All employment decisions at Synechron are based on business needs, job requirements and individual qualifications, without regard to the applicant’s gender, gender identity, sexual orientation, race, ethnicity, disabled or veteran status, or any other characteristic protected by law.

Candidate Application Notice

Synechron Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Synechron and has not been reviewed or approved by Synechron.

  • Fair & Transparent Compensation Pay is frequently characterized as competitive, particularly relative to large service-consulting peers and in certain in-demand skill areas. Compensation sentiment appears strongest when staffing is stable on strong client engagements and for market-aligned roles in major hubs.
  • Healthcare Strength Healthcare coverage is often portrayed as a strong point in the U.S., with broad coverage and relatively favorable out-of-pocket experiences. Core medical, dental, and vision options are consistently described as meeting or exceeding a baseline expectation for consulting roles.
  • Equity Value & Accessibility Equity was made broadly accessible through a company-wide RSU grant tied to a major revenue milestone. This is positioned as a notable upside even if it is framed as a one-time recognition event rather than an ongoing program.

Synechron Insights

Similar Jobs

Crunchyroll Logo Crunchyroll

Paid Search Senior Manager

Digital Media • eCommerce • Gaming • Mobile • News + Entertainment
Hybrid
Hyderabad, Telangana, IND
1300 Employees

Micron Technology Logo Micron Technology

Sr. Observability Engineer

Artificial Intelligence • Hardware • Information Technology • Machine Learning
In-Office
Hyderabad, Telangana, IND
45000 Employees

Micron Technology Logo Micron Technology

Senior Business Analyst

Artificial Intelligence • Hardware • Information Technology • Machine Learning
In-Office
Hyderabad, Telangana, IND
45000 Employees

ServiceNow Logo ServiceNow

Financial Analyst

Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Remote or Hybrid
Hyderabad, Telangana, IND
29000 Employees
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: New York, New York
12,827 Employees
Year Founded: 2001

What We Do

At Synechron, we believe in the power of digital to transform businesses for the better. Our global consulting firm combines creativity and innovative technology to deliver industry-leading digital solutions. Synechron’s progressive technologies and optimization strategies span end-to-end Artificial Intelligence, Consulting, Digital, Cloud & DevOps, Data, and Software Engineering, servicing an array of noteworthy financial services and technology firms. Through research and development initiatives in our FinLabs we develop solutions for modernization, from Artificial Intelligence and Blockchain to Data Science models, Digital Underwriting, mobile-first applications and more. Over the last 20+ years, our company has been honored with multiple employer awards, recognizing our commitment to our talented teams. With top clients to boast about, Synechron has a global workforce of 14,700+, and has 48 offices in 19 countries within key global markets. For more information on the company, please visit our website: www.synechron.com.

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account