Senior GenAI Engineer - VectorDBand MySQL
India
Job Description
Senior GenAI Engineer - VectorDBand MySQL
Pune, Maharashtra

Job Summary

Roles & ResponsibilitiesArchitect and lead the development of multi-agent AI systemsusing frameworks such as LangGraph, CrewAI, and AutoGen — enabling autonomous reasoning, tool use, inter-agent coordination, and adaptive decision-making at enterprise scale.Design and operationalize multimodal generative AI pipelines that unify text, image, tabular, and graph data using transformer-based architectures (BERT, CLIP, LLaVA, T5, Whisper, GPT-4o, Gemini) for rich, cross-modal intelligence.Build production-grade RAG and Graph-RAG systemsintegrating vector databases (Pinecone, pgvector, OpenSearch) and knowledge graphs (Neo4j, AWS Neptune) for semantic retrieval, entity-aware reasoning, and grounded generation.Lead LLM fine-tuning, prompt engineering, and model alignment strategies — including RLHF, PEFT, LoRA, and instruction tuning — to adapt foundation models for specialized enterprise use cases.Establish robust LLMOps and MLOps pipelines on Databricks (AWS) using MLflow, feature stores, prompt evaluation frameworks, model lineage tracking, and continuous retraining workflows to ensure reliable AI delivery.Develop high-performance Python backend services for LLM inference orchestration, async job handling, streaming responses, and distributed data workflows supporting high-throughput Gen AI operations.Engineer state, memory, and context management subsystems that enable agents to reason temporally, maintain session continuity, manage long-context windows, and coordinate across tools and modalities.Implement Responsible AI and AI governance practices — including bias detection, hallucination mitigation, explainability dashboards, output safety guardrails, and compliance with data ethics standards — ensuring transparency and fairness of deployed models.Apply traditional ML and statistical modeling (regression, clustering, forecasting, ensemble methods) in hybrid architectures alongside LLMs for interpretable, explainability-first decision systems.Continuously research, evaluate, and productionize advancements in generative modeling, agentic AI, multimodal transformers, and frontier foundation models — benchmarking against enterprise-scale performance and safety requirements.

Key Responsibilities

Roles & ResponsibilitiesArchitect and lead the development of multi-agent AI systemsusing frameworks such as LangGraph, CrewAI, and AutoGen — enabling autonomous reasoning, tool use, inter-agent coordination, and adaptive decision-making at enterprise scale.Design and operationalize multimodal generative AI pipelines that unify text, image, tabular, and graph data using transformer-based architectures (BERT, CLIP, LLaVA, T5, Whisper, GPT-4o, Gemini) for rich, cross-modal intelligence.Build production-grade RAG and Graph-RAG systemsintegrating vector databases (Pinecone, pgvector, OpenSearch) and knowledge graphs (Neo4j, AWS Neptune) for semantic retrieval, entity-aware reasoning, and grounded generation.Lead LLM fine-tuning, prompt engineering, and model alignment strategies — including RLHF, PEFT, LoRA, and instruction tuning — to adapt foundation models for specialized enterprise use cases.Establish robust LLMOps and MLOps pipelines on Databricks (AWS) using MLflow, feature stores, prompt evaluation frameworks, model lineage tracking, and continuous retraining workflows to ensure reliable AI delivery.Develop high-performance Python backend services for LLM inference orchestration, async job handling, streaming responses, and distributed data workflows supporting high-throughput Gen AI operations.Engineer state, memory, and context management subsystems that enable agents to reason temporally, maintain session continuity, manage long-context windows, and coordinate across tools and modalities.Implement Responsible AI and AI governance practices — including bias detection, hallucination mitigation, explainability dashboards, output safety guardrails, and compliance with data ethics standards — ensuring transparency and fairness of deployed models.Apply traditional ML and statistical modeling (regression, clustering, forecasting, ensemble methods) in hybrid architectures alongside LLMs for interpretable, explainability-first decision systems.Continuously research, evaluate, and productionize advancements in generative modeling, agentic AI, multimodal transformers, and frontier foundation models — benchmarking against enterprise-scale performance and safety requirements.

Skill Requirements

Roles & ResponsibilitiesArchitect and lead the development of multi-agent AI systemsusing frameworks such as LangGraph, CrewAI, and AutoGen — enabling autonomous reasoning, tool use, inter-agent coordination, and adaptive decision-making at enterprise scale.Design and operationalize multimodal generative AI pipelines that unify text, image, tabular, and graph data using transformer-based architectures (BERT, CLIP, LLaVA, T5, Whisper, GPT-4o, Gemini) for rich, cross-modal intelligence.Build production-grade RAG and Graph-RAG systemsintegrating vector databases (Pinecone, pgvector, OpenSearch) and knowledge graphs (Neo4j, AWS Neptune) for semantic retrieval, entity-aware reasoning, and grounded generation.Lead LLM fine-tuning, prompt engineering, and model alignment strategies — including RLHF, PEFT, LoRA, and instruction tuning — to adapt foundation models for specialized enterprise use cases.Establish robust LLMOps and MLOps pipelines on Databricks (AWS) using MLflow, feature stores, prompt evaluation frameworks, model lineage tracking, and continuous retraining workflows to ensure reliable AI delivery.Develop high-performance Python backend services for LLM inference orchestration, async job handling, streaming responses, and distributed data workflows supporting high-throughput Gen AI operations.Engineer state, memory, and context management subsystems that enable agents to reason temporally, maintain session continuity, manage long-context windows, and coordinate across tools and modalities.Implement Responsible AI and AI governance practices — including bias detection, hallucination mitigation, explainability dashboards, output safety guardrails, and compliance with data ethics standards — ensuring transparency and fairness of deployed models.Apply traditional ML and statistical modeling (regression, clustering, forecasting, ensemble methods) in hybrid architectures alongside LLMs for interpretable, explainability-first decision systems.Continuously research, evaluate, and productionize advancements in generative modeling, agentic AI, multimodal transformers, and frontier foundation models — benchmarking against enterprise-scale performance and safety requirements.

Other Requirements

null
Information at a Glance

Why HCLTech?

At HCLTech, you'll supercharge your potential. You'll find your career. And you'll find your spark. All at a place that knows that helping its customers stay on top starts by putting its people first.

HCLTech is a global technology company, home to more than 223,000 people across 60 countries, delivering industry-leading capabilities centered around digital, engineering, cloud and AI, powered by a broad portfolio of technology services and products. We work with clients across all major verticals, providing industry solutions for Financial Services, Manufacturing, Life Sciences and Healthcare, Technology and Services, Telecom and Media, Retail and CPG, and Public Services. Consolidated revenues as of 12 months ending June 2026 totaled $14.8 billion.