Agentic ai architect
Delphi Consulting · Panchkula, India
FULL TIMEpermanent
Job Description
Join Delphi - Where Innovation meets transformation
At Delphi, we believe in creating an environment where our people thrive. Our hybrid work model empowers you to choose where you work—whether it's from the office, your home, or a mix of both— so you can prioritize what matters most. We are committed to supporting your personal goals, family, and overall well- being while driving transformative results for our clients.
We welcome exceptional talent from anywhere across the globe. Interviews and onboarding are conducted virtually, reflecting our digital-first mindset.
Rooted in the region, we specialize in delivering tailored, impactful solutions in Data, Advanced Analytics and AI, Infrastructure, Cloud Security, and Application Modernization. Whether it’s enabling predictive analytics , transforming operations with automation, or driving customer engagement with intelligent platforms, we are the trusted partner for organizations ready to embrace a smarter, more efficient future.
About the Role
Senior AI/ML Technical Architect – Generative AI & Agentic Systems
Location: India
We are seeking a Senior AI/ML Solution Architect with deep expertise in Generative AI and agentic systems to lead the architecture and delivery of enterprise-scale AI solutions. This role requires a strong combination of hands-on technical capability across Large Language Models (LLMs) and Small Language Models (SLMs), along with the architectural leadership to design, integrate, and deploy AI systems across cloud, edge, and on-premises environments.
The successful candidate will architect scalable agentic platforms, build advanced RAG and fine-tuning pipelines, and design robust integration frameworks connecting AI services with enterprise applications. This role sits at the core of AI transformation initiatives, balancing cutting-edge innovation with production-grade engineering, performance optimization, and security-first design.
Experience Requirements
8+ years of experience in software engineering, data, or AI systems
2+ years of hands-on experience with Generative AI and LLM-based solutions
4+ years of experience designing and architecting enterprise-scale platforms and distributed systems
Key Responsibilities
Architecture & System Design
Architect scalable agentic systems using advanced LLM and SLM capabilities
Design multi-agent orchestration frameworks for complex, automated workflows
Define context, memory, and state-management architectures for persistent agent interactions
Implement Model Context Protocol (MCP)–based integrations across enterprise services
AI & Platform Implementation
Design and optimize Retrieval-Augmented Generation (RAG) architectures
Build agent solutions using Lang Chain, Lang Graph, Semantic Kernel, Agno, and custom frameworks
Architect and deploy model inference pipelines across cloud, edge, and on-prem environments
Develop fine-tuning strategies for LLMs and SLMs, including domain and task specialization
Lead model compression, quantization, and performance-optimization initiatives
Integration & Enterprise Connectivity
Architect secure REST, g RPC, and Graph QL APIs for AI platform services
Design event-driven architectures using message buses and webhooks
Implement authentication and authorization systems (SSO, OIDC, token-based security)
Build and govern enterprise connectors (Slack, Jira, Salesforce, ERP/CRM platforms)
Data, Models & Evaluation
Design data preprocessing and governance pipelines (cleaning, deduplication, PII handling)
Architect embedding generation, re-indexing, and retrieval optimization workflows
Define chunking, windowing, and content processing strategies
Establish model evaluation, benchmarking, and selection frameworks
Required Technical Skills
Core AI & Gen AI
Deep experience with GPT-4, Claude, LLa MA, and enterprise/open-source LLM ecosystems
Expertise deploying and optimizing SLMs (Phi-3, Gemma, Tiny Llama)
Advanced agent frameworks: Lang Chain, Lang Graph, Semantic Kernel, Agno
RAG systems, vector databases, semantic and hybrid retrieval
Fine-Tuning & Model Optimization
Parameter-efficient tuning: Lo RA, QLo RA, Do RA, Ada Lo RA
Prompt-tuning, adapters, prefix tuning, P-tuning v2
RLHF / RLAIF pipelines
Compression techniques: quantization (INT8/INT4, GPTQ, AWQ, GGML), pruning, distillation
Deployment & Performance
Multi-environment deployment (cloud, edge, on-prem)
Autoscaling, rate-limiting, and resource governance
Real-time inference, streaming pipelines, adaptive reasoning control
Engineering & Platforms
Tensor Flow, Py Torch, Hugging Face, Llama Index
API-driven system design and full-stack AI service development
AWS, Azure, GCP AI platforms
CI/CD, Docker, Kubernetes, monitoring and observability
Preferred Qualifications
Master’s or Ph D in Computer Science, AI, ML, or related field
Contributions to open-source AI projects or research publications
Experience with multi-modal and cross-modal systems
Strong grounding in MLOps and full model lifecycle management
Experience designing compliant and regulated AI systems
Demonstrated leadership in enterprise AI transformation programs
Cloud certifications (AWS, Azure, GCP – AI/ML focus)
Technical Competencies Assessed
Distributed AI system architecture and design
Production code quality, scalability, and performance optimization
Model benchmarking, evaluation, and cost optimization
Security, privacy, and AI governance engineering
Enterprise-grade deployment and scaling strategies
What We Offer:
At Delphi, we are dedicated to creating an environment where you can thrive, both professionally and personally. Our competitive compensation package, performance-based incentives, and health benefits are designed to ensure you're well-supported. We believe in your continuous growth and offer company- sponsored certifications, training programs , and skill-building opportunities to help you succeed. We foster a culture of inclusivity and support, with remote work options and a fully supported work-from- home setup to ensure your comfort and productivity. Our positive and inclusive culture includes team activities, wellness and mental health programs to ensure you feel supported.
At Delphi, we believe in creating an environment where our people thrive. Our hybrid work model empowers you to choose where you work—whether it's from the office, your home, or a mix of both— so you can prioritize what matters most. We are committed to supporting your personal goals, family, and overall well- being while driving transformative results for our clients.
We welcome exceptional talent from anywhere across the globe. Interviews and onboarding are conducted virtually, reflecting our digital-first mindset.
Rooted in the region, we specialize in delivering tailored, impactful solutions in Data, Advanced Analytics and AI, Infrastructure, Cloud Security, and Application Modernization. Whether it’s enabling predictive analytics , transforming operations with automation, or driving customer engagement with intelligent platforms, we are the trusted partner for organizations ready to embrace a smarter, more efficient future.
About the Role
Senior AI/ML Technical Architect – Generative AI & Agentic Systems
Location: India
We are seeking a Senior AI/ML Solution Architect with deep expertise in Generative AI and agentic systems to lead the architecture and delivery of enterprise-scale AI solutions. This role requires a strong combination of hands-on technical capability across Large Language Models (LLMs) and Small Language Models (SLMs), along with the architectural leadership to design, integrate, and deploy AI systems across cloud, edge, and on-premises environments.
The successful candidate will architect scalable agentic platforms, build advanced RAG and fine-tuning pipelines, and design robust integration frameworks connecting AI services with enterprise applications. This role sits at the core of AI transformation initiatives, balancing cutting-edge innovation with production-grade engineering, performance optimization, and security-first design.
Experience Requirements
8+ years of experience in software engineering, data, or AI systems
2+ years of hands-on experience with Generative AI and LLM-based solutions
4+ years of experience designing and architecting enterprise-scale platforms and distributed systems
Key Responsibilities
Architecture & System Design
Architect scalable agentic systems using advanced LLM and SLM capabilities
Design multi-agent orchestration frameworks for complex, automated workflows
Define context, memory, and state-management architectures for persistent agent interactions
Implement Model Context Protocol (MCP)–based integrations across enterprise services
AI & Platform Implementation
Design and optimize Retrieval-Augmented Generation (RAG) architectures
Build agent solutions using Lang Chain, Lang Graph, Semantic Kernel, Agno, and custom frameworks
Architect and deploy model inference pipelines across cloud, edge, and on-prem environments
Develop fine-tuning strategies for LLMs and SLMs, including domain and task specialization
Lead model compression, quantization, and performance-optimization initiatives
Integration & Enterprise Connectivity
Architect secure REST, g RPC, and Graph QL APIs for AI platform services
Design event-driven architectures using message buses and webhooks
Implement authentication and authorization systems (SSO, OIDC, token-based security)
Build and govern enterprise connectors (Slack, Jira, Salesforce, ERP/CRM platforms)
Data, Models & Evaluation
Design data preprocessing and governance pipelines (cleaning, deduplication, PII handling)
Architect embedding generation, re-indexing, and retrieval optimization workflows
Define chunking, windowing, and content processing strategies
Establish model evaluation, benchmarking, and selection frameworks
Required Technical Skills
Core AI & Gen AI
Deep experience with GPT-4, Claude, LLa MA, and enterprise/open-source LLM ecosystems
Expertise deploying and optimizing SLMs (Phi-3, Gemma, Tiny Llama)
Advanced agent frameworks: Lang Chain, Lang Graph, Semantic Kernel, Agno
RAG systems, vector databases, semantic and hybrid retrieval
Fine-Tuning & Model Optimization
Parameter-efficient tuning: Lo RA, QLo RA, Do RA, Ada Lo RA
Prompt-tuning, adapters, prefix tuning, P-tuning v2
RLHF / RLAIF pipelines
Compression techniques: quantization (INT8/INT4, GPTQ, AWQ, GGML), pruning, distillation
Deployment & Performance
Multi-environment deployment (cloud, edge, on-prem)
Autoscaling, rate-limiting, and resource governance
Real-time inference, streaming pipelines, adaptive reasoning control
Engineering & Platforms
Tensor Flow, Py Torch, Hugging Face, Llama Index
API-driven system design and full-stack AI service development
AWS, Azure, GCP AI platforms
CI/CD, Docker, Kubernetes, monitoring and observability
Preferred Qualifications
Master’s or Ph D in Computer Science, AI, ML, or related field
Contributions to open-source AI projects or research publications
Experience with multi-modal and cross-modal systems
Strong grounding in MLOps and full model lifecycle management
Experience designing compliant and regulated AI systems
Demonstrated leadership in enterprise AI transformation programs
Cloud certifications (AWS, Azure, GCP – AI/ML focus)
Technical Competencies Assessed
Distributed AI system architecture and design
Production code quality, scalability, and performance optimization
Model benchmarking, evaluation, and cost optimization
Security, privacy, and AI governance engineering
Enterprise-grade deployment and scaling strategies
What We Offer:
At Delphi, we are dedicated to creating an environment where you can thrive, both professionally and personally. Our competitive compensation package, performance-based incentives, and health benefits are designed to ensure you're well-supported. We believe in your continuous growth and offer company- sponsored certifications, training programs , and skill-building opportunities to help you succeed. We foster a culture of inclusivity and support, with remote work options and a fully supported work-from- home setup to ensure your comfort and productivity. Our positive and inclusive culture includes team activities, wellness and mental health programs to ensure you feel supported.
Details
| Company | Delphi Consulting |
| Location | Panchkula, India |
| Type | FULL TIME |
| Niche | tech |
| Experience | permanent |
