AI Agentic Platform Engineer

Data & AIHybridFull-timeID: REQ-43683
N2P Systems
Chennai, Tamil Nadu, India3-5 yearsPosted 1 week ago

Role Overview

We are looking for an AI Agentic Platform Engineer to build scalable, production-ready infrastructure for LLM and multi-agent systems. This role combines agentic AI engineering with cloud-native platform engineering, distributed systems, MLOps/LLMOps, production AI infrastructure, security, and observability. The engineer will help build the foundational platform capabilities required to develop, deploy, operate, and scale AI-agent applications in production.

Key Responsibilities

  • Build and maintain production-grade agentic AI platforms using LangChain, LangGraph, AutoGen, CrewAI, or similar orchestration frameworks.
  • Implement Model Context Protocol (MCP) integrations to connect AI agents with tools, services, and external APIs.
  • Design and implement agent memory solutions using vector databases, semantic caching, and Knowledge Graphs.
  • Build scalable cloud-native infrastructure using Kubernetes, Docker, Helm, and serverless technologies.
  • Develop and optimize model-serving infrastructure using technologies such as vLLM, TensorRT-LLM, or Ollama, including GPU resource allocation and batching.
  • Build production-grade platform services using Python, Go, or Rust.
  • Implement AgentOps, evaluation, tracing, monitoring, and observability for AI systems.
  • Build secure execution environments and implement AI guardrails, RBAC, audit logging, human-in-the-loop controls, and safety mechanisms.
  • Implement dynamic model routing and fallback mechanisms to optimize cost, performance, and latency.
  • Develop internal SDKs, CLI utilities, templates, and APIs to support AI application development.
  • Collaborate with AI/ML and engineering teams to establish reliable, scalable, and reusable platform capabilities.

Required Skills & Qualifications

  • 3+ years of backend or platform engineering experience, including at least 1+ year building production-grade LLM or agent-driven systems.
  • Strong experience in Platform Engineering, MLOps/LLMOps, DevOps/SRE, or Microservices Architecture.
  • Hands-on experience with production AI/LLM infrastructure and scalable cloud platforms.
  • Strong understanding of cloud-native infrastructure, distributed systems, containerization, and production platform engineering.
  • Bachelor's or Master's degree in Computer Science, Software Engineering, Artificial Intelligence, or a closely related technical discipline.
  • Educational background from a Tier-1 or Tier-2 engineering institution, with institutions such as IITs, NITs, BITS, or comparable nationally recognized engineering institutes preferred/expected.
  • Strong engineering fundamentals with experience building and operating production systems rather than only academic or prototype-level AI projects.
  • Experience with LangChain, LangGraph, AutoGen, CrewAI, or similar agent orchestration frameworks.
  • Experience with vLLM, TensorRT-LLM, or Ollama for production model serving.
  • Proficiency in Python, Go, or Rust for production platform services.
  • Experience with Pinecone, pgvector, or Knowledge Graphs for AI memory and retrieval architectures.
  • Familiarity with LangSmith, Langfuse, Phoenix, OpenTelemetry, Ragas, DeepEval, or Braintrust.
  • Experience with Firecracker, gVisor, or WASM-based secure execution environments.
  • Knowledge of NeMo Guardrails or Guardrails AI.
  • Experience implementing GPU scheduling, batching, model routing, evaluation pipelines, or AI safety controls.

Role Pre-Screening Criteria

Applicants will be asked to answer the following qualifying questions when applying for this position:

  • 1Do you have 3+ years of backend/platform engineering experience, including 1+ year building production LLM or agent-driven systems?
  • 2Do you have hands-on production experience with Kubernetes, Docker, and Helm?
  • 3Have you built and deployed production-grade AI/LLM infrastructure or agentic AI platforms?
  • 4Is your Bachelor's/Master's degree from a Tier-1/Tier-2 institution such as IIT, NIT, BITS, or an equivalent institute?

Target Tech Stack

Agentic AIModel Context ProtocolKubernetesDockerHelmServerlessProduction AI/LLM InfrastructureScalable Cloud PlatformsPlatform EngineeringMLOpsLLMOpsDevOpsSREMicroservices ArchitectureRBACTier-1 or Tier-2 Engineering InstitutionBackend or Platform EngineeringProduction LLM or Agent-Driven SystemsLangChainLangGraphAutoGenCrewAIvLLMTensorRT-LLMOllamaPythonGoRustPineconepgvectorKnowledge GraphsLangSmithLangfusePhoenixOpenTelemetryRagasDeepEvalBraintrustFirecrackergVisorWASMNeMo GuardrailsGuardrails AIMCP

Quick Details

Location
Chennai, Tamil Nadu, India
Employment Type
Full-time
Work Arrangement
Hybrid
Experience Level
3-5 years
Domain Focus
Data & AI
Requisition Ref
REQ-43683

Ready to Apply?

Submit your resume and contact information. Our recruitment lead for this role will review your dossier and connect with you.

Apply for this Role

Know someone who fits?

Share this role with them.