As an Applied AI Architect, you will lead technical strategies for enterprise customers, guiding their AI adoption journey and ensuring measurable business impact.
Staff AI Platform Engineer - Inference & Agentic Systems
📅 Apr 16
AI SystemsLLM ApplicationsDistributed SystemsTypeScript
📜 Description
- Own the model lifecycle: download, deploy, serve, monitor, update, swap.
- Drive inference optimization: latency, throughput, cost - including quantization, batching, caching, and routing strategies.
- Architect and build the Agentic AI Platform - runtime infrastructure, orchestration systems, and developer tooling for autonomous agents.
- Design multi-agent coordination systems enabling agents to collaborate and solve complex workflows.
🛠️ Requirements
- 8+ years of software engineering experience, with 3+ years in AI systems or LLM applications
- Strong understanding of LLM-based agent architectures: tool use, multi-step workflows, multi-agent coordination, and their failure modes
- Experience building highly reliable distributed systems
- Experience evaluating LLM systems in production: building evals, detecting regressions, and debugging non-deterministic failures
- Proficiency in TypeScript or Python, and willingness to work in both: the agent platform is TypeScript on Bun with Temporal workflows on Kubernetes and EC2, the inference platform is Python.
- Experience working with modern LLM APIs or open-source models
- Experience with or strong interest in model serving (vLLM, TensorRT-LLM, Triton)
- Understanding of distributed systems: task queues, event-driven architectures, state management, and durable long-running workflows
- Experience with cloud platforms (AWS, GCP) and containerized deployments
- Strong understanding of security risks in agentic systems (prompt injection, privilege escalation, data leakage)
✨ Benefits
- Build and operate multi-model serving across modalities (text, voice, code, vision) on shared infrastructure.
Full job description
About the Role
We are a small team of AI builders in Paytm Labs.
As a Staff AI Platform Engineer, you will work across inference and agentic systems. You will
contribute to Paytm's AI inference platform (Pi), serving internal teams and enterprise customers
- running our own coding and domain-specific models (voice, vision, risk, fintech workflows) as
well as third-party models. You will also architect and build the platform that enables
autonomous AI agents to operate safely and reliably in production - the runtime, orchestration,
and developer tooling for agents to reason, plan, use tools, and execute complex multi-step
workflows, automating both software development and business processes.
You will work at the intersection of LLMs, distributed systems, and production fintech
infrastructure, helping define how inference and agentic AI are built and deployed across
payments, risk, fraud, collections, support, and developer experience.
Go Big or Go Home!
Paytm Labs believes in diversity and equal opportunity and we will not tolerate any forms of discrimination or harassment. Our people are critical to our success and we know the more inclusive we are, the better our work will be.
We thank all applicants, however, only those selected for an interview will be contacted.
Paytm Labs is committed to meeting the accessibility needs of all individuals in accordance with the Accessibility for Ontarians with Disabilities Act (AODA) and the Ontario Human Rights Code (OHRC). Should you require accommodations during the recruitment and selection process, please let us know.
What You'll Do
What You'll Bring
Nice to Have
Similar jobs
Search more AI Engineer jobs🕒 yesterday
🕒 2 days ago
As a Staff Machine Learning Engineer, you will develop advanced AI models and systems to enhance Sentry's product capabilities, focusing on issue resolution and predictive analytics.
🕒 2 days ago
Develop and optimize AI models and training systems for cutting-edge robotics applications, focusing on large-scale model design, implementation, and rapid experimentation.
🕒 yesterday
The Forward Deployed Engineer will provide technical expertise and support for the implementation and integration of Shield AI's enterprise software products in AI and autonomy development.
🕒 yesterday
As a Senior Search Engineer on the Discovery Team, you will design and implement advanced search functionalities to enhance product discovery for leading commerce brands.
