Citi
Citi
Citi is a global banking and financial services company serving consumers, corporations, governments, and institutions. Its offerings span consumer banking, credit cards, wealth management, corporate banking, and investment banking, supported by operations in more than 100 countries. As a large international finance organization, Citi brings together teams working across banking, financial services, and fintech.

AI Lead, Python Engineering – Citi, Pune (Hybrid)

Lead the development of Python, PySpark, and generative AI platforms for Citi’s banking risk technology. Build scalable data pipelines, AI agents, and microservice integrations.

Description

  • Design and maintain enterprise AI agents that support high-volume ELT and ETL workflows
  • Develop and deploy generative AI agents with Google ADK and Google Flash 2.5+ LLMs
  • Enable application automation, advanced insights, workflow support, and human-in-the-loop designs
  • Build data federation layers for Lambda and Data Mesh architectures with tools such as Starburst
  • Apply machine learning, deep learning, and natural language processing to business use cases
  • Develop, deploy, and automate microservice integrations for data-intensive applications
  • Use cloud-native infrastructure, OpenShift or Kubernetes, and CI/CD pipelines to deliver scalable, resilient, maintainable systems
  • Apply advanced prompt engineering with agentic AI tools including Devin.AI, GitHub Copilot, and platforms such as MCP
  • Maintain data quality, integrity, and security across the full data lifecycle
  • Improve data engineering processes, standards, and best practices continuously
  • Evaluate business decision risks and protect Citi, its clients, and assets
  • Maintain regulatory compliance and transparently manage and report control issues
  • Partner with business stakeholders to advance next-generation data and analytics platforms

Requirements

  • At least 8 years of experience in large-scale application development
  • Recent experience with the required platform for securely and scalably deploying AI agents within applications
  • At least 5 years of experience as a Python and PySpark engineering lead
  • Experience building enterprise-grade, high-volume ELT and ETL processes with the PySpark and Databricks ecosystem
  • Hands-on experience developing agentic AI with YAML, JSON, FastAPI or Spring Boot, Google ADK, and LLM integrations
  • Experience using Devin.AI or GitHub Copilot
  • Experience integrating models through platforms such as MCP with advanced prompt engineering
  • Experience developing and automating microservice integrations for data-intensive applications
  • Proficiency in Python or Scala
  • Strong SQL skills and experience with relational databases
  • Deep knowledge of data modeling, data warehousing, Data Mesh architecture, and data federation
  • Experience with cloud-based big data platforms including Cloudera, Databricks, AWS, Azure, or GCP
  • Experience with frontend technologies such as Angular or React JS
  • Practical experience applying AI and machine learning techniques to business problems
  • Familiarity with Docker and Kubernetes
  • Data engineering experience in the retail banking products domain
  • Relevant industry certifications are preferred
  • Bachelor’s degree in Computer Science, Engineering, or a related field
  • A master’s degree is a plus

Benefits

  • Consideration under equal opportunity employment practices
  • Reasonable accommodation for applicants with disabilities

Related Jobs

Aleph Alpha

Senior AI Researcher, Foundation Model Pre-Training — Aleph Alpha, Heidelberg

Aleph Alpha
51 – 200 Employees
Artificial IntelligenceB2B

Lead architecture and large-scale pre-training work for Aleph Alpha’s European foundation models, shaping training methods across thousands of GPUs. Build and refine PyTorch training recipes with the Heidelberg-based hybrid team.

Open