- Location
- Hyderabad, Telangana, India
- Job type
- Full-time
Required skills
- LangChain
- Azure
- backend
- caching
- CI
- FastAPI
- frontend
- Kubernetes
- microservices
About the role
Tata Consultancy Services
Website:
tcs.com
Company:
https://www.linkedin.com/company/tata-consultancy-services
Industries: IT Services and IT Consulting
Job details:
- Build and Productionize cloud native backend services and AI/LLM inference pipelines.
- Design and develop Python-based APIs and microservices (FastAPI, async patterns) and agentic AI workflows using LangChain/LangGraph.
- Implement and optimize LLM capabilities including embeddings, RAG, vector search, prompt/context engineering, and model versioning.
- Package, serve, and monitor models for real-time and batch inference, ensuring operational readiness and performance.
- Build event driven, resilient integrations and containerized services, with hands-on Kubernetes debugging and Helm-based deployments.
- Establish observability, SLOs, CI/CD automation, testing
- Apply strong systems design principles (concurrency, caching, reliability, rate limiting) and robust data engineering practices.
- Cloud exposure preferred (Azure/AKS, managed services), with bonus experience in performance tuning, frontend collaboration, and model governance/monitoring.
Click on Apply to know more.
This page is fully interactive when JavaScript is enabled. Please enable JavaScript to apply or browse related roles.