Chaitanya Nuthalapati

Senior Technical Product Manager
Amazon Web Services
USA

About

Chaitanya Nuthalapati is a Senior Technical Product Manager at Amazon Web Services (AWS), where he works on Amazon ElastiCache for Valkey and artificial intelligence (AI) infrastructure use cases, including vector search, semantic caching, and key-value (KV) caching for large language model (LLM) inference. He has presented at AWS, Percona Live, and Valkey community events, translating distributed systems and machine learning (ML) concepts into practical architectures for enterprise AI applications.
Talk

Chaitanya Nuthalapati | Stop Recomputing: Up to 88% Lower LLM Latency with Semantic Caching

Semantic Caching, Cost Saving, Inference Optimization
Semantic caching cuts latency and reduces inference costs by reusing answers for semantically similar prompts. In this talk, Chaitanya Nuthalapati will explore how to build a production-grade semantic caching system for multi-agent systems with Valkey and Strands. Beyond the basics, this talk focuses on techniques for improving cache accuracy, including handling multi-turn interactions, applying conversation-state filters, protecting personally identifiable information (PII), and navigating the trade-offs of personal