Financial ServicesCustomer Service

How Vanguard Uses Pinecone to Boost Customer Support with 12% More Accurate Responses

Vanguard partnered with Pinecone to build Agent Assist, an internal RAG-powered AI chat tool that helps customer support representatives find answers faster and more accurately. By replacing keyword-based search with hybrid vector retrieval, Vanguard achieved 12% more accurate search results and meaningfully reduced call times — even during high-demand periods like tax season.

Outcomes

12%Search result accuracy improvement
ReducedCustomer call times
ReducedOperational overhead during peak seasons

Tools & Technologies

1P
pgvector
PostgreSQL extension enabling vector similarity search directly within relational database workloads.
2AP
AWS PrivateLink
Private network connectivity service by AWS for secure access to cloud services without internet exposure.
3F
Faiss
Open-source vector search library by Meta for efficient nearest-neighbor searches in embedding space.
4AD
Amazon DynamoDB
Managed NoSQL key-value database by AWS for high-throughput, low-latency data storage applications.
5R
Redis
In-memory data store by Redis used as a vector database for semantic similarity search in AI apps.
6PS
Pinecone Serverless
Serverless vector database by Pinecone offering scalable semantic search without infrastructure management.

AI Categories

Challenge

Vanguard's customer support teams relied on keyword-based search that returned links to lengthy documents, forcing agents to manually hunt for answers — driving up call times, reducing satisfaction, and requiring costly seasonal hiring surges. The team needed a scalable, real-time retrieval solution capable of handling a highly dynamic financial document dataset.

Solution

Vanguard's CAI team built Agent Assist, an internal RAG-powered chat assistant using Pinecone Serverless as the vector database, combining BM25 sparse embeddings with dense embeddings for hybrid retrieval, and leveraging metadata filtering to ensure agents always access the most current documents.

Full Story

Vanguard, one of the world's largest investment management firms, has long prioritized delivering exceptional client experiences — including responsive, knowledgeable customer support. With millions of clients relying on Vanguard for retirement planning, investments, and financial advice, the quality and speed of support interactions carry real financial consequences. The company's Center for Analytics and Insights (CAI) team, operating within the Chief Data Analytics office, was tasked with modernizing how customer service representatives access information during live calls.

Access 456+ AI use cases, 428+ tools, and adoption signal rankings.

Source

PINECONE
March 2026
Original case study ↗

Similar Cases

1K
How Klarna’s AI Assistant Resolves 80% of Queries in Under 2 Minutes
Klarna
80%Reduction in average customer query resolution time
2S
How Stripe Deploys Claude Code to 1,370 Engineers with Zero-Configuration Rollout
Stripe
1,370Engineers Deployed
3W
How WEX Achieved 30% Developer Productivity Gains with GitHub Copilot
WEX
~30%Developer productivity increase with GitHub Copilot
4IU
How Itaú Unibanco Uses GitHub Copilot to Ship 93% Faster
Itaú Unibanco
93%Code commit time reduction
5NB
How NBIM Uses Claude Enterprise to Save 20% Time on Investment Analysis
Norges Bank Investment Management
20%Weekly time savings per employee
6B
How Block Gives 4,000 Employees AI-Powered Data Access via Claude and Databricks
Block
75% saving 8-10+ hoursEngineers saving time weekly
7A
How Airtree Uses Claude Cowork to Automate VC Research & Reporting
Airtree
Reduced from 2 days to minutesMarket & competitor research time
8O
How OffDeal Uses Claude to Close $91M in M&A Deals with a Team of Four Bankers
OffDeal
85%Internal eval accuracy after migrating to Claude Agent SDK
9R
How Ramp Uses Claude Code to Ship 1M Lines of Code in 30 Days
Ramp
1+ million linesAI-suggested code implemented in 30 days
10CC
How Chipper Cash Uses Pinecone Vector Search to Stop Fraud in Real-Time
Chipper Cash
95%+Selfie verification accuracy
See all use cases →