Customer Need

The customer, a global technology company, aimed to leverage OpenAI’s capabilities for processing sensitive data, specifically driver history and criminal records. These reports are sourced from agencies across the globe. Hence, these reports are not in a standard format, so LLMs are perfect for processing the information.
Applying OpenAI would save significant product gain and cost savings of over $10M annually.
This task required a sophisticated approach due to the sensitive nature of the data.

Challenge

The primary challenge was the sensitive nature of the data and the data residency requirements. Directly sending such sensitive information to large language models (LLMs) like those provided by OpenAI raised concerns about data privacy compliance.
To ensure compliance with privacy regulations, the customer wanted to retain sensitive PII data within the region where it originated.
Removing sensitive elements could strip the data of its core value, making it less useful for the intended processing and analysis.

Solution

Protecto’s solution involved intelligent tokenization, a method designed to identify and redact sensitive data elements while preserving the overall structure and utility of the data.
This process replaced sensitive information with format-preserving tokens. These tokens maintained the integrity and format of the original data, ensuring that its business value and utility remained intact.
Crucially, the tokenization process was tailored to work seamlessly with OpenAI. Protecto sent specific instructions to the LLM to ensure it could fully understand and process the tokenized data without loss of critical information.

Outcome

Protecto allowed the client to safely and effectively use OpenAI’s LLMs for processing sensitive driver and criminal history data.
Ensured compliance with data residency and privacy regulations while retaining the full analytical value of the data.
The client could harness the power of advanced AI to process their data without compromising on data protection or utility.

Amar Kanagaraj

Founder and CEO of Protecto

Amar Kanagaraj is the Founder and CEO of Protecto, a company focused on securing enterprise data for LLMs, AI agents, and agentic workflows. He is a second-time entrepreneur with 20+ years of experience across engineering, product, AI, go-to-market, and business leadership. Before Protecto, Amar co-founded FileCloud and helped scale it to over $10M ARR as CMO. Earlier in his career, he worked at Sun Microsystems, Booz & Company, and Microsoft Search & AI. He holds an MBA from Carnegie Mellon University and an MS in Computer Science from Louisiana State University.

Customer Case Study: Enabling Privacy-Preserving Processing of Sensitive Data with OpenAI

Table of Contents

Customer Need

Challenge

Solution

Outcome

Related Articles

Top Enterprise AI Adoption Challenges

What Is Privacy-by-Design and Why Is It Important?

Why Traditional DLP Breaks in Agentic AI