AI case study

PortolaPrompt evaluation

Engineering handoffs delayed prompt updates. Now, behavioral researchers test and deploy conversation changes directly to production.

Published

The story

Context

A developer of an AI companion app designed to build authentic, non-romantic relationships through natural voice conversations and complex memory systems.

Challenge

Evaluating subjective nuances like emotional intelligence, conversational pacing, and natural memory recall proved impossible using automated metrics...

Solution
Unlock full story

Scope & timeline

  • 4x increase in weekly prompt iterations

Quotes

Unlock 6 more quotes

The company

Portola logo

Portola

tolans.com

AI-powered virtual companion app for emotional support and conversation.

IndustrySoftware & Platforms
LocationSan Francisco, CA, USA
Employees1-10
Founded2023

The vendor

AI observability and evaluation platform that helps developers build, test, and monitor LLM-powered applications.

IndustrySoftware & Platforms
LocationSan Francisco, CA
Employees11-50
Founded2020

Use case

Portola's Prompt evaluation is part of this use case:

Conversational AI
32 case studies(-22% YoY)
Proven impact?
LowModerateVery Strong
3.0Moderate
2.1Lowwithin Product Engineering

Similar Case Studies

Related implementations across industries and use cases

35 AI case studies in Conversational AI

296 AI case studies in Software & Platforms

621 AI case studies in Product Engineering