Skip to content

Chat Sessions & Conversation Management ​

This guide explains how VeriPrompt manages chat sessions, provider selection, and conversation continuity through summarization.

Overview ​

VeriPrompt's chat system provides:

  • Intelligent provider routing with session consistency
  • Web search augmentation for real-time queries (weather, news, prices) -- see Web Search
  • Conversation summarization for context preservation
  • Message limit management with configurable thresholds

Provider Selection ​

How Providers Are Chosen ​

When you start a conversation, VeriPrompt selects an AI provider based on several factors:

  1. Your Selection - If you manually choose a provider in the UI
  2. Category Settings - Your chat category may have specific routing rules
  3. Routing Policy - Automatic selection based on performance, cost, or availability

Session Locking ​

Once a provider is selected for a conversation, all subsequent messages use the same provider/model. This ensures:

  • Consistent response style
  • Maintained context understanding
  • Predictable behavior

Sessions automatically expire after 24 hours of inactivity, allowing the system to select a new optimal provider.

Provider Failover ​

If your selected provider experiences issues:

  1. The system automatically retries (up to 3 attempts)
  2. If needed, it falls back to an alternative provider
  3. You'll see a notification about the provider change

Message Limits ​

Limit Tiers ​

Chat sessions have configurable message limits (counted as question/answer pairs):

TierLimitUse Case
Normal40 pairsStandard conversations
Medium80 pairsExtended discussions
High160 pairsLong-form projects

Warning Thresholds ​

As you approach your message limit, you'll see warnings:

UsageBanner ColorMessage
80%Yellow"Approaching message limit"
95%Orange"Message limit almost reached" + Summarize button
100%RedInput blocked - must summarize to continue

Early Summarization ​

After 10 message pairs, a "Summarize and start new chat" button appears. This allows you to:

  • Save your conversation context
  • Start fresh with a clean message count
  • Preserve important information

Conversation Summarization ​

How It Works ​

When you click "Summarize and Start New Chat":

  1. Summary Generation - AI creates a structured summary of your conversation
  2. Context Preservation - Key facts, decisions, and artifacts are preserved
  3. New Chat Created - A new conversation opens with the summary injected
  4. Seamless Continuation - The AI remembers your previous discussion

What's Preserved ​

The summary captures:

  • Objective - What you were trying to accomplish
  • Key Context - Critical facts and decisions (up to 5 bullets)
  • Current State - Where the conversation left off
  • Pending Items - Open questions or next steps
  • Verbatim Content - Code snippets, URLs, exact values

Rolling Summary Window ​

VeriPrompt maintains a rolling window of 3 summaries per conversation chain:

  • Each summarization creates a new summary
  • If you already have 3 summaries, the oldest is automatically removed
  • This ensures you always have recent context without unlimited accumulation

Continuing from a Summary ​

When you continue from a summary:

  1. The summary is injected into the AI's system prompt
  2. The AI treats this as its "memory" of your previous conversation
  3. You can ask follow-up questions naturally
  4. The AI will reference prior context without you needing to repeat it

Best Practices ​

For Long Conversations ​

  1. Summarize proactively - Don't wait for the limit warning
  2. Review summaries - Ensure important details are captured
  3. Use clear topics - Help the AI identify key context

For Complex Projects ​

  1. Start with context - Explain your goal at the beginning
  2. Summarize at milestones - When completing major steps
  3. Reference artifacts - Include relevant code/URLs that need preservation

For Multiple Sessions ​

  1. Use categories - Organize by project or topic
  2. Continue from summaries - Don't start completely fresh
  3. Check summary depth - Monitor your continuation chain

Technical Details ​

Message Counting ​

  • Messages are counted as pairs (user message + assistant response)
  • System messages and summaries don't count toward the limit
  • Percentage is calculated as: (pairs / limit) * 100

Summary Token Limit ​

  • Summaries target 400 tokens maximum
  • This ensures efficient context injection
  • Long conversations are compressed intelligently

Classification Inheritance ​

When continuing from a summary:

  • The new conversation inherits the original classification
  • Security policies remain consistent
  • Routing rules are preserved

Troubleshooting ​

"AI says it has no memory" ​

If the AI doesn't recall prior context:

  1. Check if you used "Continue from Summary" (not "Start Fresh")
  2. Verify the summary was successfully created
  3. The new chat's system prompt should contain the summary

Provider keeps changing ​

Provider changes can occur due to:

  • Session expiration (24 hours)
  • Provider outages (automatic failover)
  • Category/policy changes

Summary not generating ​

If summarization fails:

  1. Ensure the conversation has at least 2 messages
  2. Check for network connectivity
  3. Try again - transient provider issues may resolve