Guide

Prompt Caching

Use provider-supported prompt caching to reduce latency and repeated input cost.

Base https://api.vip.lingapi.ai

Quick checklist

  1. Confirm the selected model supports this capability.
  2. Send a minimal request before enabling the feature in production.
  3. Verify usage, billing, and error handling in the console.

How to use it

Keep stable system and context prefixes unchanged to improve cache hit rate.

Safety checks

Keep retries bounded, record request ids, and make user-facing workflows resilient to provider-side failures.

Example

Example
curl https://api.vip.lingapi.ai/v1/models \
  -H "Authorization: Bearer <YOUR_API_KEY>"