Skip to main content
Each item in requests is a group: one content and the questions about it. Add groups for different content. Top-level reasoning applies to every question; grounding goes on each question. Batching saves round trips, not tokens: each question reads its content.

Request

Response

results matches requests; each answers list matches that group’s questions. Check ok on each answer: result holds the decision, error the message. meta.usage and meta.latency_ms cover the whole call.