Skip to navigation

Score a prompt

View as Markdown

Scores an agent’s prompt across 11 quality dimensions using Gemini-based analysis. Requires the prompt to have changed since the last scoring.

Input: Provide exactly one of versionId (published agent version) or draftId (agent draft). Providing both or neither returns a 400.

Credit usage: 1 credit is deducted per successful call.

Idempotency: Re-submitting the same prompt without changes returns a 400 — retrieve the cached score via the GET agent endpoint instead.

Supported agent types: Only single_prompt agents are supported. Workflow-graph agents return a 400.

Scoring model: Two sequential Gemini calls — a Platform Analyst pass followed by a Rubric Judge pass.

Scored Dimensions

TierDimensionNotes
1Role & Objective
1Personality & Voice
1Conversation Structure
1Tool Integration
1Constraints & Safety
2Conversational Naturalness
2Failure-Mode Coverage
3Information IntegrityGating — if Weak/Missing, score capped at 70
3Variable & Tool HygieneGating — if Weak/Missing, score capped at 50
3Internal Consistency
3DensityComputed from token analysis

Authentication

AuthorizationBearer

API key from the console ApiKey collection, sent as Bearer token. Also accepts session cookies for browser-based auth.

Request

This endpoint expects an object.
objectRequired
OR
objectRequired

Response

Prompt scored successfully
statusbooleanOptional
dataobjectOptional

Errors

400
Bad Request Error
401
Unauthorized Error
403
Forbidden Error
404
Not Found Error
429
Too Many Requests Error
500
Internal Server Error