response_type that determines the shape of the score Bluejay produces. Choose the type that matches how you want to reason about the result downstream, whether that’s in dashboards, alerts, or workflows.
Pass / Fail
response_type: pass_fail
The most common type. Bluejay returns either pass or fail based on whether the conversation meets your criteria.
To create this metric via API, use the Create Custom Metric endpoint with response_type set to pass_fail. The endpoint accepts optional fields such as agent_id, agent_ids, category, tags, and allow_not_applicable. See the full request schema on that page.
Quantitative
response_type: quantitative
Bluejay returns a numeric score within the range you define using min_value and max_value. Use this for nuanced scoring on a scale.
To create this metric via API, use the Create Custom Metric endpoint with response_type set to quantitative and include min_value and max_value so Bluejay knows the allowed range.
Qualitative
response_type: qualitative
Bluejay returns a free-form text summary describing its assessment. Useful for generating narrative feedback rather than a structured score.
To create this metric via API, use the Create Custom Metric endpoint with response_type set to qualitative. No extra type-specific fields are required beyond name and description.
Enum
response_type: enum
Bluejay classifies the conversation into one of the exact labels you define in enum_options. This is ideal for categorization tasks.
To create this metric via API, use the Create Custom Metric endpoint with response_type set to enum and include enum_options as an array of allowed labels.
JSON
response_type: json
Bluejay returns a structured JSON object, letting you extract multiple signals from a single metric evaluation in one pass.
To create this metric via API, use the Create Custom Metric endpoint with response_type set to json. Describe the desired object shape clearly in description so evaluations stay consistent.
Tool Call
response_type: tool_call
A deterministic check that passes when the agent called the tool(s) you name. No LLM is involved: Bluejay compares the tool_names you configure against the tool calls captured for the conversation.
To create this metric via API, use the Create Custom Metric endpoint with response_type set to tool_call and the expected names in tool_names.
Yes / No (Deprecated)
response_type: yes_no
Semantically identical to Pass / Fail, but the framing uses question-style prompts. Bluejay returns yes or no.
To create this metric via API, use the Create Custom Metric endpoint with response_type set to yes_no. Optional fields such as agent_id, category, and tags match other metric types. See the full schema on that page.
Quick Reference
Not Applicable
Any metric type can be configured withallow_not_applicable: true. When enabled, Bluejay may return Not Applicable if the criteria simply doesn’t apply to a given conversation. For example, a transfer metric on a call that never needed a transfer.
allow_not_applicable when creating or updating the metric via the Create Custom Metric or Update Custom Metric endpoints.
Resources
Prompting Guide
Write LLM-as-a-Judge prompts that score consistently for any response type.
Dynamic Variables
Inject call-specific context into your metric definitions at evaluation time.
Metrics Lab
Prototype and refine metrics against sample transcripts before going live.
Create Custom Metric API
Define a new Custom Metric programmatically with any response type.