Evaluator
Evaluate content
Configuration
Evaluation Metrics*
Evaluation Metrics
Content*
Enter the content to evaluate
Model*
Type or select a model...
API Key
••••••••
Shown when applicable to the selected model at runtime.
Azure OpenAI Endpoint
••••••••
Shown when model is one of 'azure/gpt-5.4', 'azure/gpt-5.4-mini', 'azure/gpt-5.4-nano', 'azure/gpt-5.2', 'azure/gpt-5.1', 'azure/gpt-5.1-codex', 'azure/gpt-5', 'azure/gpt-5-mini', 'azure/gpt-5-nano', 'azure/gpt-5-chat', 'azure/o3', 'azure/o4-mini', 'azure/gpt-4.1', 'azure/gpt-4.1-mini', 'azure/gpt-4.1-nano', 'azure/model-router'.
Azure API Version
2024-07-01-preview
Shown when model is one of 'azure/gpt-5.4', 'azure/gpt-5.4-mini', 'azure/gpt-5.4-nano', 'azure/gpt-5.2', 'azure/gpt-5.1', 'azure/gpt-5.1-codex', 'azure/gpt-5', 'azure/gpt-5-mini', 'azure/gpt-5-nano', 'azure/gpt-5-chat', 'azure/o3', 'azure/o4-mini', 'azure/gpt-4.1', 'azure/gpt-4.1-mini', 'azure/gpt-4.1-nano', 'azure/model-router'.
Input
| Parameter | Type | Required | Description |
|---|---|---|---|
metrics | json | Yes | Evaluation metrics configuration |
model | string | Yes | AI model to use |
apiKey | string | No | Provider API key. Shown when applicable to the selected model at runtime. |
azureEndpoint | string | No | Azure OpenAI endpoint URL. Shown when model is one of 'azure/gpt-5.4', 'azure/gpt-5.4-mini', 'azure/gpt-5.4-nano', 'azure/gpt-5.2', 'azure/gpt-5.1', 'azure/gpt-5.1-codex', 'azure/gpt-5', 'azure/gpt-5-mini', 'azure/gpt-5-nano', 'azure/gpt-5-chat', 'azure/o3', 'azure/o4-mini', 'azure/gpt-4.1', 'azure/gpt-4.1-mini', 'azure/gpt-4.1-nano', 'azure/model-router'. |
azureApiVersion | string | No | Azure API version. Shown when model is one of 'azure/gpt-5.4', 'azure/gpt-5.4-mini', 'azure/gpt-5.4-nano', 'azure/gpt-5.2', 'azure/gpt-5.1', 'azure/gpt-5.1-codex', 'azure/gpt-5', 'azure/gpt-5-mini', 'azure/gpt-5-nano', 'azure/gpt-5-chat', 'azure/o3', 'azure/o4-mini', 'azure/gpt-4.1', 'azure/gpt-4.1-mini', 'azure/gpt-4.1-nano', 'azure/model-router'. |
temperature | number | No | Response randomness level (low for consistent evaluation) |
content | string | Yes | Content to evaluate |
Output
| Parameter | Type | Description |
|---|---|---|
content | string | Evaluation results |
model | string | Model used |
tokens | json | Token usage |
cost | json | Cost information |
Usage Instructions
This is a core workflow block. Assess content quality using customizable evaluation metrics and scoring criteria. Create objective evaluation frameworks with numeric scoring to measure performance across multiple dimensions.
Notes
- Category:
tools - Type:
evaluator