List Metrics
GEThttps://api.confident-ai.com/v2/metrics
Lists all the custom metrics in your Confident AI project.
curl -X GET "https://api.confident-ai.com/v2/metrics" \
-H "CONFIDENT_API_KEY: <PROJECT-API-KEY>"{
"success": true,
"data": {
"metrics": [
{
"id": "<METRIC-ID>",
"name": "Correctness",
"algorithm": "DEFAULT",
"criteria": "Determine if the actual output is correct based on the expected output.",
"evaluationSteps": null,
"rubric": [
{
"scoreRange": [
8,
10
],
"expectedOutcome": "The answer is factually correct and complete."
}
],
"dag": {
"nodes": {
"root": {
"type": "BinaryJudgementNode",
"criteria": "Does the actual output answer the input?",
"evaluation_params": [
"input",
"actual_output"
],
"children": [
"pass",
"fail"
]
},
"pass": {
"type": "VerdictNode",
"verdict": true,
"score": 10
},
"fail": {
"type": "VerdictNode",
"verdict": false,
"score": 0
}
}
},
"multiTurn": false,
"requiredParameters": [
"actualOutput",
"expectedOutput"
]
}
]
},
"deprecated": false
}Headers
CONFIDENT_API_KEYstringRequiredThe API key of your Confident AI project.
Response
List Metrics succeeded.
successbooleanIndicates if the request was successful.
dataobjectShow 1 propertyHide 1 property
metricslist of objectsThis is the list of metrics.
Show 9 propertiesHide 9 properties
idstringThis is the unique id of the metric.
namestringThis is the name of the metric, unique within your project.
algorithmenum | nullThe algorithm the metric is evaluated with. GEVAL scores against criteria or evaluation steps, DAG walks a decision graph, CODE runs your own code, and DEFAULT is a built-in metric.
Show 4 enum valuesHide 4 enum values
DEFAULTDAGGEVALCODE
criteriastring | nullThis is the criteria the metric scores against, or null when it uses evaluation steps.
evaluationStepsarray | nullThese are the steps the metric follows to score, or null when it uses criteria.
rubricarray | nullThese are the score ranges that anchor how the metric scores, or null.
Show 2 propertiesHide 2 properties
scoreRangelist of anyThe inclusive start and end of the score range this outcome describes, each between 0 and 10, with the start no greater than the end.
expectedOutcomestringWhat a response scoring in this range looks like.
dagobject | nullThe decision graph of a DAG metric. It is validated on write, and metric references are resolved against the metrics in your project.
Show 1 propertyHide 1 property
nodesobjectThe graph's nodes keyed by node id, as serialized by deepeval's DAG metric. Judgement nodes list their
childrenby id; verdict nodes carry averdictand ascore, or point at another metric bymetric_name.
multiTurnbooleanThis is true when the metric evaluates conversations rather than single test cases.
requiredParameterslist of enumsThe test case fields the metric needs to run.
Show 14 enum valuesHide 14 enum values
inputactualOutputexpectedOutputcontextexpectedToolscontentrolescenarioexpectedOutcometurnstoolsCalledretrievalContextmetadatatags
deprecatedbooleanIndicates if this endpoint is deprecated.