str
Required. Python function for parsing results. The function
should be defined within this string.
The function takes a list of strings (LLM responses) and
should return either a list of dictionaries (for rubrics) or
a single dictionary (for a metric result).
Example function signature: def parse(responses: list[str])
-> list[dict[str, Any]] \| dict[str, Any]:
When parsing rubrics, return a list of dictionaries, where
each dictionary represents a Rubric. Example for rubrics: [
{ "content": {"property": {"description": "The response is
factual."}}, "type": "FACTUALITY", "importance": "HIGH" }, {
"content": {"property": {"description": "The response is
fluent."}}, "type": "FLUENCY", "importance": "MEDIUM" } ]
When parsing critique results, return a dictionary
representing a MetricResult. Example for a metric result: {
"score": 0.8, "explanation": "The model followed most
instructions.", "rubric_verdicts": [...] }
... code for result extraction and aggregation
This field is a member of oneof_ _parsing_function.
[[["Easy to understand","easyToUnderstand","thumb-up"],["Solved my problem","solvedMyProblem","thumb-up"],["Other","otherUp","thumb-up"]],[["Hard to understand","hardToUnderstand","thumb-down"],["Incorrect information or sample code","incorrectInformationOrSampleCode","thumb-down"],["Missing the information/samples I need","missingTheInformationSamplesINeed","thumb-down"],["Other","otherDown","thumb-down"]],["Last updated 2026-08-01 UTC."],[],[]]