openapi: 3.1.1 info: version: 1.0.0 title: Braintrust Acls Datasets API description: 'API specification for the backend data server. The API is hosted globally at https://api.braintrust.dev or in your own environment. You can access the OpenAPI spec for this API at https://github.com/braintrustdata/braintrust-openapi.' license: name: Apache 2.0 servers: - url: https://api.braintrust.dev security: - bearerAuth: [] - {} tags: - name: Datasets paths: /v1/dataset: post: tags: - Datasets security: - bearerAuth: [] - {} operationId: postDataset description: Create a new dataset. If there is an existing dataset in the project with the same name as the one specified in the request, will return the existing dataset unmodified summary: Create dataset requestBody: description: Any desired information about the new dataset object required: false content: application/json: schema: $ref: '#/components/schemas/CreateDataset' responses: '200': description: Returns the new dataset object content: application/json: schema: $ref: '#/components/schemas/Dataset' '400': description: The request was unacceptable, often due to missing a required parameter content: text/plain: schema: type: string application/json: schema: nullable: true '401': description: No valid API key provided content: text/plain: schema: type: string application/json: schema: nullable: true '403': description: The API key doesn’t have permissions to perform the request content: text/plain: schema: type: string application/json: schema: nullable: true '429': description: Too many requests hit the API too quickly. We recommend an exponential backoff of your requests headers: Retry-After: schema: type: string content: text/plain: schema: type: string application/json: schema: nullable: true '500': description: Something went wrong on Braintrust's end. (These are rare.) content: text/plain: schema: type: string application/json: schema: nullable: true get: operationId: getDataset tags: - Datasets description: List out all datasets. The datasets are sorted by creation date, with the most recently-created datasets coming first summary: List datasets security: - bearerAuth: [] - {} parameters: - $ref: '#/components/parameters/AppLimitParam' - $ref: '#/components/parameters/StartingAfter' - $ref: '#/components/parameters/EndingBefore' - $ref: '#/components/parameters/Ids' - $ref: '#/components/parameters/DatasetName' - $ref: '#/components/parameters/ProjectName' - $ref: '#/components/parameters/ProjectIdQuery' - $ref: '#/components/parameters/OrgName' responses: '200': description: Returns a list of dataset objects content: application/json: schema: type: object properties: objects: type: array items: $ref: '#/components/schemas/Dataset' description: A list of dataset objects required: - objects additionalProperties: false '400': description: The request was unacceptable, often due to missing a required parameter content: text/plain: schema: type: string application/json: schema: nullable: true '401': description: No valid API key provided content: text/plain: schema: type: string application/json: schema: nullable: true '403': description: The API key doesn’t have permissions to perform the request content: text/plain: schema: type: string application/json: schema: nullable: true '429': description: Too many requests hit the API too quickly. We recommend an exponential backoff of your requests headers: Retry-After: schema: type: string content: text/plain: schema: type: string application/json: schema: nullable: true '500': description: Something went wrong on Braintrust's end. (These are rare.) content: text/plain: schema: type: string application/json: schema: nullable: true /v1/dataset/{dataset_id}: get: operationId: getDatasetId tags: - Datasets description: Get a dataset object by its id summary: Get dataset security: - bearerAuth: [] - {} parameters: - $ref: '#/components/parameters/DatasetIdParam' responses: '200': description: Returns the dataset object content: application/json: schema: $ref: '#/components/schemas/Dataset' '400': description: The request was unacceptable, often due to missing a required parameter content: text/plain: schema: type: string application/json: schema: nullable: true '401': description: No valid API key provided content: text/plain: schema: type: string application/json: schema: nullable: true '403': description: The API key doesn’t have permissions to perform the request content: text/plain: schema: type: string application/json: schema: nullable: true '429': description: Too many requests hit the API too quickly. We recommend an exponential backoff of your requests headers: Retry-After: schema: type: string content: text/plain: schema: type: string application/json: schema: nullable: true '500': description: Something went wrong on Braintrust's end. (These are rare.) content: text/plain: schema: type: string application/json: schema: nullable: true patch: operationId: patchDatasetId tags: - Datasets description: Partially update a dataset object. Specify the fields to update in the payload. Any object-type fields will be deep-merged with existing content. Currently we do not support removing fields or setting them to null. summary: Partially update dataset security: - bearerAuth: [] - {} parameters: - $ref: '#/components/parameters/DatasetIdParam' requestBody: description: Fields to update required: false content: application/json: schema: $ref: '#/components/schemas/PatchDataset' responses: '200': description: Returns the dataset object content: application/json: schema: $ref: '#/components/schemas/Dataset' '400': description: The request was unacceptable, often due to missing a required parameter content: text/plain: schema: type: string application/json: schema: nullable: true '401': description: No valid API key provided content: text/plain: schema: type: string application/json: schema: nullable: true '403': description: The API key doesn’t have permissions to perform the request content: text/plain: schema: type: string application/json: schema: nullable: true '429': description: Too many requests hit the API too quickly. We recommend an exponential backoff of your requests headers: Retry-After: schema: type: string content: text/plain: schema: type: string application/json: schema: nullable: true '500': description: Something went wrong on Braintrust's end. (These are rare.) content: text/plain: schema: type: string application/json: schema: nullable: true delete: operationId: deleteDatasetId tags: - Datasets description: Delete a dataset object by its id summary: Delete dataset security: - bearerAuth: [] - {} parameters: - $ref: '#/components/parameters/DatasetIdParam' responses: '200': description: Returns the deleted dataset object content: application/json: schema: $ref: '#/components/schemas/Dataset' '400': description: The request was unacceptable, often due to missing a required parameter content: text/plain: schema: type: string application/json: schema: nullable: true '401': description: No valid API key provided content: text/plain: schema: type: string application/json: schema: nullable: true '403': description: The API key doesn’t have permissions to perform the request content: text/plain: schema: type: string application/json: schema: nullable: true '429': description: Too many requests hit the API too quickly. We recommend an exponential backoff of your requests headers: Retry-After: schema: type: string content: text/plain: schema: type: string application/json: schema: nullable: true '500': description: Something went wrong on Braintrust's end. (These are rare.) content: text/plain: schema: type: string application/json: schema: nullable: true /v1/dataset/{dataset_id}/insert: post: operationId: postDatasetIdInsert tags: - Datasets description: Insert a set of events into the dataset summary: Insert dataset events security: - bearerAuth: [] - {} parameters: - $ref: '#/components/parameters/DatasetIdParam' requestBody: description: An array of dataset events to insert required: false content: application/json: schema: $ref: '#/components/schemas/InsertDatasetEventRequest' responses: '200': description: Returns the inserted row ids content: application/json: schema: $ref: '#/components/schemas/InsertEventsResponse' '400': description: The request was unacceptable, often due to missing a required parameter content: text/plain: schema: type: string application/json: schema: nullable: true '401': description: No valid API key provided content: text/plain: schema: type: string application/json: schema: nullable: true '403': description: The API key doesn’t have permissions to perform the request content: text/plain: schema: type: string application/json: schema: nullable: true '429': description: Too many requests hit the API too quickly. We recommend an exponential backoff of your requests headers: Retry-After: schema: type: string content: text/plain: schema: type: string application/json: schema: nullable: true '500': description: Something went wrong on Braintrust's end. (These are rare.) content: text/plain: schema: type: string application/json: schema: nullable: true /v1/dataset/{dataset_id}/fetch: post: operationId: postDatasetIdFetch tags: - Datasets description: Fetch the events in a dataset. Equivalent to the GET form of the same path, but with the parameters in the request body rather than in the URL query. For more complex queries, use the `POST /btql` endpoint. summary: Fetch dataset (POST form) security: - bearerAuth: [] - {} parameters: - $ref: '#/components/parameters/DatasetIdParam' requestBody: description: Filters for the fetch query required: false content: application/json: schema: $ref: '#/components/schemas/FetchEventsRequest' responses: '200': description: Returns the fetched rows content: application/json: schema: $ref: '#/components/schemas/FetchDatasetEventsResponse' '400': description: The request was unacceptable, often due to missing a required parameter content: text/plain: schema: type: string application/json: schema: nullable: true '401': description: No valid API key provided content: text/plain: schema: type: string application/json: schema: nullable: true '403': description: The API key doesn’t have permissions to perform the request content: text/plain: schema: type: string application/json: schema: nullable: true '429': description: Too many requests hit the API too quickly. We recommend an exponential backoff of your requests headers: Retry-After: schema: type: string content: text/plain: schema: type: string application/json: schema: nullable: true '500': description: Something went wrong on Braintrust's end. (These are rare.) content: text/plain: schema: type: string application/json: schema: nullable: true get: operationId: getDatasetIdFetch tags: - Datasets description: Fetch the events in a dataset. Equivalent to the POST form of the same path, but with the parameters in the URL query rather than in the request body. For more complex queries, use the `POST /btql` endpoint. summary: Fetch dataset (GET form) security: - bearerAuth: [] - {} parameters: - $ref: '#/components/parameters/DatasetIdParam' - $ref: '#/components/parameters/FetchLimitParam' - $ref: '#/components/parameters/MaxXactId' - $ref: '#/components/parameters/MaxRootSpanId' - $ref: '#/components/parameters/Version' responses: '200': description: Returns the fetched rows content: application/json: schema: $ref: '#/components/schemas/FetchDatasetEventsResponse' '400': description: The request was unacceptable, often due to missing a required parameter content: text/plain: schema: type: string application/json: schema: nullable: true '401': description: No valid API key provided content: text/plain: schema: type: string application/json: schema: nullable: true '403': description: The API key doesn’t have permissions to perform the request content: text/plain: schema: type: string application/json: schema: nullable: true '429': description: Too many requests hit the API too quickly. We recommend an exponential backoff of your requests headers: Retry-After: schema: type: string content: text/plain: schema: type: string application/json: schema: nullable: true '500': description: Something went wrong on Braintrust's end. (These are rare.) content: text/plain: schema: type: string application/json: schema: nullable: true /v1/dataset/{dataset_id}/feedback: post: operationId: postDatasetIdFeedback tags: - Datasets description: Log feedback for a set of dataset events summary: Feedback for dataset events security: - bearerAuth: [] - {} parameters: - $ref: '#/components/parameters/DatasetIdParam' requestBody: description: An array of feedback objects required: false content: application/json: schema: $ref: '#/components/schemas/FeedbackDatasetEventRequest' responses: '200': description: Returns a success status content: application/json: schema: $ref: '#/components/schemas/FeedbackResponseSchema' '400': description: The request was unacceptable, often due to missing a required parameter content: text/plain: schema: type: string application/json: schema: nullable: true '401': description: No valid API key provided content: text/plain: schema: type: string application/json: schema: nullable: true '403': description: The API key doesn’t have permissions to perform the request content: text/plain: schema: type: string application/json: schema: nullable: true '429': description: Too many requests hit the API too quickly. We recommend an exponential backoff of your requests headers: Retry-After: schema: type: string content: text/plain: schema: type: string application/json: schema: nullable: true '500': description: Something went wrong on Braintrust's end. (These are rare.) content: text/plain: schema: type: string application/json: schema: nullable: true /v1/dataset/{dataset_id}/summarize: get: operationId: getDatasetIdSummarize tags: - Datasets description: Summarize dataset summary: Summarize dataset security: - bearerAuth: [] parameters: - $ref: '#/components/parameters/DatasetIdParam' - $ref: '#/components/parameters/SummarizeData' responses: '200': description: Dataset summary content: application/json: schema: $ref: '#/components/schemas/SummarizeDatasetResponse' '400': description: The request was unacceptable, often due to missing a required parameter content: text/plain: schema: type: string application/json: schema: nullable: true '401': description: No valid API key provided content: text/plain: schema: type: string application/json: schema: nullable: true '403': description: The API key doesn’t have permissions to perform the request content: text/plain: schema: type: string application/json: schema: nullable: true '429': description: Too many requests hit the API too quickly. We recommend an exponential backoff of your requests headers: Retry-After: schema: type: string content: text/plain: schema: type: string application/json: schema: nullable: true '500': description: Something went wrong on Braintrust's end. (These are rare.) content: text/plain: schema: type: string application/json: schema: nullable: true components: parameters: EndingBefore: schema: $ref: '#/components/schemas/EndingBefore' required: false description: 'Pagination cursor id. For example, if the initial item in the last page you fetched had an id of `foo`, pass `ending_before=foo` to fetch the previous page. Note: you may only pass one of `starting_after` and `ending_before`' name: ending_before in: query DatasetIdParam: schema: $ref: '#/components/schemas/DatasetIdParam' required: true description: Dataset id name: dataset_id in: path StartingAfter: schema: $ref: '#/components/schemas/StartingAfter' required: false description: 'Pagination cursor id. For example, if the final item in the last page you fetched had an id of `foo`, pass `starting_after=foo` to fetch the next page. Note: you may only pass one of `starting_after` and `ending_before`' name: starting_after in: query AppLimitParam: schema: $ref: '#/components/schemas/AppLimitParam' required: false description: Limit the number of objects to return name: limit in: query DatasetName: schema: $ref: '#/components/schemas/DatasetName' required: false description: Name of the dataset to search for name: dataset_name in: query allowReserved: true FetchLimitParam: schema: $ref: '#/components/schemas/FetchLimitParam' required: false description: 'limit the number of traces fetched Fetch queries may be paginated if the total result size is expected to be large (e.g. project_logs which accumulate over a long time). Note that fetch queries only support pagination in descending time order (from latest to earliest `_xact_id`. Furthermore, later pages may return rows which showed up in earlier pages, except with an earlier `_xact_id`. This happens because pagination occurs over the whole version history of the event log. You will most likely want to exclude any such duplicate, outdated rows (by `id`) from your combined result set. The `limit` parameter controls the number of full traces to return. So you may end up with more individual rows than the specified limit if you are fetching events containing traces.' name: limit in: query Version: schema: $ref: '#/components/schemas/Version' required: false description: 'Retrieve a snapshot of events from a past time The version id is essentially a filter on the latest event transaction id. You can use the `max_xact_id` returned by a past fetch as the version to reproduce that exact fetch.' name: version in: query Ids: schema: $ref: '#/components/schemas/Ids' required: false description: Filter search results to a particular set of object IDs. To specify a list of IDs, include the query param multiple times name: ids in: query MaxXactId: schema: $ref: '#/components/schemas/MaxXactId' required: false description: 'DEPRECATION NOTICE: The manually-constructed pagination cursor is deprecated in favor of the explicit ''cursor'' returned by object fetch requests. Please prefer the ''cursor'' argument going forwards. Together, `max_xact_id` and `max_root_span_id` form a pagination cursor Since a paginated fetch query returns results in order from latest to earliest, the cursor for the next page can be found as the row with the minimum (earliest) value of the tuple `(_xact_id, root_span_id)`. See the documentation of `limit` for an overview of paginating fetch queries.' name: max_xact_id in: query ProjectIdQuery: schema: $ref: '#/components/schemas/ProjectIdQuery' required: false description: Project id name: project_id in: query OrgName: schema: $ref: '#/components/schemas/OrgName' required: false description: Filter search results to within a particular organization name: org_name in: query allowReserved: true ProjectName: schema: $ref: '#/components/schemas/ProjectName' required: false description: Name of the project to search for name: project_name in: query allowReserved: true SummarizeData: schema: $ref: '#/components/schemas/SummarizeData' required: false description: Whether to summarize the data. If false (or omitted), only the metadata will be returned. name: summarize_data in: query MaxRootSpanId: schema: $ref: '#/components/schemas/MaxRootSpanId' required: false description: 'DEPRECATION NOTICE: The manually-constructed pagination cursor is deprecated in favor of the explicit ''cursor'' returned by object fetch requests. Please prefer the ''cursor'' argument going forwards. Together, `max_xact_id` and `max_root_span_id` form a pagination cursor Since a paginated fetch query returns results in order from latest to earliest, the cursor for the next page can be found as the row with the minimum (earliest) value of the tuple `(_xact_id, root_span_id)`. See the documentation of `limit` for an overview of paginating fetch queries.' name: max_root_span_id in: query schemas: InsertDatasetEvent: type: object properties: input: nullable: true description: The argument that uniquely define an input case (an arbitrary, JSON serializable object) expected: nullable: true description: The output of your application, including post-processing (an arbitrary, JSON serializable object) metadata: type: object nullable: true properties: model: type: string nullable: true description: The model used for this example additionalProperties: nullable: true description: A dictionary with additional data about the test example, model outputs, or just about anything else that's relevant, that you can use to help find and analyze examples later. For example, you could log the `prompt`, example's `id`, or anything else that would be useful to slice/dice later. The values in `metadata` can be any JSON-serializable type, but its keys must be strings tags: type: array nullable: true items: type: string description: A list of tags to log id: type: string nullable: true description: A unique identifier for the dataset event. If you don't provide one, Braintrust will generate one for you created: type: string nullable: true format: date-time description: The timestamp the dataset event was created origin: $ref: '#/components/schemas/ObjectReferenceNullish' facets: type: object nullable: true additionalProperties: nullable: true description: Facets for categorization (dictionary from facet id to value) _object_delete: type: boolean nullable: true description: Pass `_object_delete=true` to mark the dataset event deleted. Deleted events will not show up in subsequent fetches for this dataset _is_merge: type: boolean nullable: true description: 'The `_is_merge` field controls how the row is merged with any existing row with the same id in the DB. By default (or when set to `false`), the existing row is completely replaced by the new row. When set to `true`, the new row is deep-merged into the existing row, if one is found. If no existing row is found, the new row is inserted as is. For example, say there is an existing row in the DB `{"id": "foo", "input": {"a": 5, "b": 10}}`. If we merge a new row as `{"_is_merge": true, "id": "foo", "input": {"b": 11, "c": 20}}`, the new row will be `{"id": "foo", "input": {"a": 5, "b": 11, "c": 20}}`. If we replace the new row as `{"id": "foo", "input": {"b": 11, "c": 20}}`, the new row will be `{"id": "foo", "input": {"b": 11, "c": 20}}`' _merge_paths: type: array nullable: true items: type: array items: type: string description: 'The `_merge_paths` field allows controlling the depth of the merge, when `_is_merge=true`. `_merge_paths` is a list of paths, where each path is a list of field names. The deep merge will not descend below any of the specified merge paths. For example, say there is an existing row in the DB `{"id": "foo", "input": {"a": {"b": 10}, "c": {"d": 20}}, "output": {"a": 20}}`. If we merge a new row as `{"_is_merge": true, "_merge_paths": [["input", "a"], ["output"]], "input": {"a": {"q": 30}, "c": {"e": 30}, "bar": "baz"}, "output": {"d": 40}}`, the new row will be `{"id": "foo": "input": {"a": {"q": 30}, "c": {"d": 20, "e": 30}, "bar": "baz"}, "output": {"d": 40}}`. In this case, due to the merge paths, we have replaced `input.a` and `output`, but have still deep-merged `input` and `input.c`.' _array_delete: type: array nullable: true items: type: object properties: path: type: array items: type: string delete: type: array items: nullable: true required: - path - delete description: 'The `_array_delete` field allows removing specific values from array fields. It is an array of objects with `path` and `delete` properties. For example, to remove tags "foo" and "bar" from an existing row: `{"_is_merge": true, "_array_delete": [{"path": ["tags"], "delete": ["foo", "bar"]}]}`. For nested fields like `metadata.categories`, use `[{"path": ["metadata", "categories"], "delete": ["value"]}]`. This will remove those specific values from the array while preserving others.' _parent_id: type: string nullable: true description: 'DEPRECATED: The `_parent_id` field is deprecated and should not be used. Support for `_parent_id` will be dropped in a future version of Braintrust. Log `span_id`, `root_span_id`, and `span_parents` explicitly instead. Use the `_parent_id` field to create this row as a subspan of an existing row. Tracking hierarchical relationships are important for tracing (see the [guide](https://www.braintrust.dev/docs/instrument) for full details). For example, say we have logged a row `{"id": "abc", "input": "foo", "output": "bar", "expected": "boo", "scores": {"correctness": 0.33}}`. We can create a sub-span of the parent row by logging `{"_parent_id": "abc", "id": "llm_call", "input": {"prompt": "What comes after foo?"}, "output": "bar", "metrics": {"tokens": 1}}`. In the webapp, only the root span row `"abc"` will show up in the summary view. You can view the full trace hierarchy (in this case, the `"llm_call"` row) by clicking on the "abc" row. If the row is being merged into an existing row, this field will be ignored.' span_id: type: string nullable: true description: 'Use `span_id`, `root_span_id`, and `span_parents` instead of `_parent_id`, which is now deprecated. The span_id is a unique identifier describing the row''s place in the a trace, and the root_span_id is a unique identifier for the whole trace. See the [guide](https://www.braintrust.dev/docs/instrument) for full details. For example, say we have logged a row `{"id": "abc", "span_id": "span0", "root_span_id": "root_span0", "input": "foo", "output": "bar", "expected": "boo", "scores": {"correctness": 0.33}}`. We can create a sub-span of the parent row by logging `{"id": "llm_call", "span_id": "span1", "root_span_id": "root_span0", "span_parents": ["span0"], "input": {"prompt": "What comes after foo?"}, "output": "bar", "metrics": {"tokens": 1}}`. In the webapp, only the root span row `"abc"` will show up in the summary view. You can view the full trace hierarchy (in this case, the `"llm_call"` row) by clicking on the "abc" row. If the row is being merged into an existing row, this field will be ignored.' root_span_id: type: string nullable: true description: 'Use `span_id`, `root_span_id`, and `span_parents` instead of `_parent_id`, which is now deprecated. The span_id is a unique identifier describing the row''s place in the a trace, and the root_span_id is a unique identifier for the whole trace. See the [guide](https://www.braintrust.dev/docs/instrument) for full details. For example, say we have logged a row `{"id": "abc", "span_id": "span0", "root_span_id": "root_span0", "input": "foo", "output": "bar", "expected": "boo", "scores": {"correctness": 0.33}}`. We can create a sub-span of the parent row by logging `{"id": "llm_call", "span_id": "span1", "root_span_id": "root_span0", "span_parents": ["span0"], "input": {"prompt": "What comes after foo?"}, "output": "bar", "metrics": {"tokens": 1}}`. In the webapp, only the root span row `"abc"` will show up in the summary view. You can view the full trace hierarchy (in this case, the `"llm_call"` row) by clicking on the "abc" row. If the row is being merged into an existing row, this field will be ignored.' span_parents: type: array nullable: true items: type: string description: 'Use `span_id`, `root_span_id`, and `span_parents` instead of `_parent_id`, which is now deprecated. The span_id is a unique identifier describing the row''s place in the a trace, and the root_span_id is a unique identifier for the whole trace. See the [guide](https://www.braintrust.dev/docs/instrument) for full details. For example, say we have logged a row `{"id": "abc", "span_id": "span0", "root_span_id": "root_span0", "input": "foo", "output": "bar", "expected": "boo", "scores": {"correctness": 0.33}}`. We can create a sub-span of the parent row by logging `{"id": "llm_call", "span_id": "span1", "root_span_id": "root_span0", "span_parents": ["span0"], "input": {"prompt": "What comes after foo?"}, "output": "bar", "metrics": {"tokens": 1}}`. In the webapp, only the root span row `"abc"` will show up in the summary view. You can view the full trace hierarchy (in this case, the `"llm_call"` row) by clicking on the "abc" row. If the row is being merged into an existing row, this field will be ignored.' description: A dataset event InsertDatasetEventRequest: type: object properties: events: type: array items: $ref: '#/components/schemas/InsertDatasetEvent' description: A list of dataset events to insert required: - events OrgName: type: string description: Filter search results to within a particular organization DataSummary: type: object nullable: true properties: total_records: type: integer minimum: 0 description: Total number of records in the dataset required: - total_records description: Summary of a dataset's data FetchEventsRequest: type: object properties: limit: $ref: '#/components/schemas/FetchLimit' cursor: $ref: '#/components/schemas/FetchPaginationCursor' max_xact_id: $ref: '#/components/schemas/MaxXactId' nullable: true max_root_span_id: $ref: '#/components/schemas/MaxRootSpanId' nullable: true version: $ref: '#/components/schemas/Version' nullable: true FetchPaginationCursor: type: string nullable: true description: 'An opaque string to be used as a cursor for the next page of results, in order from latest to earliest. The string can be obtained directly from the `cursor` property of the previous fetch query' FeedbackResponseSchema: type: object properties: status: type: string enum: - success required: - status DatasetEvent: type: object properties: id: type: string description: A unique identifier for the dataset event. If you don't provide one, Braintrust will generate one for you _xact_id: type: string description: The transaction id of an event is unique to the network operation that processed the event insertion. Transaction ids are monotonically increasing over time and can be used to retrieve a versioned snapshot of the dataset (see the `version` parameter) created: type: string format: date-time description: The timestamp the dataset event was created _pagination_key: type: string nullable: true description: A stable, time-ordered key that can be used to paginate over dataset events. This field is auto-generated by Braintrust and only exists in Brainstore. project_id: type: string format: uuid description: Unique identifier for the project that the dataset belongs under dataset_id: type: string format: uuid description: Unique identifier for the dataset input: nullable: true description: The argument that uniquely define an input case (an arbitrary, JSON serializable object) expected: nullable: true description: The output of your application, including post-processing (an arbitrary, JSON serializable object) metadata: type: object nullable: true properties: model: type: string nullable: true description: The model used for this example additionalProperties: nullable: true description: A dictionary with additional data about the test example, model outputs, or just about anything else that's relevant, that you can use to help find and analyze examples later. For example, you could log the `prompt`, example's `id`, or anything else that would be useful to slice/dice later. The values in `metadata` can be any JSON-serializable type, but its keys must be strings tags: type: array nullable: true items: type: string description: A list of tags to log span_id: type: string description: A unique identifier used to link different dataset events together as part of a full trace. See the [tracing guide](https://www.braintrust.dev/docs/instrument) for full details on tracing root_span_id: type: string description: A unique identifier for the trace this dataset event belongs to is_root: type: boolean nullable: true description: Whether this span is a root span origin: $ref: '#/components/schemas/ObjectReferenceNullish' comments: type: array nullable: true items: nullable: true description: Optional list of comments attached to this event audit_data: type: array nullable: true items: nullable: true description: Optional list of audit entries attached to this event facets: type: object nullable: true additionalProperties: nullable: true description: Facets for categorization (dictionary from facet id to value) classifications: type: object nullable: true additionalProperties: type: array items: type: object properties: id: type: string description: Stable classification identifier label: type: string description: Original label of the classification item, which is useful for search and indexing purposes confidence: type: number nullable: true description: Optional confidence score for the classification metadata: type: object nullable: true additionalProperties: nullable: true description: Optional metadata associated with the classification source: $ref: '#/components/schemas/SavedFunctionId' required: - id description: Classifications for this event (dictionary from classification name to items) required: - id - _xact_id - created - project_id - dataset_id - span_id - root_span_id CreateDataset: type: object properties: project_id: type: string format: uuid description: Unique identifier for the project that the dataset belongs under name: type: string minLength: 1 description: Name of the dataset. Within a project, dataset names are unique description: type: string nullable: true description: Textual description of the dataset tags: type: array nullable: true items: type: string description: A list of tags for the dataset metadata: type: object nullable: true additionalProperties: nullable: true description: User-controlled metadata about the dataset required: - project_id - name FetchLimit: type: integer nullable: true minimum: 0 description: 'limit the number of traces fetched Fetch queries may be paginated if the total result size is expected to be large (e.g. project_logs which accumulate over a long time). Note that fetch queries only support pagination in descending time order (from latest to earliest `_xact_id`. Furthermore, later pages may return rows which showed up in earlier pages, except with an earlier `_xact_id`. This happens because pagination occurs over the whole version history of the event log. You will most likely want to exclude any such duplicate, outdated rows (by `id`) from your combined result set. The `limit` parameter controls the number of full traces to return. So you may end up with more individual rows than the specified limit if you are fetching events containing traces.' FetchLimitParam: type: integer nullable: true minimum: 0 description: 'limit the number of traces fetched Fetch queries may be paginated if the total result size is expected to be large (e.g. project_logs which accumulate over a long time). Note that fetch queries only support pagination in descending time order (from latest to earliest `_xact_id`. Furthermore, later pages may return rows which showed up in earlier pages, except with an earlier `_xact_id`. This happens because pagination occurs over the whole version history of the event log. You will most likely want to exclude any such duplicate, outdated rows (by `id`) from your combined result set. The `limit` parameter controls the number of full traces to return. So you may end up with more individual rows than the specified limit if you are fetching events containing traces.' MaxRootSpanId: type: string description: 'DEPRECATION NOTICE: The manually-constructed pagination cursor is deprecated in favor of the explicit ''cursor'' returned by object fetch requests. Please prefer the ''cursor'' argument going forwards. Together, `max_xact_id` and `max_root_span_id` form a pagination cursor Since a paginated fetch query returns results in order from latest to earliest, the cursor for the next page can be found as the row with the minimum (earliest) value of the tuple `(_xact_id, root_span_id)`. See the documentation of `limit` for an overview of paginating fetch queries.' FeedbackDatasetItem: type: object properties: id: type: string description: The id of the dataset event to log feedback for. This is the row `id` returned by `POST /v1/dataset/{dataset_id}/insert` comment: type: string nullable: true description: An optional comment string to log about the dataset event metadata: type: object nullable: true additionalProperties: nullable: true description: A dictionary with additional data about the feedback. If you have a `user_id`, you can log it here and access it in the Braintrust UI. Note, this metadata does not correspond to the main event itself, but rather the audit log attached to the event. source: type: string nullable: true enum: - app - api - external - null description: The source of the feedback. Must be one of "external" (default), "app", or "api" tags: type: array nullable: true items: type: string description: A list of tags to log required: - id PatchDataset: type: object properties: name: type: string nullable: true description: Name of the dataset. Within a project, dataset names are unique description: type: string nullable: true description: Textual description of the dataset tags: type: array nullable: true items: type: string description: A list of tags for the dataset metadata: type: object nullable: true additionalProperties: nullable: true description: User-controlled metadata about the dataset FetchDatasetEventsResponse: type: object properties: events: type: array items: $ref: '#/components/schemas/DatasetEvent' description: A list of fetched events cursor: type: string nullable: true description: 'Pagination cursor Pass this string directly as the `cursor` param to your next fetch request to get the next page of results. Not provided if the returned result set is empty.' required: - events AppLimitParam: type: integer nullable: true minimum: 0 description: Limit the number of objects to return SummarizeDatasetResponse: type: object properties: project_name: type: string description: Name of the project that the dataset belongs to dataset_name: type: string description: Name of the dataset project_url: type: string format: uri description: URL to the project's page in the Braintrust app dataset_url: type: string format: uri description: URL to the dataset's page in the Braintrust app data_summary: $ref: '#/components/schemas/DataSummary' required: - project_name - dataset_name - project_url - dataset_url description: Summary of a dataset DatasetIdParam: type: string format: uuid description: Dataset id Ids: anyOf: - type: string format: uuid - type: array items: type: string format: uuid description: Filter search results to a particular set of object IDs. To specify a list of IDs, include the query param multiple times ProjectName: type: string description: Name of the project to search for SummarizeData: type: boolean nullable: true description: Whether to summarize the data. If false (or omitted), only the metadata will be returned. FunctionTypeEnum: type: string enum: - llm - scorer - task - tool - custom_view - preprocessor - facet - classifier - tag - parameters - sandbox - null default: scorer description: The type of global function. Defaults to 'scorer'. Version: type: string description: 'Retrieve a snapshot of events from a past time The version id is essentially a filter on the latest event transaction id. You can use the `max_xact_id` returned by a past fetch as the version to reproduce that exact fetch.' StartingAfter: type: string format: uuid description: 'Pagination cursor id. For example, if the final item in the last page you fetched had an id of `foo`, pass `starting_after=foo` to fetch the next page. Note: you may only pass one of `starting_after` and `ending_before`' ProjectIdQuery: type: string format: uuid description: Project id Dataset: type: object properties: id: type: string format: uuid description: Unique identifier for the dataset project_id: type: string format: uuid description: Unique identifier for the project that the dataset belongs under name: type: string description: Name of the dataset. Within a project, dataset names are unique description: type: string nullable: true description: Textual description of the dataset created: type: string nullable: true format: date-time description: Date of dataset creation deleted_at: type: string nullable: true format: date-time description: Date of dataset deletion, or null if the dataset is still active user_id: type: string nullable: true format: uuid description: Identifies the user who created the dataset tags: type: array nullable: true items: type: string description: A list of tags for the dataset metadata: type: object nullable: true additionalProperties: nullable: true description: User-controlled metadata about the dataset url_slug: type: string description: URL slug for the dataset. used to construct dataset URLs required: - id - project_id - name - url_slug ObjectReferenceNullish: type: object nullable: true properties: object_type: type: string enum: - project_logs - experiment - dataset - prompt - function - prompt_session description: Type of the object the event is originating from. object_id: type: string format: uuid description: ID of the object the event is originating from. id: type: string description: ID of the original event. _xact_id: type: string nullable: true description: Transaction ID of the original event. created: type: string nullable: true description: Created timestamp of the original event. Used to help sort in the UI required: - object_type - object_id - id description: Indicates the event was copied from another object. DatasetName: type: string description: Name of the dataset to search for MaxXactId: type: string description: 'DEPRECATION NOTICE: The manually-constructed pagination cursor is deprecated in favor of the explicit ''cursor'' returned by object fetch requests. Please prefer the ''cursor'' argument going forwards. Together, `max_xact_id` and `max_root_span_id` form a pagination cursor Since a paginated fetch query returns results in order from latest to earliest, the cursor for the next page can be found as the row with the minimum (earliest) value of the tuple `(_xact_id, root_span_id)`. See the documentation of `limit` for an overview of paginating fetch queries.' SavedFunctionId: anyOf: - type: object properties: type: type: string enum: - function id: type: string version: type: string description: The version of the function required: - type - id title: function - type: object properties: type: type: string enum: - global name: type: string function_type: $ref: '#/components/schemas/FunctionTypeEnum' required: - type - name title: global - type: 'null' description: Optional function identifier that produced the classification FeedbackDatasetEventRequest: type: object properties: feedback: type: array items: $ref: '#/components/schemas/FeedbackDatasetItem' description: A list of dataset feedback items required: - feedback EndingBefore: type: string format: uuid description: 'Pagination cursor id. For example, if the initial item in the last page you fetched had an id of `foo`, pass `ending_before=foo` to fetch the previous page. Note: you may only pass one of `starting_after` and `ending_before`' InsertEventsResponse: type: object properties: row_ids: type: array items: type: string description: The ids of all rows that were inserted, aligning one-to-one with the rows provided as input required: - row_ids securitySchemes: bearerAuth: type: http scheme: bearer bearerFormat: API key or JWT description: 'Most Braintrust endpoints are authenticated by providing your API key as a header `Authorization: Bearer [api_key]` to your HTTP request. You can create an API key in the Braintrust [organization settings page](https://www.braintrustdata.com/app/settings?subroute=api-keys).'