> ## Documentation Index
> Fetch the complete documentation index at: https://docs.veri.studio/llms.txt
> Use this file to discover all available pages before exploring further.

# List experiment items



## OpenAPI

````yaml /api-reference/openapi.json get /v1/experiments/{experiment_id}/items
openapi: 3.1.0
info:
  title: Veri API
  description: >-
    REST API for the Veri RL post-training platform. All requests require a
    Bearer API key (`vk_` prefix).
  license:
    name: ''
  version: 0.1.0
servers:
  - url: https://api.veri.studio
    description: Production
security: []
tags:
  - name: Training jobs
    description: Create, monitor, and manage training jobs.
  - name: Datasets
    description: Upload and connect training datasets.
  - name: Deployments
    description: Serve trained models and run inference.
  - name: Volumes
    description: Persistent file storage mounted into jobs.
  - name: Models
    description: Custom model registry deployments serve from.
  - name: Regions
    description: Discover available launch regions.
  - name: GPU
    description: Live GPU availability by provider and region.
  - name: Code artifacts
    description: Upload custom training script bundles.
  - name: Billing
    description: Credit balance and transaction history.
  - name: API keys
    description: Create and revoke API keys.
  - name: Account
    description: The authenticated caller's identity.
  - name: Settings
    description: Account-level integrations (Weights & Biases).
  - name: Metrics
    description: Prometheus metrics export for your own observability stack.
  - name: Evaluators
    description: 'Evaluators: versioned scoring rules (LLM judge, code, human).'
  - name: Experiments
    description: >-
      Experiments: offline runs of a pinned dataset snapshot through a target,
      scored by pinned evaluators.
  - name: Annotation queues
    description: >-
      Human review: queue traces, threads and experiment items, reserve one at a
      time, score them and feed corrections back into datasets.
  - name: Monitors
    description: >-
      Monitors: evaluators scoring a sampled share of a deployment's live
      traffic.
  - name: Observability
    description: Agent conversations, agents and their traffic.
  - name: Public runs
    description: Unauthenticated reads of runs their owners published to Explore.
paths:
  /v1/experiments/{experiment_id}/items:
    get:
      tags:
        - Experiments
      summary: List experiment items
      operationId: items
      parameters:
        - name: experiment_id
          in: path
          description: Experiment id (exp_...)
          required: true
          schema:
            type: string
        - name: status
          in: query
          description: pending | running | done | error
          required: false
          schema:
            type: string
        - name: pending_local
          in: query
          description: |-
            Only finished items that still owe a local code evaluator's score,
            each with the owed scores and their args (what the SDK harness runs
            before POST /v1/scores).
          required: false
          schema:
            type: boolean
        - name: limit
          in: query
          description: Page size, 1..=500 (default 100).
          required: false
          schema:
            type: integer
            format: int64
        - name: after
          in: query
          description: 'Cursor: items after this item id (`<row_id>.<trial>`), in row order.'
          required: false
          schema:
            type: string
      responses:
        '200':
          description: A page of items in row order, each with its scores
          content:
            application/json:
              schema:
                $ref: '#/components/schemas/PaginatedList_ExperimentItemResponse'
        '400':
          description: Bad cursor, or evals disabled
        '404':
          description: Not found in the caller's workspace
      security:
        - bearerAuth: []
components:
  schemas:
    PaginatedList_ExperimentItemResponse:
      type: object
      required:
        - object
        - data
        - has_more
      properties:
        object:
          type: string
        data:
          type: array
          items:
            type: object
            description: One (row, trial) with its output and scores.
            required:
              - object
              - experiment_id
              - item_id
              - row_id
              - trial
              - kind
              - status
              - request_ids
              - attempts
              - scores
            properties:
              object:
                type: string
                description: Always "experiment_item".
              experiment_id:
                type: string
              item_id:
                type: string
                description: '`<row_id>.<trial>`'
              row_id:
                type: integer
                format: int64
              trial:
                type: integer
                format: int32
              kind:
                type: string
                description: single_turn | thread
              turns:
                type:
                  - integer
                  - 'null'
                format: int32
                description: 'Thread items: the user turns replayed.'
              status:
                type: string
                description: pending | running | done | error
              input: {}
              expected: {}
              metadata: {}
              output:
                description: The target's answer (the last one, for a thread).
              transcript:
                description: The conversation, when the item has one.
              latency_ms:
                type:
                  - integer
                  - 'null'
                format: int32
              prompt_tokens:
                type:
                  - integer
                  - 'null'
                format: int32
              completion_tokens:
                type:
                  - integer
                  - 'null'
                format: int32
              request_ids:
                type: array
                items:
                  type: string
                description: The deployment requests this item made.
              error:
                type:
                  - string
                  - 'null'
              attempts:
                type: integer
                format: int32
              scores:
                type: array
                items:
                  $ref: '#/components/schemas/ScoreResponse'
              pending_local:
                type:
                  - array
                  - 'null'
                items:
                  $ref: '#/components/schemas/PendingLocalScore'
                description: >-
                  Only with `?pending_local=true`: the local code scores this
                  item

                  still owes, each with the args to run the evaluator on.
        has_more:
          type: boolean
    ScoreResponse:
      type: object
      description: One score.
      required:
        - object
        - id
        - evaluator_id
        - evaluator_version
        - name
        - target_type
        - value
        - source
        - status
        - created_at
      properties:
        object:
          type: string
          description: Always "score".
        id:
          type: string
          description: '`scr_...`'
        evaluator_id:
          type: string
        evaluator_version:
          type: integer
          format: int32
        name:
          type: string
          description: The evaluator's name at write time.
        target_type:
          type: string
          description: trace | thread
        request_id:
          type:
            - string
            - 'null'
          description: The scored request (trace targets). Trace id = `veri-<request_id>`.
        deployment_id:
          type:
            - string
            - 'null'
        thread_id:
          type:
            - string
            - 'null'
          description: |-
            The thread (thread targets; also stamped on trace targets whose
            request carried a thread id).
        target_revision:
          type:
            - string
            - 'null'
          description: 'Thread targets: the last request id the score covers.'
        experiment_id:
          type:
            - string
            - 'null'
          description: 'Experiment-item targets: the experiment.'
        item_id:
          type:
            - string
            - 'null'
          description: |-
            Experiment-item targets: `<row_id>.<trial>`, or
            `<row_id>.<trial>.t<turn>` for one turn of a replayed thread.
        value:
          description: boolean | label string | number; null for error and skipped scores.
        passed:
          type:
            - boolean
            - 'null'
          description: >-
            The evaluator version's pass rule applied to the value; null when
            the

            version has no pass rule or the score carries no value.
        reasoning:
          type:
            - string
            - 'null'
        source:
          type: string
          description: api | sdk | human | experiment | monitor
        status:
          type: string
          description: ok | error | skipped
        error:
          type:
            - string
            - 'null'
          description: Why an error / skipped score has no value.
        author:
          type:
            - string
            - 'null'
          description: The user who wrote the score.
        model_version_id:
          type:
            - string
            - 'null'
          description: The model version that answered the scored request, when known.
        monitor_id:
          type:
            - string
            - 'null'
          description: The monitor that wrote the score (source monitor).
        confidence:
          type:
            - number
            - 'null'
          format: double
          description: |-
            The judge's probability for its verdict, 0..1; null unless a
            decision-mode judge wrote the score.
        probabilities:
          type:
            - object
            - 'null'
          description: |-
            The Jev judge's distribution behind the value: noul
            `{"true": p, "false": 1 - p}`, choice `{option: p}`, score
            `{"<level>": p}`. Null for every other score.
        created_at:
          type: string
          format: date-time
    PendingLocalScore:
      type: object
      description: |-
        A score a local code evaluator still owes: run the pinned version's
        source with `args` (bindings applied) and POST /v1/scores with this
        item_id, the version and the source's sha256.
      required:
        - evaluator_id
        - evaluator_version
        - item_id
        - args
      properties:
        evaluator_id:
          type: string
        evaluator_version:
          type: integer
          format: int32
        item_id:
          type: string
        args:
          type: object
          description: '`{input, output, expected, metadata, transcript}` after bindings.'
  securitySchemes:
    bearerAuth:
      type: http
      scheme: bearer
      bearerFormat: API key
      description: API key with the `vk_` prefix. Create one from the dashboard.

````

This documentation is built and hosted on [Mintlify](https://mintlify.com), a developer documentation platform.