v1

latestOpenAPI 3.0.3BSL2026-07-17133415718.8 KB
Message

Regenerate message

Regenerate the assistant response to the last user message of a topic. This will delete the last message and replace it with a new message. The response will include Chunks first on the stream if the topic is using RAG. The structure will look like [chunks]||mesage. See docs.trieve.ai for more information. Auth'ed user or api key must have an admin or owner role for the specified dataset's organization.

delete/api/message

Headers

TR-Datasetstring uuid required

The dataset id or tracking_id to use for the request. We assume you intend to use an id if the value is a valid uuid.

Request body

concat_user_messages_queryboolean nullable

If concat user messages query is set to true, all of the user messages in the topic will be concatenated together and used as the search query. If not specified, this defaults to false. Default is false.

currencystring nullable

The currency symbol to use for the completion. If not specified, this defaults to "$".

{"stackTrail":"components:schemas:RegenerateMessageReqPayload:properties:metadata","oasType":"schema","type":"unknown","description":"Metadata is any metadata you want to associate w/ the event that is created from this request","nullable":true}
modelstring nullable

Model name to use for the completion. If not specified, this defaults to the dataset's model.

no_result_messagestring nullable

No result message for when there are no chunks found above the score threshold.

number_of_messages_to_includeinteger nullable

Number of messages to include in the context window. If not specified, this defaults to 10.

only_include_docs_usedboolean nullable

Only include docs used is a boolean that indicates whether or not to only include the docs that were used in the completion. If true, the completion will only include the docs that were used in the completion. If false, the completion will include all of the docs.

page_sizeinteger nullable

Page size is the number of chunks to fetch during RAG. If 0, then no search will be performed. If specified, this will override the N retrievals to include in the dataset configuration. Default is None.

rag_contextstring nullable

Overrides what the way chunks are placed into the context window

remove_stop_wordsboolean nullable

If true, stop words (specified in server/src/stop-words.txt in the git repo) will be removed. Queries that are entirely stop words will be preserved.

score_thresholdnumber float nullable

Set score_threshold to a float to filter out chunks with a score below the threshold. This threshold applies before weight and bias modifications. If not specified, this defaults to 0.0.

search_querystring nullable

Query is the search query. This can be any string. The search_query will be used to create a dense embedding vector and/or sparse vector which will be used to find the result set. If not specified, will default to the last user message or HyDE if HyDE is enabled in the dataset configuration. Default is None.

search_type'fulltext' | 'semantic' | 'hybrid' | 'bm25'
topic_idstring uuid required

The id of the topic to regenerate the last message for.

use_agentic_searchboolean nullable

If true, the search will be conducted using llm tool calling. If not specified, this defaults to false.

use_group_searchboolean nullable

If use_group_search is set to true, the search will be conducted using the search_over_groups api. If not specified, this defaults to false.

use_quote_negated_termsboolean nullable

If true, quoted and - prefixed words will be parsed from the queries and used as required and negated words respectively. Default is false.

user_idstring nullable

The user_id is the id of the user who is making the request. This is used to track user interactions with the RAG results.

Example request

{
  "filters": {
    "must": [
      {
        "field": "tag_set",
        "match_all": [
          "A",
          "B"
        ]
      },
      {
        "field": "num_value",
        "range": {
          "gte": 10,
          "lte": 25
        }
      }
    ]
  },
  "llm_options": {
    "image_config": {
      "images_per_chunk": 1,
      "use_images": true
    }
  }
}

Response

This will be a JSON response of a string containing the LLM's generated inference. Response if not streaming.