Text Splitter

processing.text-splitter Processing v0.1.0

Splits long text into overlapping chunks for RAG ingestion — by characters (a sliding window), paragraphs, or sentences — carrying an overlap between chunks. Reads the text from an input field (blank = the whole payload) and emits a string array that feeds straight into the Embeddings and Vector Store nodes.

Finding it in the library

Search the builder's node library for Text Splitter (it lives under Processing). A single click opens the in-editor docs panel shown here — description, ports, and every property, without leaving the canvas. Double-click (or drag) to add it to the workflow.

Library capture pending — regenerate with npm run shots:nodes.

Wired up in the builder

Text Splitter in a real, runnable flow — captured live from the Studio editor, exactly as it looks on your canvas. This is the same workflow used for the example input & output below.

Canvas capture pending — regenerate with npm run shots:nodes.

How it’s configured

The node’s Configure panel as it opens in the builder when you select the step — every setting laid out with real values. Click any field to edit it.

Config capture pending — regenerate with npm run shots:nodes. See the full property reference below.

Ports

Ports are the node’s contract with its neighbours. In the editor a port label renders bold when wired and italic when optional; ports accept attachment carriers rather than data wires.

DirectionPortLabelWhat flows through it
InputinputInput
OutputoutputChunks

How data flows through it

Text Splitter consumes the content of the incoming envelope — when it is fed directly by a trigger, the trigger’s wrapper is unwrapped at the node boundary so the node sees the actual data, not the metadata shell. Its output becomes the payload for the next node, while the envelope (trace ids, correlation, binary refs) rides along untouched. In the Runs view you always see the whole envelope for both sides of this node.

Expressions in the config

None of this node’s properties are string-typed, so {{ }} expressions don’t apply here — JSON- and code-typed fields are always taken literally.

Build it with AI

Every node in this reference is reachable through Flowdrome’s AI Copilot and the MCP tools — say what you want, and the graph surgery happens server-side. Node types resolve fuzzily, so the catalog label (Text Splitter) works as well as the exact type id (processing.text-splitter).

In the Copilot panel (or any connected AI):

add a text splitter node after the trigger

As a step in a create_chain_workflow call:

{"type":"Text Splitter","config":{}}
Raw MCP call — add this node to a workflow with add_node
curl -s -X POST http://localhost:48170/mcp -H "content-type: application/json" -d '{ "jsonrpc": "2.0", "id": "1", "method": "tools/call", "params": { "name": "add_node", "arguments": { "workflowId": "<id>", "type": "Text Splitter" } } }'

Example input & output

Captured from a real test run of the workflow above — this is what you see in the run data panel after pressing Test workflow.

Input — what the node received

{
  "kind": "json",
  "contentType": "application/json",
  "payload": {
    "text": "abcdefghijklmnopqrstuvwxyz"
  },
  "body": {
    "text": "abcdefghijklmnopqrstuvwxyz"
  }
}

Output — what the node produced

{
  "chunks": [
    "abcdefghij",
    "hijklmnopq",
    "opqrstuvwx",
    "vwxyz"
  ],
  "count": 4
}

Property reference

Every setting, with its type and default — the same fields shown configured in the panel above.

PropertyTypeDefaultDescription
Strategy
strategy
select "characters" How to split: characters (a sliding window of characters), paragraphs (group blank-line paragraphs), or sentences (group by sentence terminators).
charactersparagraphssentences
Chunk size
chunkSize
int 1000 Maximum number of characters per chunk.
Chunk overlap
chunkOverlap
int 200 Characters carried from the end of one chunk into the start of the next (clamped below the chunk size).
Text field
textField
field "" Dot-path to the text to split. Blank = the whole payload (stringified).
Output field
outputField
field "chunks" Key the chunk array is written to on the output.

Related nodes

The rest of the Processing group — the same folder you’d scan in the editor’s library.

This page is generated from the node registry by gen-node-docs.mjs on every site build — ports, properties, defaults and visibility rules cannot drift from the code. The screenshots and example data are captured from a live Flowdrome by npm run shots:nodes and npm run gen:examples.