Docs
Tutorials
How-to guides
Reference
Explanation
Login
Free Sign Up
Docs
/
Flows
/
Node types
/
Text
Text
Templates, splitting, parsing, and content conversion.
Reference
Text - Content to Text
Converts rich content to serialised output. By default it serialises to plain text, and it can optionally serialise to HTML. Any included media is converted to pre-signed URLs for access.
Text - Extract List Items
The Text - Extract List Items node parses Markdown-formatted text and extracts items from the first list found in the document. It supports both ordered and unordered lists, including GitHub Flavored Markdown (GFM) syntax. Each list item is returned as a trimmed string in an array.
Text - Parse Document to Content
Parses PDF files into structured content. Extracts text with page-level granularity, generates page images, and splits content into chunks.
Text - Parse HTML to Content
Parses HTML string into rich content. Converts HTML elements into structured content nodes including paragraphs, headings, lists, links, and more. Optionally resolves relative URLs in the content.
Text - Parse Markdown to Content
Parses Markdown text into rich content. Converts Markdown elements into structured content nodes. Handles code blocks, including automatic stripping of Markdown code fences if wrapping the entire content.
Text - Recursive Split
The Text - Recursive Split node divides text into smaller chunks suitable for processing by language models or other text analysis tools. It splits on natural boundaries (paragraphs, lines, words) where possible, while respecting the configured chunk size. Overlapping between chunks ensures context is preserved across boundaries.
Text - Template
Renders a rich-content template with variable substitution. Insert variables through the template editor, then connect each generated input handle to the value it should render. It generates dynamic content such as messages, emails, or prompts.
Text HTML Query and Transform
The Text HTML Query and Transform node manipulates HTML through a sequence of configurable operations. It can query elements using CSS selectors to extract content (HTML, text, or attributes) and remove unwanted elements from the document. Each query operation creates a dynamic output handle for its results.