NEWVectors or files. Pick a path.Start →

    Feature Extractors

    After your data is connected, extractors run in parallel to pull out structured features, embeddings, entities, transcripts, and more.

    11 production extractors, each with a README, a live schema, and a Studio path

    Web Scraper + Multimodal Embeddings

    Crawl sites (docs, job boards, news, SPAs) and extract text, code & image embeddings in one pass.

    Text

    Text Embeddings (E5-Large)

    Multilingual dense text embeddings with E5-Large: semantic search & RAG out of the box.

    Text

    Image Embeddings (SigLIP)

    Dense 768-D image embeddings with Google SigLIP: text-to-image search in one contrastive space.

    Image

    Multimodal Video/Audio/Image (Vertex v1 · Gemini v2)

    Unified embeddings for video, audio, image & text: FFmpeg scene/silence chunking, Whisper transcription, thumbnails.

    Multimodal

    Universal All-in-One (Gemini)

    One extractor for image, video, audio & documents: auto-detects modality and applies the right pipeline.

    Multimodal

    Multi-File Object Embeddings (Gemini)

    Embed ALL files of an object (images, PDFs, video, audio, text) into one 3072-D Gemini vector.

    Multimodal

    Document Layout Graph

    Decompose PDFs into spatial blocks: paragraphs, tables, forms, headers: with layout classification & confidence.

    Document

    Passthrough (Storage Only)

    Store and canonicalize objects with zero ML: metadata-only ingestion.

    Utility

    Scrolling/Marquee Text OCR

    Reads scrolling video text via phase-correlation band detection, panoramic stitching, and VLM OCR.

    Video

    Face Identity (SCRFD + ArcFace)

    Production face recognition: detect, align, and embed faces to 512-D ArcFace vectors across image, video & PDF.

    Image

    Audio Fingerprinting (CLAP)

    512-D audio embeddings with CLAP: content-based audio search and matching from files or video tracks.

    Audio

    What's new in extractors

    Full changelog