NEWVectors or files. Pick a path.Start →
    Models/Image Text To Text/stepfun-ai/step3
    Image Text To Texttransformersapache-2.0

    step3

    by stepfun-ai

    Identifier
    Model ID
    stepfun-ai/step3

    Deploy step3

    Single-tenant

    Mixpeek has no managed extractor for this model. On a single-tenant deployment you upload the weights and a custom plugin serves them next to the rest of your pipeline.

    Tags

    transformerssafetensorsstep3_vltext-generationimage-text-to-textconversationalcustom_codearxiv:2507.19427license:apache-2.0endpoints_compatibleregion:us

    Use step3 on Mixpeek

    Build multimodal processing pipelines with this model and others. Extract features, run inference, and set up retrieval in Mixpeek Studio.

    Open Studio

    How It Runs on Mixpeek

    On Mixpeek, step3 runs as a managed extractor inside a processing pipeline. Point a bucket of image text to text data at it, and Mixpeek handles GPU provisioning, batching, retries, and writing the outputs into a vector store you can query.

    Extractor outputs land in the Mixpeek Vector Store (MVS), where you can combine them with retrieval, reranking, and filter stages to build end-to-end search and agent-perception pipelines, no model-serving infrastructure to maintain.