# Mixpeek: v1 Documentation

## Documentation

### Get started

- [Introduction](https://docs.mixpeek.com/docs/overview/introduction.md): Mixpeek is the semantic retrieval layer for unstructured and multimodal data
- [Quickstart](https://docs.mixpeek.com/docs/overview/quickstart.md): Start searching in under 60 seconds with BYO vectors, or index multimodal content with the managed platform
- [Explore the sample data](https://docs.mixpeek.com/docs/overview/sample-data.md): Run real multimodal queries against a live sample namespace — no account, no API key, no setup
- [Core Concepts](https://docs.mixpeek.com/docs/overview/concepts.md): Understand the building blocks that power Mixpeek
- [Data Model & Lineage](https://docs.mixpeek.com/docs/overview/data-model.md): Understand how objects transform into documents and how lineage tracks provenance
- [Agents](https://docs.mixpeek.com/docs/overview/agents.md): Give AI agents tools to search, ingest, classify, and monitor multimodal content

#### Migrate to Mixpeek

- [Migrate from Elasticsearch](https://docs.mixpeek.com/docs/overview/from-elasticsearch.md): From keyword search to multimodal warehouse retrieval
- [Migrate from Pinecone](https://docs.mixpeek.com/docs/overview/from-pinecone.md): Move from single-vector search to multi-stage warehouse retrieval
- [Migrate from Weaviate](https://docs.mixpeek.com/docs/overview/from-weaviate.md): Upgrade from vector database to multimodal data warehouse

#### Using Studio

- [Quickstart](https://docs.mixpeek.com/docs/studio/quickstart.md)
- [Namespaces](https://docs.mixpeek.com/docs/studio/namespaces.md)
- [Buckets](https://docs.mixpeek.com/docs/studio/buckets.md)
- [Collections](https://docs.mixpeek.com/docs/studio/collections.md)
- [Retrievers](https://docs.mixpeek.com/docs/studio/retrievers.md)
- [Clusters](https://docs.mixpeek.com/docs/studio/clusters.md)
- [Taxonomies](https://docs.mixpeek.com/docs/studio/taxonomies.md)
- [Explorer](https://docs.mixpeek.com/docs/studio/explorer.md)
- [Research](https://docs.mixpeek.com/docs/studio/research.md)

### Connect your data

- [Ingest Data](https://docs.mixpeek.com/docs/platform/data-model.md): Get your files into Mixpeek — namespaces, buckets, objects, uploads, and batching
- [Syncs](https://docs.mixpeek.com/docs/platform/syncs.md): Automatically ingest files from external storage into Mixpeek buckets on a schedule

#### Object storage

- [Overview](https://docs.mixpeek.com/docs/integrations/object-storage/overview.md): Connect Mixpeek to your cloud object storage to ingest multimodal data in place.
- [AWS S3](https://docs.mixpeek.com/docs/integrations/object-storage/s3.md): Connect Mixpeek to your Amazon S3 buckets to ingest and process your data.
- [Google Cloud Storage](https://docs.mixpeek.com/docs/integrations/object-storage/gcs.md): Connect Mixpeek to your Google Cloud Storage buckets to ingest and process your data.
- [Google Drive](https://docs.mixpeek.com/docs/integrations/object-storage/google-drive.md): Sync files from a Google Drive folder or shared drive into Mixpeek
- [SharePoint](https://docs.mixpeek.com/docs/integrations/object-storage/sharepoint.md): Connect SharePoint sites and OneDrive for Business document libraries to a Mixpeek bucket.
- [Azure Blob Storage](https://docs.mixpeek.com/docs/integrations/object-storage/azure-blob.md): Connect Mixpeek to your Azure Blob Storage containers to ingest and process your data.
- [Cloudflare R2](https://docs.mixpeek.com/docs/integrations/object-storage/r2.md): Connect Mixpeek to your Cloudflare R2 buckets to ingest and process your data.
- [Wasabi Hot Cloud Storage](https://docs.mixpeek.com/docs/integrations/object-storage/wasabi.md): Connect Mixpeek to your Wasabi Hot Cloud Storage buckets to ingest and process your data.
- [Tigris](https://docs.mixpeek.com/docs/integrations/object-storage/tigris.md): Connect Mixpeek to your Tigris buckets to ingest and process your data.
- [Backblaze B2](https://docs.mixpeek.com/docs/integrations/object-storage/backblaze.md): Sync files from Backblaze B2 Cloud Storage into Mixpeek buckets using the S3-compatible API.
- [Box](https://docs.mixpeek.com/docs/integrations/object-storage/box.md): Connect Mixpeek to your Box account to sync and process files from Box folders.
- [Mux](https://docs.mixpeek.com/docs/integrations/object-storage/mux.md): Connect Mux video infrastructure to Mixpeek for automated video asset processing
- [Iconik](https://docs.mixpeek.com/docs/integrations/object-storage/iconik.md): Connect Iconik DAM to Mixpeek for automated media asset processing and search
- [Supabase Storage](https://docs.mixpeek.com/docs/integrations/object-storage/supabase.md): Sync files from Supabase Storage buckets into Mixpeek using the S3-compatible API.

#### Other sources

- [RTSP](https://docs.mixpeek.com/docs/integrations/streaming/rtsp.md): Capture segments from a live RTSP camera or encoder into a Mixpeek bucket.
- [Mixpeek + Snowflake](https://docs.mixpeek.com/docs/integrations/snowflake-warehouse.md): Use Mixpeek for unstructured multimodal data, Snowflake for structured analytics
- [Mixpeek + Databricks](https://docs.mixpeek.com/docs/integrations/databricks-warehouse.md): Multimodal data warehouse meets data lakehouse -- complementary layers for the modern data stack
- [PostgreSQL](https://docs.mixpeek.com/docs/integrations/databases/postgresql.md): Sync PostgreSQL table rows into a Mixpeek bucket and run SQL lookups from a retriever.
- [BrightData](https://docs.mixpeek.com/docs/integrations/web-data/brightdata.md): Collect web datasets and scraper results from BrightData and ingest them into Mixpeek.
- [HTTP API](https://docs.mixpeek.com/docs/integrations/web-data/http-api.md): Sync items from any JSON REST endpoint into a Mixpeek bucket, one object per item.
- [RSS](https://docs.mixpeek.com/docs/integrations/web-data/rss.md): Sync RSS and Atom feed entries into a Mixpeek bucket, one object per entry.
- [Overview](https://docs.mixpeek.com/docs/integrations/social-media/overview.md): Connect Mixpeek to social media platforms to ingest and analyze public media at scale.
- [Instagram](https://docs.mixpeek.com/docs/integrations/social-media/instagram.md): Monitor public Instagram Business and Creator accounts using Mixpeek's Instagram integration.
- [TikTok](https://docs.mixpeek.com/docs/integrations/social-media/tiktok.md): Sync videos from a TikTok account into a Mixpeek bucket through the TikTok Content API.
- [Email](https://docs.mixpeek.com/docs/integrations/email.md): Receive emails as bucket objects with a dedicated inbound address for document intake

### Extract features

- [Extract Features](https://docs.mixpeek.com/docs/platform/processing.md): Turn raw files into searchable documents by picking what you want to search by per file type — collections run the pipeline for you
- [Deduplication & Re-processing](https://docs.mixpeek.com/docs/processing/deduplication.md): Content-hash dedup decides what work is skipped, replaced, or forced when objects are processed into collections — and lineage decides what re-runs when a source changes
- [Features](https://docs.mixpeek.com/docs/processing/features.md): Features are what you want to search by — pick them per file type; Mixpeek resolves the pipeline, defaults the wiring, and prices per natural unit
- [Pipeline Configuration (Advanced)](https://docs.mixpeek.com/docs/processing/feature-extractors.md): Advanced knobs behind a collection's processing pipeline — input mappings, field passthrough, parameters, and feature URIs
- [Decomposition](https://docs.mixpeek.com/docs/processing/decomposition.md): How Mixpeek turns raw objects into searchable documents and queryable features

#### Built-in extractors

- [Multimodal Extractor](https://docs.mixpeek.com/docs/processing/extractors/multimodal.md): Unified embeddings for video, image, audio, text, and GIF with transcription, OCR, thumbnails, and structured extraction
- [Image Extractor](https://docs.mixpeek.com/docs/processing/extractors/image.md): Dense vector embeddings for images using Google SigLIP (768D) for visual similarity search
- [Document Graph Extractor](https://docs.mixpeek.com/docs/processing/extractors/document.md): Extract spatial blocks from PDFs with layout classification, confidence scoring, optional VLM correction, and 1024-d E5 text embeddings
- [Text Extractor](https://docs.mixpeek.com/docs/processing/extractors/text.md): Dense multilingual (E5-Large) text embeddings for semantic and cross-lingual search — query in one language, match content in 100+ others
- [Face Identity Extractor](https://docs.mixpeek.com/docs/processing/extractors/face-identity.md): Production-grade face recognition (SCRFD + ArcFace, 99.8%+ accuracy) — track a specific person or person-of-interest (POI) across images and video
- [Audio Fingerprint Extractor](https://docs.mixpeek.com/docs/processing/extractors/audio-fingerprint.md): Audio fingerprinting with CLAP — 512-d embeddings from audio files or video audio tracks for sound-mark matching and audio similarity
- [Gemini Multifile Extractor](https://docs.mixpeek.com/docs/processing/extractors/gemini-multifile.md): Embed multiple files from one object into a single 3072-d vector using Gemini Embedding 2 — one embedding per object, not per file
- [Scrolling Text Extractor](https://docs.mixpeek.com/docs/processing/extractors/scrolling-text.md): Extract scrolling/marquee text from video using phase-correlation band detection, panoramic stitching, and VLM OCR
- [Web Scraper Extractor](https://docs.mixpeek.com/docs/processing/extractors/web-scraper.md): Recursive website crawling with multimodal content extraction and semantic embeddings
- [Passthrough Extractor](https://docs.mixpeek.com/docs/processing/extractors/passthrough.md): Copy source fields without processing for metadata propagation and vector passthrough
- [Universal Extractor](https://docs.mixpeek.com/docs/processing/extractors/universal.md): All-in-one multimodal extractor — image, video, audio, and documents — producing 3072-d Gemini embeddings plus text descriptions, OCR, and transcription
- [Video Transcode Extractor](https://docs.mixpeek.com/docs/processing/extractors/video-transcode.md): Convert RED R3D, ARRI RAW and mixed-codec video to a web-playable H.264 or H.265 MP4 mezzanine in object storage

#### Custom extractors

- [Custom Extractors](https://docs.mixpeek.com/docs/processing/custom-extractors.md): Build, test, and deploy your own feature extractors and real-time inference endpoints on Mixpeek infrastructure
- [Custom Extractor API (Dedicated Infrastructure)](https://docs.mixpeek.com/docs/processing/custom-extractor-api.md): HTTP reference for the custom-extractor upload, deploy, real-time, and version-management endpoints — available on dedicated Enterprise deployments
- [Extractor Submissions](https://docs.mixpeek.com/docs/processing/extractor-marketplace.md): Submit custom extractors for review and inclusion in the Mixpeek extractor catalog
- [Multi-Tier Feature Extraction](https://docs.mixpeek.com/docs/processing/multi-tier-extractors.md): How chained collections, dependency tiers, and cross-tier lineage turn extraction into a composable DAG

#### Models & migration

- [Model Registry](https://docs.mixpeek.com/docs/processing/model-registry.md): Load HuggingFace models, built-in models, or your own fine-tuned weights inside custom extractors
- [Migrate Embedding Models](https://docs.mixpeek.com/docs/processing/model-migration.md): Switch a namespace to a new embedding model by re-extracting into a target namespace — safely, with validation and dry-run
- [Migrating from Extractor Names](https://docs.mixpeek.com/docs/processing/extractor-migration.md): Built-in extractor names are deprecated aliases — map your existing feature_extractor configs to features keys

### Build retrievers

- [Retrievers](https://docs.mixpeek.com/docs/retrieval/retrievers.md): Compose stage-based search pipelines over your collections
- [Retrieval Cookbook](https://docs.mixpeek.com/docs/retrieval/cookbook.md): Ready-to-copy, runnable multi-stage retriever configurations for common multimodal patterns

#### Retriever Stages

- [Retriever Stages](https://docs.mixpeek.com/docs/retrieval/stages/overview.md): The core of the warehouse query layer: composable retriever stages for multi-stage search pipelines

##### Filter

- [Feature Search](https://docs.mixpeek.com/docs/retrieval/stages/feature-search.md): Unified semantic and hybrid search across multiple embedding features with configurable fusion strategies
- [Attribute Filter](https://docs.mixpeek.com/docs/retrieval/stages/attribute-filter.md): Filter documents by metadata field conditions with boolean logic support
- [LLM Filter](https://docs.mixpeek.com/docs/retrieval/stages/llm-filter.md): Filter documents using LLM-based content evaluation and criteria matching
- [Agent Search](https://docs.mixpeek.com/docs/retrieval/stages/agent-search.md): LLM-driven multi-step retrieval with iterative reasoning and tool orchestration
- [Query Expand](https://docs.mixpeek.com/docs/retrieval/stages/query-expand.md): Generate query variations using LLMs and fuse results for improved recall

##### Sort

- [Sort Relevance](https://docs.mixpeek.com/docs/retrieval/stages/sort-relevance.md): Reorder documents by their relevance scores from previous stages
- [Sort Attribute](https://docs.mixpeek.com/docs/retrieval/stages/sort-attribute.md): Reorder documents by any metadata field value
- [MMR (Maximal Marginal Relevance)](https://docs.mixpeek.com/docs/retrieval/stages/mmr.md): Diversify search results by balancing relevance with result variety
- [Rerank](https://docs.mixpeek.com/docs/retrieval/stages/rerank.md): Re-score and reorder search results using cross-encoder models for higher precision
- [Score Normalize](https://docs.mixpeek.com/docs/retrieval/stages/score-normalize.md): Rescale document scores to a common range for consistent comparison

##### Reduce

- [Aggregate](https://docs.mixpeek.com/docs/retrieval/stages/aggregate.md): Compute statistical aggregations and metrics across document results
- [Temporal](https://docs.mixpeek.com/docs/retrieval/stages/temporal.md): Group documents by time windows and compute per-window aggregations with drift detection
- [Sample](https://docs.mixpeek.com/docs/retrieval/stages/sample.md): Select a random or stratified sample of documents from results
- [Summarize](https://docs.mixpeek.com/docs/retrieval/stages/summarize.md): Generate LLM-powered summaries from document sets
- [Limit](https://docs.mixpeek.com/docs/retrieval/stages/limit.md): Truncate results to a maximum count with optional offset for pagination
- [Deduplicate](https://docs.mixpeek.com/docs/retrieval/stages/deduplicate.md): Remove duplicate documents by field match or content similarity
- [Moment Group](https://docs.mixpeek.com/docs/retrieval/stages/moment-group.md): Merge contiguous temporal intervals into consolidated video moments, grouped by parent object
- [Score Threshold](https://docs.mixpeek.com/docs/retrieval/stages/score-threshold.md): Drop results below an absolute score and return no results when nothing qualifies

##### Group

- [Group By](https://docs.mixpeek.com/docs/retrieval/stages/group-by.md): Aggregate documents by shared field values into logical groups
- [Cluster](https://docs.mixpeek.com/docs/retrieval/stages/cluster.md): Group documents by embedding similarity into semantic clusters

##### Apply

- [JSON Transform](https://docs.mixpeek.com/docs/retrieval/stages/json-transform.md): Transform document structure using Jinja2 templates for API payloads or custom schemas
- [RAG Prepare](https://docs.mixpeek.com/docs/retrieval/stages/rag-prepare.md): Prepare documents for LLM context windows with token management and source citations — evidence/provenance per result so every claim traces back to a source clip
- [External Web Search](https://docs.mixpeek.com/docs/retrieval/stages/external-web-search.md): Augment results with real-time web search using Exa's neural search API
- [API Call](https://docs.mixpeek.com/docs/retrieval/stages/api-call.md): Enrich documents with external API calls (Stripe, GitHub, weather APIs, etc.)
- [SQL Lookup](https://docs.mixpeek.com/docs/retrieval/stages/sql-lookup.md): Enrich documents with data from SQL databases using parameterized queries
- [Cross Compare](https://docs.mixpeek.com/docs/retrieval/stages/cross-compare.md): Multi-tier cross-collection content matching with configurable classification
- [Traverse Edge](https://docs.mixpeek.com/docs/retrieval/stages/traverse-edge.md): Follow typed relationships (edges) from documents to their linked documents
- [Web Scrape](https://docs.mixpeek.com/docs/retrieval/stages/web-scrape.md): Extract and parse web content from URLs using Firecrawl
- [Unwind](https://docs.mixpeek.com/docs/retrieval/stages/unwind.md): Decompose array fields into separate documents for per-element processing
- [Code Execution](https://docs.mixpeek.com/docs/retrieval/stages/code-execution.md): Execute custom code to transform, filter, or enrich documents

##### Enrich

- [LLM Enrich](https://docs.mixpeek.com/docs/retrieval/stages/llm-enrich.md): Extract structured data from documents using language model analysis
- [Classify](https://docs.mixpeek.com/docs/retrieval/stages/classify.md): Classify documents at query time using built-in tasks (like NSFW content safety) or a custom model deployed as a plugin
- [Taxonomy Enrich](https://docs.mixpeek.com/docs/retrieval/stages/taxonomy-enrich.md): Classify and tag documents using predefined taxonomies
- [Document Enrich](https://docs.mixpeek.com/docs/retrieval/stages/document-enrich.md): Join documents across collections for cross-reference enrichment
- [Agentic Enrich](https://docs.mixpeek.com/docs/retrieval/stages/agentic-enrich.md): Classify documents using a multi-turn reasoning agent with tool access

#### Tune relevance

- [Improve Relevance](https://docs.mixpeek.com/docs/platform/improve-relevance.md): Make search better over time with interactions, fusion strategies, and evaluations
- [Auto-Tune](https://docs.mixpeek.com/docs/retrieval/auto-tune.md): Self-improving relevance that adapts per user
- [Rollout & Safety](https://docs.mixpeek.com/docs/retrieval/auto-tune-rollout.md): Safely deploy learned fusion with traffic splitting, shadow mode, and kill switches
- [Reward Signals](https://docs.mixpeek.com/docs/retrieval/reward-signals.md): Configure how different interaction types influence learned fusion weights
- [Evaluations](https://docs.mixpeek.com/docs/retrieval/evaluations.md): Measure and compare retriever quality with ground truth datasets and standard IR metrics
- [Interactions](https://docs.mixpeek.com/docs/retrieval/interactions.md): Turn user behavior into better retrieval through automated feedback loops
- [Typeahead](https://docs.mixpeek.com/docs/retrieval/typeahead.md): Prefix suggestions for a search box, drawn concurrently from field values, collection names, and the caller's own recent searches
- [Query Optimization & Explain](https://docs.mixpeek.com/docs/retrieval/query-optimization.md): How Mixpeek automatically optimizes retriever pipelines, and how to inspect the execution plan with explain
- [Multi-Stage Retrieval](https://docs.mixpeek.com/docs/retrieval/multi-stage-deep-dive.md): The composable pipeline architecture that makes Mixpeek a warehouse, not a database
- [Lineage Traversal](https://docs.mixpeek.com/docs/retrieval/lineage-traversal.md): Walk a document's lineage chain — its source, provenance, and evidence trail from the original object through every transformation — efficiently, without N+1 round-trips.
- [Filters](https://docs.mixpeek.com/docs/retrieval/filters.md): Compose filter conditions with logical operators
- [Deep Research Patterns](https://docs.mixpeek.com/docs/retrieval/research.md): Compose multi-stage retrievers for investigations, literature reviews, and analysis
- [Architecture](https://docs.mixpeek.com/docs/relevance/architecture.md): How Mixpeek's two-tower design decouples ingestion from retrieval and learns from usage
- [Fusion Strategies](https://docs.mixpeek.com/docs/relevance/fusion-strategies.md): How multiple search results are combined into a single ranked list using RRF, DBSF, Weighted, Max, or Learned fusion
- [Learned Fusion](https://docs.mixpeek.com/docs/relevance/learned-fusion.md): How Thompson Sampling adapts fusion weights from user interactions to personalize search results
- [Analytics](https://docs.mixpeek.com/docs/relevance/analytics.md): Monitor retriever performance, identify slow queries, and get AI-powered tuning recommendations

### Enrich & organize

- [Classify Content](https://docs.mixpeek.com/docs/platform/enrichment.md): Auto-label documents with taxonomies, retriever enrichments, and annotations
- [Taxonomies](https://docs.mixpeek.com/docs/enrichment/taxonomies.md): Enrich documents with similarity-based joins
- [Clusters](https://docs.mixpeek.com/docs/enrichment/clusters.md): Group documents by semantic similarity or metadata attributes, then label, visualize, and enrich
- [Alerts](https://docs.mixpeek.com/docs/enrichment/alerts.md): Monitor ingested content with retriever-powered alerts and real-time notifications
- [Triggers](https://docs.mixpeek.com/docs/platform/triggers.md): Schedule clustering, taxonomy enrichment, and batch reruns on cron, intervals, or events

#### Annotations & enrichments

- [Retriever Enrichments](https://docs.mixpeek.com/docs/enrichment/retriever-enrichments.md): Run retriever pipelines on documents at ingestion time
- [Annotations](https://docs.mixpeek.com/docs/relevance/annotations.md): Record human decisions on documents for review workflows, compliance, and model improvement

### Integrate & operate

- [Search Widget](https://docs.mixpeek.com/docs/integrations/search-widget.md): Drop-in React component for AI-powered multimodal search — add search to any site in minutes
- [Operate](https://docs.mixpeek.com/docs/platform/operations.md): Run Mixpeek in production — security, webhooks, manifests, and infrastructure
- [Billing & Pricing](https://docs.mixpeek.com/docs/platform/billing.md): Three questions set your price: what kind of files, how much content, and what you want to search by — tiers include a monthly usage pool
- [Rate Limits & Quotas](https://docs.mixpeek.com/docs/operations/rate-limits-quotas.md): Understand API rate limits, usage pools, and strategies for scaling under constraints

#### Agents

- [MCP Server](https://docs.mixpeek.com/docs/agent-integrations/mcp.md): Give AI agents access to video, image, and audio search through the Model Context Protocol
- [LangChain](https://docs.mixpeek.com/docs/agent-integrations/langchain.md): Give your AI agents eyes, ears, and memory with Mixpeek's LangChain integration
- [OpenAI Function Calling](https://docs.mixpeek.com/docs/agent-integrations/openai-function-calling.md): Wire Mixpeek retrievers into OpenAI's function calling API so GPT models can search multimodal content

#### SDKs & tools

- [SDK Overview](https://docs.mixpeek.com/docs/integrations/developer-tools/sdk-usage.md): Official SDKs for Python, JavaScript/TypeScript, and more
- [Python SDK](https://docs.mixpeek.com/docs/integrations/developer-tools/python-sdk.md): Official Python SDK for the Mixpeek API
- [JavaScript SDK](https://docs.mixpeek.com/docs/integrations/developer-tools/javascript-sdk.md): Official TypeScript/JavaScript SDK for the Mixpeek API
- [Mixpeek CLI](https://docs.mixpeek.com/docs/integrations/developer-tools/mixpeek-cli.md): Command-line interface for building, testing, and deploying custom extractors
- [MCP Server](https://docs.mixpeek.com/docs/integrations/developer-tools/mcp-server.md): Connect Claude and AI assistants to Mixpeek via the Model Context Protocol
- [Claude Code Skill](https://docs.mixpeek.com/docs/integrations/developer-tools/claude-skill.md): Stand up a complete Mixpeek namespace, buckets, collections, retrievers, taxonomies, clusters, alerts, and triggers from a single slash command

#### Apps

- [Apps](https://docs.mixpeek.com/docs/canvas/apps.md): Deploy your own frontend code — React, vanilla JS, or any static site — connected to your Mixpeek retrievers. Served on your custom domain with auth built in.
- [Deploy from Code](https://docs.mixpeek.com/docs/canvas/apps/deploy.md): Upload a zip of your frontend code and Mixpeek builds and hosts it automatically. Deploy to staging or production with full versioning.
- [Build Logs](https://docs.mixpeek.com/docs/canvas/apps/build-logs.md): Stream real-time build output during deploys and diagnose failures with persistent log history.
- [Preview Deploys](https://docs.mixpeek.com/docs/canvas/apps/preview-deploys.md): Get a unique preview URL for every pull request, automatically built and cleaned up when the PR closes.
- [GitHub Integration](https://docs.mixpeek.com/docs/canvas/apps/github.md): Connect a GitHub repo to your Canvas app for automatic deploys on push and preview URLs on pull requests.
- [CLI](https://docs.mixpeek.com/docs/canvas/apps/cli.md): Deploy, stream logs, manage previews, and run a local dev server from the command line with @mixpeek/cli.
- [Server Functions](https://docs.mixpeek.com/docs/canvas/apps/server-functions.md): Run server-side JavaScript and TypeScript in your Canvas app — access environment variables, key-value storage, and the Mixpeek API without exposing secrets.
- [Environment Variables](https://docs.mixpeek.com/docs/canvas/apps/environment-variables.md): Configure environment variables for your Canvas app — available in both the browser runtime and server functions.
- [Runtime Logs & Analytics](https://docs.mixpeek.com/docs/canvas/apps/logs.md): View request logs, server function output, and performance metrics for your Canvas app.
- [Monitoring](https://docs.mixpeek.com/docs/canvas/apps/monitoring.md): Automatic error tracking, web vitals, and custom event reporting for Canvas apps — with optional Sentry and PostHog integration.
- [Version History](https://docs.mixpeek.com/docs/canvas/apps/versions.md): Every change to a Canvas app is versioned with a content hash, commit message, and diffable snapshot — like git for deployed applications.
- [Authentication](https://docs.mixpeek.com/docs/canvas/apps/authentication.md): Add Clerk-based sign-in to your Canvas app with zero configuration. Users authenticate via Google, GitHub, or email — the SDK is auto-injected.
- [User Management](https://docs.mixpeek.com/docs/canvas/apps/users.md): Manage users for your Canvas app with Clerk organizations. Invite members, assign roles, and control access — all through the API or Studio.
- [Custom Domains](https://docs.mixpeek.com/docs/canvas/apps/domains.md): Point your own subdomain at a Mixpeek App. TLS is provisioned automatically.

#### Security & access

- [Security & Tenancy](https://docs.mixpeek.com/docs/operations/security.md): Authentication, authorization, isolation, and operational safeguards
- [API Keys](https://docs.mixpeek.com/docs/operations/api-keys.md): Create, scope, rotate, and monitor API keys — permissions, resource scopes, per-key usage, and end-user keys
- [Permissions](https://docs.mixpeek.com/docs/platform/permissions.md): Authorize retrieval per end-user — with Mixpeek's built-in document ACL or your own external authorization system (OpenFGA).
- [Document-Level ACL](https://docs.mixpeek.com/docs/operations/document-acl.md): Row-level security for multi-user applications with automatic access control on documents
- [Using Mixpeek on corporate networks](https://docs.mixpeek.com/docs/operations/corporate-networks.md): If Mixpeek won't load behind a corporate proxy or firewall, allowlist these domains — a one-page checklist for IT administrators

#### Automation

- [Triggers vs Alerts](https://docs.mixpeek.com/docs/operations/triggers-vs-alerts.md): Triggers run work on a schedule. Alerts watch for a condition and tell you. Which one you want, and where the seam is.
- [Webhooks](https://docs.mixpeek.com/docs/operations/webhooks.md): Respond to Mixpeek events without polling
- [Change Feed](https://docs.mixpeek.com/docs/operations/change-feed.md): Pull an ordered, resumable log of everything that changed in your organization, with a 90-day retention window
- [Manifests](https://docs.mixpeek.com/docs/operations/manifests.md): Declarative resource configuration with YAML manifests
- [Environment Branching](https://docs.mixpeek.com/docs/operations/environment-branching.md): Clone namespaces, collections, retrievers, and taxonomies to create isolated staging environments and run experiments without re-processing data

#### Monitoring & scale

- [Analytics & Performance](https://docs.mixpeek.com/docs/operations/analytics-overview.md): Interpret metrics, optimize pipelines, and tune retriever performance
- [Observability](https://docs.mixpeek.com/docs/operations/observability.md): Monitor API, Engine, storage, and asynchronous jobs
- [Batch ingestion at scale](https://docs.mixpeek.com/docs/operations/batch-ingestion-at-scale.md): How to size, submit, and monitor large batch ingestions — chunking, concurrency, automatic recovery, and the limits that apply at each plan tier.
- [Batch Diagnostics & Troubleshooting](https://docs.mixpeek.com/docs/troubleshoot/batch-diagnostics.md): Use the API to diagnose, troubleshoot, and fix batch processing issues without accessing infrastructure.
- [Deployment](https://docs.mixpeek.com/docs/operations/deployment.md): Run Mixpeek locally, in Kubernetes, or on managed Ray
- [Deployment Models](https://docs.mixpeek.com/docs/resources/deployment-models.md): Three ways to run Mixpeek, and exactly what each party owns in every one
- [Single Tenant](https://docs.mixpeek.com/docs/resources/single-tenant.md): Dedicated infrastructure with isolated compute, storage, and data — deploy on any cloud, in any region

#### Customer-Hosted Mixpeek

- [Customer-Hosted Mixpeek](https://docs.mixpeek.com/docs/customer-hosted/overview.md): Run the Mixpeek runtime inside your own Kubernetes environment
- [End-to-End Data Flow](https://docs.mixpeek.com/docs/customer-hosted/data-flow.md): One object, from arrival to a moderation decision: what runs where, what it costs, and what leaves your account.
- [What you provide](https://docs.mixpeek.com/docs/customer-hosted/kuberay.md): The substrate Customer-Hosted Mixpeek expects from your Kubernetes cluster
- [Running inside your existing Ray cluster](https://docs.mixpeek.com/docs/customer-hosted/existing-raycluster.md): Why Mixpeek brings its own Ray, and what a shared-cluster deployment would cost
- [Ray infrastructure lifecycle](https://docs.mixpeek.com/docs/customer-hosted/lifecycle.md): What happens to a Customer-Hosted install when your platform team deletes, drains or upgrades underneath it
- [Qualification](https://docs.mixpeek.com/docs/customer-hosted/qualification.md): What Mixpeek checks on your cluster before deploying anything, and how to read the report
- [Security and RBAC](https://docs.mixpeek.com/docs/customer-hosted/security-rbac.md): Every permission Mixpeek asks for in your cluster, and how to verify the claim yourself
- [Supported environments](https://docs.mixpeek.com/docs/customer-hosted/compatibility.md): The Kubernetes, GPU, storage and KubeRay configurations Customer-Hosted Mixpeek supports
- [Troubleshooting](https://docs.mixpeek.com/docs/customer-hosted/troubleshooting.md): Every failure the qualification suite reports, what it means, and what to do

### Resources

- [Best Practices](https://docs.mixpeek.com/docs/resources/best-practices.md): Schema design, feature selection, caching, and cost optimization
- [Troubleshooting](https://docs.mixpeek.com/docs/resources/troubleshooting.md): Common errors, rate limits, and debugging tips
- [Caching & Signatures](https://docs.mixpeek.com/docs/overview/caching.md): Keep retrieval fast without serving stale data
- [Tasks](https://docs.mixpeek.com/docs/processing/tasks.md): Track asynchronous jobs across Mixpeek
