Multimodal AI Glossary
Definitions for multimodal tables, RAG, video search, embeddings, agents, and what an open-source stack costs.
Multimodal data
What is a multimodal database?
A multimodal database stores structured values and typed media in one schema. An image, a video, an audio clip, or a document is a column the engine understands, not a file path hidden in a string.
What is a multimodal table?
A multimodal table is a row-and-column object that holds structured values and typed media together, with derived columns and history on that same object.
What is a typed media column?
A typed media column stores an image, video, audio clip, or document as a value the engine can decode, iterate, and index, instead of an opaque blob.
What is an image column?
An image column stores a still picture as a first-class value a schema can caption, embed, or transform.
What is a video column?
A video column stores a clip with a timeline, so frames, audio, and timestamps stay addressable from the row.
What is an audio column?
An audio column stores a waveform or audio container the schema can transcribe, embed, or synthesize from.
What is a document column?
A document column stores a PDF or similar file so pages and passages can be split, embedded, and cited from the row.
What is a TableModel?
A TableModel is a Python class that declares tables, computed columns, and indexes. HTTP routes are a FastAPIRouter next to the class, not fields on it.
What is a view?
A view is a derived table whose rows are produced from a base table, often one row per frame, chunk, or other iterator output, and stay aligned when the base changes.
What is an iterator?
An iterator expands one source row into many child rows — frames, sentences, pages — and keeps a pointer back to the source.
What is a multimodal data plane?
A multimodal data plane is the layer that stores media, runs transforms, and serves retrieval as one system of record, rather than glue between a bucket, a warehouse, and an index.
What is an application database?
An application database stores the scalar rows of an app — users, orders, documents as JSON — and serves CRUD over them.
Retrieval
What is RAG?
Retrieval-augmented generation answers from a corpus you supply. The model sees retrieved passages, images, or other snippets at question time, not only what it learned in training.
What is an embedding?
An embedding is a numeric vector that places a passage, frame, or other piece in a space where similar items sit near each other.
What is an embedding index?
An embedding index ranks nearest neighbors for a query and stays attached to the table that owns those rows.
How does an embedding index differ from a vector database?
A vector database stores vectors and answers similarity queries. An embedding index is that search capability declared on the same table that holds the source pieces.
What is chunking?
Chunking splits a source into retrieval units — sentences, pages, token windows — small enough to embed and cite.
What is semantic search?
Semantic search ranks items by meaning, using embeddings, not only by exact keyword match.
What is cross-modal search?
Cross-modal search queries one modality with another — a sentence against frames, a still against a video library — because they share an embedding space.
What is multimodal RAG?
Multimodal RAG retrieves pages, frames, or transcripts — not only plain-text passages — and passes those pieces to the model.
What is a retrieval UDF?
A retrieval UDF packages a similarity query so an agent or another column can call retrieval as a function.
Video
How does video search work?
Video search finds a moment inside a clip, not only a filename. The usual path is to sample frames, embed each frame, and rank those frames by similarity to a text phrase or a still image.
What is video RAG?
Video RAG retrieves a moment — a frame, a timestamp, and maybe a transcript line — and gives that evidence to the model.
What is a frame iterator?
A frame iterator walks a video at a chosen rate and emits one row per frame, with an index and a timestamp on the source clip.
What is keyframe extraction?
Keyframe extraction selects representative frames — fixed rate or scene changes — so search and models run on stills instead of every encoded frame.
What is video intelligence?
Video intelligence is a pipeline that turns clips into structured outputs you can search: frames, detections, transcripts, and embeddings.
What is a video moment?
A video moment is a search hit that names the source clip plus a time offset, and often a frame, so you can jump to the match.
Documents
What is PDF RAG?
PDF RAG keeps the file, splits it into passages or pages, embeds those units, and generates an answer that can cite them.
What is a document splitter?
A document splitter cuts a document into retrieval units — sentences, paragraphs, or pages — as rows over the source file.
What is OCR?
OCR, optical character recognition, turns pixels of text — scans, screenshots, photos — into strings.
What is document extraction?
Document extraction pulls text, layout, or tables out of a file so later columns can chunk, embed, or summarize.
What is document Q&A?
Document Q&A answers questions from stored documents using retrieved passages, not a model that only saw the file once in its context window.
Audio
What is speech-to-text?
Speech-to-text transcribes spoken audio into text that can be searched, chunked, or passed to a model.
What is Whisper?
Whisper is a speech-recognition model family used to transcribe audio and video soundtracks into text.
What is text-to-speech?
Text-to-speech synthesizes spoken audio from text.
What is audio intelligence?
Audio intelligence turns calls, podcasts, or voiceovers into chapters, search, and structured fields on top of a transcript.
What is an audio embedding?
An audio embedding is a vector for a clip or a transcript window, used to find similar speech or sound rather than only exact words.
Vision
What is computer vision?
Computer vision is the set of models and pipelines that detect, classify, segment, or search visual content.
What is CLIP?
CLIP is a model that maps images and text into one space, so a sentence can retrieve a matching picture or frame.
What is object detection?
Object detection finds instances of objects in an image or frame and returns boxes, usually with a class and a score.
What is YOLO?
YOLO is a family of real-time object detectors, including YOLOX, that predict boxes and classes on stills or video frames.
What is image segmentation?
Image segmentation labels the pixels or regions that belong to an object or a prompt, rather than only a bounding box.
What is visual search?
Visual search finds similar images or frames from a picture or a text description, using visual embeddings.
What is image annotation?
Image annotation is human or model labels — boxes, captions, classes — stored on visual rows so training and review can reuse them.
What is image generation?
Image generation synthesizes or edits a picture from a prompt or another picture, and stores the result as an image.
Agents
What is agent memory?
Agent memory is durable, queryable state an agent reads and writes across turns: messages, retrieved facts, and tool results, not a hidden JSON file.
What is MCP?
MCP, the Model Context Protocol, is a standard way for models and editors to call tools and read resources from a server.
What is a stateful agent?
A stateful agent chooses its next action from persisted history and artifacts, not only from the current prompt.
What is an agent harness?
An agent harness is the data and tool layer an agent runs against — tables, memory, retrieval, and HTTP — so the loop is not just a prompt file.
What is an agent session log?
An agent session log is a table of turns, tool calls, and outcomes you can query later.
What is a context graph?
A context graph links what an agent retrieved, chose, and produced, so you can inspect why it acted.
Orchestration
What is a declarative pipeline?
A declarative pipeline is a schema of transforms. The engine decides when rows run, instead of you scheduling every step.
What is a computed column?
A computed column is a value produced by an expression or a model when the row or its dependencies change, then cached.
What is a UDF?
A UDF, a user-defined function, is your own logic the engine can call from a computed column or a query.
What is a UDA?
A UDA, a user-defined aggregate, reduces many rows — for example the frames of one clip — into a summary with logic you write.
What is incremental computation?
Incremental computation recomputes only the rows and downstream columns a change actually affects, instead of rerunning the whole pipeline.
What is a dependency graph?
A dependency graph records which columns depend on which, so an edit invalidates only descendants.
What is an AI function?
An AI function is a model call expressed in the schema — caption, transcribe, generate — rather than a separate pipeline product.
How does rate limiting work for model APIs?
Rate limiting throttles and retries provider calls so a burst or an HTTP 429 does not fail the rows that could have waited.
What is structured output?
Structured output constrains a model to JSON or a schema, so downstream columns get typed fields instead of free prose.
What is a model provider?
A model provider is the API that serves an embedding, a transcription, a vision model, or a chat model a column can call.
What is local inference?
Local inference runs a model on hardware you control, instead of calling a hosted model API.
What is an LLM framework?
An LLM framework is a library for chaining model calls, tools, and prompts in application code.
Storage
What is a media store?
A media store holds the bytes for images, video, audio, and documents. The catalog stores a typed pointer, not the bytes themselves.
What is object storage?
Object storage is a bucket API that holds blobs addressed by key. A database may point at those keys instead of copying every byte.
What is a local catalog?
A local catalog is the on-disk database, caches, and media home where tables live when you run the engine on your own machine.
What is bring-your-own-bucket?
Bring-your-own-bucket means media URIs point at a bucket you already pay for, so the catalog stores typed pointers instead of a second copy of the bytes.
What is a file cache?
A file cache keeps a local copy of media bytes so repeated transforms do not download the same object again.
What is lakehouse export?
Lakehouse export ships table data, often as Iceberg or Arrow, into an analytics lake without making the lake the system of record for media pipelines.
What is a data warehouse?
A data warehouse stores analytical tables for SQL over large scans. Unstructured files are usually paths or external tables, not typed media with model columns.
Versioning
What is time travel?
Time travel queries a prior version of a table, so you can see rows and computed outputs as they were after an earlier change.
What is data lineage?
Data lineage traces an output back to the source rows, the model, and the column definition that produced it.
What is dataset versioning?
Dataset versioning pins which rows, labels, and derived features a training or eval run used.
What is reproducibility in a data pipeline?
Reproducibility means you can inspect the same inputs and column definitions and get the same cached outputs, or a known diff.
Serving
What is HTTP serving for tables?
HTTP serving exposes insert, query, and computed results as routes from the same schema that defines the tables.
What is a query decorator?
A query decorator names a reusable query you can call from an agent, from HTTP, or from another column.
How does the local-to-cloud loop work?
The local-to-cloud loop runs the same schema on a laptop catalog and on a hosted database, so you change it locally and then apply it remotely.
What is a CLI catalog?
A CLI catalog is the same tables and services, operated from the terminal, that a dashboard can also show.
What is a multimodal backend?
A multimodal backend is storage, orchestration, retrieval, and HTTP for media and models in one deployable schema.
Cost
What does an open-source multimodal AI stack cost?
The bill is the engine, the hosted database if you use one, the media bytes, and the model APIs those pipelines call. A free license does not make the model calls free, and a low subscription does not make you a content-delivery network.
What is egress?
Egress is network transfer of stored bytes out of a provider, usually billed per gigabyte.
How does incremental computation change AI cost?
Incremental compute cost means you pay model APIs mainly for new or changed rows. Cached columns are not sent again.
What is model API cost?
Model API cost is the provider’s invoice for embeddings, transcription, vision, and chat. It is usually the surprise next to hosting.
Evaluation
What is ML evaluation?
ML evaluation scores model or pipeline outputs against labels or a judge, so you can compare versions.
What is a feature store?
A feature store serves consistent numeric or categorical features to training and to inference, usually for tabular models.
What is active learning?
Active learning chooses which unlabeled examples a person should label next, so the model improves faster than if you labeled at random.
What is semantic deduplication?
Semantic deduplication drops near-duplicate items by embedding similarity, not only by exact hash equality.
What is dataset curation?
Dataset curation selects, cleans, and versions the rows that train or evaluate a model.
Blog tags
Each tag used on the blog, and the concept it belongs to. 410 tags, 88 concepts.
- @pxt.query → What is a retrieval UDF?
- Active Learning → What is active learning?
- Advertising → What is visual search?
- Agent Architecture → What is a stateful agent?
- Agent Dashboard → What is a CLI catalog?
- Agent Data Plane → What is a multimodal data plane?
- Agent Engineering → What is a stateful agent?
- Agent Framework → What is a stateful agent?
- Agent Memory → What is agent memory?
- Agent State → What is agent memory?
- Agent-as-Tool → What is an agent harness?
- Agents → What is a stateful agent?
- Aggregations → What is a UDA?
- AI Agent Systems → What is an agent harness?
- AI Agents → What is a stateful agent?
- AI Architecture → What is a multimodal backend?
- AI Art → What is image generation?
- AI Automation → What is a declarative pipeline?
- AI Course → What is video intelligence?
- AI Data Infrastructure → What is a multimodal data plane?
- AI Data Processing → What is incremental computation?
- AI Data Stack → What is a multimodal data plane?
- AI Data Storage → What is a media store?
- AI Database → What is a multimodal database?
- AI Deployment → How does the local-to-cloud loop work?
- AI Evaluation → What is ML evaluation?
- AI Framework → What is an LLM framework?
- AI Functions → What is an AI function?
- AI Governance → What is data lineage?
- AI Infrastructure → What is a multimodal data plane?
- AI Integration → What is a model provider?
- AI Metrics → What is ML evaluation?
- AI Production → What is a multimodal backend?
- AI Project Failure → What is a multimodal data plane?
- AI Provider Comparison → What is model API cost?
- AI Testing → What is ML evaluation?
- AI Transformation → What is an AI function?
- AI Workflow → What is a declarative pipeline?
- AI Workflows → What is a declarative pipeline?
- AI/ML → What is ML evaluation?
- Airflow Alternative → What is a declarative pipeline?
- Amazon Nova → What is a model provider?
- Amazon S3 → What is object storage?
- Analytics → What is a data warehouse?
- Annotation Tools → What is image annotation?
- Anthropic → What is a model provider?
- Anthropic Claude → What is a model provider?
- Apache 2.0 → What does an open-source multimodal AI stack cost?
- API Integration → What is HTTP serving for tables?
- API Keys → What is HTTP serving for tables?
- API Management → How does rate limiting work for model APIs?
- APIs → What is HTTP serving for tables?
- Architecture → What is a multimodal backend?
- Array → What is a typed media column?
- Arrow Batches → What is lakehouse export?
- Audio → What is an audio column?
- Audio Processing → What is an audio column?
- Audio Transcription → What is speech-to-text?
- Audio Translation → What is text-to-speech?
- Automated Translation → What is text-to-speech?
- Automation → What is a declarative pipeline?
- AWS Bedrock → What is a model provider?
- Backend Architecture → What is HTTP serving for tables?
- Backend Infrastructure → What is a multimodal data plane?
- Bedrock → What is a model provider?
- Built-in Vector Search → What is an embedding index?
- Classification → What is computer vision?
- Claude → What is a model provider?
- Claude 3.5 Sonnet → What is a model provider?
- Claude AI → What is a model provider?
- Claude Vision → What is a model provider?
- CLI → What is a CLI catalog?
- CLIP → What is CLIP?
- Cloud AI → How does the local-to-cloud loop work?
- Code Generation → What is an AI function?
- Compliance → What is data lineage?
- Computed Columns → What is a computed column?
- Computer Vision → What is computer vision?
- Computer Vision Training → What is dataset curation?
- Context Engineering → What is a context graph?
- Context Graphs → What is a context graph?
- Convex → What is an application database?
- Cost Comparison → How does incremental computation change AI cost?
- Cost Optimization → How does incremental computation change AI cost?
- Cost Reduction → How does incremental computation change AI cost?
- Cost-Effective Hosting → How does incremental computation change AI cost?
- CPU Inference → What is local inference?
- Cross-Modal Search → What is cross-modal search?
- Custom Aggregations → What is a UDA?
- Custom Functions → What is a UDF?
- DAG Management → What is a declarative pipeline?
- DALL-E → What is image generation?
- Dashboard → What is a CLI catalog?
- Data Aggregation → What is a UDA?
- Data Annotation → What is image annotation?
- Data Annotation Platforms → What is image annotation?
- Data Architecture → What is a multimodal backend?
- Data Curation → What is dataset curation?
- Data Engineering → What is a multimodal data plane?
- Data Export → What is lakehouse export?
- Data Freshness → What is time travel?
- Data Friction → What is a multimodal data plane?
- Data Infrastructure → What is a multimodal data plane?
- Data Labeling → What is image annotation?
- Data Lineage → What is data lineage?
- Data Pipelines → What is a declarative pipeline?
- Data Plumbing → What is dataset curation?
- Data Sharing → What is lakehouse export?
- Data Silos → What is a multimodal data plane?
- Data Versioning → What is time travel?
- Data Wrangling → What is dataset curation?
- Data-Centric AI → What is a multimodal data plane?
- Databricks → What is a data warehouse?
- Databricks Alternative → What is a data warehouse?
- Dataset Management → What is dataset versioning?
- Dataset Preparation → What is dataset versioning?
- Dataset Versioning → What is dataset versioning?
- Datasets → What is dataset versioning?
- Debugging → What is a dependency graph?
- Decision Traces → What is a context graph?
- Declarative → What is a declarative pipeline?
- Declarative AI → What is a declarative pipeline?
- Declarative Pipelines → What is a declarative pipeline?
- Declarative Schema → What is a TableModel?
- Deepseek → What is a model provider?
- DeepSeek → What is a model provider?
- DeepSeek-V3 → What is a model provider?
- Deployment → How does the local-to-cloud loop work?
- DETR → What is object detection?
- Docker → What is a local catalog?
- Document → What is a document column?
- Document Processing → What is a document column?
- Document RAG → What is PDF RAG?
- Document Understanding → What is document extraction?
- Durable Execution → What is incremental computation?
- Durable State → What is agent memory?
- E-commerce → What is visual search?
- Efficient LLM → What is a model provider?
- Embedded Postgres → What is an application database?
- Embedding Analytics → What is an embedding?
- Embedding Index → What is an embedding index?
- Embedding Indexes → What is an embedding index?
- Embedding Management → What is an embedding?
- Embeddings → What is an embedding?
- Encord → What is image annotation?
- ETL → What is a declarative pipeline?
- Evals → What is ML evaluation?
- Evaluation Infrastructure → What is ML evaluation?
- Event Sourcing → What is an agent session log?
- Excel → What is dataset curation?
- Experimentation → What is incremental computation?
- fal.ai → What is image generation?
- Fast Inference → What is a model provider?
- FastAPI → What is HTTP serving for tables?
- Feast → What is a feature store?
- Feature Store → What is a feature store?
- File Management → What is a media store?
- FILE Type → What is a typed media column?
- Fine-Tuning → What is dataset curation?
- FinOps → How does incremental computation change AI cost?
- Fireworks AI → What is a model provider?
- FLUX → What is image generation?
- Frame Iterator → What is a frame iterator?
- frame_iterator → What is a frame iterator?
- Full-Stack → What is a multimodal backend?
- Full-Stack AI → What is a multimodal backend?
- Gemini → What is a model provider?
- GGUF → What is local inference?
- Google Gemini → What is a model provider?
- Governance → What is data lineage?
- GPT-4 → What is a model provider?
- GPT-4o → What is a model provider?
- GPU Inference → What is local inference?
- Groq → What is a model provider?
- Guardrails → What is structured output?
- HTAP → What is a data warehouse?
- HTTP Serving → What is HTTP serving for tables?
- Hugging Face → What is a model provider?
- Hugging Face Buckets → What is object storage?
- HuggingFace → What is a model provider?
- Iceberg → What is lakehouse export?
- Image Analysis → What is an image column?
- Image Editing → What is an image column?
- Image Generation → What is image generation?
- Image to Video Search → What is cross-modal search?
- Imagen → What is image generation?
- Incremental AI → What is a computed column?
- Incremental Computation → What is incremental computation?
- Incremental Compute → What is a computed column?
- Incremental Updates → What is a computed column?
- Infrastructure → What is a multimodal data plane?
- IVM → What is incremental computation?
- Jev → What is ML evaluation?
- Jina → What is an embedding?
- Json → What is structured output?
- JSON → What is structured output?
- Keyframes → What is keyframe extraction?
- Knowledge Base → What is document Q&A?
- Label Studio → What is image annotation?
- Labelbox → What is image annotation?
- Lakehouse → What is lakehouse export?
- LanceDB Integration → How does an embedding index differ from a vector database?
- LangChain → What is an LLM framework?
- LangChain Alternative → What is an LLM framework?
- LangChain Alternatives → What is an LLM framework?
- Lineage → What is data lineage?
- list_iterator → What is an iterator?
- Llama → What is a model provider?
- Llama 3.3 → What is a model provider?
- llama.cpp → What is local inference?
- LLM → What is a model provider?
- LLM Agents → What is a stateful agent?
- LLM Inference → What is a model provider?
- LLM Providers → What is a model provider?
- LLMs → What is a model provider?
- Local LLM → What is local inference?
- Local-First Development → What is a local catalog?
- LTAP → What is a data warehouse?
- Machine Learning → What is ML evaluation?
- MCP → What is MCP?
- Media Columns → What is a typed media column?
- Media Store → What is a media store?
- Memory → What is agent memory?
- Memory Architecture → What is agent memory?
- Microsoft Fabric → What is a data warehouse?
- Mistral AI → What is a model provider?
- Mistral Small → What is a model provider?
- Mixtral → What is a model provider?
- ML Data Pipeline → What is a declarative pipeline?
- ML Engineering → What is ML evaluation?
- ML Infrastructure → What is a multimodal data plane?
- ML Orchestration → What is a declarative pipeline?
- ML Pipeline → What is a declarative pipeline?
- ML Production → What is a multimodal backend?
- ML Training → What is dataset curation?
- MLOps → What is a declarative pipeline?
- Model Context Protocol → What is MCP?
- Model Lineage → What is data lineage?
- Model Marketplace → What is a model provider?
- Model Training → What is dataset curation?
- MongoDB Alternative → What is an application database?
- Multi-Agent Systems → What is an agent harness?
- Multi-Provider → What is model API cost?
- Multi-Provider AI → What is model API cost?
- Multimodal → What is a multimodal database?
- Multimodal AI → What is a multimodal database?
- Multimodal Chatbot → What is a multimodal database?
- Multimodal Data → What is a multimodal table?
- Multimodal Data Table → What is a multimodal table?
- Multimodal LLM → What is multimodal RAG?
- Multimodal RAG → What is multimodal RAG?
- Multimodal Search → What is semantic search?
- Nebius → What is a model provider?
- Neon → What is an application database?
- Object Detection → What is object detection?
- Object Storage → What is object storage?
- Observability → What is a dependency graph?
- OCR → What is OCR?
- OLAP → What is a data warehouse?
- Ollama → What is local inference?
- OLTP → What is an application database?
- Open Source → What does an open-source multimodal AI stack cost?
- Open Source AI → What does an open-source multimodal AI stack cost?
- Open Source LLM → What is local inference?
- OpenAI → What is a model provider?
- OpenAI Whisper → What is Whisper?
- OpenAPI → What is HTTP serving for tables?
- OpenRouter → What is a model provider?
- Optimization → What is incremental computation?
- Orchestration → What is a declarative pipeline?
- Orchestration Framework → What is a declarative pipeline?
- Pandas → What is dataset curation?
- Pandas vs Pixeltable → What is dataset curation?
- PDF → What is a document column?
- Performance → What is incremental computation?
- Performance Optimization → What is incremental computation?
- Persistence → What is agent memory?
- Pinecone Alternative → How does an embedding index differ from a vector database?
- Pipelines → What is a declarative pipeline?
- Pixeltable Storage → What is a media store?
- Pixeltable vs Airflow → What is a declarative pipeline?
- Pixeltable vs LangChain → What is an LLM framework?
- Pixeltable vs Pinecone → How does an embedding index differ from a vector database?
- PIXELTABLE_HOME → What is a local catalog?
- Podcast → What is audio intelligence?
- Post-AI Data Stack → What is a multimodal data plane?
- Postgres → What is an application database?
- PostgreSQL → What is an application database?
- Private AI → What is local inference?
- Product Catalog → What is visual search?
- Production AI → What is a multimodal backend?
- Production AI Infrastructure → What is a multimodal data plane?
- Production RAG → What is PDF RAG?
- Promptable Segmentation → What is image segmentation?
- PyArrow → What is lakehouse export?
- Pydantic → What is structured output?
- Python → What is a TableModel?
- Python REPL → What is a CLI catalog?
- Python SDK → What is a CLI catalog?
- Python UDFs → What is a UDF?
- PyTorch → What is dataset curation?
- PyTorch Integration → What is dataset curation?
- Queries → What is a retrieval UDF?
- Query Decorator → What is a query decorator?
- Query Patterns → What is a retrieval UDF?
- Qwen → What is a model provider?
- RAG → What is RAG?
- RAG Comparison → What is PDF RAG?
- RAG Deployment → What is PDF RAG?
- RAG Infrastructure → What is PDF RAG?
- RAG Performance → What is PDF RAG?
- RAG Pipeline → What is PDF RAG?
- RAG Systems → What is PDF RAG?
- Rate Limiting → How does rate limiting work for model APIs?
- Reasoning → What is a model provider?
- Replicate → What is a model provider?
- Reproducibility → What is reproducibility in a data pipeline?
- Rerun Alternative → What is computer vision?
- Retrieval → What is RAG?
- Retrieval UDF → What is a retrieval UDF?
- Reusable Queries → What is a retrieval UDF?
- Reve → What is image generation?
- RunwayML → What is image generation?
- S3 Files → What is object storage?
- Sales Intelligence → What is audio intelligence?
- SAM 3 → What is image segmentation?
- Scale AI → What is image annotation?
- Schema Design → What is a TableModel?
- Schema Validation → What is a TableModel?
- Schema-Driven → What is a TableModel?
- Segment Anything Model 3 → What is image segmentation?
- Segmentation → What is image segmentation?
- Self-Hosted → What is a local catalog?
- Self-Hosted AI → What is a local catalog?
- Semantic Layer → What is a data warehouse?
- Semantic Search → What is semantic search?
- Serverless Functions → What is an application database?
- Serving → What is HTTP serving for tables?
- Session Log → What is an agent session log?
- Session State → What is an agent session log?
- Similarity Search → What is semantic search?
- Simple Deployment → How does the local-to-cloud loop work?
- Smart Image Organizer → What is an image column?
- Speech Recognition → What is speech-to-text?
- Stable Diffusion → What is image generation?
- State Management → What is agent memory?
- Stateful Agents → What is a stateful agent?
- Storage → What is a media store?
- Storage Architecture → What is a media store?
- Structured Data → What is structured output?
- Supabase → What is an application database?
- SuperAnnotate → What is image annotation?
- TableModel → What is a TableModel?
- Tecton → What is a feature store?
- Templates → What is an agent harness?
- Testing → What is ML evaluation?
- Text Extraction → What is a document splitter?
- Text Generation → What is an AI function?
- Text-to-Speech → What is text-to-speech?
- Time Travel → What is time travel?
- Together AI → What is a model provider?
- Tool Calling → What is an agent harness?
- Training Automation → What is dataset curation?
- Training Data → What is dataset versioning?
- Training Infrastructure → What is dataset curation?
- Twelve Labs → What is video intelligence?
- Type Safety → What is a typed media column?
- Type System → What is a typed media column?
- TypeSafe → What is a typed media column?
- UDA → What is a UDA?
- UGC → What is visual search?
- Unified AI Infrastructure → What is a multimodal data plane?
- Unified Backend → What is a multimodal backend?
- Unified Infrastructure → What is a multimodal data plane?
- Unified Multimodal AI Infrastructure → What is a multimodal database?
- User-Defined Aggregates → What is a UDA?
- V7 → What is image annotation?
- Vector Database → How does an embedding index differ from a vector database?
- Vector Database Alternative → How does an embedding index differ from a vector database?
- Vector Database Comparison → How does an embedding index differ from a vector database?
- Vector Search → What is an embedding index?
- Veo → What is image generation?
- Version Control → What is time travel?
- Versioning → What is time travel?
- Video → How does video search work?
- Video Agent → What is video RAG?
- Video AI → What is video intelligence?
- Video Analysis → What is a video column?
- Video Analytics → What is a video column?
- Video Finder → What is a video moment?
- Video Generation → What is video intelligence?
- Video Intelligence → What is video intelligence?
- Video Processing → What is a video column?
- Video Search → How does video search work?
- Video Segmentation → What is video intelligence?
- Video Translation → What is video intelligence?
- Video Understanding → What is a video column?
- VideoRAG → What is video RAG?
- Vision AI → What is computer vision?
- Visual AI → What is computer vision?
- Visual Search → What is visual search?
- vLLM → What is local inference?
- Voiceover Automation → What is text-to-speech?
- Whisper → What is Whisper?
- Whisper API → What is speech-to-text?
- Whisper Translation → What is Whisper?
- Workflow → What is a declarative pipeline?
- Workflow Automation → What is a declarative pipeline?
- YOLO Family → What is YOLO?
- YOLOX → What is YOLO?
Not a concept page
These tags stay on the blog. Some are the brand, a release, or an event. The rest are general words, such as a team, an industry, or a menu label, that are not a category this glossary defines.
2025 Trends, Advanced Pixeltable, AI, AI Applications, AI Assistant, AI Assistants, AI Development, AI for Beginners, AI Production Deployment, AI Strategy, AI Tools, AI Tutorial, Anchor-Free, Autonomous Vehicles, AWS, Best Practices, Building Blocks, Changelog, Clipping, Community, Comparison, Conference, Consensus, Content Localization, Core Concepts, Creative, Cursor, Data Access, Data Consistency, Data Council, Data Import, Data Management, Data Management Crisis, Data Modeling, Data Validation, Database Integration, Database Landscape, Database Optimization, Design Principles, Developer Experience, Developer Tools, Development, DevOps, Directory Organization, Documentation, Education, Enterprise AI, Finance, Foundations, Generative AI, Getting Started, Google AI, Hands-on, Healthcare, Indie Developer, Instance Masks, Insurance, Kubernetes, Launch, Low-Ops Infrastructure, LPU, Maintained Fork, Migration, Moderation, Multi-Team Projects, Multilingual Video, Mutations, Namespace Management, Natural Language, Next.js, Open Vocabulary, Pagination, Pixelagent, PixelGolf, PixelSearch, Pixeltable, Pixeltable 101, Pixeltable Cloud, Production Challenges, Project Structure, PyCon, PyData, Quantization, React, Release, Reliability, Sandbox, Seattle, Solo Developer, Sports Analytics, SQL, Starter Kit, Startup Row, Strategy, System One, Table Packaging, Team Collaboration, Team Workflow, Tech Talk, Technical Deep Dive, Transformers, Triage, Tutorial, TypeScript, Use Case, Vibe Coding, Video Localization, YOLOX License