---
title: "What is computer vision?"
description: "Computer vision is the set of models and pipelines that detect, classify, segment, or search visual content."
url: "https://pixeltable.com/learn/what-is-computer-vision"
updated: "2026-09-29"
vertical: "Vision"
doc: "https://docs.pixeltable.com/howto/cookbooks/video/video-extract-frames"
---

# What is computer vision?

Computer vision is the set of models and pipelines that detect, classify, segment, or search visual content.

Updated: 2026-09-29


## On this page

- [How it works](https://pixeltable.com/learn/what-is-computer-vision#how-it-works)
- [What it is not](https://pixeltable.com/learn/what-is-computer-vision#what-it-is-not)
- [Comparison](https://pixeltable.com/learn/what-is-computer-vision#comparison)
- [Where Pixeltable fits](https://pixeltable.com/learn/what-is-computer-vision#where-pixeltable-fits)
- [Questions](https://pixeltable.com/learn/what-is-computer-vision#questions)

## How it works {#how-it-works}


- Store images or frames as typed media.
- Run detection, classification, or embedding columns.
- Query the labels or the neighbors.

## What it is not {#what-it-is-not}

It is not a labeling UI with no inference, and it is not a warehouse of image URLs.

## computer vision: this, and the thing it is confused with {#comparison}

|  | This | Not this |
| --- | --- | --- |
| Input | Images or frames | URLs in a spreadsheet |
| Output | Boxes, classes, masks, neighbors | A folder listing |
| Labels | Stored on the row | Trapped in another tool’s export |

## Where Pixeltable fits {#where-pixeltable-fits}

Pixeltable runs those models as computed columns on image and frame rows, and can store human labels on the same rows.

## Questions {#questions}

### How does computer vision work? {#faq-1}

Store images or frames as typed media. Run detection, classification, or embedding columns. Query the labels or the neighbors.

### What is computer vision often confused with? {#faq-2}

It is not a labeling UI with no inference, and it is not a warehouse of image URLs.

## In the blog

- [What Is Jev? TypeSafe's System One Model — and Why the Decision Belongs in the Table](https://pixeltable.com/blog/jev-system-one-model)
- [Best Video Intelligence APIs in 2026](https://pixeltable.com/blog/best-video-intelligence-apis-2026)
- [Find a Video from an Image with Pixeltable](https://pixeltable.com/blog/find-video-from-image-pixelsearch)
- [SAM 3 Promptable Segmentation in Pixeltable](https://pixeltable.com/blog/sam3-promptable-segmentation-pixeltable)
- [One Swing, Many Frames: How We Aggregated Video into a Summary with Pixeltable](https://pixeltable.com/blog/one-swing-many-frames-pixelgolf-uda-aggregation)
- [Build a Complete Video Intelligence Pipeline in 20 Minutes](https://pixeltable.com/blog/video-intelligence-pipeline-tutorial)

## Related

- [Documentation](https://docs.pixeltable.com/howto/cookbooks/video/video-extract-frames)
- [Rerun vs Pixeltable: From 450 Lines to 15 in Computer Vision Pipelines](https://pixeltable.com/blog/rerun-vs-pixeltable-computer-vision)
- [Computer vision pipeline](https://pixeltable.com/use-cases/computer-vision-pipeline-optimization)
- [Pixeltable vs Voxel51](https://pixeltable.com/compare/pixeltable-vs-voxel51)
- [Pixeltable vs Label Studio](https://pixeltable.com/compare/pixeltable-vs-labelstudio)
