Skip to content

Scan Operation

The Scan Operation runs a datastore's data quality checks against its containers (tables, views, or file patterns) and writes every identified anomaly to the linked Enrichment Destination. Defaults for source examples and record-anomaly limits come from the platform's Scan defaults and can be overridden in the scan form.

Note

On the datastore overview, the Get started in three steps widget keeps Scan disabled until the Sync and Profile operations have completed. Other entry points, such as the toolbar Run menu and the API, can start a scan at any time. A scan only applies active quality checks that already exist, so before your first scan, consider running a Profile with AI Effort enabled and activating any generated checks that remain in Draft.

A scan identifies two kinds of anomaly:

  • Record Anomalies: A single record (row) flagged as anomalous, with details on why. The simplest example is a row missing an expected value for a field.

  • Shape Anomalies: Structural issues at the column or schema level, such as missing fields or inconsistent patterns across the dataset.

Within the wizard you can:

  • Choose between an incremental load and a full load.
  • Automatically resolve previously open anomalies that no longer flag on a Full scan.
  • Limit the number of records scanned.
  • Pick which tables or file patterns to include.
  • Schedule the scan to run later.

To open the Scan Operation modal, navigate to a source datastore from the side menu, click the Run dropdown in the top-right corner of the datastore page, and select Scan. The modal opens at Step 1, Select Tables on a database datastore or Select File Patterns on a file-based one, and the stepper at the top shows the full configuration flow.

Scan Operation modal overview

Show me how

The app can walk you through this. Click Show me how in the Scan dialog's header, or press H while it is open, and the Run a Scan walkthrough highlights each step while you fill in the real form. See The How-tos Scope for every way to start one.

Deep Dive

  • Read Strategies


    Incremental vs Full and Auto-Resolve behavior.

    Read Strategies

  • Scan Settings


    Conceptual reference for every setting in the scan form.

    Scan Settings

  • Permissions


    Who can run, schedule, and configure scans.

    Permissions

How-tos

The numbered cards walk through each step of the wizard. The two unnumbered cards cover post-scan analysis and the API helper for runtime variables.

  • Select Tables / File Patterns


    Choose the containers to scan: All, Specific, or by Tag.

    Select Tables

  • Select Check Categories


    Choose Metadata, Data Integrity, or both.

    Select Check Categories

  • Read Settings


    Pick Incremental or Full, set an optional starting threshold, and the record limit.

    Read Settings

  • Scan Settings


    Anomaly Options (including Auto Resolve Anomalies, shown only when the read strategy is Full), record-anomaly limits, and source examples.

    Scan Settings

  • Schedule Options


    Set up a recurring run, or skip this step and use Run Now.

    Schedule Options

  • Follow the Run


    Follow the Scan Run on the Activity tab, from the row anatomy to its states, actions, and permissions.

    Runs

  • Use Runtime Variables


    Supply check variable values per scan, per schedule, or as container defaults.

    Use Runtime Variables

Reference

  • Troubleshooting


    Resolution steps for known errors.

    Troubleshooting

  • API


    Payload examples for run, schedule, and retrieve.

    API

  • FAQ


    Common questions.

    FAQ