Skip to main content
The dbt connector reads dbt artifact files (manifest.json, run_results.json, catalog.json, sources.json) and indexes them for RAG search. Pavo supports two data sources for dbt artifacts:

Option A: S3 Data Source

What You Need to Provide

Step 1: Upload Artifacts to S3

Customers already have a dbt project that produces artifacts. These files are being uploaded to S3 (e.g., via dbt Cloud artifact export or a CI/CD step):

Step 2: Create a Read-Only IAM User

Create a dedicated IAM user in your AWS account with minimal S3 read permissions:
Minimum IAM policy (attach to the user):

Step 3: Add the Connector in Pavo

Navigate to Settings → Data sources and click Add source. Data sources page Select dbt from the connector list. Connector picker Then: 3. Add S3 URI of the location where dbt artifacts are stored 4. Add AWS access key ID and secret key 5. Click Connect This adds the connection and automatically dispatches a sync job. You can see the progress of data ingestion on the same connectors page.

Option B: Snowflake Data Source

If your dbt artifacts are stored in Snowflake tables (e.g., via dbt Cloud’s Snowflake artifact storage), use this option.

What You Need to Provide

Setup

  1. Create a read-only Snowflake user with access to the artifact tables
  2. In Pavo, go to Connectors → dbt
  3. Select Snowflake as the data source type
  4. Enter your Snowflake credentials
  5. Click Connect