Getting Started with Export Pipelines
Export Pipelines let you schedule automated, recurring data exports from Clarisights to your own cloud storage (Amazon S3 or Google Cloud Storage). You define what data to export, where to send it, and when — and Clarisights handles the rest.
Key Concepts
An export pipeline is composed of four building blocks:
Concept | What it does |
Pipeline | The top-level container. Defines a name and which data sources (channels) to include. |
Manifest | Defines what data to export: which metrics, dimensions, and filters to apply. Versioned — you can update your data shape without losing history. |
Destination | Defines where to send the data: an S3 bucket or GCS bucket with the necessary credentials. |
Schedule | Defines when to export: a cron expression, date range, format, and compression settings. Links a manifest to a destination. |
Each time a schedule fires, it creates an Export Run — a record of the execution with status, file URLs, and byte counts.
What's Supported
Destinations: Amazon S3, Google Cloud Storage
Format: CSV
Compression: gzip (default) or none
Scheduling: Any cron expression with a minimum interval of 1 hour
Date ranges: Presets like
today,yesterday,last_7d,last_30d, or explicit start/end dates
Quick Walkthrough
Here's the typical flow to set up your first export pipeline:
Create a Destination — Configure your S3 or GCS bucket with the appropriate credentials and permissions. See Configuring Export Pipeline Destinations & Permissions.
Create a Pipeline — Give it a name and select which data sources (channels) to include (e.g., Google Ads, Facebook, TikTok).
Create a Manifest — Define one or more extracts, each specifying the metrics, dimensions, and optional filters you want. The available metrics and dimensions are shown dynamically based on your selected data sources.
Create a Schedule — Set a cron expression (e.g.,
0 12 * * *for daily at noon), choose a date range preset, link it to your destination, and activate it.Monitor Runs — Check the export runs list to see status (running, succeeded, failed), file sizes, and download links.
Pipeline Lifecycle
Active: Schedules fire normally and create export runs.
Paused: No exports are triggered. Resume at any time by setting status back to active.
Deleted: Soft-deleted. The pipeline and its data are retained for 30 days before permanent removal.
You can pause and resume both individual schedules and entire pipelines independently.
Manifest Versioning
When you create a new manifest for a pipeline, the previous manifest enters a 30-day grace period. During this time, both the old and new manifests are usable. After 30 days, the old manifest expires and is eventually cleaned up.
This means you can update your export shape (add/remove metrics, change filters) without breaking downstream consumers that still expect the old format.
What Gets Exported
Each export run produces:
A CSV data file (optionally gzip-compressed) per extract, chunked into 100,000-row files if needed
A metadata file (
.meta.json) alongside each data file containing column mappings, date range, and compression info
Files are organized in your bucket based on your set filename template as:{path_prefix}/from_clarisights/{company_name}/{channel_name}/{extract_id}/{timestamp}_{start_date}_{end_date}.csv.gz
How is this different from Legacy Data Exporters / Widget Exports?
Feature | Export Pipelines | Data Exporters (Legacy) | Widget Exports (Legacy) |
Self-serve setup | Yes | No (requires Clarisights team) | No (requires Clarisights team) |
Choose metrics & dimensions | Yes | Fixed per channel | Linked to widget config |
Filters | Yes (flexible filter syntax) | No | Linked to widget config |
Custom schedule | Any cron (min 1hr) | Daily only | Daily only |
Row limit | No limit | No limit | 100,000 rows |
Manifest versioning | Yes (30-day grace) | No | No |
Note: Data Exporters and Widget Exports will be sunset in August 2026. We recommend migrating to Export Pipelines before this date.
For detailed setup instructions, see: