Documentation Index

Fetch the complete documentation index at: https://docs.dataddo.com/llms.txt

Use this file to discover all available pages before exploring further.

Similarweb

Prev Next

SimilarWeb is a web analytics platform that offers insights into website traffic, audience engagement, and online market intelligence. It provides businesses with data on their own website performance as well as competitor analysis, helping them make informed decisions

Refer to our website for the list of metrics and attributes available in Dataddo.

Refer to SimilarWeb's official documentation to see all available endpoints from the SimilarWeb API.

Authorize Connection to SimilarWeb

In SimilarWeb

To authorize your SimilarWeb account, you will need an API key (= API token).

  1. In your SimilarWeb account, navigate to the Settings page.
  2. Continue to Account and under API select Standard API / Batch API.
  3. Click on Generate a new API key and copy the API key.

In Dataddo

  1. On the Authorizers page, click on Authorize New Service and select SimilarWeb.
  2. Fill in the SimilarWeb API Key.
  3. Rename your authorizer for easier identification and click on Save.

Data Coverage

Similarweb exposes the following datasets. Each dataset maps to a table you can extract. Example fields are a representative sample; each dataset returns more columns.

Dataset Description Example fields Date range
Average visit duration - all traffic Returns the average visit duration for a given domain (in seconds) for all traffic (desktop + mobile web) Average Visit Duration, Country, Date, Domain, End Date, Start Date Yes
Bounce rate - all traffic Returns the bounce rate for a given domain (in seconds) for all traffic (desktop + mobile web) Bounce Rate, Country, Date, Domain, End Date, Start Date Yes
Monthly Visit Data Provides detailed data about visits to a domain, including traffic, engagement, bounce rates, and geographic information, for the specified months. Average Time, Bounce Rate, Country, Country Code, Date, Domain (+6 more) Yes
Pages per visit - all traffic Returns the average page views for a given domain (in seconds) for all traffic (desktop + mobile web) Country, Date, Domain, End Date, Pages Per Visit, Start Date Yes
Traffic Sources Overview - Desktop Returns the estimated desktop traffic volume per marketing channel. Country, Date, Domain, End Date, Organic Visits, Paid Visits (+2 more) Yes
Traffic Sources Overview - Mobile Web Returns the estimated mobile web traffic volume by source. Country, Date, Domain, End Date, Source Type, Start Date (+2 more) Yes
Visits - all traffic Returns estimated number of visits for a given domain (in seconds) for all traffic (desktop + mobile web) Country, Date, Domain, End Date, Start Date, Visits Yes
Visits - all traffic by Country Returns estimated number of visits for a given domain (in seconds) for all traffic (desktop + mobile web) Country, Date, Domain, End Date, Start Date, Visits Yes

How Data Extraction Works

Every dataset for this connector uses a relative date range: the source reads a relative window (for example "last 7 days"), and that window slides forward with the current date. Every run re-reads the window, so a range of "1 day ago" always pulls the previous day (D-1). Each run replaces the window's data rather than adding older history. To load records from before the window, run a full data re-sync with a wider range. See Data Backfilling.

Set the relative date range when you create the source.

Metadata Columns

When you create a source, you can add these Dataddo metadata columns to the extracted data:

  • dataddo_hash - a fingerprint built from each record's key fields. It works as a natural key, so it is ideal for upserts (updating existing rows in your destination instead of creating duplicates).
  • dataddo_extraction_timestamp - the date and time the row was extracted. Use it to track how records change over time, for example to build slowly changing dimensions.

How to Create a SimilarWeb Data Source

Creating a data source takes you through six steps, shown in the progress bar at the top of the wizard. Each step is explained below.

1. Pick the connector

On the Sources page, click Create Source, then select the connector from the catalog. Use the search bar or the category tabs if you do not see it right away. You can rename the source at any time using the pencil icon next to its name.

2. Select the dataset

A dataset defines the shape of your data: which fields you get and how they relate. Select the dataset you want; you can still fine-tune the exact fields later.

  • Each dataset has a short description of what it contains. Use the search box to find a dataset, attribute, or metric by name.
  • The panel on the right previews the selected dataset's fields. For each field you can see its data type, whether it holds sensitive data (personal fields such as name or email are flagged), and which other datasets it links to, so you can see how the datasets relate.

3. Choose the account

This step selects what Dataddo reads from.

  • Authorizer: Select an account you have already authorized from the drop-down. If you have none yet, choose Add new account and follow the prompts. If no authorizer is selected, Dataddo asks you to authorize before you continue.
  • What to extract from: Select the exact entity you want to pull data from. Depending on the service this may be labelled an account, property, profile, workspace, or similar, sometimes with a sub-level to choose as well.
  • Multiple accounts: To pull the same data from every entity you can access, turn on Automatically collect data from all .... This is multi-account extraction. Leave it off to choose them by hand.

4. Refine the attributes and metrics

The dataset already sets the structure. Here you fine-tune it: tick or untick the specific attributes and metrics you want to keep, and use the search box to find a field quickly. Click Test on Sample Data at any point to preview the result before you continue.

5. Add metadata columns (optional)

Two optional columns help your destination handle the data.

  • Dataddo Hash (Include Row Hash): a fingerprint built from the columns you pick. It works as a natural key, so your destination can deduplicate rows and run upserts instead of creating duplicates. Turn it on, then select the columns that uniquely identify a row.
  • Dataddo Extraction Timestamp: the time each row was extracted. Use it to watermark the data, for example to build slowly changing dimensions or to track when a value last changed.

6. Set the schedule

Decide how often Dataddo runs the extraction.

  • Frequency: how often the pipeline runs, for example daily. Click Show advanced settings to also set the exact hour and minute (UTC).
  • Date range: the relative window each run extracts, for example "Yesterday". The window moves forward on every run.
  • Historical data: a new source starts from the current window. To load older data, run a full data re-sync after the source is created.
  • Allow Empty Data Extractions: when on, a run that returns no data records zero rows instead of failing. Turn it on if the source can legitimately have periods with no data.

Click Save. Your data source is ready.

Troubleshooting

Data Preview Unavailable

No data preview when you click on Test Data might be caused by an issue with your source configuration. The most common causes are:

  • Date range: Try a smaller date range. You can load the rest of your data afterward via manual data load.
  • Insufficient permissions: Please make sure your authorized account has at least admin-level permissions.

Related Articles

Now that you have successfully created a data source, see how you can connect your data to a dashboarding app or a data storage.

Sending Data to Dashboarding Apps

Sending Data to Data Storages

Other Resources