Frequently Asked Questions

Frequently Asked Questions

Last updated August 3, 2026

General

What use cases does Dataddo serve?

Dataddo supports ETL, ELT, reverse ETL, database replication, event-based integrations, and end-to-end integration of online sources with dashboarding apps. It also offers access to a full REST API, so any of Dataddo’s data integration functionality can be deployed in a headless scenario.

For more details on specific use cases, see our other use case pages:

Is Dataddo deployable within the Google Cloud Platform/AWS/Microsoft Azure ecosystems?

Absolutely. Dataddo subscriptions can be managed through accounts with any of the three major cloud providers — and the data plane can also run in European sovereign clouds (STACKIT, OVHcloud, Scaleway, Hetzner, Exoscale) or on-premises when data residency requirements demand it. See our:

How does Dataddo ensure data quality?

Dataddo includes several built-in mechanisms for data quality that apply across every workload:

  • The Data Quality Firewall (rule-based, configurable per column)
  • Detailed monitoring and notifications
  • Format harmonization, so data from disparate sources is analytics-ready
  • Data blending/union
  • The ability to exclude personally identifiable information (PII) from extractions

Specific workloads add their own controls. Reverse ETL adds flexible write modes, easy data mapping, and syncs as frequent as every 5 minutes. Database replication adds the Truncate Insert write mode, detailed flow logs, and automatic data type conversion.

See our documentation for more information about how Dataddo approaches data quality.

Where can I learn more about Dataddo's SOC 2 Type II certificate?
How is Dataddo priced?

Dataddo’s standard pricing is based on the number of data flows. A data flow is the connection between a data source (or sources) and a destination - for example, sending data from Facebook Ads to Google BigQuery counts as one flow. This keeps budgeting predictable: costs do not escalate with data volume, extraction frequency, or number of data sources.

Headless and embedded deployments are priced individually using integration units, which account for rows and actions per month, per connector.

For current tiers and enterprise pricing, see our pricing page.

Platform & Capabilities

What is Dataddo?

Dataddo is an enterprise data movement platform. A single control plane orchestrates data planes that run where your data lives - in your cloud, on-premises, or both - so data moves as managed infrastructure without being routed through a vendor’s cloud. Dataddo was founded in 2019 and is headquartered in Prague, Czech Republic.

Explore the Dataddo platform.

What is the difference between Dataddo's control plane and data plane?

Management - scheduling, configuration, and monitoring - runs in Dataddo’s cloud control plane. The data movement itself runs in a data plane. By default, Dataddo operates as a fully managed cloud-to-cloud service, moving data directly between cloud sources and destinations. When data must stay in place, the data plane can instead execute inside your own environment - your cloud, on-premises, or both - so sensitive workloads never traverse public infrastructure. The same control plane orchestrates either model.

Learn more about a single control plane with multiple data planes.

What types of data movement does Dataddo support?

Dataddo supports ETL, ELT, CDC, streaming, reverse ETL, batch file delivery, and zero-copy Apache Arrow - all in one platform.

See all data transport types.

Can I use Dataddo through a UI, API, CLI, or MCP?

All four, with one governance layer. Every action available in the web UI is also available via REST API and CLI, so pipelines can be version-controlled as YAML and deployed as code (CI/CD, GitOps). An MCP server gives AI agents native access.

See UI, API, CLI, and MCP.

How many connectors does Dataddo have, and what if mine isn't supported?

Dataddo maintains 400+ connectors across SaaS APIs, databases, files, and object storage. Dataddo owns the connector contract end-to-end, so API changes, schema drift, and connector maintenance are Dataddo’s responsibility, not yours. If you need a connector that does not exist yet, Dataddo builds custom connectors, typically within about four weeks.

Browse the connector catalog.

How does Dataddo get data into AI systems?

Dataddo keeps the storage your AI systems read from current, validated, and PII-safe. There are three patterns: a warehouse, lake, or object storage for RAG corpora and fine-tuning; a real-time CDC replica for latency-sensitive cases, such as a support chatbot that needs live order state without loading your production database; or direct retrieval through the MCP server, REST, or Apache Arrow, where an agent gets the latest extracted data with no storage layer to operate. Governance is applied before data lands - PII is excluded or hashed at extraction, and the Data Quality Firewall can block failing records outright. Dataddo does not train or host models and does not generate embeddings; where a vector database is part of your stack, your own embedding pipeline populates it from the storage Dataddo keeps current.

See how Dataddo is operated via UI, API, CLI, and MCP.

Evaluating Dataddo

What should I evaluate when choosing a data integration platform?

A useful evaluation goes beyond the connector count. The questions that separate platforms in practice are:

  • Coverage across source types. Can one platform handle SaaS APIs, databases (with CDC), files, and object storage, or will you need separate tools that each fail differently?

  • Who maintains the connectors. When a source API changes or a schema drifts, is that the vendor’s responsibility or yours? Unmaintained connectors are the hidden cost of most “self-serve” tools.

  • Deployment model. Does your data have to transit the vendor’s cloud, or can the platform run inside your own environment when residency or security rules require it?

  • Security and compliance. SOC 2 Type II, ISO 27001, GDPR alignment, encryption, role-based access control, and audit logging - confirmed, not just claimed.

  • Governance and auditability. Can you show where a piece of data came from and who accessed it?

  • Pricing predictability. Does cost scale with data volume and frequency, or is it predictable as you grow?

  • AI readiness. Can the platform deliver governed data into the systems AI depends on, and expose data to agents through a controlled interface?

Dataddo is built to answer each of these. See the Dataddo platform.

What makes Dataddo a good fit for large enterprises?

Dataddo is an enterprise data movement platform designed for organizations that treat data pipelines as managed infrastructure rather than self-serve tooling. Four things make it suited to enterprise scale:

  • One layer for every source and destination. SaaS APIs, databases with change data capture, files, and object storage move through a single platform, so teams stop stitching together tools that each break differently.

  • It runs in your environment. A single cloud control plane orchestrates data planes that execute where your data lives - in your cloud, on-premises, or both - so sensitive data does not have to transit a vendor’s cloud.

  • Dataddo absorbs the maintenance. API changes, schema drift, and connector upkeep are Dataddo’s responsibility, backed by SLAs and a dedicated solutions architect for enterprise accounts.

  • It passes the security review. SOC 2 Type II, ISO 27001, GDPR alignment, encryption with bring-your-own-key support, role-based access control, and audit logging.

Explore the Dataddo platform.

Should we build our own data pipelines instead of buying a tool?

Writing a script for one integration is easy. The cost shows up later, and it is rarely one-off. Each source you connect by hand becomes something your team has to maintain forever: APIs change without notice, schemas drift, authentication tokens expire, and rate limits shift. With a handful of sources this is manageable; across dozens it becomes a recurring engineering job that competes with the work you actually hired those engineers to do.

Buying a managed platform changes who owns that maintenance. With Dataddo, API changes, schema drift, and connector upkeep are handled on Dataddo’s side - your team configures pipelines and keeps control, without owning the break-fix work. Custom connectors for sources that do not exist yet are built by Dataddo, typically within about four weeks.

The honest question is not whether the integration is unique enough to justify building — it is whether you have a person whose job it is to maintain it when things break. If the answer is no, you are not building a solution; you are creating an on-call incident waiting for an owner.

Read more about the Dataddo platform.

Who maintains my data pipelines when a source API changes?

Dataddo does. This is the core of how Dataddo differs from self-serve tooling: Dataddo owns the connector contract end to end, so when a source API changes or a schema drifts, fixing it is Dataddo’s responsibility, not your team’s. In most cases API changes are handled before they affect live pipelines; schema changes are detected and absorbed automatically so pipelines do not break when sources change.

Concretely, that means your team is not on call for someone else’s API. You get:

  • Proactive monitoring of connections, with notifications when something needs attention
  • Automatic handling of source schema changes
  • SLA-backed incident response, with a dedicated solutions architect for enterprise accounts
  • Custom connector builds for sources that do not exist yet, typically within about four weeks

This is why customers describe Dataddo as a pipeline team they do not have to hire - the operational burden moves to Dataddo while control stays with them.

Data Sovereignty & Deployment

Where is my data processed - does it leave my environment?

It depends on how you deploy. For fully managed cloud-to-cloud pipelines, data moves directly between your sources and destinations. For sensitive workloads, the data plane can execute inside your own environment, so your data stays within your perimeter and does not pass through Dataddo’s cloud.

Learn more about control plane and data planes.

Where can Dataddo be deployed?

Anywhere your data needs it to run. Dataddo deploys in any public cloud (AWS, Azure, Google Cloud), in European sovereign and regional clouds (such as STACKIT, OVHcloud, Scaleway, Hetzner, and Exoscale), on-premises, or in a hybrid setup with multi-region agents. It is portable across Kubernetes, OpenShift, and VMware Tanzu. Both the data plane and the control plane can run in the EU region or provider you choose.

See the versatile platform architecture.

Does Dataddo support EU data residency and GDPR requirements?

Yes. Dataddo is GDPR-aligned, and because the data plane runs in the region you choose, data can stay within your jurisdiction to meet EU data residency requirements. The control plane can also be EU-hosted, so both planes can run under EU law.

Learn more about control plane and data planes.

Does Dataddo meet EU data sovereignty requirements?

Yes. Dataddo is built by an EU-domiciled company (HQ: Prague, Czech Republic) with no foreign parent, and it is structured so your data stays under EU control end to end. The data plane runs in your own environment, so data never transits Dataddo’s cloud. The control plane can also be EU-hosted, keeping both planes under EU jurisdiction. As an EU legal entity with no non-EU parent, Dataddo cannot be compelled to grant foreign access to your data.

This maps directly to SOV-3, the data and AI criterion in the EU Cloud Sovereignty Framework: pipelines must be developed, hosted, and governed under EU control. Dataddo meets all three. It carries no SEAL (Sovereignty Effectiveness Assurance Level) rating of its own, because it is not a cloud provider — instead it preserves the SEAL level of whatever cloud you run it on.

This is where US-based integration tools fall short: they route your data through their own US-controlled clouds, capping your effective sovereignty at the vendor’s SEAL level no matter how sovereign your destination is. Because Dataddo keeps the data plane in your environment, it removes that ceiling — and the same architecture positions you for the Cloud and AI Development Act (CADA), proposed June 2026, which will require data pipelines to be developed, hosted, and governed under EU control.

Read more about built for high-security environments.

Is Dataddo suitable for regulated industries?

Yes. Dataddo is used by enterprises in banking, insurance, healthcare, and the public sector, where the in-environment data plane combined with SOC 2 Type II, ISO 27001, and GDPR alignment meets strict regulatory requirements.

Read more about built for high-security environments.

Reverse ETL (Data Activation)

What are the benefits of reverse ETL?

Generally speaking, reverse ETL gives business teams like sales and marketing insights from information that only organizations know about themselves.

To illustrate, let’s look at CRMs. CRMs collect a lot of customer data. Payment amounts, support tickets, acquisition information—the data’s there. But it’s not all there. Companies have their own specific way of calculating certain metrics, in particular metrics based on first-party data, and it’s difficult for CRMs to display these without heavy—and costly—customizations.

Imagine your company sells software as a service, and that customers with three or more user accounts tend to have a lower risk of churn. There is no way your CRM could know this straight out of the box. But, if you have the right customer data in your data warehouse, your engineers can run computations there, then send the risk scores back into your CRM via a reverse ETL tool, making them visible to your sales and support teams alongside all other customer data. With little to no effort, these teams will then be able to identify who is at risk of churn.

This is why reverse ETL is often referred to as the “last mile” of the modern data stack—it enables organizations to display any information in any business app.

Database Replication

What are the benefits of database replication?

Generally speaking, database replication helps keep data accessible across locations and platforms for the whole organization. More specifically, it is used for:

  • Analytics. Replicating data from a production database to a data warehouse, i.e. a safe sandbox for analytics teams.

  • Data migration. Simplifies the data migration processes during system upgrades or transitions to new database platforms.

  • Disaster recovery. Redundancy and backup solutions ensure data availability and reliability in case of system failures or disasters.

  • Data consistency. Keeps data consistent across databases and other systems, preventing discrepancies and ensuring uniformity.

  • Performance optimization. Distributes data load across multiple databases, improving read performance and reducing latency.

  • Improving scalability. Distributing data across multiple servers or cloud environments enables systems to scale better.

  • Automated real-time data sync & updates. Keeps data automatically updated and synchronized in real-time and between systems, reducing manual efforts and errors.

Powering Analytics

What are the benefits of integrating data with Dataddo?
  • No more manual CSV uploads. Automate all your data connections.
  • No more switching between platforms to see data. By syncing data from all your platforms to an analytics tool, you can monitor important cross-platform metrics from a single place.
  • Analytics-ready data. Dataddo automatically unifies the format of data from your various apps, so that it’s ready to analyze by the time it gets to your dashboard.
  • Long-term scalability. Dataddo is a comprehensive, any-to-any data integration tool designed to meet the needs of any professional or organization that works with data—from solo marketers to global enterprises. This means you can start by using it to send data from online services to analytics tools and then, when your organization is ready, use it to centralize data in a warehouse, replicate data between databases, or send data from warehouses into business apps like CRMs and marketing automation platforms. A central screen for managing all data connections, multiple users per account, and multi-tenant deployment makes it easy for various departments and teams to adopt Dataddo for their own use cases.
  • Predictable pricing. Always know what you’re paying and never get a bad surprise at the end of the month. Since Dataddo’s pricing is based on number of data connections (i.e., flows) costs do not escalate with data volume, extraction frequency, or number of data sources.
  • Proactive pipeline monitoring and maintenance. Our engineers proactively monitor connections and manage all API changes behind the curtain. This means you don’t have to worry about your connections breaking in the middle of the night.
  • Inbuilt data quality tools. Dataddo automatically unifies the formats of data it sends to dashboarding tools, so your data will always be ready to analyze. We also enable rule-based monitoring and data quality checks, to help you prevent inaccuracies and errors in any data you transfer with Dataddo.
  • Connects any A to any B. Never worry that you might be “stuck” with a tool that can’t connect to one of your services. In addition to offering a massive connector portfolio of apps and databases, we build new connectors for clients all the time.
How does Dataddo compare to Supermetrics?

See our Supermetrics comparison page for the full breakdown.

Data Security, Governance & SOC 2

What security certifications and compliance does Dataddo hold?

Dataddo is SOC 2 Type II certified, ISO 27001 certified, GDPR-aligned, and PCI DSS compliant.

Read more about built for high-security environments.

How does Dataddo encrypt data and control access?

Dataddo uses end-to-end encryption with bring-your-own-key support (AWS KMS, Azure Key Vault, HSM), network isolation, and PII detection, masking, and tokenization at ingestion. Access is governed by SSO (SAML 2.0, OIDC), role-based access control, and immutable audit logging.

Read more about built for high-security environments.

Does Dataddo provide data lineage and audit logging?

Yes. Dataddo gives you end-to-end traceability for the data it moves. Every flow records its source, its destination, and its run history, so you can show where a piece of data came from and how it reached its destination. Access is recorded through immutable audit logging - which identity or agent accessed which data, and when - and detailed flow logs and monitoring track each pipeline run.

For data exposed to AI agents, access is governed per token and tied to specific SQL transformations of a data flow, so the audit trail captures exactly what each agent was permitted to see. These controls fall within the scope of Dataddo’s SOC 2 Type II report.

Read more about built for high-security environments.

What governance controls apply when adding a new data source?

Adding a source in Dataddo happens within the same governance layer that covers every pipeline, so a new connection does not open a new gap. The controls that apply:

  • Access control. Single sign-on (SAML 2.0, OIDC) and role-based access control determine who can create and manage the connection.

  • Scoped credentials. Tokens can be tied to specific SQL transformations of a data flow, giving effectively row- and column-level control over what each consumer or agent can read.

  • Sensitive data handling. Personally identifiable information can be detected, masked, tokenized, or excluded at ingestion, before it reaches any destination.

  • Audit trail. The new flow is recorded with immutable audit logging and run-level flow logs from its first execution.

  • Data quality checks. The rule-based Data Quality Firewall can be applied per column as data starts flowing.

Read more about built for high-security environments.

What is the Dataddo SOC 2 report?

Dataddo has obtained an SOC 2 Type II report for its data integration platform, which outlines the security controls implemented for the platform and evaluates their appropriateness and effectiveness in meeting the AICPA Trust Service Criteria. The report serves as an independent assessment of Dataddo’s ability to manage data with regard to security, availability, and confidentiality.

Which Dataddo services are covered by the SOC 2 Type II report?

The scope of the SOC 2 Type II report includes all products and features within the Dataddo platform.

What regions are covered by the Dataddo SOC 2 Type II report?

The report covers all regions in which Dataddo is available for use.

Who performs the independent third-party audit of Dataddo for SOC reports?

BDO Czech Republic performs the Dataddo SOC 2 audits.

How often are Dataddo SOC 2 audits performed?

Annually.

Is an NDA required to receive Dataddo SOC reports?

Yes, an NDA is required to review the Dataddo SOC 2 Type II report. Please contact us to begin the process.

Headless

Who should use Dataddo's Headless Data Integration for data products?

Any organization building a data product whose main focus is to generate insights from data; for example, CDPs or data analytics platforms—these need data from various sources to work properly, but their main functionality is analytics. By connecting to the unified Dataddo API, such organizations can put all of Dataddo’s integration functionality under the hood of their product, and focus instead on developing the product’s insight-generation functionality.

Doing this will:

  • Shorten time to market
  • Improve market adaptability
  • Save money and engineering resources
Who should use Dataddo's Headless Data Integration for custom workloads?

Any organization that wants to go beyond our user interface to achieve more control over data integrations — for example, to automate repetitive jobs like historical data loads, or to implement client-specific configurations.

Doing this via the Dataddo API allows you to:

  • Automate any complex or repetitive integration workload
  • Implement custom authentication and authorization flows
  • Subscribe and manage Dataddo via AWS, Azure, or GCP marketplaces

Still have questions?

Talk to our team or start a free trial - see how Dataddo moves your data with the governance, security, and SLA an enterprise expects.