Load your databases into Databricks, live.
Describe the sync in plain English, approve it once, and rsync.ai streams every change from SQL Server, Postgres, MySQL, or MongoDB into Databricks — as Delta tables, registered in Unity Catalog.
Bring every source into the lakehouse.
Point rsync.ai at your databases and apps. It builds and runs the pipeline into Databricks — you approve it before anything moves.
Databases
Log-based CDC keeps operational databases mirrored as Delta tables continuously, not in nightly batches. Deletes are applied as soft-deletes.
SaaS & APIs
Pull business objects from the tools your team runs on, on a schedule you set and approve.
Files & storage
Ingest Parquet, JSON, or CSV from object storage into Delta tables.
Delta tables, ready for SQL and Spark.
Four steps. Nothing runs until you approve.
Describe it
Say which source and tables to load into Databricks in plain English — no connector config.
Review the plan
rsync.ai discovers the schema, proposes the Unity Catalog path and Delta types, and scans columns for likely personal data before a row moves.
Approve
Nothing runs until you say yes. Later schema changes are applied and recorded; only a dropped table waits for you.
Stream
CDC keeps Delta tables up to date; incremental MERGE loads send only the rows that changed.
Then work with the data without leaving rsync.ai
Synced tables become something you can query, schedule and share in the same workspace.
SQL Explorer
Run SQL against your connections, with personal columns masked in previews; writes are gated by role and every write is audited.
Models
Turn a query into a table that rebuilds on a schedule or when its inputs refresh, with version history and data checks.
Charts and Dashboards
Draw a model or saved query as a chart, then put charts on one live dashboard.
Workflows and Ask rsync.ai
Schedule reports and data alerts to Slack or email. Describe one in plain English and Ask rsync.ai drafts it.
Why teams load Databricks this way.
| What you care about | Notebooks + jobs | Fivetran | Hand-built ETL | rsync.ai |
|---|---|---|---|---|
| Set up without an engineer | Spark required | Connector forms | No | Plain English |
| Real-time, log-based CDC | Batch | Yes | Varies | Yes |
| You approve before it runs | n/a | No | No | Yes |
| Cost as volume grows | Cluster time | Per MAR | Your time | Per GB, not per row |
| Run in your own VPC | Yes | No | Varies | Yes |
Straight about status: the Databricks destination and the SQL Server, MongoDB, Postgres, and MySQL sources are live in production. The Stripe and GitHub sources are in preview. Self-hosting in your own VPC is available now — source-available under the rsync.ai Source-Available License — or run rsync.ai as a managed cloud.
Loading Databricks — common questions
How fresh is the data in Databricks?
Do tables register in Unity Catalog?
Do I need to write Spark or configure a connector?
How are you priced?
Who controls what moves?
Get your data into Databricks.
Connect a source free and land your first live Delta tables in Databricks today — registered in Unity Catalog.
Live on rsync.ai Cloud today · self-hosted (rsync.ai Source-Available License)