Apache Cassandra integration

Apache Cassandra data, in your warehouse.

A built-in ingestr connector: add credentials, pick tables, schedule it.

name: raw.table
type: ingestr
parameters:
  source_connection: apache_cassandra
  source_table: '<table>'
  destination: snowflake
  incremental_strategy: merge

$ bruin run assets/raw/apache_cassandra.asset.yml

  1. extract · Apache Cassandra dataincremental
  2. load · snowflake raw.datamerged
  3. checks · not_null, unique

Loaded and checked. Downstream models can run.

How it connects

Connected in three steps.

Apache Cassandra is a distributed wide-column database. ingestr supports Cassandra as both a source and a destination.

  1. 01

    Add a Apache Cassandra connection with its credentials.

  2. 02

    Pick the tables to load and how: replace, append or merge.

  3. 03

    Bruin runs it on your schedule and checks every load.

Connection parameters

user
the username to connect with, optional
password
the password for the user, optional
host
one or more seed hosts. Multiple hosts can be provided with ?hosts=host1,host2
port
the Cassandra native transport port, default is 9042
keyspace
the default keyspace, either in the URI path or as ?keyspace=...
consistency
one of one, quorum, localquorum, all, and other Cassandra consistency names
pagesize
source read page size
ssl=true
enable TLS
disableinitialhostlookup=true
useful for single-node Docker or NAT setups
replicationfactor
destination keyspace creation replication factor, default is 1

Every way Apache Cassandra works with Bruin

  • SourceBuilt-in ingestr sourceDocs →
  • DestinationBuilt-in ingestr destinationDocs →

The platform

Part of the Bruin platform.

Data in, ready for everything downstream: the models, the checks, the lineage and the AI layer.

01 · Move

Data Ingestion

02 · Model & govern

SQL & Python
Data Quality
Data Governance

03 · Use

AI Data Analyst
AI Dashboards
Data Apps
Self-Healing Pipelines
Bruin Cloudorchestration · governance · observability

Frequently asked

Questions about Apache Cassandra.

Does Bruin have a built-in Apache Cassandra integration?

Yes. Built-in ingestr source. Built-in ingestr destination.

Where can Apache Cassandra data go?

Snowflake, BigQuery, Databricks, Redshift, ClickHouse, Postgres, DuckDB, MotherDuck, Microsoft Fabric and more, plus files on S3 and GCS.

How fresh is the data?

As fresh as your schedule. Incremental loads append, merge or replace a time window, every few minutes if you like.

Do we need Bruin Cloud?

No. The Bruin CLI and ingestr run locally, in CI or in your own orchestrator. Bruin Cloud adds scheduling, lineage, alerts and the AI data analyst on top.

Ready to connect Apache Cassandra?

$100 in credits and 50 AI tasks. No credit card.

A demo walks through your own data.

Sign up to our newsletter

Practical updates on open-source data pipelines, AI analysts, governance, and what we are shipping at Bruin.

The signup form is hosted by Brevo. Allow marketing cookies to load it.