Open-source Python library that moves data from any source into your warehouse, handling schema and normalization for you.
Results for “data-warehouse”
18 tools found
Real-time data movement platform: streaming, log-based CDC and batch through 200+ managed connectors, sub-100ms latency.
Open-source feature store that manages the machine-learning data teams use for both model training and real-time predictions.
Fully-managed change-data-capture platform that streams database changes into data warehouses in real time, without building Kafka/Debezium pipelines yourself.
A data platform that pulls data from 200+ sources into a warehouse and lets you query it in plain English via an AI assistant.
An open-source AI agent that logs into your data warehouse overnight, writes its own SQL, double-checks every finding against the real data, and hands you a ranked list of things worth looking at — before you've asked a single question.
Data-syncing platform that moves data both ways between a company's warehouse and the everyday tools (Salesforce, Google Sheets, Stripe) that teams actually work in.
A data quality platform that uses AI to automatically learn what 'normal' looks like in your tables and flags anomalies — without anyone having to write manual validation rules for every column.
An enterprise data management platform that bundles data quality, observability, cataloging, and lineage into one product, with a specific focus on making sure the data fed into AI systems is trustworthy.
A data catalog and governance platform that builds a formal, certified map of what your business data actually means, then feeds that context to AI models so they answer questions correctly instead of guessing or hallucinating.
An all-in-one data platform that replaces a Fivetran + Snowflake + dbt + Looker stack with a single tool for ingesting, storing, and querying business data, plus an AI analyst that answers questions in plain English.
The most widely used open-source tool for writing data quality checks in Python — you define what your data should look like as code, and it validates every pipeline run against those expectations automatically.
An AI agent that watches your Snowflake, BigQuery, or Databricks bill and your data quality at the same time, flagging wasted spend and broken pipelines before they become a surprise invoice or a bad dashboard.
A data quality platform built around 'data contracts' — explicit, collaborative agreements about what a dataset should look like — with AI that detects, explains, and can automatically fix anomalies when they break.
Syncs Stripe, HubSpot, Postgres, and 40+ other business tools into one encrypted data lake you can query directly from Claude or Codex, without waiting on engineers or building a BI stack.
A free terminal (TUI) app for inspecting and monitoring Snowflake Tasks without leaving the command line.
A fully-managed, GPU-accelerated engine that runs your existing Spark and SQL workloads faster and cheaper on data already sitting in BigQuery, Snowflake, or Redshift.
Python library that captures streaming data durably into Iceberg tables with no broker, daemon, or catalog service to run.