Bauplan - The Git-for-Data Lakehouse
Bauplan is a Git-for-data lakehouse that runs your Python and SQL pipelines directly on Apache Iceberg tables in your own object storage, with no servers or clusters to manage.
Learn how Bauplan fits into your architecture here.
Get Started
Install Bauplan
CLI, SDK, and your API key
Examples
Learn by example
Reference
CLI and SDK
Videos
Demos and webinars
Go Deeper
Core Concepts
Models, tables, pipelines
Git for Data
Branch, commit, roll back
Common Workflows
Real-world patterns
Integrations
Connect your stack
AI Agents
Use AI Agents to build pipelines, explore data, and manage your lakehouse with MCP Server and Skills.
Watch
Git-for-Data and AI Agents on Production Data
Import from S3 on a branch, merge to main, and roll back a bad write.
Getting started with the Bauplan CLI and SDK
Install, set your API key, and work with your first data branch.
Platform primitives: Python, Iceberg, and Git-for-Data
Write a Python model, query Iceberg tables, and version a pipeline run.
More walkthroughs and webinars on the Bauplan YouTube channel.