
Move your data from GitHub with your free account
Capture data from GitHub and deliver it to your destinations using Estuary's pre-built connectors. Build pipelines for real-time, incremental, or batch data movement based on connector capabilities.
- <100ms to batch
- 200+ connectors
- Managed data pipelines



GitHub connector details
The GitHub connector continuously captures repository and organization data from GitHub into Estuary collections using the GitHub REST API, enabling right-time visibility across code, collaboration, and DevOps activities.
- Comprehensive coverage: Captures a wide range of GitHub resources including commits, pull requests, issues, workflows, releases, stargazers, and more, spanning both batch and incremental data.
- Right-time synchronization: Continuously ingests new commits, issues, and discussions as they occur, providing developers and data teams with an up-to-date view of repository activity.
- Flexible authentication: Supports OAuth2 for secure browser-based access or Personal Access Tokens (PATs) for command-line or managed integration setups.
- Granular configuration: Allows selective repository capture, branch-level filtering, and adjustable page sizes for large projects.
- Scalable for enterprise teams: Efficiently handles multi-repository or organization-wide synchronization while respecting GitHub API rate limits.
- Schema-aligned structure: Each GitHub resource maps to a separate data collection, simplifying downstream analysis, metrics tracking, or data lake ingestion.
💡 Tip: For organizations with many repositories, use wildcard patterns (like org/*) to automatically capture all repositories under one organization, ensuring comprehensive and future-proof coverage of your GitHub data.
How to connect GitHub to your destination in 3 easy steps
- 1
Connect GitHub as your data source
Securely connect GitHub and choose the objects, tables, or collections you need to sync.
- 2
Prepare and transform your data
Apply transformations and schema mapping as data moves whether you are streaming in real time or loading in batches.
- 3
Deliver to your destination
Continuously or periodically deliver your data to the destination you choose, based on the capabilities and configuration of your pipeline.
Trusted by data teams worldwide
All data connections are fully encrypted in transit and at rest. Estuary also supports private cloud and BYOC deployments for maximum security and compliance.
Read success storyTogether AI
How Together AI Uses Near-Real-Time Data to Understand Inference Economics
Read success storyEnvoy
Right-Time Visibility for a High-Velocity Business: Envoy’s Estuary Story
Read success storyColtene
From SAP on SQL Server to Snowflake: How Coltene Built Reliable Replication with Estuary
Read success storyGlossier
Glossier Runs Real-Time Supply Chain and Marketing Analytics with Estuary

HIGH THROUGHPUT
Distributed, event-driven architecture scales for demanding data workloads.
DURABLE COLLECTIONS
Store data as it moves so you can transform, replay, and deliver it downstream.
FLEXIBLE LATENCY
From <100ms CDC and real-time streaming to scheduled batch, depending on connector capabilities.
From <100ms to batch
Estuary supports CDC, real-time streaming, incremental syncs, and scheduled batch across its connector ecosystem. Capture data from GitHub using the cadence supported by the connector, then transform and deliver it to the destinations your team uses.
- Connect GitHub to warehouses, databases, data lakes, search platforms, and event systems through Estuary-supported destinations.
- Capture once and reuse GitHub data across transformations and multiple downstream destinations without rebuilding source extraction.
Don't see a connector?Request and our team will get back to you in 24 hours
Pipelines as fast as Kafka, easy as managed ELT/ETL, cheaper than building it.
Feature Comparison
| Estuary | Batch ELT/ETL | DIY Python | Kafka | |
|---|---|---|---|---|
| Price | $ | $$-$$$$ | $-$$$$ | $-$$$$ |
| Latency | <100ms to scheduled | 5min+ | Varies | <100ms |
| Ease | Analysts can manage | Analysts can manage | Data Engineer | Senior Data Engineer |
| Scale | ||||
| Maintenance Effort | Low | Medium | High | High |
One platform for real-time and batch data pipelines

Deliver real-time and batch data from DBs, SaaS, APIs, and more

Popular sources/destinations you can sync your data with
Choose from more than 100 supported databases and SaaS applications. Click any source/destination below to open the integration guide and learn how to sync your data in real time or batches.
![Apache Iceberg Logo]()
Apache Iceberg
![Databricks Logo]()
Databricks
![MotherDuck Logo]()
MotherDuck
![MySQL Logo]()
MySQL
![Amazon Redshift Logo]()
Amazon Redshift
![PostgreSQL Logo]()
PostgreSQL
![Snowflake Logo]()
Snowflake
![Elastic Logo]()
Elastic
![Google Bigquery Logo]()
Google Bigquery
![HubSpot Logo]()
HubSpot
![Dremio Logo]()
Dremio
![AWS OpenSearch Logo]()
AWS OpenSearch
![Amazon EventBridge Logo]()
Amazon EventBridge
![Amazon SNS Logo]()
Amazon SNS
![Google Bigtable Logo]()
Google Bigtable
![ClickHouse Logo]()
ClickHouse
![Bauplan Logo]()
Bauplan
![Google Spanner Logo]()
Google Spanner
![SingleStore Logo]()
SingleStore
![Supabase Logo]()
Supabase
![Azure Blob Storage Parquet Logo]()
Azure Blob Storage Parquet
![RisingWave Logo]()
RisingWave
![Materialize Logo]()
Materialize
![Imply Polaris Logo]()
Imply Polaris
![ClickHouse Kafka API Logo]()
ClickHouse Kafka API
![Bytewax Logo]()
Bytewax
![SingleStore Dekaf Logo]()
SingleStore Dekaf
![StarTree Logo]()
StarTree
![Tinybird Logo]()
Tinybird
![Azure Fabric Warehouse Logo]()
Azure Fabric Warehouse
![Dekaf Logo]()
Dekaf
![Apache Kafka Logo]()
Apache Kafka
![Amazon S3 Iceberg (delta updates) Logo]()
Amazon S3 Iceberg (delta updates)
![Google Cloud Storage CSV Logo]()
Google Cloud Storage CSV
![Google GCS Parquet Logo]()
Google GCS Parquet
![Amazon S3 CSV Logo]()
Amazon S3 CSV
![Amazon RDS for SQL Server Logo]()
Amazon RDS for SQL Server
![Amazon RDS for PostgreSQL Logo]()
Amazon RDS for PostgreSQL
![Amazon RDS for MariaDB Logo]()
Amazon RDS for MariaDB
![Amazon RDS for MySQL Logo]()
Amazon RDS for MySQL
![Google Cloud SQL for SQL Server Logo]()
Google Cloud SQL for SQL Server
![Google Cloud SQL for PostgreSQL Logo]()
Google Cloud SQL for PostgreSQL
![Google Cloud SQL for MySQL Logo]()
Google Cloud SQL for MySQL
![Oracle MySQL Heatwave Logo]()
Oracle MySQL Heatwave
![Amazon Aurora for MySQL Logo]()
Amazon Aurora for MySQL
![MariaDB Logo]()
MariaDB
![Amazon DynamoDB Logo]()
Amazon DynamoDB
![SQL Server Logo]()
SQL Server
![HTTP Webhook Logo]()
HTTP Webhook
![Pinecone Logo]()
Pinecone
![Slack Logo]()
Slack
![Azure Cosmos DB Logo]()
Azure Cosmos DB
![Amazon Aurora for Postgres Logo]()
Amazon Aurora for Postgres
![SQLite Logo]()
SQLite
![MongoDB Logo]()
MongoDB
![Alloy DB for Postgres Logo]()
Alloy DB for Postgres
![Timescale Logo]()
Timescale
![Google PubSub Logo]()
Google PubSub
![Google Sheets Logo]()
Google Sheets
![Amazon S3 Parquet Logo]()
Amazon S3 Parquet





























































