# Neo4j + Amazon S3

> Store data as Parquet in S3 and Iceberg with Erathos. Sync Neo4j data with Erathos, an incremental, no-code pipeline with observability.

Source: https://www.erathos.com/en/pipelines/neo4j-amazon-s3
Em português: https://www.erathos.com/pipelines/neo4j-amazon-s3

Neo4j is a graph database management system designed to store and query highly connected data, efficiently representing relationships between entities. Postgres, on the other hand, is a highly robust and scalable open-source relational database management system. It supports a wide range of data types and advanced features.

With Erathos, Neo4j data lands in Amazon S3 in organized, query-ready files, without your team having to write a single ETL script. This paves the way for queries via Athena or Spark and for archiving full history without the cost of a traditional data warehouse.

### What Neo4j data does Erathos sync with PostgreSQL?

The integration automatically syncs key Neo4j objects:

- **Selected tables** incremental replication of any configured table
- **Schema drift** new columns automatically detected and added to the destination
- **Primary keys and timestamps** used for efficient incremental syncing
- **Historical data** full initial load followed by incremental updates

### Why sync Neo4j with Amazon S3?

In Amazon S3, Neo4j data is available in an open format, ready to be queried via Athena, Spark, or any analytics engine, without relying on a specific data warehouse.

### How it works

Erathos connects to Neo4j via the official API and syncs your data incrementally: only new or updated records are processed in each run, keeping pipelines fast and Amazon S3 costs predictable. You choose the sync frequency (from 5 minutes to daily), the objects to sync, and the destination dataset. Every run is logged with full observability: execution time, processed rows, errors with context, and instant alerts via Slack or email if anything goes wrong.

## Neo4j data in Amazon S3 in minutes

### Ready-to-use Neo4j connector

Connect Neo4j to Amazon S3 and automatically export data. Centralize database information for analysis — no spreadsheets, no scripts.

### Full control over your Neo4j pipelines

Configure the schedule, frequency, and sync type at the table level. Configure partitioning, file format, and write frequency at the table level. The Iceberg format ensures ACID compliance and schema evolution — without full bucket rewrites.

### End-to-end observability

Stop finding out about Neo4j issues when the business team complains. Every run is logged with runtime, rows processed, and error context. Automatic alerts via Slack, Discord, or email as soon as something goes off track — so your data stays fresh for analysis.

## Centralizing Neo4j data in Amazon S3 has never been so simple

Erathos is a data ingestion platform built for data and engineering teams. With the Neo4j connector, you automatically centralize nodes, relationships, and graph data in Amazon S3 — always up to date, with full observability into every run and zero maintenance.

## FAQ

### What is Erathos, and how can it help my business?

Erathos is a data ingestion platform built for reliability, transparency, and control. We help data teams connect tools like Neo4j to their data warehouse—with full observability into every run, zero maintenance, and none of the black-box opacity of traditional market tools.

### What Neo4j data does Erathos sync to Amazon S3?

Erathos performs incremental or full replication from Neo4j to your target data warehouse, including tables, schemas, and relationships. Sync frequency and schedule are configurable—no code required.

### How often does Erathos sync data from Neo4j to Amazon S3?

You can configure sync frequency from every 5 minutes up to daily, at the table level. Erathos uses incremental sync—only new or updated records are processed on each run, keeping the Neo4j pipeline efficient and Amazon S3 costs predictable.

### What happens if a Neo4j sync fails?

Erathos automatically detects failures and sends alerts to your email, Slack, or Discord with full context — not just “job failed.” Smart retries handle transient errors, and every run is logged with runtime, rows processed, and error context so your team can debug in minutes, not hours.

### Is there a free trial for the Neo4j connector?

Yes. Every Erathos connector includes a 14-day free trial. Connect Neo4j to Amazon S3 and start syncing immediately — no credit card required.
