Definite runs as a private deployment in your own cloud (AWS, GCP, Azure, or anywhere with Kubernetes); you keep control of data, compute, and networking. Explore private deployment →
Explore with AI
ChatGPTClaudeGeminiPerplexity
← All connectors/Cloud Infrastructure
S3
§ Connector · Popular
Amazon S3

Analyze your S3 data with AI today.

Build dashboards, automate reports, and ask questions in plain English — all from your S3 data, no complex infrastructure to maintain.

Have multiple S3 accounts? Analytics across multiple S3 accounts →

Want it to watch your S3 data and act on its own? Meet the S3 agent →

§ Live with
§ What you get

Everything S3 exposes, modeled and queryable.

Reads files stored in Amazon S3 (and S3‑compatible object stores) such as CSV, JSON Lines, Parquet, and Avro, applying glob/prefix filters and optional schema inference to turn files into tabular records. This enables centralizing raw data dumps, exports, logs, and other file-based datasets for downstream analytics.

Standard on every Definite connector
Sync cadence
Hourly or faster
CDC
Native where supported
Auth
OAuth / API key
Row-level security
Yes

Tables & streams

4 objects
◆Dataset

A logical table composed of files in a bucket matched by a prefix or glob with optional schema inference. Enables analysis of dataset freshness, completeness, row counts, and ingestion latency across file-based exports and logs.

general_data_storagecustomerengagement
◆File Object

An individual S3 object (file) and its metadata such as size and last modified time. Supports monitoring of file delivery SLAs, volume growth, partition coverage, and file-level ingestion errors.

general_data_storagecustomerengagement
◆Record

Row-level records parsed from file contents (CSV, JSON Lines, Parquet, Avro). Powers downstream KPIs and aggregations like volumes, trends, and cohort analyses once ingested.

general_data_storagecustomerengagement
◆Schema

The inferred columns and data types derived from the dataset’s files. Enables tracking of schema drift, type conformance, and data quality validation.

general_data_storagecustomerengagement
Authentication

Uses your AWS Access Key ID and Secret (or an assumed IAM role) to authenticate; can read public buckets without credentials

Requirements

Requires a S3 account to connect.

Domainsgeneral_dataoperations
§ How it works

Three steps. One afternoon.

01
Connect

Authenticate S3 in a few clicks. OAuth, API key, or IAM role — we handle secrets and rotation.

definite connect s3
02
Sync

We pull every stream into your warehouse. CDC where the API supports it; full + incremental otherwise. Hourly-or-faster, row-level secure.

→ s3.raw (synced hourly)
03
Query

SQL, dashboards, or ask Fi in plain English. Your S3 data lives next to every other source — ready to join.

SELECT * FROM s3.*
Explore analytics for S3

Watch how Definite connects your sources and turns data into answers.

Watch the demo →
*
Don't see yours?
Any API becomes a Definite connector.

Build your own with the Definite SDK, or request it. Most go live in days.

Request a connector →
§ Combine with your stack

Pair S3 with the rest of your data.

Join S3 with the rest of your data, then ask Fi questions across all of it.

Your answer engine
is one afternoon away.

Watch the walkthrough to see how Definite connects your data and answers questions.