DATASETS · Parquet · CSV · S3

Blockchain datasets as files. Buy ready-made or request a custom extract.

Buy a ready-made dataset in the Bitquery Data Store, or tell us the chains, tables and date range you need and we run a custom extract (Data on Demand). Files arrive as Parquet (custom extracts can also be CSV), as download links or straight into your own bucket.

ParquetCSVS3ready-made or custom
pipeline · datasets
⛓
On-chain data
40+ chains, decoded from genesis
BQ
Bitquery datasets
ready-made or cut to your spec
↘
Your files
download links or your S3 bucket
ParquetCSVS3download links
Two ways to get the files

Ready-made from the store, or cut to your spec.

Both deliver decoded, typed Parquet you keep. Pick the store when a listed dataset and window fit; send a spec when you need other tables, chains or ranges.

Self-serve · Data Store

Buy a ready-made dataset

Browse the catalog, open the free sample, check every column, then pay by card or bank transfer. Each dataset page lists its periods and price.

Browse the Data Store →
Custom · Data on Demand

Get a slice cut to your spec

Name the chains, tables and date range. We scope it, quote a fixed price and deliver exactly that slice, as a one-off.

Request a custom extract →
Always current · Datashares

Keep a warehouse up to date

Need the same tables refreshed every day instead of a one-time file? That is Datashares, delivered into Snowflake, BigQuery, S3 or Azure.

See Datashares →
Trusted by 40,000+ developers & teams like
Binance logoChainalysis logoTRM Labs logoNansen logo0x logoCoinMarketCap logoCoin Metrics logoBybit logoLukka logo3Commas logoNexo logoTether logo
40+
Chains supported
1PB+
Blockchain data indexed
10B+
API calls / month
99.9%
Production uptime
01
Output formats & delivery

Delivered in the format and destination you choose.

Pick your output: Parquet for a lake or warehouse, or CSV on a custom extract for a no-dependency export. Take the files as download links, or we write them straight into your own S3 bucket.

Pq

Parquet

Partitioned, compressed Apache Parquet. Point Athena, Spark, DuckDB or any lake straight at the files.

Cv

CSV

Plain CSV when you want a no-dependency export your analysts can open anywhere, with no API access required.

S3

Your S3 bucket

We write straight into a bucket you own. Your storage, your access control, no copy back to us.

Dl

Download links

Prefer not to share a bucket? Every order can also arrive as download links, ready to load wherever you work.

02
Write an extract spec

Pick a dataset, hand us a spec.

Every dataset is a typed, partitioned table with the same schema across chains. Name the dataset, chains and range in a short extract spec and we deliver exactly that slice as files you own.

dex_trades
transfers
balances
blocks
transactions
nft_trades
extract.dex_trades.json
{
  "dataset": "dex_trades",
  "chains": ["ethereum", "bsc"],
  "range": { "from": "2021-01-01",
             "to": "2024-12-31" },
  "format": "parquet",
  "destination":
    "s3://your-bucket/bitquery/dex_trades"
}
block_timeTIMESTAMPblock time, UTC
dexSTRINGprotocol name
pairSTRINGbase / quote symbol
sideSTRINGbuy or sell
price_usdDOUBLEUSD price at trade
amount_usdDOUBLEtrade size in USD
03
Forget scraping nodes for a one-off backfill

A one-off backfill shouldn't mean standing up an indexing stack.

Pulling years of decoded history for a single project means an archive node per chain, a decoder per protocol, and a backfill job you babysit, all to throw it away when the extract is done. We already indexed and decoded the chains; you buy a ready-made dataset or hand us a spec, and get clean, partitioned files.

What you're doing
Scrape it yourself
Bitquery Datasets
Years of decoded history
Replay an archive node per chain
Extracted from our indexed history
Trades, transfers, balances
A decoder and schema per dataset
One schema, ready files
A specific date range
Backfill and prune by hand
Bounded to the range you specify
Where it lands
Stand up storage and a loader
Download links or your own bucket
Solana and EVM together
Two stacks, two schemas
One extract spec across chains
No infra after delivery
Tear down nodes and jobs
Nothing to run, just your files
04
What teams request on demand

One spec in, clean files out.

Each one is a single extract spec over the same decoded datasets, delivered as files you own and ready to load.

ML · Training data

Build an ML training set

Pull years of clean, labelled on-chain history bounded to your window: a ready feature store or RAG corpus, delivered as Parquet you can load straight into your pipeline.

Read the docs →
Backfill · New product

Backfill a new product

Launch with full history on day one. Request a one-off extract of trades, transfers or balances for your chains and seed your database without scraping a single node.

Read the docs →
Research · Audit

One-off research & audit

Hand analysts or auditors a bounded slice (point-in-time balances and full transfer history for a set of addresses and dates) as a CSV they can open anywhere, no API access required.

Read the docs →

What teams say about our data

"We did a thorough search of the market for the best onchain data. Bitquery came out on top — and now powers all live prices across Nansen. We don't think of them as a vendor. They're a partner."

A
Alexander Karsten
Nansen
1PB+
Decoded history to extract from
40+
Chains available on demand
Parquet
& CSV, as links or in your bucket

Bitquery does the hard work of parsing blockchain transaction data into a usable form so that we don't have to. We use their interface to diagnose issues with complex transactions and their analytics as a starting point for our own.

0x Protocol logo
Alex Knaggs
0x Protocol

They proved they had the technology to deliver sophisticated data solutions. We extended our support through the Binance X fellowship — building an open-source library of visualization widgets on their blockchain data.

Director, Binance X logo
Flora Sun
Director, Binance X

The complex raw data is available at different levels of detail and from different viewpoints — whether we need simple aggregated transfers or parameters for failed contract calls. The support is responsive, friendly and quick.

Backend Developer, Blockpit logo
Jan Dreske
Backend Developer, Blockpit

Partnering with Bitquery has been highly cost-effective — leveraging their established infrastructure rather than building our own let us rapidly expand our blockchain support and reach a much broader segment of on-chain users.

Co-Founder, Syla logo
Nick Christie
Co-Founder, Syla

Bitquery's products are very intuitive and easy to use. We currently use their products to obtain DEX-related trading and liquidity information, which saves us the manpower and tedious technical details required to develop our own system. Their excellent technical team deserves special praise; they provide near-24/7 support and resolve issues quickly. I greatly appreciate their products and work ethic.

Ourbit logo
Data Team
Ourbit

Bitquery provides the infrastructure we rely on every day. Fast, reliable, and comprehensive across the chains that matter to our business.

Webacy
Webacy
webacy.com
FAQ

Datasets and custom extracts, answered.

How do I get blockchain datasets as files?
Two ways. Buy a ready-made dataset in the Bitquery Data Store: pick a dataset and a period, pay by card or bank transfer, and receive Parquet files. Or send a spec for a custom extract (Data on Demand): the chains, tables and date range you need, which we run against our indexed, decoded history. Either way there's no node, indexer or pipeline on your side. See the cloud docs for tables and layouts.
What formats and destinations do you support?
Files are Apache Parquet (compressed, partitioned, ready for Athena, Spark, DuckDB or any lake), and custom extracts can also be plain CSV. They arrive as download links, or we write them straight into an Amazon S3 bucket you own, under your storage and access control.
How is this different from Datashares?
Datashares is a continuous, daily-refreshed warehouse share: live tables in Snowflake, BigQuery, S3 or Azure that stay current. Data Store datasets and custom extracts are one-time files of a specific range that you keep. Choose Datashares for an always-fresh feed, and files for a defined backfill or slice.
Which datasets and chains can I extract?
The Data Store lists every ready-made dataset with its tables, columns, periods and a free sample. For anything else (DEX trades, transfers, balances, blocks and transactions, NFT events, or data decoded from any contract or program across EVM chains, Solana, Tron and Bitcoin) send us a spec. See the dataset docs for fields and coverage.
What's the turnaround on an extract?
Turnaround depends on the chains, datasets and range. A focused slice can land quickly, while a multi-chain, multi-year backfill takes longer. We scope the size and timeline with you up front. Send us your requirements for an estimate.
Can I get fresh data every day instead of a one-time file?
Yes, through Datashares, which keeps the same tables refreshed daily in your warehouse or bucket. Data Store datasets and custom extracts are one-time files, a good fit for a backfill before the daily feed starts.

Get the blockchain data you need as files.

Buy a ready-made dataset in the Data Store, or send us a spec for DEX trades, transfers, balances, blocks, NFT events or any decoded data. Parquet or CSV, as download links or in your own bucket.

Data Store · custom extracts · Parquet · CSV · S3 · 40+ chains