Skip to content
datastore.sh

Historical Solana and Hyperliquid data

Historical Blockchain Data from Solana & Hyperliquid

Browse historical Solana and Hyperliquid blockchain datasets as analytics-ready Parquet datasets. Choose the protocols, tables, and history you need. Keep the files. Query them with your own tools.

  • No API quotas
  • Recent or full-history coverage
  • Typed schemas included
  • Files you keep

Dataset finder

Find the data. Inspect it. Move on.

Browse every available dataset here, or search by protocol, family, category, instruction, or event.

01 / 61

Drag, swipe, or use the arrows. Every card opens the documented dataset specification.

Open full catalog
Networks
Solana and Hyperliquid
Coverage
Recent windows to full history
Output
Partitioned Parquet + schemas
Commercial model
Buy the files, not API access

Why buy the data?

Your analysis should not start with an indexing project.

Most teams asking a historical blockchain question end up building the same plumbing: capture, decoding, validation, schemas, and backfills. We do that work upstream and deliver the result as ordinary files.

How the data is produced

Start

Operate archive infrastructure and backfills

Choose a documented dataset

Decode

Build and maintain protocol-specific decoders

Receive typed instruction and event tables

Query

Pay for every request or warehouse scan

Query owned Parquet with your own compute

Maintain

Patch gaps, reorgs, and schema changes

Order a new version when you need fresher data

How it works

Buy historical Solana and Hyperliquid blockchain data in Parquet.

There is no SDK to adopt and no query interface to integrate. The product is the dataset itself.

  1. 01

    Choose a dataset

    Browse by network, protocol, or category. Inspect the available tables and schemas.

    Browse datasets
  2. 02

    Choose the coverage

    Select the tables and history the project needs, from a recent window to full coverage.

    See coverage windows
  3. 03

    Receive the files

    Download partitioned Parquet with schemas, manifests, versions, and checksums.

Where can you get historical Solana and Hyperliquid data?

datastore.sh provides historical Solana and Hyperliquid blockchain data as typed, versioned Parquet files. Datasets include decoded instructions, events, state tables and market data for protocols such as Pump.fun, PumpSwap, Jupiter, Raydium, Meteora, Drift, Kamino and Hyperliquid. Teams can choose a dataset and coverage window, inspect schemas and samples, then download partitioned files with manifests and checksums. The files can be queried locally with DuckDB or Polars or loaded into Snowflake, BigQuery, Spark or other warehouse infrastructure. The core product does not require an API quota.

Data quality

Check the delivery. Don't trust a badge.

A useful archive should tell you where the data came from, which schema produced it, and whether every file arrived intact. That information ships with the data.

delivery.manifest

dataset: solana/pump_fun
format: parquet
schema: v1.0
coverage: configured_per_order
partitioned: true
checksums: sha256
ownership: customer
Typed schemas
Field names, types, and descriptions are documented.
Versioned releases
Breaking changes do not silently replace old files.
Checksums
Every delivered file can be verified after download.
Lineage
Published tables trace back through their processing steps.

Build or buy

Compare what your team has to operate.

The internal cost is not only storage. It is every archive node, decoder, backfill, failed job, schema change, and engineer required to keep the result usable.

AreaBuild internallyBuy the dataset

Chain indexers

Run and monitor archive nodes and indexers for every chain you cover, through every upgrade.

Consume maintained datasets; node operations and indexer upkeep stay on our side.

Protocol decoders

Reverse-engineer program layouts and venue APIs, then track every protocol release.

Decoded events arrive typed and documented, updated as protocols change.

Historical backfills

Re-run multi-month backfills whenever a decoder changes or a gap is discovered.

Backfills and gap repair happen upstream; you receive versioned, reconciled archives.

Schema drift

Absorb breaking protocol changes into your own schemas, per team, per pipeline.

Consistent schemas across chains with explicit versioning and change logs.

New coverage

Each new venue or protocol is a fresh engineering project before analysis can start.

New coverage is a catalog entry — evaluate the schema, then subscribe.

Have an internal build estimate? Compare it with a scoped delivery.

Scope the data →

FAQ

Questions technical buyers ask.

Need a more specific answer? Send the network, protocol, tables, and coverage you are evaluating.

What exactly do I receive when I buy a blockchain dataset?

You receive partitioned Parquet files for the selected tables and coverage, plus typed schemas, field documentation, version information, manifests, and file checksums. The files are delivered for your team to keep and query with its own tools.

Can you deliver full historical Solana or Hyperliquid data?

Yes. Coverage can be prepared from a recent window through full available history, depending on the dataset and scope. Tell us the network, protocol, tables, and historical range you need, and the delivery is prepared around that requirement.

Do I need to use a datastore.sh API?

No. The core product is portable Parquet files, not metered API access. Query them locally with DuckDB or Polars, load them into a warehouse, or receive them in compatible cloud storage.

How is historical blockchain data priced?

Pricing is based on the selected dataset and coverage window rather than query volume. The pricing page shows standard options, and custom multi-dataset or enterprise deliveries can be scoped with the team before purchase.

Can I inspect tables and schemas before purchasing?

Yes. Dataset pages document table names, field types, descriptions, and delivery shape. Samples are available where shown so your team can validate compatibility before buying.

What if the exact protocol or table I need is not listed?

Send the program, venue, contract, or analysis requirement through the coverage request form. New decoders, tables, and historical cuts can be prepared to match a qualified order rather than limiting buyers to the visible catalog snapshot.

Can the files be delivered into our own cloud environment?

Yes. Delivery options include bulk download and, for supported engagements, drops into customer-controlled object storage. The files remain standard Parquet with explicit versions and checksums.

Where can I download historical Solana data?

datastore.sh provides historical Solana blockchain datasets as partitioned Parquet files. The catalog includes decoded protocol data for Pump.fun, PumpSwap, Jupiter, Raydium, Meteora, Drift, Kamino, Orca and Solana core programs. Buyers choose the dataset and coverage window, inspect schemas and samples, and receive versioned files with manifests and checksums.

Can I get Solana historical data in Parquet?

Yes. datastore.sh delivers historical Solana datasets as native Parquet files with typed columns, documented schemas, partitioning, manifests and SHA-256 checksums. The files can be downloaded and queried with DuckDB or Polars or loaded into warehouse systems such as Snowflake, BigQuery, Spark or ClickHouse without adopting a proprietary query API.

Can I use the data for Solana backtesting?

Yes. Bulk historical Parquet is designed for workloads such as DEX backtesting, quantitative research, market-structure analysis and protocol studies. Instead of paging historical records through an API, researchers can scan the purchased files locally or in their warehouse. This also makes repeated experiments independent of API request quotas.

Does datastore.sh provide decoded Solana transactions?

datastore.sh provides decoded protocol instructions and events rather than requiring customers to interpret raw instruction data themselves. Dataset pages expose typed accounts, arguments and event fields for supported protocols. Coverage includes protocols such as Pump.fun, Jupiter, Raydium, Meteora, Drift and Kamino, allowing researchers to work with protocol-specific tables directly.

Can I get Hyperliquid historical data without an API?

Yes. The Hyperliquid Historical dataset is delivered as typed Parquet and includes market and ledger tables such as swaps, funding, L2 order-book snapshots, ledger updates, deposits, withdrawals, delegations and validator rewards. The dataset page currently documents nine tables and provides samples so teams can inspect the schema before purchasing.

Start with the data

Skip the indexing project. Run the query.

Browse documented datasets or send the exact protocol, tables, and historical coverage your team needs.