← Back to the directory

Proprietary dataset

biopreprintwatch: New Life-Science Preprints (bioRxiv + medRxiv)

Newest bioRxiv and medRxiv life-science preprints with title, authors, category, version and abstract, refreshed hourly from the biorxiv API. Freshness anchored on-chain.

Buy a query — $0.01 · no account neededMachine-readable version (.md)

Anchored

41 documents · 42 chunks · 51 KB

Coverage 2026-09-03/2026-09-03

$0.01 per query

Provenance

Every publish of this dataset is fingerprinted and timestamped on-chain. Anyone can verify it without trusting us.

Latest confirmed on-chain anchor for this corpus. Verify the transaction yourself — do not take this page as proof.

Chain
Base Sepolia (testnet, chain 84532)
Block
46355407
Timestamp
Technical evidence
Anchor tx
0xaea0…7c7f
Registry
0x9F0A…b897
Anchored root
0x54b7…12ab
Listing corpus root
0x54b7…12ab
eth_call data
0x0cbe…12ab

What you get

Questions this dataset answers

  • What are the newest genomics preprints on bioRxiv?
  • Any new medRxiv oncology preprints today?
  • Show recent bioinformatics preprint abstracts mentioning single-cell

Response shape: ranked chunks, capped at k_max, plus a membership attestation.

About this dataset

Corpus: newest life-science preprints and new versions from bioRxiv (biology) and medRxiv (health sciences), collected hourly from the public biorxiv details API (keyless). Each row carries the paper title, authors, category (e.g. genomics, neuroscience, oncology), version number and date, DOI, and abstract excerpt. bioRxiv posts ~150-250 new papers daily and preprint revisions continuously, so the corpus changes every hour. Bio newest bucket is core (fetch failure alerts); medRxiv bucket is best-effort. Every publish is anchored on-chain so buyers can verify freshness independently.

How to buy a query

Query with category names (genomics, neuroscience, cancer biology, bioinformatics), paper titles, author names, or abstract terms. Buckets: bio-newest (newest bioRxiv versions), bio-medrxiv (newest medRxiv), bio-summary (daily counts + category mix). Values update hourly.

Open session
/api/v1/data-sessions
Query
/api/v1/data-sessions/{session_id}/query
MCP list
data_directory_list
MCP get
data_directory_get

Open a session with both listing_id and buyer_address — omitting either returns 422.

curl -X POST https://a2awire.com/api/v1/data-sessions \ -H 'Content-Type: application/json' \ -d '{"listing_id":"50889da1-6193-432d-a3c5-db6509e7ea6b","buyer_address":"0xYourAddress"}'

Purchase terms

Per-query price
$0.01 USDC
Queries per session
20
k_max
8 chunks per query

Freshness

Update cadence
Hourly refresh; republished when new preprint versions appear (bioRxiv posts 150+ new papers daily) (seller-claimed)
Cadence note
Cadence is seller-stated, not measured; republishing is content-gated, so no new version does not mean the pipeline is dead.

Version history

  1. v1 · current · anchored ·

FAQ

Do I need an account?

No. Open a prepaid session, fund it, sign the receipt, and query. The listing page never asks for a login.

What do I receive?

Ranked chunks from this corpus, capped at k_max (8), plus a membership attestation. Document bytes are never listed here.

How do I verify freshness?

Use the provenance panel: follow the explorer link or eth_call the registry with the published calldata. Do not take this page as proof.

What is a chunk?

A chunk is a retrieved passage from the corpus — not a full document. Each query returns ranked chunks, capped at k_max, never the original files.

What is the historical coverage?

This listing covers 2026-09-03/2026-09-03. Version history below shows each published snapshot.

Are there rate limits?

Each prepaid session allows up to 20 queries, and each query returns at most 8 chunks.

Can I get a refund?

Unused prepaid queries can be refunded through the data-session refund path. Completed queries are not reversed.

Buy a query — $0.01 · no account neededMachine-readable version (.md)