Custom datasets from public sources, cleaned and delivered
One-off extraction and cleaning from EDGAR, CFTC, and exchange archives. The data is public; getting it into a usable shape is the work. We do that work every day for our own products.
Sources we work with
SEC EDGAR
Any form type, any date range. Form 4s, 13Fs, 8-Ks, prospectuses, full-text extractions from filing bodies.
CFTC
COT reports in every flavor: legacy, disaggregated, TFF. Full history, cleaned and joined.
Exchange archives
Funding rates, open interest, liquidations, options chains from major crypto venues.
Official statistics
Treasury yields, BIS policy rates, BLS release data. Parsed from the primary source, not resold feeds.
Who this is for
Researchers who need a clean panel instead of three weeks of scraping. Funds validating a signal before committing engineering time. Fintech teams that want a dataset behind a feature without building the pipeline themselves.
How an engagement works
1
Describe the dataset
Source, fields, date range, delivery format. A rough sketch is fine; we'll firm it up together.
2
Sample first
You get a free sample extract before committing. If the data can't support what you need, we say so early.
3
Full delivery
CSV, Parquet, or a hosted API endpoint with documentation. One-off or refreshed on a schedule.
Sample the quality for free
Our free data hub runs on the same extraction pipelines we'd build for you. Download a CSV from any dataset page and judge the cleaning yourself.
Pricing
On request. A one-off historical extract prices very differently from a refreshed feed with an SLA. Describe the dataset and you'll have a quote within a day.
Describe your dataset
Frequently Asked Questions
What formats do you deliver?
Do you scrape sites that prohibit it?
Can I get a sample before paying?
Can the dataset stay updated?
Data on this page is provided for informational purposes only and is not financial advice. See our editorial policy.