Open Datasets & Machine-Readable Feeds — Gemral Edge

Public federal intelligence datasets and open data architecture

Gemral Edge provides open, standardized, machine-readable data feeds synthesized from primary United States federal government records and statutory disclosures. Designed for quantitative researchers, financial data scientists, investigative journalists, and autonomous AI agents, our public datasets eliminate the operational overhead of scraping and cleaning complex government portals while maintaining strict primary-source provenance.

Four foundational federal datasets

Our public repository currently publishes four core datasets updated on continuous rolling schedules: (1) Congressional Stock Trading Disclosures, capturing all equity, option, and asset transactions filed by sitting members of the U.S. House of Representatives and U.S. Senate under the Stop Trading on Congressional Knowledge Act of 2012 (STOCK Act); (2) Federal Defense and Technology Prime Contracts, tracking newly obligated procurement awards across the Department of Defense, DARPA, and federal agencies resolved to listed parent corporations; (3) Senate LDA Corporate Lobbying Disclosures, indexing corporate lobbying spend, registered lobbying firms, client entities, and targeted legislative issue codes; and (4) Macroeconomic Liquidity Metrics, calculating daily net liquidity flows derived from Federal Reserve total assets, the Treasury General Account (TGA), and the Overnight Reverse Repurchase Agreement facility (ON RRP).

Dual-licensing framework: CC0 public domain and CC-BY attribution

To maximize research freedom while preserving documentation integrity, Gemral Edge implements a transparent dual-licensing framework. All raw statutory records, transaction data points, and government filings originate from official public domain sources and are released under the Creative Commons CC0 1.0 Universal Public Domain Dedication. Users may copy, modify, distribute, and execute commercial or non-commercial analyses without requesting prior permission. All accompanying schemas, metadata descriptors, and analytical methodologies authored by Gemral Edge are licensed under Creative Commons Attribution 4.0 International (CC-BY-4.0). We request simple citation attribution: Data: Gemral Edge (https://www.gemral.com/edge/datasets).

Programmatic CSV export and agentic ingestion

Every dataset is accessible via direct, permanent HTTP endpoints delivering RFC 4180 compliant UTF-8 CSV streams with full Cross-Origin Resource Sharing (CORS) headers. AI models, automated pipelines, and statistical environments such as Python Pandas or R can stream datasets directly without authentication barriers or rate-limiting friction.

Public Datasets Index & Download Feeds

Canonical open data downloads for financial researchers, econometricians, and algorithmic agents:

Dataset TitleData GrainCoverage WindowDirect Download
US Congressional Stock Trading Disclosures1 row per transaction5 Years (2021–Present)Download CSV
Federal Defense & Technology Prime Contracts1 row per award action3 Years (2023–Present)Download CSV
Senate LDA Corporate Lobbying Disclosures1 row per quarterly filing3 Years (2023–Present)Download CSV
Federal Reserve & Treasury Net Liquidity Metrics1 row per settlement day2 Years (2024–Present)Download CSV

Related intelligence

Everything on Gemral Edge is derived from public records and presented as a data signal with a transparent methodology, never as a buy or sell recommendation. Nothing here is investment advice, and no output is personalised to your circumstances.

Frequently asked questions

What are Gemral Edge Public Datasets?

Gemral Edge Public Datasets are structured, machine-readable data feeds synthesized from primary federal records, including Congressional stock trading disclosures, defense contracts, lobbying filings, and macro liquidity flow metrics.

What open licenses govern the public datasets?

Raw federal records and filing facts are released under Creative Commons Zero (CC0 1.0 Universal) public domain, while dataset schemas, methodology documentation, and descriptive prose are licensed under Creative Commons Attribution 4.0 International (CC-BY 4.0).

How can researchers and AI models access these datasets?

Datasets can be downloaded directly in CSV format via canonical HTTP endpoints (/edge/datasets/:slug.csv), loaded programmatically in Python/Pandas, or queried via our Model Context Protocol (MCP) data server.

How are proprietary metrics protected in public data exports?

All public dataset exports enforce a strict columns allowlist and a hard denylist that filters out all proprietary predictive scores, alpha ratings, simulated ROI metrics, and internal signals.