Open Datasets & Machine-Readable Feeds — Gemral Edge
Public federal intelligence datasets and open data architecture
Gemral Edge provides open, standardized, machine-readable data feeds synthesized from primary United States federal government records and statutory disclosures. Designed for quantitative researchers, financial data scientists, investigative journalists, and autonomous AI agents, our public datasets eliminate the operational overhead of scraping and cleaning complex government portals while maintaining strict primary-source provenance.
Four foundational federal datasets
Our public repository currently publishes four core datasets updated on continuous rolling schedules: (1) Congressional Stock Trading Disclosures, capturing all equity, option, and asset transactions filed by sitting members of the U.S. House of Representatives and U.S. Senate under the Stop Trading on Congressional Knowledge Act of 2012 (STOCK Act); (2) Federal Defense and Technology Prime Contracts, tracking newly obligated procurement awards across the Department of Defense, DARPA, and federal agencies resolved to listed parent corporations; (3) Senate LDA Corporate Lobbying Disclosures, indexing corporate lobbying spend, registered lobbying firms, client entities, and targeted legislative issue codes; and (4) Macroeconomic Liquidity Metrics, calculating daily net liquidity flows derived from Federal Reserve total assets, the Treasury General Account (TGA), and the Overnight Reverse Repurchase Agreement facility (ON RRP).
Dual-licensing framework: CC0 public domain and CC-BY attribution
To maximize research freedom while preserving documentation integrity, Gemral Edge implements a transparent dual-licensing framework. All raw statutory records, transaction data points, and government filings originate from official public domain sources and are released under the Creative Commons CC0 1.0 Universal Public Domain Dedication. Users may copy, modify, distribute, and execute commercial or non-commercial analyses without requesting prior permission. All accompanying schemas, metadata descriptors, and analytical methodologies authored by Gemral Edge are licensed under Creative Commons Attribution 4.0 International (CC-BY-4.0). We request simple citation attribution: Data: Gemral Edge (https://www.gemral.com/edge/datasets).
Programmatic CSV export and agentic ingestion
Every dataset is accessible via direct, permanent HTTP endpoints delivering RFC 4180 compliant UTF-8 CSV streams with full Cross-Origin Resource Sharing (CORS) headers. AI models, automated pipelines, and statistical environments such as Python Pandas or R can stream datasets directly without authentication barriers or rate-limiting friction.
Public Datasets Index & Download Feeds
Canonical open data downloads for financial researchers, econometricians, and algorithmic agents:
| Dataset Title | Data Grain | Coverage Window | Direct Download |
|---|---|---|---|
| US Congressional Stock Trading Disclosures | 1 row per transaction | 5 Years (2021–Present) | Download CSV |
| Federal Defense & Technology Prime Contracts | 1 row per award action | 3 Years (2023–Present) | Download CSV |
| Senate LDA Corporate Lobbying Disclosures | 1 row per quarterly filing | 3 Years (2023–Present) | Download CSV |
| Federal Reserve & Treasury Net Liquidity Metrics | 1 row per settlement day | 2 Years (2024–Present) | Download CSV |
Related intelligence
- Federal contracts radar
- Congressional trading directory
- Government contractors directory
- Calculation methodology and standards
Everything on Gemral Edge is derived from public records and presented as a data signal with a transparent methodology, never as a buy or sell recommendation. Nothing here is investment advice, and no output is personalised to your circumstances.
Frequently asked questions
What are Gemral Edge Public Datasets?
Gemral Edge Public Datasets are structured, machine-readable data feeds synthesized from primary federal records, including Congressional stock trading disclosures, defense contracts, lobbying filings, and macro liquidity flow metrics.
What open licenses govern the public datasets?
Raw federal records and filing facts are released under Creative Commons Zero (CC0 1.0 Universal) public domain, while dataset schemas, methodology documentation, and descriptive prose are licensed under Creative Commons Attribution 4.0 International (CC-BY 4.0).
How can researchers and AI models access these datasets?
Datasets can be downloaded directly in CSV format via canonical HTTP endpoints (/edge/datasets/:slug.csv), loaded programmatically in Python/Pandas, or queried via our Model Context Protocol (MCP) data server.
How are proprietary metrics protected in public data exports?
All public dataset exports enforce a strict columns allowlist and a hard denylist that filters out all proprietary predictive scores, alpha ratings, simulated ROI metrics, and internal signals.