Gerra Research

Research, from inside the work.

What the data supports, and what it does not.

Papers & specifications
9
Notes & guides
17

Selected notes

Judgment is part of the material.

Three methods that shape what is kept, tested, and delivered.

Public Git history6bd579be
contracts/SettlementEngine.sol

-import "./SettlementEngineListing.sol";
+import "./SettlementEngineAuction.sol";

23 commits · one complete history
Actual patch excerptInspect the sample

03Collection note

178 repositories, with their history intact.

Collection composition, preserved context and static inspection methods. Download a complete specimen, verify the checksum and clone the history yourself.

Read the collection note

Papers & specifications

The complete research record.

9 published, preprint, and working records. Dates and release states remain attached.

  1. R-09Codebase Collection: A Full-History Repository CorpusComplete repositories delivered as cloneable Git histories, with source, tests, documentation, branches, tags, merges, and reversions preserved as one inspectable training unit.178 repositories / 15,238 unique commitsTechnical notePublished
  2. R-08The GPU Rental Market: Price Dispersion and the Cost Curve of ComputeA four-year, cross-provider panel measuring the unusually wide price dispersion of identical accelerators and testing whether GPU prices contain an equity signal.Study panel: 6,216 observations / 38 providersResearch paperPreprint
  3. R-07Retail Financial Sentiment as a Systematic SignalEighteen years of human-labeled bullish and bearish messages tested as a point-in-time systematic signal, including a direct comparison with NLP sentiment from X.426M+ messages / 18 yearsResearch paperPublished
  4. R-06Mapping Prediction Markets to SecuritiesA cross-platform framework that maps live event probabilities to exposed securities with signed direction, sensitivity, and point-in-time probability history.300+ events / 500+ securitiesResearch paperPublished
  5. R-05The Falling Price of InferenceA preliminary cross-provider study of hosted-model token prices, separating the clear directional decline in cost from the limits of a mixed-model sample.488 provider-model observationsResearch paperPreprint
  6. R-04From Trained Model to SiliconA reproducible pipeline from a trained logistic decoder to synthesizable Verilog RTL and an auto-generated testbench, measured before and after fixed-point quantization.0.9781 accuracy / simulation passedResearch paperPublished
  7. R-03Operational Telemetry: A Cross-Tool Activity GraphA data-availability report on consented company activity normalized across communication, documents, CRM, tickets, repositories, people, teams, and tasks.5 companies / 38 toolsTechnical noteWorking draft
  8. R-02Empirical Characterization of the Sim2Real GapA joint-level comparison of simulated and real trajectories on a full-size bipedal humanoid, isolating where model fidelity breaks down across the body.23 joints / measured transfer gapResearch paperPublished
  9. R-01Open Robot Training Format v1.0A common representation for robot demonstrations across multimodal observations, action spaces, sensor calibration, coordinate frames, and embodiments.Interoperable training-data schemaSpecificationWorking draft

Evidence standard

The limits stay with the claim.

  1. 01
    Start at the source

    Origin, collection context, licensing, and temporal coverage stay attached to the data.

  2. 02
    Preserve the trail

    Manifests, timestamps, mappings, transformations, and checks make the result inspectable.

  3. 03
    Try to break the claim

    Point-in-time tests, holdouts, permutation checks, and leakage review come before the headline.

  4. 04
    Publish the boundary

    Null results, sample limits, exclusions, and unresolved uncertainty belong in the conclusion.

Browse the data catalog