Open-source peer
exchange-core
150 ns per matching and 5M operations/sec on a single order book
Published README benchmark on older Xeon hardware; excludes network interface latency, IPC, and journaling from the latency figure.
FerroMatch is in the same broad core-latency conversation, but its current public end-to-end snapshot is still below this published top-end throughput claim.
exchange-core README