Working notes on distributed data pipelines and model evaluation.
Benchmarking throughput for streaming ingestion under variable batch sizes. Early results suggest the flush_interval parameter dominates tail latency more than buffer capacity does.
Reproducibility harness is being rewritten. The previous version relied on wall-clock seeding, which made cross-machine comparison unreliable.
This site hosts static notes only. No public API is exposed.