Computer Science editorial
Engine-Transfer-Bench: An Evidence-Based Benchmark for Document Compilation Engine Selection
The core problem
Innovation
ETB evaluates engines across **GitHub Actions** hosts running **macOS, Ubuntu, and Windows**, with **N = 4,211 compiles per host**. The benchmark defines four tasks: reliability, latency, text consistency, and failures. Reliability is measured as successful compilation rate; latency captures time-to-PDF; text consistency uses the **S_pdf** metric to detect real content divergence; and failure analysis classifies the architectural causes of unsuccessful compiles. The harness is pinned so that engine versions, package sets, and host images remain comparable across runs. A key methodological choice is the separation between **portable LaTeX documents** and **engine-specific templates**. On **702 portable LaTeX documents**, the tested engines succeed at **100%**, which isolates latency as the primary selection factor. Failures concentrate in **107 engine-specific templates**, where font, layout, and asset assumptions differ. The paper also validates the text-consistency metric with a **50-pair validation**, achieving **94% precision** for real content divergence. Formally, if denotes the text-consistency score between PDF outputs and , the validation targets the decision rule
Why it matters
Who should read this
Opening member contentโฆ