Batch-invariant inference nodes for guaranteed reproducibility in ComfyUI. ThinkingMachines + ECHO 2.0 + Nemotron patterns.
-
Updated
Jan 14, 2026 - Python
Batch-invariant inference nodes for guaranteed reproducibility in ComfyUI. ThinkingMachines + ECHO 2.0 + Nemotron patterns.
In bfloat16 a trainer gives the same token a different log probability depending on its batch shape; measured on 8 models, with the controls that survived
Batch-invariance verifier for llama.cpp continuous batching. Per-cell diffing, a five-verdict output contract, and signed GREEN or RED certificates. Mock-sourced passes are non-promotable by construction. 138 tests, CI green.
Bitwise batch-invariant matmul and attention kernels for MLX on Apple Silicon: identical logits at any batch size, with the harness that proves it.
量測 LLM 推論的決定性:同一 prompt 重複執行時,輸出從第幾個 token 開始分歧。
Measures whether an LLM inference engine returns the same tokens when requests share engine steps. Real results for OpenVINO GenAI on an Intel CPU and Arc iGPU.
To associate your repository with the batch-invariance topic, visit your repo's landing page and select "manage topics."