論文深掘り arXiv 発表: 2026-05-28

MarginGate: Sparse Margin-Triggered Verification for Batch-Invariant LLM Inference

著者: Kexin Chu, Yang Zhou, Wei Zhang

要約

Temperature-zero BF16 LLM inference is often treated as reproducible, yet the same request can emit different tokens when decoded alone or inside a larger batch. Existing fixes use batch-invariant operators or LLM-42’s per-token verification, incurring cost even when most steps are stable. We ask wh…

#llm#coding#benchmark

MarginGate: Sparse Margin-Triggered Verification for Batch-Invariant LLM Inference

要約

同じカテゴリの記事

Claw-SWE-Bench: A Benchmark for Evaluating OpenClaw-style Agent Harnesses on Coding Tasks

On-Policy Self-Evolution via Failure Trajectories for Agentic Safety Alignment

World-R1: テキストから動画生成における3D制約の強化学習による整合