Skip to content
Dispatch
Support
Send feedback
Revision history
Apple proposes ARBITRAGE, a step-level speculative decoding framework to reduce LLM reasoning latency by up to twofold
Original publish · no revisions.
← Back to article
Tweaks