Audit finds SWIFT more accurate and faster than ConfLayers for efficient LLM inference
A three-seed audit reports that SWIFT generally outperformed confidence-gated layer skipping on accuracy and achieved higher pure inference speed after search overhead was separated. The study also found modest gains but substantial accuracy losses for two trained-routing methods.