arXiv Open Access 2025

EdgeRunner 20B: Military Task Parity with GPT-5 while Running on the Edge

Jack FitzGerald Aristotelis Lazaridis Dylan Bates Aman Sharma Jonnathan Castillo +15 lainnya

Lihat Sumber

Abstrak

We present EdgeRunner 20B, a fine-tuned version of gpt-oss-20b optimized for military tasks. EdgeRunner 20B was trained on 1.6M high-quality records curated from military documentation and websites. We also present four new tests sets: (a) combat arms, (b) combat medic, (c) cyber operations, and (d) mil-bench-5k (general military knowledge). On these military test sets, EdgeRunner 20B matches or exceeds GPT-5 task performance with 95%+ statistical significance, except for the high reasoning setting on the combat medic test set and the low reasoning setting on the mil-bench-5k test set. Versus gpt-oss-20b, there is no statistically-significant regression on general-purpose benchmarks like ARC-C, GPQA Diamond, GSM8k, IFEval, MMLU Pro, or TruthfulQA, except for GSM8k in the low reasoning setting. We also present analyses on hyperparameter settings, cost, and throughput. These findings show that small, locally-hosted models are ideal solutions for data-sensitive operations such as in the military domain, allowing for deployment in air-gapped edge devices.

Topik & Kata Kunci

cs.AI

Penulis (20)

Jack FitzGerald

Aristotelis Lazaridis

Dylan Bates

Aman Sharma

Jonnathan Castillo

Yousif Azami

Sean Bailey

Jeremy Cao

Peter Damianov

Kevin de Haan

Luke Kerbs

Vincent Lu

Joseph Madigan

Jeremy McLaurin

Jonathan Tainer

Dave Anderson

Jonathan Beck

Jamie Cuticello

Colton Malkerson

Tyler Saltsman

Format Sitasi

APA MLA BibTeX

FitzGerald, J., Lazaridis, A., Bates, D., Sharma, A., Castillo, J., Azami, Y. et al. (2025). EdgeRunner 20B: Military Task Parity with GPT-5 while Running on the Edge. https://arxiv.org/abs/2510.26550

Akses Cepat

Lihat di Sumber

Informasi Jurnal

Tahun Terbit: 2025
Bahasa: en
Sumber Database: arXiv
Akses: Open Access ✓