Publiora

Menghubungkan ke Publiora...

Publiora

Large language models for pattern recognition in text data

Kosayakova, AknurIldar, KurmashevSpada, Luigi LaZeeshan, NidaBakyt, MakhabbatKhuralay, MoldamuratAbdirashev, Omirzak
IAES International Journal of Artificial Intelligence (IJ-AI) (Sinta 1)Vol. 0 No. 01 Desember 2025
DOI10.11591/ijai.v14.i6.pp5311-5332

Abstrak

Large language models (LLMs) are widely deployed in settings where both reliability and efficiency matter. We present a calibrated, seed‑robust empirical comparison of an encoder fine‑tuned model (bidirectional encoder representations from transformers (BERT)‑base) and a decoder in‑context model (generative pre-trained transformer (GPT)‑2 small) across Stanford question answering dataset v2.0 (SQuAD v2.0) and general language understanding evaluation (GLUE)-multi-genre natural language inference (MNLI), Stanford sentiment treebank 2 (SST‑2). Beyond accuracy, we assess reliability (expected calibration error with reliability diagrams and confidence–coverage analysis) and efficiency (latency, memory, throughput) under matched conditions and three fixed seeds. BERT‑base yields higher accuracy and lower calibration error, while GPT‑2 narrows gaps under few‑shot prompting but remains more sensitive to prompt design and context length. Efficiency benchmarks show that decoder‑only prompting incurs near‑linear latency/memory growth with k‑shot exemplars, whereas fine‑tuned encoders maintain stable per‑example cost. These findings offer practical guidance on when to prefer fine‑tuning versus prompting and demonstrate that reliability must be evaluated alongside accuracy for risk‑aware deployment.

Kata Kunci

BERT-baseComputational efficiencyExpected calibration errorGPT-2In context learningLarge language modelsQuestion answering

Cari jurnal yang tepat untuk naskah Anda

MatchMind AI mencocokkan abstrak naskah Anda dengan ribuan jurnal terakreditasi dan menampilkan rekomendasi terbaik beserta alasannya.

Coba MatchMind

Lihat profil lengkap jurnal ini

Waktu review, biaya APC, statistik sitasi, indeksasi Scopus, dan banyak lagi.

Buka IAES International Journal of Artificial Intelligence (IJ-AI)

Artikel ini juga tersedia di situs resmi jurnal.

Large language models for pattern recognition in text data | IAES International Journal of Artificial Intelligence (IJ-AI) | Publiora