cs.CL 2412.08905

Phi-4 Technical Report

phi-4 is a 14B parameter model leveraging synthetic data, surpassing GPT-4 in STEM QA, with innovative training strategies.

Marah Abdin, Jyoti Aneja, Harkirat Behl et al.

2024-12-12 32
cs.CL 2411.15594

A Survey on LLM-as-a-Judge

Defines LLM-as-a-Judge with focus on reliability, bias mitigation, and multi-scenario adaptation, proposing a standardized evaluation framework.

Jiawei Gu, Xuhui Jiang, Zhichao Shi et al.

2024-11-24 30
cs.CL 2411.13676

Hymba: A Hybrid-head Architecture for Small Language Models

Hymba employs a hybrid-head architecture combining transformer attention and state space models, achieving superior efficiency and performance with fewer parameters, surpassing comparable small models.

Xin Dong, Yonggan Fu, Shizhe Diao et al.

2024-11-21 109 citations 27
cs.CL 2411.10636

Mitigating Extrinsic Gender Bias for Bangla Classification Tasks

Proposes RandSymKL, combining symmetric KL and cross-entropy, to mitigate extrinsic gender bias in Bangla classification tasks, achieving 90.66% accuracy and bias reduction.

Sajib Kumar Saha Joy, Arman Hassan Mahy, Meherin Sultana et al.

2024-11-16 61