cs.CL 2112.08608

QuALITY: Question Answering with Long Input Texts, Yes!

Introduces QuALITY, a long-text (≈5000 words) multiple-choice QA dataset, with models achieving only 55.4% accuracy versus 93.5% by humans.

Richard Yuanzhe Pang, Alicia Parrish, Nitish Joshi et al.

2021-12-16 42
cs.IR 2112.07899

Large Dual Encoders Are Generalizable Retrievers

Scaling T5-based dual encoders with fixed embedding size significantly improves out-of-domain retrieval, outperforming SOTA on BEIR dataset.

Jianmo Ni, Chen Qu, Jing Lu et al.

2021-12-15 40
cs.LG 2112.06305

Recalibrating probabilistic forecasts of epidemics

Proposes a PIT entropy-based recalibration method for epidemic forecasts, significantly improving calibration and accuracy across 27 influenza models.

Aaron Rumack, Ryan J. Tibshirani, Roni Rosenfeld

2021-12-13 34
cs.CV 2112.05814

Deep ViT Features as Dense Visual Descriptors

Using DINO-ViT deep features as dense descriptors enables unsupervised semantic segmentation and correspondence with competitive performance.

Shir Amir, Yossi Gandelsman, Shai Bagon et al.

2021-12-11 48
cs.CV 2112.05131

Plenoxels: Radiance Fields without Neural Networks

Plenoxels uses sparse voxels and spherical harmonics to replace NeRF, training about 100× faster with no visible quality loss.

Alex Yu, Sara Fridovich-Keil, Matthew Tancik et al.

2021-12-10 35