AI translation of literary texts is "fine", but readers still prefer human translations Paper • 2606.26040 • Published Jun 24 • 9
Error Span Annotation: A Balanced Approach for Human Evaluation of Machine Translation Paper • 2406.11580 • Published Oct 18, 2024
AI use in American newspapers is widespread, uneven, and rarely disclosed Paper • 2510.18774 • Published Oct 21, 2025 • 1
CaLMQA: Exploring culturally specific long-form question answering across 23 languages Paper • 2406.17761 • Published Jun 25, 2024
People who frequently use ChatGPT for writing tasks are accurate and robust detectors of AI-generated text Paper • 2501.15654 • Published Jan 26, 2025 • 16
Open Pangram Collection Open models and datasets based on Pangram's ICLR 2026 EditLens paper licensed for noncommercial use ONLY under CC BY-NC-SA 4.0 • 4 items • Updated Apr 24 • 19
view article Article Supercharge your OCR Pipelines with Open Models +5 merve, ariG23498, davanstrien, hynky, andito, reach-vb, pcuenq • Oct 21, 2025 • 319
CLIPPER: Compression enables long-context synthetic data generation Paper • 2502.14854 • Published Feb 20, 2025 • 11
One Thousand and One Pairs: A "novel" challenge for long-context language models Paper • 2406.16264 • Published Jun 24, 2024 • 1
princeton-nlp/Llama-3-8B-ProLong-64k-Instruct Text Generation • 8B • Updated Oct 31, 2024 • 8.33k • • 13