arXivLLMs & Text
A Multi-Task Evaluation of LLMs' Processing of Academic Text Input
Large language models like Google's Gemini struggle with processing academic text, showing reliable summarizing and paraphrasing skills but poor text grading and reflection abilities, hence their unchecked use in peer reviews is discouraged.
Featured in No. 110 on 20 Aug 2025 · 5 days after release · 6 citations today
- Released
- 15 Aug 2025
- First featured
- No. 110 · 20 Aug 2025
- Citations (Semantic Scholar)
- 6
- Influential citations
- 0
- Published in
- Not yet, as far as Semantic Scholar knows
- Shares when featured
- 8
- Identifier
- arXiv:2508.11779
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).