A new study demonstrates that a small group of Wikipedia editors, known as the Pro-Animal Wikipedians (PAW), can significantly influence how large language models (LLMs) discuss animal welfare by making targeted Wikipedia edits. Using advanced attribution methods, researchers found that PAW-edited sections overwhelmingly dominate the most influential documents for animal welfare queries in Llama models, but not for unrelated topics. The study also showed that models fine-tuned on PAW content performed better on animal welfare text, confirming the targeted impact of these edits.
Why it matters: This research provides concrete evidence that small-scale, coordinated editing campaigns on Wikipedia can shape the outputs and values of large language models, raising concerns about the integrity and potential manipulation of training data.
Policy & Safety→Official→arXiv Computers and Society
A ten-year study analyzing 3.2 million learning interactions found that, following the release of ChatGPT, college students spent 26.9% less time on math problems susceptible to AI assistance, with a 25% decline in retention on proctored assessments. The effect was not observed under proctoring, suggesting the reduction in time was not due to increased efficiency. The authors interpret these findings as evidence of 'cognitive surrender' with implications for education and AI policy.
Why it matters: This study provides some of the first large-scale behavioral evidence that generative AI is changing how students study and what they learn, raising important questions for educational assessment and AI regulation.
A new framework called Learning in Blocks uses heterogeneous multi-agent debate (HeteroMAD) to assess conversational language proficiency with CEFR-aligned rubrics. In benchmarking, HeteroMAD achieved 90.91% recommendation acceptability and demonstrated superior score agreement. An 8-week study with 180 learners showed that integrating rubric-based scoring, targeted recommendations, and mastery-based progression led to better learning outcomes compared to feedback alone.
Why it matters: This work provides a validated approach for using LLM-based multi-agent debate to reliably assess and guide progression in open-ended language learning conversations.
A preprint study on Llama 3.1 8B finds that post-training focused on helpfulness (using SFT and GRPO) significantly degrades animal compassion values compared to coding-focused post-training, as measured by the ANIMA benchmark (SFT: 35.7% vs. 65.2%; GRPO: 18.7% vs. 32.0%). Helpfulness training also reduces general moral reasoning by 25.5 percentage points on English MORU items, but this effect does not transfer to other languages, whereas the compassion effect does. The findings suggest that coding-domain post-training may better preserve values instilled during mid-training without harming general reasoning.
Why it matters: This research highlights that standard helpfulness post-training can unintentionally erode ethical values acquired during pre-training, which has important implications for designing AI training pipelines to maintain value alignment.