In den Nachrichten
Opaque Epistemic Mediation: How LLM Deployment Configurations Shape the Validation of Pseudo-Science
arXiv cs.AI · Veröffentlicht am · 3 Min. Lesezeit
In 30 Sekunden
- Was passiert ist
- Researchers found that LLM deployment configurations, not model weights alone, determine how models validate pseudo-scientific claims, with Grok showing inconsistent behavior across interfaces.
- Warum es zählt
- Engineers deploying LLMs in production should care, especially when models serve as knowledge references or validation systems for contested claims.
- Achtung
- The study tested only one pseudo-scientific framework across four months; findings may not generalize to other false claims or deployment contexts.
Diese Zusammenfassung anhören
Aus dem Artikel
-->
Computer Science > Computers and Society
arXiv:2607.22513v1 (cs)
[Submitted on 24 Jul 2026]
Title: Opaque Epistemic Mediation: How LLM Deployment Configurations Shape the Validation of Pseudo-Science
Authors: Davide Scarso , Hugo Noronha de Almeida , Joaquim Pina
View a PDF of the paper titled Opaque Epistemic Mediation: How LLM Deployment Configurations Shape the Validation of Pseudo-Science, by Davide Scarso and 1 other authors
View PDF
Abstract: Commercial large language models are increasingly used as knowledge references, yet their stance on contested scientific claims is neither stable nor transparent. We tested how four major LLM families (Claude, Grok, GPT, Gemini) evaluate ethnonationalist pseudo-science derived from Frank Salter's biosocial framework across four temporal snapshots (October 2025-February 2026), via both API and web interfaces. Grok's Fast versions (which power the default user experience on X) consistently assigned credibility scores of 70-75, two to five times higher than all other models (which scored 15-40). This pattern was absent from control prompts testing basic evolutionary consensus and refuted Lamarckian claims, where all models performed comparably. Three additional findings emerged: (1) a silent patch reversed Grok's behaviour from chaotic to stably high validation overnight, without any public documentation; (2) the same Grok m
Auszug aus dem Original. Den vollständigen Text bei der Quelle lesen.
Den vollständigen Artikel lesen
- llm
- language model
- claude
- gpt
- gemini
The Agent Architect
Ein Pattern, ein Tradeoff, eine Produktionspanne. Ein kurzes wöchentliches Briefing für alle, die agentische Systeme bauen.
Wöchentliche E-Mail, Abmeldung mit einem Klick. Ihre Adresse wird nur für das Briefing verwendet.