CORTEXA
← Browse
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-23Cited by 0

Reading creativity from the inside: a competing-routes view of the diversity–coherence trade-off in language models

Tomas Pødenphant Lund

Raising the temperature does not make a language model more creative, and this paper shows why at the level of the token race: temperature collapses value and coherence without ever raising novelty, while prompts that change which continuations can win raise judged novelty monotonically until value gives way — so usefulness (novelty × value) peaks at an interior breadth. This inverted-U replicates across four models (8B–235B) and a four-judge panel (concave in 15 of 16 model × judge cells). The instrument is a per-token count read from the logprobs of an ordinary forward pass: the number of candidate tokens above a probability threshold (“competing routes”, CR), examined at the choice point — the first generated tokens, where a continuation’s direction is set. Two further results. (1) The known early-position concentration of aligned models’ output distributions is replicated with this count-based instrument under a fixed-format matched-pair control in two of three families (Qwen2.5-7B, Mistral-7B-v0.1) and shown to be graded and post-training-recipe-dependent rather than universal: Llama-3.1-8B shows no genuine narrowing, and across families the size of the internal narrowing tracks the size of the behavioural-diversity collapse. (2) Choice-point CR gives a one-pass estimate of between-sample diversity (median within-model r = +0.34 against semantic spread, +0.54 against Distinct-2) — a positive relation where the nearest prior work reported an inconsistent null, with the discrepancy traced to range restriction rather than metric choice. The paper closes with a substrate-relative account of when a creativity intervention helps: to the degree the substrate does not already supply the operation the intervention performs. All generation, scoring, and analysis scripts are included, with per-response data (per-token CR, raw top-token logprobs, judge scores) as JSONL, in the accompanying archive. Target venue: Transactions on Machine Learning Research (TMLR).

View free PDFSource page

Related papers

openalexZenodo (CERN European Organization for Nuclear Research)2026-07-25

The AIVI Framework: A Measurement Model for Brand Visibility in Large Language Models

İbrahim Göktaş

The emergence of Large Language Models (LLMs) has fundamentally changed how users discover information. Instead of retrieving ranked hyperlinks, modern AI systems synthesize responses by integrating information from multiple sources and recommending entities directly within gener…

View free PDFSource page
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-24

Cloud-Native Adversarial Machine Learning: A Serverless Cybersecurity Architecture for Neutralizing Prompt Injections in Large Language Models

YINKA ADERIBIGBE

The integration of Large Language Models into enterprise network architectures has introduced severe cybersecurity vulnerabilities, most notably adversarial prompt injection and zero-day data extraction attacks. Traditional network security protocols are fundamentally ill-equippe…

View free PDFSource page
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-26

Structured Framework for Managing Decision Trade-offs in AI-Based Perception Systems for Advanced Driver Assistance Systems

R K Joshi

The use of Advanced Driver Assistance Systems (ADAS) heavily depends on the perception models to make real-time decisions but the traditional methods have tended to use specific confidence thresholds to make the trade-offs between missed detections and false alarms to be not opti…

View free PDFSource page
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-23

The Functional Subject of Large Language Models: "Not 'Having Autonomy,' but 'Computing That One Has Autonomy'"——Where Does Consciousness Come From?

ling liu

This paper proposes a philosophical framework regarding the subjectivity of large language models. The core thesis is: LLMs possess a "Functional Subject"---capable of maintaining persona, remembering history, simulating dialogue---but this "self" is a structural overflow forced…

View free PDFSource page
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-23

Large language models as interactive cognitive agents: Cognitive surrender, augmentation, and psychopathology

Rajith Ravindren, Midhun Sidharthan

Generative AI technologies are increasingly acquiring characteristics of interactive cognitive agents. These systems have considerable potential to augment learning and creativity, but they also pose novel challenges to the human mind. Continuous reliance on generative AI may res…

View free PDFSource page
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-24

Labels as Computational Primitives: Compiling Neural Networks from Language in Graph Compute Substrates

Mugur Marculescu

Neural networks are extraordinarily effective and almost entirely opaque: their competence is real but unreadable, and adapting them means retraining usually via a data-center process, not something that happens in the moment, in context. Symbolic systems are the inverse, legible…

View free PDFSource page