Researchers have created a unique language model called Talkie that is "stuck" in 1930. This project acts as a digital time capsule because the artificial intelligence has zero awareness of events, technologies, and historical turning points that occurred after December 31, 1930. The model was trained exclusively on public domain texts, completely isolating the AI from 20th and 21st-century information. The choice of this unusual time barrier is driven by US copyright laws, according to which works enter the public domain after 95 years.

Background and technical features of the experiment

Creating the Talkie model required colossal resources. The training dataset is equivalent to approximately 234 billion pages of English text digitized using modern Optical Character Recognition (OCR) systems. The main challenge for the authors was protecting the database from "contamination": introducing even a single document created after 1930 into the system could irreversibly ruin the purity of the experiment. The scientists aimed to obtain a pure digital interlocutor whose worldview is formed solely on the literature of the early 20th century and earlier eras.

During large-scale tests, researchers checked the AI's reaction to 5,000 key historical milestones that remain in the future for it. The model proved completely incapable of predicting global conflicts, including World War II or the rise of the Nazis to power. Moreover, when the bot was told about modern technologies, its reaction was filled with genuine technological shock: Talkie had never heard of the internet, smartphones, television, or the space race, perceiving such concepts as something purely fantastic.

Contradictory data

Despite strict isolation from modern information, the experiments with the model revealed interesting cognitive paradoxes. On the one hand, researchers note the absolute purity of the historical context and the bot's inability to foresee technological progress. On the other hand, skeptics and security experts raise questions about the limits of dataset "legal purity": critics argue that hidden anachronisms or indirect indications of future discoveries could still persist within the massive 234-billion-page dataset, which the model interprets through complex linguistic analogies to create the illusion of genuine 1930s human thinking.

Unusual abilities and practical benefits

One of the most striking results was a programming experiment. When the bot was asked to write code in Python (created in 1991), the algorithm—having never seen computers in its life and not even knowing the word "computer"—managed to generate syntactically plausible code using deep linguistic analogies. In another test, the model was trained on pre-1911 materials and tested to see if it could independently derive the general theory of relativity, much like Albert Einstein did in 1915. Currently, the bot is available for public chat online, where scientists and enthusiasts can observe its discussions despite the AI's tendency toward classic "hallucinations" and inventing nonexistent facts about the past.