Efi's Demo Lab
Research showcase

When training data changes the story

A literary model study separates style corruption from factual knowledge poisoning, using GPT-2 and Llama.

Inside the project

Research showcase
tolkien gpt training
Original artifact ↗

How it works.

Define the attack

Modify literary training material or factual question-answer data.

From the original project.

Saved artifacts · click to inspect

tolkien gpt training
Original artifact ↗
model response categories
Original artifact ↗
figur 6 stats analysis
Original artifact ↗

Continue in the source.

Open the source notebook in Jupyter, Colab, or the environment described in the README. Data and model downloads may be required.

git clone https://github.com/eforus-overseer/Literary-LLM-Knowledge-Data-Poisoning.git
Read the setup and requirements ↗