FireTofu
Researchers can now reverse-engineer LLM prompts from output text with near-perfect accuracy

Technology · en

Researchers can now reverse-engineer LLM prompts from output text with near-perfect accuracy

The Decoder · Aug 12, 2026, 5:32 PM UTC

Researchers at IIT Bombay and Adobe Research have built an inverse language model that reconstructs the original prompt from an LLM's output with near-perfect accuracy. Their method, called "Previous-Token Prediction," doesn't need access to model weights and works across different models.…