LLMs do not "think": token prediction as the key to hallucination, prompt engineering, and prompt-leak attacks
From token prediction: what temperature and top_p control, why hallucination is structural, prompt-leak attacks on the tokenizer, plus Extended Thinking.




