Generative Finetuning

base model → chatbot

AI generated illustrations

chain of thought (CoT)

AI generated illustration



  user input     self-generated prompt (thinking)    response 
Revisiting LLM Reasoning via Information Bottleneck (2025)

robot laws


  system prompt    user input    chain of thought    response 

customized AIs

AI psychology

[Yao et.al, 2023]

Tree of Thoughts:
Deliberate Problem Solving with Large Language Models



foundation models

GPT - generative pretrained transformer

automatic pretraining by predicting next word
$ \begin{array}{rl} {\color{red}\equiv} & \mathrm{base\ model\ (GPT\!-\!X)} \\[0.0ex] {\color{red}+} & \mathrm{human\ supervised\ finetuning\ (SFT\ model)} \\[0.0ex] {\color{red}+} & \mathrm{human\ supervised\ reinforcement\ learning} \\[0.0ex] {\color{red}+} & \mathrm{chain{-}of{-}thought\ training} \\[0.0ex] {\color{red}+} & \mathrm{agentic{-}AI\ training} \\[0.0ex] {\color{red}\Rightarrow} & \mathrm{chat\ assistant\ (foundation\ model)} \end{array} $


a short history of the future