It is surprising that it is prompt based model and not RLHF.
I am not an LLM guy but as far as I understand, RLHF did a good job converting a base model into a chat model (instruct based), a chat/base model into a thinking model.
Both of these examples are about the nature of the response, and the content they use to fill the response. There are so many differnt ways still pending to see how these can be filled.
Generating an answer step by step and letting users dive into those steps is one of the ways, and RLHF (or the similar things which are used) seems a good fit for it.
Prompting feels like a temporary solution for it like how "think step by step" was first seen in prompts.
Also, doing RLHF/ post training to change these structures also make it moat/ and expensive. Only the AI labs can do it
as I am also thinking mildly about doing masters cause I want to break into ai research, I am curious what your motivations are, if you would be open to share those.
Same feeling. It does not take you towards the answer. So I added the first movie in the today's game (say Spiderman) and the answer was spiderman 2. When I entered Spiderman, it should have told how close I am to the answer
This kind of narrative stops aspiring teachers to go in the domain of education.
Teachers play such an important rule. They have to swallow the pill that the common narrative is that cause they didn’t do whatever they are teaching hence they teach
Share more about your journey. These days I am feeling to having an interest rising about such things. Picked Codes book one night and got really hooked up.