Matches my superficial experiments with trying to tweak Ollama's "modelfile" using some LLaMa- or gpt-oss-based instruction-tuned model as "base".
I need to experiment more with base models. The time from the end of 2019 onwards, when I first came across talktotransformer, it felt so magical.
Getting meaningful things out of these things can feel so... restraining.
And on the other hand: I'm tbh freshly stuck in the stage of being amazed at what current frontier coding models and apps can do.
Matches my superficial experiments with trying to tweak Ollama's "modelfile" using some LLaMa- or gpt-oss-based instruction-tuned model as "base".
I need to experiment more with base models. The time from the end of 2019 onwards, when I first came across talktotransformer, it felt so magical.
Getting meaningful things out of these things can feel so... restraining.
And on the other hand: I'm tbh freshly stuck in the stage of being amazed at what current frontier coding models and apps can do.