@realeigenvalues @RealTjDunham @teortaxesTex their inference endpoint is just llama.cpp serving quantized mistral 7b with --control-vector-scaled yapping.gguf 10.0
cited on: mistral-7b
Reproduced against link rot, credited and linked to its original. Yours and you’d rather it weren’t here? Open an issue.