I am checking for other reliable models to replace it with, but I have not found one better than it. Model 1 is perfect and does not have any issues that is why I am now giving options to users on whether they want to use Model 2.
Maybe you can add an option for the users to communicate with their local LLM as I mentioned before? For example I have an Ollama with a gemma model. I can use the website to scrap texts from the forum, then throw them to my LLM to check, and show the result on the website in real time. I know gemma isn't the best model for testing AI texts, but this is just an example. This means users don't have to download the model manually and have more control over which models they wanted to use.