Ollama runs local models. Hermes is an agent that can use local models. Magnitude combines both in one: it profiles your hardware, recommends and downloads models, then configures and runs them inside the agent. Nothing else to set up.
No. When using Magnitude’s built-in local models, your prompts and files stay on your machine.When using your own inference server, model requests go to the endpoint you configure.