Ollama vs MS Foundry Local: Benchmarking Local LLMs on an RTX 5070
Following up on Foundry Local in VS Code, I wanted hard numbers. I built Ollama BenchRig to benchmark llama.cpp against ONNX Runtime GenAI on real developer workloads.
Following up on Foundry Local in VS Code, I wanted hard numbers. I built Ollama BenchRig to benchmark llama.cpp against ONNX Runtime GenAI on real developer workloads.