Local AI: Stress-Testing Local Models on Real Tasks
Part 3 of the “Local AI” series — Coding, reasoning, medical summarization, and clinical decision support Benchmarks tell you how fast a model runs. They don’t tell you if it can actually do anything useful. In Part 1 we set up Ollama, and in Part 2 we measured generation speeds. Now we push these models…