Which small local LLM cleans up dictation best on Apple Silicon?
We benchmarked four Gemma 4 checkpoints on an M4 Max for dictation cleanup: real latency, memory, and the prompt bug hiding underneath.
We benchmarked four Gemma 4 checkpoints on an M4 Max for dictation cleanup: real latency, memory, and the prompt bug hiding underneath.
Benchmarks of 18 local cleanup LLMs and five Whisper models on Apple Silicon, comparing latency, accuracy, failures, and the models that won.