Following is model that seems to perform best on MacBook here atm....running as part of a Koog agent (showing telemetry using @langfuse.bsky.social). Full agent run takes about 2s with LLM calls taking (in this case) 1.52s and 0.31s. Memory footprint ~20GB. qwen3.5:35b-a3b-coding-nvfp4
2 likes 0 replies
?