IBM Granite 4.2 llama.cpp and Claude Code make hybrid local agents shippable
Builders spent August chasing cheaper, longer-running agents. Today the stack snapped into focus. IBM put native reasoning and sandbox-trained tool use into downloadable dense models you can run on-prem or at the edge. ggml’s llama.cpp finally treated vision and audio as…