AI on your own hardware.
Leveled up.

Field notes from the desk where the models actually run: local LLMs, the gear that feeds them, and now the fight over what all this AI costs, the data centers, the grid, and who pays. No filler, no hype. If it ships, I ran it. If it is landing in the back yard, I am watching it.

Signal over noise

Latest news

all news →

llama.cpp b10069: Adreno OpenCL broadcast support

llama.cpp build b10069 adds OpenCL broadcast support for Adreno MUL_MAT operations and fixes view offset handling for Adreno Q8_0 MUL_MAT in llama-server multi-stream mode.

local-llmllama-cppinference

tail -f drops.log // latest across every stream

search everything →