The Most-Downloaded Vision Model You Haven't Heard Of Runs Without Python
mudler's locate-anything.cpp-gguf packages NVIDIA's LocateAnything-3B as a C++/ggml engine for open-vocabulary detection and visual grounding with no Python at inference time. Quantizations run from 9.15GB (F16) to 4.72GB (Q4_K), with Q8_0 and Q6_K reported box-identical to full precision and 3.9–5.5× speedups. Downloads rose 55% in a day to 1.42 million.