OpenSquilla Open-Source Project Uses Smart Routing and Local Retrieval to Cut LLM Costs
开源项目OpenSquilla:智能路由与本地检索,大幅降低LLM使用成本
OpenSquilla combines local model routing, vector retrieval, incremental sending, and cache hits to reduce transmitted tokens by more than 90%, while routing simple tasks to cheaper models and complex tasks to stronger models without spending tokens on the routing decision.
Why it matters: HKR-H/K/R all pass, but the source appears to be a single X project post; repo traction, test setup, and limits are not disclosed. Score lands at the featured threshold for practical open-source cost tooling.