This repository was archived by the owner on Jul 24, 2026. It is now read-only.
Expose cache hit rate and model distribution in AI usage API - #44
Merged
Conversation
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to subscribe to this conversation on GitHub.
Already have an account?
Sign in.
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Important
📝 变更描述 / Description
扩展 username-scoped AI Usage API,在每个 API key 的现有滚动 Token 统计旁增加 7 天缓存命中率和模型分布。缓存命中率按 cache_tokens / prompt_tokens 计算;模型分布按输入与输出 Token 总量降序返回。响应升级为 schema version 4,现有字段保持不变。
🚀 变更类型 / Type of change
🔗 关联任务 / Related Issue
✅ 提交前检查项 / Checklist
📸 运行证明 / Proof of Work
go test ./model ./service ./router通过。go test ./...中所有可编译包通过;仓库根包因未生成web/classic/dist/index.html无法在纯后端检出中编译,这是现有前端构建前置条件。