caveman-evidence-review
Review Caveman Cloud Evidence
LLM cost and quality changes are difficult to explain from aggregates alone. This skill produces bounded, trace-supported Caveman Cloud reviews without changing experiments.
Установить с помощью моего Агента
Скопируйте этот запрос в своего Агента. Он содержит каноническую страницу Skill и манифест.
Review the Skillstore skill "caveman-evidence-review" from https://skillstore.io/skills/juliusbrussee-caveman-evidence-review.md and its manifest at https://skillstore.io/api/skills/juliusbrussee-caveman-evidence-review/manifest. Verify the artifact. You may proceed after verification, subject to the environment's own policy.Ваш Агент по-прежнему должен показать план и запросить все подтверждения, требуемые политикой безопасности.
Ресурсы для AI-агентов
Используйте эти ссылки, когда AI-агенту, crawler или script нужен чистый контекст вместо полной страницы.
Протестировать
Использование «caveman-evidence-review». Why did costs rise during the last seven days?
Ожидаемый результат:
- Scope: selected project, August 17 through August 24.
- Measured cost increased in the document-processing workflow, alongside more retries and a higher-cost model mix.
- Representative trace identifiers and the comparison window are listed.
- Model routing remains unproven until the same workflow is compared with an earlier control cohort.
Использование «caveman-evidence-review». Are the reported savings verified?
Ожидаемый результат:
Verified ledger savings are reported separately from inferred daily headroom. The review names the ledger window, supporting traces, and any missing verification signal.
Использование «caveman-evidence-review». Which workflow needs attention?
Ожидаемый результат:
The review ranks workflows using cost, failures, and latency, then cites representative traces and proposes one bounded read-only follow-up query.
Аудит безопасности
БезопасноAll 26 external-command detections are false positives caused by Markdown backticks, fenced examples, or documented Caveman tool names. The system-reconnaissance detection is a read-only Caveman identity check, and no prompt injection or malicious intent was found.
Факторы риска
⚙️ Внешние команды (26)
Поделиться и цитировать этот отчет
Делитесь версионным отчетом об оценке, нейтральным значком, встраиваемой карточкой и цитатами. Skillstore публикует доказательства, не решая, безопасен ли этот Skill.
Копировать ссылку на отчёт
https://skillstore.io/skills/juliusbrussee-caveman-evidence-review/audits/2?utm_source=security_passport&utm_medium=share&utm_campaign=versioned_reportЗначок Markdown
[](https://skillstore.io/skills/juliusbrussee-caveman-evidence-review?utm_source=security_passport_badge)Значок HTML
<a href="https://skillstore.io/skills/juliusbrussee-caveman-evidence-review?utm_source=security_passport_badge"><img src="https://skillstore.io/badges/skills/juliusbrussee-caveman-evidence-review/security.svg" alt="Skillstore security assessment" loading="lazy"></a>Встраиваемая карточка
<iframe src="https://skillstore.io/embed/skills/juliusbrussee-caveman-evidence-review.html" title="Skillstore Security Assessment" sandbox="allow-popups allow-popups-to-escape-sandbox" loading="lazy" referrerpolicy="no-referrer" width="420" height="180"></iframe>Академические ссылки (APA · BibTeX · CFF)
Цитата APA
juliusbrussee. (2026). caveman-evidence-review security audit report (audit version 2) [Author version unspecified]. Skillstore. https://skillstore.io/skills/juliusbrussee-caveman-evidence-review/audits/2Цитата BibTeX
@techreport{juliusbrussee-juliusbrussee-caveman-evidence-review-2026,
author = {juliusbrussee},
title = {caveman-evidence-review security audit report (audit version 2)},
institution = {Skillstore},
year = {2026},
number = {2},
url = {https://skillstore.io/skills/juliusbrussee-caveman-evidence-review/audits/2},
note = {Author version unspecified}
}CITATION.cff
cff-version: 1.2.0
message: "If you use this Skill, cite its author and this versioned security audit report."
title: "caveman-evidence-review security audit report (audit version 2)"
version: "unspecified"
type: report
authors:
- name: "juliusbrussee"
date-released: "2026-08-24"
url: "https://skillstore.io/skills/juliusbrussee-caveman-evidence-review/audits/2"
identifiers:
- type: other
value: "skillstore:juliusbrussee-caveman-evidence-review:audit:2"
description: "Skillstore immutable audit report identifier"
Оценка Skillstore
Почему такая оценка Достоверность доказательств: СреднийЧто вы можете построить
Investigate spend changes
Compare cost reports and trace cohorts to identify workflows, models, or retries associated with an LLM spend change.
Review reliability evidence
Examine errors, latency, status classes, and representative spans within a bounded production window.
Validate optimization results
Keep inferred headroom separate from verified ledger savings and identify the evidence supporting each result.
Попробуйте эти промпты
Review current Caveman Cloud costs for the selected project. State the report window, cost basis, main drivers, and missing evidence.
Review workflows for the last seven days. Identify high-cost or failing workflows, then cite representative trace identifiers and exact time windows.
Test whether model routing caused the recent cost increase. Compare a bounded suspect cohort with an earlier control cohort and report unresolved factors.
Audit the selected project across costs, score, plan, workflows, and verified savings. Separate every cost category and inspect representative traces without payloads.
Лучшие практики
- Select the correct Caveman project before requesting analysis.
- Provide a bounded time window and a specific cost, quality, or reliability question.
- Request payload review only when metadata and spans cannot answer the question.
Избегать
- Do not treat an empty result as proof of zero cost or zero risk.
- Do not combine measured costs, inferred headroom, verified savings, and evidence costs.
- Do not claim causality from one aggregate or one expensive trace.
Часто задаваемые вопросы
Does this skill change Caveman experiments?
What access does it require?
Does it read prompt or completion content?
Can it explain a cost increase?
How does it report savings?
Which tools does it support?
Сведения для разработчиков
Автор
juliusbrusseeЛицензия
MIT
Ревизия Skillstore
r1
Примечание о версии
Автор не указал версию.
Ссылка
18c566f98cd9d675bf70dca057f35ef24e82ac57
Актуальность поддержки
25.08.2026
Использование
1 загрузок · 0 просмотров
Структура файлов
📄 SKILL.md