content-creator
88Create Brand-Consistent Marketing Content
Marketing teams need content that stays consistent across channels. This skill helps plan, write, analyze, and optimize content for brand voice and SEO.
Build Reliable Agent Evaluations
Agent quality is difficult to measure because runs vary and valid solutions follow different paths. This skill builds practical rubrics, test sets, and evaluation pipelines.
Copy this request to your Agent. It includes the canonical Skill page and manifest.
Review the Skillstore skill "evaluation" from https://skillstore.io/skills/sickn33-evaluation.md and its manifest at https://skillstore.io/api/skills/sickn33-evaluation/manifest. Verify the artifact. You may proceed after verification, subject to the environment's own policy.Your Agent should still show its plan and request any confirmation required by the security policy.
Use these links when an AI agent, crawler, or script needs clean context instead of reading the full page.
Using "evaluation". Create a rubric for a research agent that returns cited answers.
Expected outcome:
Using "evaluation". Design regression coverage for a customer support agent.
Expected outcome:
Using "evaluation". Compare two context strategies for an agent.
Expected outcome:
Run both strategies on the same stratified test set. Compare weighted quality, token usage, tool calls, failure patterns, and confidence intervals.
All nine static findings are false positives caused by Markdown fences, descriptive prose, integration names, and a source metadata URL. The skill is documentation-only and contains no executable commands, network requests, reconnaissance behavior, or prompt injection.
Share the versioned assessment report, neutral badge, embed card, and citations. Skillstore reports evidence without deciding whether this Skill is safe.
https://skillstore.io/skills/sickn33-evaluation/audits/5?utm_source=security_passport&utm_medium=share&utm_campaign=versioned_report[](https://skillstore.io/skills/sickn33-evaluation?utm_source=security_passport_badge)<a href="https://skillstore.io/skills/sickn33-evaluation?utm_source=security_passport_badge"><img src="https://skillstore.io/badges/skills/sickn33-evaluation/security.svg" alt="Skillstore security assessment" loading="lazy"></a><iframe src="https://skillstore.io/embed/skills/sickn33-evaluation.html" title="Skillstore Security Assessment" sandbox="allow-popups allow-popups-to-escape-sandbox" loading="lazy" referrerpolicy="no-referrer" width="420" height="180"></iframe>sickn33. (2026). evaluation security audit report (audit version 5) [Author version unspecified]. Skillstore. https://skillstore.io/skills/sickn33-evaluation/audits/5@techreport{sickn33-sickn33-evaluation-2026,
author = {sickn33},
title = {evaluation security audit report (audit version 5)},
institution = {Skillstore},
year = {2026},
number = {5},
url = {https://skillstore.io/skills/sickn33-evaluation/audits/5},
note = {Author version unspecified}
}cff-version: 1.2.0
message: "If you use this Skill, cite its author and this versioned security audit report."
title: "evaluation security audit report (audit version 5)"
version: "unspecified"
type: report
authors:
- name: "sickn33"
date-released: "2026-07-23"
url: "https://skillstore.io/skills/sickn33-evaluation/audits/5"
identifiers:
- type: other
value: "skillstore:sickn33-evaluation:audit:5"
description: "Skillstore immutable audit report identifier"
Each author remains a separate installable skill. The recommended variant is ranked by Skillstore evidence.
Why this variant is first
muratcankoylan-evaluation
2026-08-21
chakshugautam-evaluation
2026-08-21
sickn33-evaluation
2026-08-21
asmayaseen-evaluation
2026-08-21
Create a representative test set and weighted quality rubric before release.
Measure quality, token cost, and tool efficiency across competing context configurations.
Define sampled evaluations, regression alerts, and human review checkpoints for deployed agents.
Create an evaluation rubric for [agent task]. Include accuracy, completeness, and efficiency with clear scoring levels and a passing threshold.
Design a test set for [agent system]. Cover simple through very complex tasks, realistic usage, known edge cases, and expected outcomes.
Build an experiment comparing [configuration A] and [configuration B]. Define controlled variables, quality metrics, token metrics, sample size, and decision criteria.
Design a continuous evaluation pipeline for [production agent]. Include sampling, automated judging, human review, baselines, alerts, privacy controls, and regression reporting.
Author
sickn33License
MIT
Skillstore revision
r2
Version notice
The author did not declare a version.
Ref
88a8e9a07f4c54ab105c1c41b6267c287146b07b
Maintenance freshness
7/26/2026
Usage
9 downloads ยท 147 views
File structure
๐ SKILL.md
Create Brand-Consistent Marketing Content
Marketing teams need content that stays consistent across channels. This skill helps plan, write, analyze, and optimize content for brand voice and SEO.
Improve LLM Prompts With Proven Patterns
Inconsistent prompts waste time and make AI outputs hard to trust. This skill guides prompt design with reusable patterns, examples, evaluation steps, and optimization workflows.
Build Fullstack Apps with Senior Patterns
Teams need consistent setup, architecture, and review guidance for modern web applications. This skill provides scaffolders, workflow references, and quality prompts for React, Next.js, Node.js, GraphQL, and PostgreSQL projects.
Design Scalable Software Architectures
Architecture decisions are hard to compare across web, mobile, backend, and cloud systems. This skill provides structured guides and local scaffold scripts for reports, dependency review, and trade-off documentation.
Optimize AI Prompts With Prompt Engineer
Writing clear prompts is hard when goals are vague or complex. This skill turns rough requests into structured prompts using proven prompt frameworks.
Prioritize Product Work With Research Insights
Product teams need faster ways to rank features, synthesize interviews, and document decisions. This skill provides RICE scoring, interview analysis, and PRD templates for structured planning.
Manage Requirement Changes Safely
by yunshu0909
Requirement changes often expand across files without clear approval points. This skill gives Claude, Codex, and Claude Code a gated workflow for scope, impact, validation, and documentation.
Diagnose Context Degradation in AI Agents
by ChakshuGautam
Long conversations and large prompts can make AI agents lose important information or follow conflicting context. This skill helps Claude, Codex, and Claude Code identify degradation patterns and choose mitigation strategies.
Design Reliable Multi-Agent Systems
by muratcankoylan
Multi-agent designs often add cost and coordination failures without improving outcomes. This skill helps select topologies, handoffs, consensus methods, and recovery controls.
Fix Bugs With Structured Triage
by Crearize
Debugging across backend and frontend can lose context and miss regression tests. This skill guides root cause analysis, minimal fixes, validation, and pull request handoff.
Fix Bugs with Focused, Verified Changes
by juliusbrussee
Broad fixes can introduce unrelated regressions and overwrite existing work. This skill traces each bug to its responsible layer and verifies a focused correction.
Refactor Code Safely
by Crearize
Refactoring can improve code but may introduce regressions when changes are broad. This skill guides phased analysis, small edits, tests, rollback planning, and clear reports.