By JudgmentLabs
Skills for working with Judgment — the continuous-improvement stack for agents. Add tracing, evaluations, code judges, MCP server workflows, and monitoring with best practices.
Build, inspect, or execute tenant-safe Judgment Query Language (JQL) through the public Judgeval Python or TypeScript SDK. Use for querying project traces, spans, or sessions; aggregations and pipelines; chart or table presentations; discovery; public response handling; and JQL troubleshooting.
Use Judgment for agent tracing, evaluations, code judges, datasets, and monitoring. Use when integrating Judgment or judgeval, adding tracing to agents/workflows, creating evaluations or scorers, debugging traces, or looking up Judgment SDK usage and docs.
Best practices for using the Judgment MCP server effectively. Covers when to use MCP vs other tools, how to use search_traces with batching, and general usage patterns.
Agent Skills that teach AI coding assistants how to work with Judgment for tracing, evaluations, code judges, datasets, MCP server workflows, and agent performance workflows.
| Skill | Description |
|---|---|
| judgment | Main skill for adding Judgment tracing and evaluations, choosing scorer patterns, using code judges, and following Judgment SDK best practices. |
| mcp-server-best-practices | Standalone skill for using the Judgment MCP server effectively, including full-text trace search, batched queries, and production data workflows. |
| judgeval-jql | Query and analyze project traces, spans, and sessions with tenant-safe JQL, including aggregations, pipelines, presentations, and discovery. |
Use your coding agent with this instruction so it can install the Judgment skill and apply it to your task.
Install the Judgment skill from github.com/JudgmentLabs/skills
and use it to add tracing to this application
following Judgment best practices.
For MCP-specific workflows:
Install the mcp-server-best-practices skill from github.com/JudgmentLabs/skills.
For public Judgeval JQL:
Install the judgeval-jql skill from github.com/JudgmentLabs/skills.
Detect whether this project uses Judgeval Python or TypeScript, then use the
matching public SDK to build tenant-safe JQL for: <request>.
Install as a Cursor plugin:
/add-plugin judgment
Or via the skills CLI:
npx skills add JudgmentLabs/skills --skill "judgment" --agent cursor
npx skills add JudgmentLabs/skills --skill "mcp-server-best-practices" --agent cursor
npx skills add JudgmentLabs/skills --skill "judgeval-jql" --agent cursor
Add the marketplace and install:
claude plugin marketplace add JudgmentLabs/skills
claude plugin install judgment@judgment-skills
Or via the skills CLI:
npx skills add JudgmentLabs/skills --skill "judgment" --agent claude-code
npx skills add JudgmentLabs/skills --skill "mcp-server-best-practices" --agent claude-code
npx skills add JudgmentLabs/skills --skill "judgeval-jql" --agent claude-code
npx skills add JudgmentLabs/skills --skill "judgment"
npx skills add JudgmentLabs/skills --skill "mcp-server-best-practices"
npx skills add JudgmentLabs/skills --skill "judgeval-jql"
Agents that cannot install a skill can fetch the stable docs-hosted files directly:
The immutable JQL source commit and SHA-256 hashes for this publication are
recorded in source.json and verified in
CI.
Set your Judgment credentials before asking an agent to run traces or evaluations:
export JUDGMENT_API_KEY=...
export JUDGMENT_ORG_ID=...
Once installed, your agent can use these skills when you ask it to:
Own this plugin?
Verify ownership to unlock analytics, metadata editing, and a verified badge. GitHub access is read-only (username + org membership).
Sign in to claimOwn this plugin?
Verify ownership to unlock analytics, metadata editing, and a verified badge. GitHub access is read-only (username + org membership).
Sign in to claimBased on adoption, maintenance, documentation, and repository signals. Not a security audit or endorsement.
npx claudepluginhub judgmentlabs/skills --plugin judgmentAccess Judgment traces, sessions, behaviors, judges, prompts, datasets, and tests directly from your AI coding tool, with official skills for tracing, evaluation, and MCP usage best practices.
Enables AI agents to use Judgeval for LLM evaluation, logging, and observability. Provides correct API usage, working examples, and helper scripts for common operations.
A growing collection of Claude-compatible academic workflow bundles. Covers scientific figures, manuscript writing and polishing, reviewer assessment, citation retrieval, data availability, paper reading, literature search, response letters, paper-to-PPTX conversion, and evidence-grounded Chinese invention patent drafting. Rules are organized as reusable skill folders with explicit workflows and quality checks.
Core skills library for Claude Code: TDD, debugging, collaboration patterns, and proven techniques
Harness-native ECC plugin for engineering teams - 67 agents, 279 skills, 94 legacy command shims, reusable hooks, rules, MCP conventions, and operator workflows for Claude Code plus adjacent agent harnesses
Plugin-safe Claude Code distribution of Agentic Awesome Skills with 1,933 supported skills.
Comprehensive skill pack with 66 specialized skills for full-stack developers: 12 language experts (Python, TypeScript, Go, Rust, C++, Swift, Kotlin, C#, PHP, Java, SQL, JavaScript), 10 backend frameworks, 6 frontend/mobile, plus infrastructure, DevOps, security, and testing. Features progressive disclosure architecture for 50% faster loading.
This skill should be used when users need to generate ideas, explore creative solutions, or systematically brainstorm approaches to problems. Use when users request help with ideation, content planning, product features, marketing campaigns, strategic planning, creative writing, or any task requiring structured idea generation. The skill provides 30+ research-validated prompt patterns across 14 categories with exact templates, success metrics, and domain-specific applications.