26,170 packages matching “skill-evaluation”
@skillgauge/cli
v0.1.2 · 12 days ago
SkillGauge command-line interface for coding-agent skill evaluation
No known vulnerabilities
skillave
v0.1.6 · 4 months ago
Skill evaluation pipeline: generate test cases, execute them, and verify results
No known vulnerabilities
skillbench
v2.1.1 · 4 months ago
Skill evaluation framework for Claude agents — static analysis + live agent testing
No known vulnerabilities
@mizchi/waxa
v0.1.1 · 2 months ago
Skill evaluation CLI — waza-schema-compatible runner with empirical-prompt-tuning iteration loop, structured self-report grader, and LLM-as-Judge.
No known vulnerabilities
skills
v1.5.21 · 3 days ago
The open agent skills ecosystem
No known vulnerabilities
postcss-custom-media
v12.0.1 · 5 months ago
Use Custom Media Queries in CSS
No known vulnerabilities
quickjs-wasi
v3.3.0 · 1 day ago
Snapshotable JavaScript runtime via WebAssembly. QuickJS-NG compiled to WASM with snapshot/restore support.
No known vulnerabilities
orizu
v0.5.16 · 10 days ago
Orizu is a platform that helps you build continually learning agents and other LLM applications. It does so in a scientific, measurable manner by first helping you build evals and then helping you hill climb on them.
No known vulnerabilities
@cortex-js/compute-engine
v0.99.0 · 3 days ago
Symbolic computing and numeric evaluations for JavaScript and Node.js
No known vulnerabilities
@ai-sdk/gateway
v4.0.37 · 18 hours ago
The Gateway provider for the [AI SDK](https://ai-sdk.dev/docs) allows the use of a wide variety of AI models and providers.
No known vulnerabilities
quality.md
v0.35.2 · 17 days ago
Companion CLI for the QUALITY.md file format and /quality agent skill, used to evaluate and improve AI assistant projects and harnesses.
No known vulnerabilities
ai
v7.0.48 · 18 hours ago
AI SDK by Vercel - build apps like ChatGPT, Claude, Gemini, and more with a single interface for any model using the Vercel AI Gateway or go direct to OpenAI, Anthropic, Google, or any other model provider.
No known vulnerabilities
@openfeature/flagd-core
v4.0.0 · 2 months ago
flagd-core contain the core logic of flagd [in-process evaluation](https://flagd.dev/architecture/#in-process-evaluation) provider. This package is intended to be used by concrete implementations of flagd in-process providers.
No known vulnerabilities
@ai-sdk/google
v4.0.31 · 1 day ago
The **[Google provider](https://ai-sdk.dev/providers/ai-sdk-providers/google)** for the [AI SDK](https://ai-sdk.dev/docs) contains language model support for the [Google Generative AI](https://ai.google/discover/generativeai/) APIs.
No known vulnerabilities
@apollo/client
v4.2.9 · 2 days ago
A fully-featured caching GraphQL client.
No known vulnerabilities
@ai-sdk/anthropic
v4.0.27 · 1 day ago
The **[Anthropic provider](https://ai-sdk.dev/providers/ai-sdk-providers/anthropic)** for the [AI SDK](https://ai-sdk.dev/docs) contains language model support for the [Anthropic Messages API](https://docs.anthropic.com/claude/reference/messages_post).
No known vulnerabilities
@tloncorp/tlon-skill
v0.4.5 · 3 days ago
Tlon/Urbit skill for OpenClaw agents
No known vulnerabilities
@ai-sdk/openai
v4.0.27 · 1 day ago
The **[OpenAI provider](https://ai-sdk.dev/providers/ai-sdk-providers/openai)** for the [AI SDK](https://ai-sdk.dev/docs) contains language model support for the OpenAI chat and completion APIs and embedding model support for the OpenAI embeddings API.
No known vulnerabilities
@panguard-ai/scan-core
v1.8.26 · 14 days ago
Unified skill scanning engine - shared between CLI Auditor and Website
No known vulnerabilities
@pierre/diffs
v1.3.1 · 1 day ago
Docs at [https://diffs.com](https://diffs.com)
No known vulnerabilities