15,698 packages matching “eval-gate”
@keboola/ui-gen-bench
v0.2.13 · 22 hours ago
Scoring model and reproducible benchmark for agent-generated Keboola UI — turns @keboola/validate-ui verdicts into a tunable 0-100 BenchScore and runs a fixed fixture target set to record a committed baseline
No known vulnerabilities
@intentsolutions/jrig-cli
v0.2.0 · 1 month ago
J-Rig seven-layer binary eval CLI for Claude Skills — the `j-rig` command: package integrity, trigger/functional/regression/baseline scoring, optimizer, and rollout-gate evidence. Self-contained (bundles the internal eval engine).
No known vulnerabilities
@intentsolutions/refiner-core
v0.3.0 · 1 month ago
Skill Refiner pure core: bounded-edit apply transform, deterministic synthetic eval-set bootstrap, the Pareto-dominant acceptance gate (DR-028 P0-RATIFY-1), and the swappable RefinerStrategy interface (AC-13).
No known vulnerabilities
@intentsolutions/core
v0.10.0 · 1 month ago
Canonical contracts kernel for the Intent Eval Platform — TypeScript types, JSON Schemas, Zod validators, and state machines for the 16 canonical entities.
No known vulnerabilities
expr-eval
v2.0.2 · 6 years ago
Mathematical expression evaluator
2 high
@arthurcarlson/evalgate
v1.0.0 · 1 month ago
LLM prompt eval gate for CI: run YAML-defined eval cases against your prompts on every PR, post a pass/fail report, and fail the build when outputs regress.
No known vulnerabilities
werift-ice
v0.2.3 · 1 day ago
ICE(Interactive Connectivity Establishment) Implementation for TypeScript (Node.js)
No known vulnerabilities
agentic-flow
v2.1.2 · 13 days ago
The agentic meta-harness — freeze the model, evolve the harness. An open runtime that routes each query to the cost-optimal model, evolves its own harness (planner/context/reviewer/retry/tool/memory/score policy) and autonomously repairs code, then orches
No known vulnerabilities
@pixi/unsafe-eval
v7.4.3 · 1 year ago
Adds support for environments that disallow support of new Function
No known vulnerabilities
ts-node
v10.9.2 · 2 years ago
TypeScript execution environment and REPL for node.js, with source map support
No known vulnerabilities
cycle
v1.0.3 · 12 years ago
decycle your json
No known vulnerabilities
content-security-policy-parser
v0.6.0 · 2 years ago
Parse Content Security Policy directives.
No known vulnerabilities
@tangle-network/agent-eval
v0.145.0 · 5 hours ago
Evaluate and improve AI agents from runs, traces, judges, and feedback. Compare candidates, cluster failures, measure lift, and gate releases.
No known vulnerabilities
safe-eval
v0.4.1 · 8 years ago
Safer version of eval()
4 critical · 1 high
tamedevil
v0.1.1 · 3 months ago
Build and evaluate JavaScript strings safely via tagged template literals
No known vulnerabilities
@metaharness/darwin
v0.9.1 · 12 hours ago
Freeze the model, evolve the harness. Two measured applications: (1) SWE-bench code-repair — conformant GLM->Opus empty-patch cascade resolves 51.3% Lite (n=300) and 55.6% Verified (278/500, Wilson 95% CI [51.2, 59.9], official swebench gold eval, no gold
No known vulnerabilities
@atlaskit/feature-gate-js-client
v6.0.2 · 8 days ago
Atlassians wrapper for the Statsig js-lite client.
No known vulnerabilities
@dak-os/optimize
v0.2.0 · 1 month ago
The Dak Stack deepen generator — the self-improving loop's ORIGINATOR (the eval gate is the validator; this proposes what the gate judges). proposeRevision is a GEPA/DSPy-style hill-climb over a skill's natural-language scaffolding (the SKILL.md body): a
No known vulnerabilities
oci-goldengate
v2.139.1 · 13 hours ago
OCI NodeJS client for Golden Gate Service
No known vulnerabilities
built-in-math-eval
v0.3.2 · 3 months ago
Evaluate mathematical expression with the built-in math object
No known vulnerabilities