477,079 packages matching “judge-d”
judge-d
v1.5.1 · 4 years ago
CLI for publishing and validating contract tests using judge-d API
No known vulnerabilities
vitest-evals
v0.16.1 · 13 days ago
Harness-backed AI testing on top of Vitest.
No known vulnerabilities
openevals
v0.2.2 · 3 days ago
Much like tests in traditional software, evals are an important part of bringing LLM applications to production. The goal of this package is to help provide a starting point for you to write evals for your LLM applications, from which you can write more c
No known vulnerabilities
@presentation-md/render
v1.20.9 · 16 days ago
Render a deck JSON spec to a self-contained HTML slide deck.
No known vulnerabilities
@turf/quadrat-analysis
v7.4.0 · 18 days ago
Quadrat analysis lays a set of equal-size areas(quadrat) over the study area and counts the number of features in each quadrat and creates a frequency table.
No known vulnerabilities
hydrooj
v5.0.4 · 1 month ago
No description provided.
No known vulnerabilities
werewolf-judge-cdn
v0.0.0-g19305ab0 · 21 hours ago
Runtime asset loader and manifest for the WerewolfJudge web app — fonts, audio sprites, and image bundles
No known vulnerabilities
@metaharness/redblue
v0.1.6 · 7 days ago
AI red-teaming for the AI agents & LLM apps you own: stress-test them with adversarial models to find security failures (prompt injection, tool misuse / excessive agency, data leakage, jailbreaks, denial-of-wallet), auto-patch (blue team), retest, and get
No known vulnerabilities
@tagma/completion-llm-judge
v0.2.93 · 3 days ago
LLM-as-judge completion plugin for tagma-sdk pipelines
No known vulnerabilities
@rulvar/evals
v1.248.0 · 4 hours ago
Rulvar evals: eval cases, golden outputs, rubric and judge graders, matrix sweeps, canary fingerprint.
No known vulnerabilities
@tangle-network/agent-bench
v0.8.22 · 5 hours ago
Benchmark adapters and execution for agent-runtime across coding, tool-use, RAG, memory, browser, and terminal tasks.
No known vulnerabilities
@tarpit/judge
v2.1.1 · 4 months ago
Runtime type validation and assertion utilities
No known vulnerabilities
@halooj/oj
v5.0.14 · 25 days ago
Copyright [@hydro-dev](https://github.com/hydro-dev) team.
No known vulnerabilities
@looprun-ai/eval
v0.20.0 · 10 days ago
looprun eval harness: run a generated subject's cases against a model target (governed or ungoverned variant), dump per-case traces for the LLM judge, fold verdicts, and certify against the bar.
No known vulnerabilities
@akagilnc/pi-workflow-roles
v0.1.2150 · 3 hours ago
Soul-bound workflow roles for Pi
No known vulnerabilities
@exercode/problem-utils
v2.0.0 · 15 hours ago
:100: A set of utilities for judging programs on Exercode (https://exercode.willbooster.com/).
No known vulnerabilities
egg-logrotator
v3.2.0 · 1 year ago
logrotator for egg
No known vulnerabilities
@su-record/vibe
v3.2.44 · 21 hours ago
AI Coding Framework for Claude Code — 7+ agents, 52 skills, multi-LLM orchestration
No known vulnerabilities
@handsealed/engine
v0.25.0 · 13 days ago
Handsealed's deterministic trust rules: lanes, mandate binding, scope ceilings, evidence classes, verdicts.
No known vulnerabilities
vue-count-to
v1.0.13 · 8 years ago
It's a vue component that will count to a target number at a specified duration
No known vulnerabilities