31,982 packages matching “llm-rubric”
@eva-llm/llm-as-a-jest
v1.0.5 · 5 months ago
Jest plugin with LLM-as-a-Judge matchers basing on G-Eval, B-Eval, and LLM-Rubric
No known vulnerabilities
earn-autonomy
v0.1.0 · 1 month ago
Eval-gated autonomy: deterministic rules + LLM rubric + an N-clean-runs gate that decides when an AI agent has earned unattended runs.
No known vulnerabilities
llm-as-a-jest
v1.0.5 · 4 months ago
Jest plugin with LLM-as-a-Judge matchers basing on G-Eval, B-Eval, and LLM-Rubric
No known vulnerabilities
@pie-element/rubric
v8.2.3 · 4 days ago
Rubric Scoring Interaction
No known vulnerabilities
@pie-element/complex-rubric
v7.2.3 · 4 days ago
Complex Rubric Scoring Interaction
No known vulnerabilities
@pie-element/multi-trait-rubric
v8.2.3 · 4 days ago
A [pie][pie]choice component.
No known vulnerabilities
@tagma/completion-llm-judge
v0.2.96 · 1 hour ago
LLM-as-judge completion plugin for tagma-sdk pipelines
No known vulnerabilities
llm-governance-gateway
v0.15.0 · 5 days ago
Governed structured-output LLM pipeline: rate limit → spend caps (per-user + global circuit breaker) → cache → provider-chain failover → Zod validation → usage logging → LLM-judge, with a deterministic mock provider for CI.
No known vulnerabilities
@rulvar/evals
v1.253.0 · 13 days ago
Rulvar evals: eval cases, golden outputs, rubric and judge graders, matrix sweeps, canary fingerprint.
No known vulnerabilities
@cejel/cejel
v0.4.10 · 4 days ago
Offline trust certificate for your codebase — scores tests, secrets, isolation, claim-vs-reality, and CI discipline; especially useful for AI-written code. Supports JS/TS, Python, Go, Rust, Java, Ruby, PHP, C#, C/C++ and more — see README for the full lan
No known vulnerabilities
@llm-ports/capabilities
v0.1.0-alpha.34 · 3 days ago
Reusable cognitive operation factories for llm-ports: classify, score, draft, summarize, extract, plan, analyze. Configure once at definition time, call many times.
No known vulnerabilities
@deepseek-ai/dsh-llm
v0.0.1-rc.1 · 1 month ago
Provider-neutral LLM service interface for the DeepSeek Harness
No known vulnerabilities
vitest-evals
v0.17.0 · 10 days ago
Harness-backed AI testing on top of Vitest.
No known vulnerabilities
@deepseek-ai/dsh-llm-retry
v0.0.1-rc.1 · 1 month ago
Provider-routed LLM request retry policy for the DeepSeek Harness
No known vulnerabilities
@deepseek-ai/dsh-llm-pi-ai
v0.0.1-rc.1 · 1 month ago
pi-ai-backed DeepSeek adapter for the DeepSeek Harness LLM seam (design-verification twin of dsh-llm-deepseek)
No known vulnerabilities
@deepseek-ai/dsh-llm-deepseek
v0.0.1-rc.1 · 1 month ago
DeepSeek chat-completions adapter for the DeepSeek Harness LLM seam
No known vulnerabilities
promptfoo
v0.123.1 · 3 days ago
LLM eval & testing toolkit
No known vulnerabilities
@pwtap/plugin-ai-judge
v0.2.0 · 1 month ago
AI/LLM judge matchers for Playwright — toPassRubric/toScoreAtLeast/toMatchImage over Ollama, OpenAI-compatible endpoints (OpenAI/OpenRouter/NVIDIA/Groq/…), and native Claude
No known vulnerabilities
@mlc-ai/web-llm
v0.2.85 · 13 days ago
Hardware accelerated language model chats on browsers
No known vulnerabilities
langchain
v1.5.11 · 11 days ago
Typescript bindings for langchain
No known vulnerabilities