31,985 packages matching “llm-rubric”
@deepseek-ai/dsh-deepseek-llm-api-extensions
v0.1.2-alpha.2 · 21 days ago
Additive request-field registry for the official DeepSeek LLM API adapter
No known vulnerabilities
@llm-ui/react
v0.13.3 · 2 years ago
Display language model outputs in your React project.
No known vulnerabilities
@bryel/evals
v0.2.0 · 27 days ago
Export an eval session's artifacts — screenshots and/or text outputs — to bryel for judging. Framework-agnostic, zero dependencies.
No known vulnerabilities
@deepseek-ai/dsh-session-title-first-prompt-llm
v0.0.1-rc.3 · 1 month ago
First-message LLM provider plugin for DeepSeek Harness session titles
No known vulnerabilities
@unotest/judge
v0.40.0 · 1 day ago
LLM-judge service for the @unotest ecosystem: judges free-form text (chat replies, generated content) against a natural-language rubric and returns a structured pass/fail verdict with reasoning. Runs as a small HTTP service (`npx @unotest/judge`) or in-pr
No known vulnerabilities
crosscheck-mcp
v0.2.24 · 17 hours ago
Multi-LLM MCP server: confer / debate / coordinate / audit / orchestrate across Anthropic + OpenAI + xAI + Gemini + Mistral + Groq + DeepSeek + Kimi + Qwen, with scoreboard-driven router, canary-leak detection, sandboxed shell verifiers, and a tier-aware
No known vulnerabilities
micromark-extension-llm-math
v3.1.1-20250610 · 1 year ago
micromark extension to support math (`$C_L$`, `\(C_L\)`)
No known vulnerabilities
@llm-ui/markdown
v0.13.3 · 2 years ago
[llm-ui](https://llm-ui.com) markdown block.
No known vulnerabilities
@llm-ports/eval
v0.1.0-alpha.34 · 3 days ago
Durable storage for post-hoc evaluations (LLM-judge scores, human annotations, rule-based verdicts) keyed on the @llm-ports/observability-contract EvaluationRef shape. Ships an in-memory store (default) and an opt-in SQLite writer (peer-dep on better-sqli
No known vulnerabilities
@deepseek-ai/dsh-session-title-llm
v0.0.1-rc.1 · 1 month ago
Shared LLM generation policy for DeepSeek Harness session-title providers
No known vulnerabilities
@llm-ui/json
v0.13.3 · 2 years ago
[llm-ui](https://llm-ui.com) JSON blocks for building custom components.
No known vulnerabilities
@wix-pilot/detox
v1.0.13 · 1 year ago
Detox driver for Wix Pilot usage
No known vulnerabilities
openevals
v0.2.2 · 1 month ago
Much like tests in traditional software, evals are an important part of bringing LLM applications to production. The goal of this package is to help provide a starting point for you to write evals for your LLM applications, from which you can write more c
No known vulnerabilities
@llm-ui/code
v0.13.3 · 2 years ago
[llm-ui](https://llm-ui.com) code block.
No known vulnerabilities
rubric-x402-screen
v0.5.0 · 14 days ago
Screen the wallets that pay your x402 endpoints against OFAC. Free, local, sub-millisecond. Optionally anchor each screening as permanent evidence.
No known vulnerabilities
openrubric
v0.1.1 · 8 days ago
Rubric-scored feedback on any conversation transcript. Bring your own LLM client, bring your own rubric.
No known vulnerabilities
@grafana/llm
v1.0.9 · 3 months ago
A library for working with LLMs in Grafana plugins
No known vulnerabilities
@stlw/val
v0.1.2 · 17 hours ago
External eval harness that proves your product does what it claims — deterministic L0 checks + LLM-judged L2 rubric scoring
No known vulnerabilities
@promptster/rubric
v0.7.0 · 28 days ago
Promptster AI-fluency rubric — the public data artifact (4 process dimensions + sub-facets & behavioral anchors, 5 tiers, methodology sources). Anchors and citations only; criteria and scoring weights are not part of this package.
No known vulnerabilities
opik
v2.2.71 · 1 day ago
Opik TypeScript and JavaScript SDK
No known vulnerabilities