279 packages matching “evaluators”
@drift-ci/core
v1.1.4 · 3 months ago
drift-ci core engine: types, providers, evaluators, storage, runner
No known vulnerabilities
@goodbones/core
v0.1.0-beta.12 · 12 days ago
Architecture policy as one manifest of your repository, language-agnostic: the manifest schema, the evaluators for imports, exports, members, surface, structure and the import graph, the ports a language pack implements, and the loader that turns a manife
No known vulnerabilities
@habenula-ai/governance
v1.0.0 · 1 month ago
Habenula governance kernel — the pure permission and spending evaluators. Leaf package: no runtime dependencies.
No known vulnerabilities
genkit-cli
v1.43.0 · 18 days ago
CLI for interacting with the Google Genkit AI framework
No known vulnerabilities
pzzld-wasm
v0.0.2 · 11 months ago
Various evaluators for the smartshop platform.
No known vulnerabilities
@genkit-ai/evaluator
v1.42.0 · 1 month ago
Genkit AI framework plugin for RAG evaluation.
No known vulnerabilities
@gixcopilot/evals
v0.2.4 · 5 days ago
Evaluation framework for AI Copilot applications - versioned datasets, execution records derived from recorded diagnostics, deterministic evaluators (tools, permission compliance, RAG, citations, memory, agents, workflows, UI, latency, tokens, cost), base
No known vulnerabilities
@luma.gl/gpgpu
v9.4.2 · 16 days ago
General-purpose GPU data and computation for luma.gl
No known vulnerabilities
tracecase
v0.1.3 · 2 months ago
Version-controlled, replayable eval cases for AI agents: ingest traces, store YAML cases, score with pluggable evaluators, fail CI on regression.
No known vulnerabilities
openevals
v0.2.2 · 1 month ago
Much like tests in traditional software, evals are an important part of bringing LLM applications to production. The goal of this package is to help provide a starting point for you to write evals for your LLM applications, from which you can write more c
No known vulnerabilities
@hadron-memory/access-control
v0.3.0 · 2 months ago
Hadron's authorization core — the pure, isomorphic access-control evaluators used by hadron-server and every Hadron client. The server is always the enforcement authority; clients use this to predict its decisions.
No known vulnerabilities
universal-turing-machine
v1.0.2 · 3 years ago
Generic class to process and serialise universal Turing machines and evaluators
No known vulnerabilities
evalkit
v0.2.0 · 7 months ago
Lightweight deterministic evaluators for AI agents. Binary pass/fail checks, zero dependencies, no LLM cost.
No known vulnerabilities
comm-sense-vlm
v1.0.0 · 7 months ago
Visual Oracle for Agentic Evaluators - Semantic UI testing using LLMs.
No known vulnerabilities
@versaprotocol/semval
v0.10.0 · 11 months ago
This is a collection of rules, rule-evaluators, and tests for semantic validation of Versa receipts. Written in Rust, it uses [napi-rs](https://napi.rs/) to compile to native modules for use in NodeJS environments. It can also be used in Rust backends, an
No known vulnerabilities
@operor/testing
v0.5.0 · 7 months ago
Testing utilities for Agent OS — CSV test runner, evaluators, and test suite tools
No known vulnerabilities
@types/multisort
v0.5.2 · 2 years ago
TypeScript definitions for multisort
No known vulnerabilities
@found-in-space/journey
v0.2.0-alpha.0 · 4 months ago
Plain authored journey graphs, timed evaluators, cues, and retiming helpers for Found in Space
No known vulnerabilities
@arizeai/phoenix-evals
v2.6.0 · 13 days ago
A library for running evaluations for AI use cases
No known vulnerabilities
com.echo.melodyquest
v1.1.1 · 27 days ago
Daily Mission & Check-in event module for Echo games. Players complete missions to earn Sound Waves that fill stage chests, progressing toward a Grand Prize. Supports duration and weekly-schedule modes, auto/manual start, pluggable mission-type evaluators
No known vulnerabilities