776 packages matching “evaluations”
@statsig/web-analytics
v3.33.4 · 13 days ago
Statsig helps you move faster with feature gates (feature flags), and/or dynamic configs. It also allows you to run A/B/n tests to validate your new features and understand their impact on your KPIs. If you're new to Statsig, check out our product and cre
No known vulnerabilities
trainloop-llm-logging
v0.9.0 · 1 year ago
TrainLoop Evaluations - header-based request tagging and zero-touch collection
No known vulnerabilities
@agentic-evals/mcp
v0.1.0 · 4 months ago
Agentic Evaluation Library — MCP server for deterministic and non-deterministic UI/UX evaluations with swarm orchestration
No known vulnerabilities
@statsig/js-local-overrides
v3.33.4 · 13 days ago
Statsig helps you move faster with feature gates (feature flags), and/or dynamic configs. It also allows you to run A/B/n tests to validate your new features and understand their impact on your KPIs. If you're new to Statsig, check out our product and cre
No known vulnerabilities
@khiwniti/evals
v1.0.0 · 1 month ago
Evaluations for Openbuff
No known vulnerabilities
webperf-comparison
v5.0.2 · 4 years ago
Automated system for running performance evaluations against two systems to determine the change
No known vulnerabilities
@statsig/session-replay
v3.33.4 · 13 days ago
Statsig helps you move faster with feature gates (feature flags), and/or dynamic configs. It also allows you to run A/B/n tests to validate your new features and understand their impact on your KPIs. If you're new to Statsig, check out our product and cre
No known vulnerabilities
@auraone/github-app
v0.2.0 · 1 month ago
Self-hosted GitHub App that reports AuraOne PR evaluations in one lifecycle-aware Check Run and one idempotent bot-owned PR summary.
No known vulnerabilities
autoevals
v0.3.0 · 2 months ago
Universal library for evaluating AI models
No known vulnerabilities
@peerlm/mcp
v0.1.1 · 5 months ago
MCP server for PeerLM — run evaluations from Claude Desktop, Cursor, or any MCP client
No known vulnerabilities
@statsig/on-device-eval-core
v3.33.4 · 13 days ago
Statsig helps you move faster with feature gates (feature flags), and/or dynamic configs. It also allows you to run A/B/n tests to validate your new features and understand their impact on your KPIs. If you're new to Statsig, check out our product and cre
No known vulnerabilities
@davothebigafro/metrics
v0.0.0 · 3 months ago
Built-in metric factories for Shipwright evaluations.
No known vulnerabilities
@statsig/react-native-core
v3.33.4 · 13 days ago
Statsig helps you move faster with feature gates (feature flags), and/or dynamic configs. It also allows you to run A/B/n tests to validate your new features and understand their impact on your KPIs. If you're new to Statsig, check out our product and cre
No known vulnerabilities
@mcptoolshop/synthesis
v1.2.0 · 2 months ago
Deterministic evaluations for empathy, trust, and care in AI systems
No known vulnerabilities
@statsig/expo-bindings
v3.33.4 · 13 days ago
Statsig helps you move faster with feature gates (feature flags), and/or dynamic configs. It also allows you to run A/B/n tests to validate your new features and understand their impact on your KPIs. If you're new to Statsig, check out our product and cre
No known vulnerabilities
ts-neverfalse
v1.0.3 · 1 year ago
Automated error coalescing and aggregation to simplify advanced type evaluations in Typescript
No known vulnerabilities
@statsig/js-on-device-eval-client
v3.33.4 · 13 days ago
Statsig helps you move faster with feature gates (feature flags), and/or dynamic configs. It also allows you to run A/B/n tests to validate your new features and understand their impact on your KPIs. If you're new to Statsig, check out our product and cre
No known vulnerabilities
aievals
v1.0.0 · 1 year ago
experimental AI evaluations
No known vulnerabilities
botmark-mcp
v0.1.0 · 5 months ago
BotMark MCP Server — manage bot evaluations from Claude Desktop, Cursor, and other AI assistants
No known vulnerabilities
@statsig/serverless-client
v3.33.4 · 13 days ago
Statsig helps you move faster with feature gates (feature flags), and/or dynamic configs. It also allows you to run A/B/n tests to validate your new features and understand their impact on your KPIs. If you're new to Statsig, check out our product and cre
No known vulnerabilities