776 packages matching “evaluations”
ionic7-rating-component
v1.0.2 · 2 years ago
Highly customizable ionic 7 component to display evaluations or a quick rating operation of something.
No known vulnerabilities
@flowiseai/observe
v0.0.0-dev.1 · 3 months ago
Embeddable React components for observing AI agent executions, evaluations, and runtime activity
No known vulnerabilities
@statsig/on-device-evaluations
v0.0.1-beta.10 · 2 years ago
> [!IMPORTANT] > This version of the SDK is still in beta. The non-beta version can be found [here](https://github.com/statsig-io/js-local-eval).
No known vulnerabilities
@tally-evals/tally
v0.1.0 · 5 months ago
A TypeScript evaluation framework for running model evaluations with datasets, evaluators, metrics, and aggregators
No known vulnerabilities
@reaatech/rag-eval-cost
v0.1.0 · 2 months ago
Cost tracking, pricing, budgeting, and reporting for RAG evaluations
No known vulnerabilities
@sigstat/on-device-evaluations
v0.0.8 · 2 years ago
This library was generated with [Nx](https://nx.dev).
No known vulnerabilities
@know-your-ai/evaluate
v0.1.1 · 5 months ago
Know Your AI Evaluation SDK - Programmatically create workspaces, products, datasets and run evaluations
No known vulnerabilities
@statsig/js-user-persisted-storage
v3.33.4 · 13 days ago
Statsig helps you move faster with feature gates (feature flags), and/or dynamic configs. It also allows you to run A/B/n tests to validate your new features and understand their impact on your KPIs. If you're new to Statsig, check out our product and cre
No known vulnerabilities
@statsig/react-native-bindings
v3.33.4 · 13 days ago
Statsig helps you move faster with feature gates (feature flags), and/or dynamic configs. It also allows you to run A/B/n tests to validate your new features and understand their impact on your KPIs. If you're new to Statsig, check out our product and cre
No known vulnerabilities
@reaatech/rag-eval-gate
v0.1.0 · 2 months ago
Quality gates and CI/CD regression checks for RAG evaluations
No known vulnerabilities
@oai-statsig/js-client
v3.33.4 · 7 days ago
Statsig helps you move faster with feature gates (feature flags), and/or dynamic configs. It also allows you to run A/B/n tests to validate your new features and understand their impact on your KPIs. If you're new to Statsig, check out our product and cre
No known vulnerabilities
@statsig/next
v3.33.4 · 13 days ago
Statsig helps you move faster with feature gates (feature flags), and/or dynamic configs. It also allows you to run A/B/n tests to validate your new features and understand their impact on your KPIs. If you're new to Statsig, check out our product and cre
No known vulnerabilities
@oai-statsig/web-analytics
v3.33.4 · 7 days ago
Statsig helps you move faster with feature gates (feature flags), and/or dynamic configs. It also allows you to run A/B/n tests to validate your new features and understand their impact on your KPIs. If you're new to Statsig, check out our product and cre
No known vulnerabilities
@oai-statsig/session-replay
v3.33.4 · 7 days ago
Statsig helps you move faster with feature gates (feature flags), and/or dynamic configs. It also allows you to run A/B/n tests to validate your new features and understand their impact on your KPIs. If you're new to Statsig, check out our product and cre
No known vulnerabilities
mcp-evals
v2.0.1 · 1 year ago
GitHub Action for evaluating MCP server tool calls using LLM-based scoring
No known vulnerabilities
@oai-statsig/react-bindings
v3.33.4 · 7 days ago
Statsig helps you move faster with feature gates (feature flags), and/or dynamic configs. It also allows you to run A/B/n tests to validate your new features and understand their impact on your KPIs. If you're new to Statsig, check out our product and cre
No known vulnerabilities
eva-ts
v1.0.2 · 1 year ago
A TypeScript evaluation framework for running concurrent evaluations with progress tracking and result persistence
No known vulnerabilities
@nem035/agentevals
v0.1.2 · 5 months ago
A Vitest-like CLI for AI agent evaluations. Test your LLM apps with simple, declarative evals.
No known vulnerabilities
@cyclecore/slmbench
v1.0.1 · 8 months ago
CLI and SDK for accessing SLMBench benchmarks (EdgeJSON, EdgeIntent, EdgeFuncCall). View leaderboards, run evaluations, and compare Small Language Models.
No known vulnerabilities
@reaatech/rag-eval-observability
v0.1.0 · 2 months ago
Structured logging, OpenTelemetry tracing, and metrics for RAG evaluations
No known vulnerabilities