17 packages matching “grpo”
@elizaos/training
v2.0.0-alpha.77 · 6 months ago
ElizaOS RL training pipeline with benchmarking and model publishing support
No known vulnerabilities
stz-foundry
v1.16.1 · 2 months ago
STZ Foundry: the evolution of slice-tournament-zoo into a standalone BYO-LLM foundry — a contract-bounded slice pipeline that implements each slice adversarially via an N-specimen tournament with frozen sealed tests, GRPO-style selection, layered anti-rew
No known vulnerabilities
slice-tournament-zoo
v0.9.6 · 3 months ago
STZ: a contract-bounded slice pipeline that implements each slice adversarially via an N-specimen tournament with frozen sealed tests, GRPO-style selection, layered anti-reward-hacking, a replayable markdown audit trail, and (0.9.0) a bounded harness-leve
No known vulnerabilities
@mlx-node/trl
v0.0.15 · 14 days ago
No description provided.
No known vulnerabilities
@orchestra-research/ai-research-skills
v1.7.2 · 3 months ago
Install AI research engineering skills to your coding agents (Claude Code, OpenCode, Cursor, Gemini CLI, Hermes Agent, and more)
No known vulnerabilities
@ignitionai/agent-trainer-environment
v0.1.0-alpha.1 · 3 months ago
Prototype state/action/reward environment primitives for Ignition Agent Trainer.
No known vulnerabilities
@classytic/stage
v0.6.0 · 3 days ago
Domain-neutral visual primitive engine for the web. Declarative SVG diagram/math primitives, draggable handles, a math↔pixel coordinate system, a motion core + energy effect defs, theming, pure sim cores, and a small kit of neutral glyph primitives. React
No known vulnerabilities
@holoscript/absorb-service
v6.1.5 · 23 days ago
Codebase intelligence, recursive self-improvement pipeline, and daemon system for HoloScript
No known vulnerabilities
posttrain
v0.1.8 · 4 hours ago
PostTrain: run and track post-training (data, RL environments, SFT, preference tuning, RL, evals, deploy) from the terminal or the browser.
No known vulnerabilities
@belvedir/cli
v0.4.3 · 1 day ago
Command-line access to the Belvedir platform: projects, keys, traces, training, review, benchmarks, environments, organizations, routing, and billing.
No known vulnerabilities
@mlx-node/core
v0.0.15 · 14 days ago
Internal native bindings - import from @mlx-node/lm or @mlx-node/trl instead
No known vulnerabilities
@ignitionai/agent-trainer-rl
v0.1.0-alpha.1 · 3 months ago
Experimental RL-inspired utilities for Ignition Agent Trainer.
No known vulnerabilities
wasmtune
v0.2.0 · 7 days ago
Fine-tune a small LLM on your website's own pages (local LoRA/QLoRA via Unsloth on NVIDIA or MLX on Apple Silicon) and embed it as a WebGPU chat window that serves the model tier matching each visitor's hardware. Generic — works with any static site folde
No known vulnerabilities
@moolam/learning
v1.1.0 · 2 months ago
Internal component of Sutra SDK — applications should install sutra-sdk. Learning substrate contracts: turn-trajectory schema (B9 metadata + C0 training fields), parse boundary for corpus/gym/critic pipelines.
No known vulnerabilities
@timps-ai/timps-code
v2.0.3 · 1 month ago
Open-source CLI coding agent with 22-layer persistent memory. Free with Ollama. Drop-in Claude Code alternative. Works with any model.
No known vulnerabilities
ah-my-openresearch
v0.1.5 · 3 months ago
Research-lab layer for OpenCode: six personas, 17 curated skills, and a typed claim-level lab record with provenance that survives the chat.
No known vulnerabilities
@daydreamsai/synthetic
v0.3.9-alpha.1 · 1 year ago
Synthetic data generation for AI agent training
No known vulnerabilities