Skip to main content
Avatar for Gurbaksh Chahal from gravatar.com

Gurbaksh Chahal

Username    auraoneai
Date joined   Joined

26 projects

auraone-sdk

Last released

Python SDK and CLI for AuraOne hosted AI evaluations, with sync and async clients for templates, runs, analytics, training, and governance.

auraone-agent-studio-open

Last released

Headless Agent Studio Open protocol, trace-store, sidecar, and export CLI.

rubric-studio

Last released

Install guidance and package-name reservation for Rubric Studio Open.

lerobot-quality-gates

Last released

Check LeRobot-style metadata, episodes, sensors, videos, labels, and dataset-card disclosures.

robostudio-engine

Last released

Inspect, index, QA, cluster, probe, and export robotics datasets from a headless CLI.

robot-recovery-bench

Last released

Validate robot recovery segment JSONL and summarize intervention outcomes and timing.

vla-robustness-kit

Last released

Run deterministic, simulator-free VLA perturbation diagnostics over episode metadata and instructions.

embodiment-card

Last released

Validate and render robot embodiment metadata for dataset and VLA release review.

tool-call-replay

Last released

Normalize recorded tool-call traces into deterministic local assertions and pytest regressions.

agent-trace-card

Last released

Generate local Markdown, HTML, and JSON review cards from agent trace JSON.

datasheet-ci

Last released

Validate required Datasheet, Model Card, and Data Card headings with warning-only PII pattern checks.

mcp-risk-linter

Last released

Statically lint MCP server source and docs for capability, permission, and disclosure risks.

a2a-contract-test

Last released

Validate A2A-style agent cards and recorded task lifecycle transcripts without contacting an agent.

prompt-rubric-drift

Last released

Detect deterministic prompt and rubric drift and produce pull request review reports.

eval-adapter

Last released

Export rubric-spec rubrics and normalize evaluation result shapes across common eval frameworks.

otel-eval-bridge

Last released

Convert OTLP and Phoenix GenAI trace exports into redacted local eval cases and manifests.

judge-bench

Last released

Synthetic bias and calibration probes for testing LLM-as-judge reliability.

failure-gallery

Last released

Validate and build a synthetic agent and robotics failure gallery with reproducible review records.

synthetic-disagreement

Last released

Generate deterministic synthetic reviewer disagreement and inter-annotator agreement stress curves.

judge-card

Last released

Generate, validate, and render inspectable LLM judge disclosure cards from diagnostic results.

iaa-kit

Last released

Inter-annotator agreement metrics and bootstrap confidence intervals for review and labeling workflows.

contamination-audit

Last released

Generate item-level eval contamination signals from lexical, canary, pattern, hash, and optional embedding checks.

eval-run-manifest

Last released

Build and validate eval-run provenance manifests with deterministic directory digests.

eval-conformance-suite

Last released

Run rubric-spec adapter conformance checks and render deterministic status badges.

rubric-spec

Last released

Validate, lint, diff, and convert portable LLM evaluation rubrics with AuraOne Rubric Schema v1.

auraone-evalkit

Last released

Local-first Python CLI for AI and LLM evaluation rubrics, deterministic scoring, reviewer QA, leakage checks, and evidence reports.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page