Open Issues Need Help
View All on GitHub good first issue
Open, local-first behavioral evidence and evaluation infrastructure for AI systems, benchmarks, and agents.
TypeScript
#ai#ai-agents#ai-evaluation#behavioral-evaluation#benchmarking#evidence#llm#local-first#observability#open-source#reproducibility#research-infrastructure
documentation good first issue
Open, local-first behavioral evidence and evaluation infrastructure for AI systems, benchmarks, and agents.
TypeScript
#ai#ai-agents#ai-evaluation#behavioral-evaluation#benchmarking#evidence#llm#local-first#observability#open-source#reproducibility#research-infrastructure
documentation good first issue
Open, local-first behavioral evidence and evaluation infrastructure for AI systems, benchmarks, and agents.
TypeScript
#ai#ai-agents#ai-evaluation#behavioral-evaluation#benchmarking#evidence#llm#local-first#observability#open-source#reproducibility#research-infrastructure