Open Issues Need Help
View All on GitHub Add licensed public multimodal dataset adapters about 1 month ago
enhancement help wanted
Local-first, auditable evaluation framework for AI quality, safety, compliance evidence, and serving performance.
Python
#ai-evaluation#ai-safety#benchmarking#llm-evaluation#performance-testing#python
Document the offline demo JSON output contract about 1 month ago
documentation good first issue
Local-first, auditable evaluation framework for AI quality, safety, compliance evidence, and serving performance.
Python
#ai-evaluation#ai-safety#benchmarking#llm-evaluation#performance-testing#python
Improve offline demo report accessibility checks about 1 month ago
enhancement good first issue
Local-first, auditable evaluation framework for AI quality, safety, compliance evidence, and serving performance.
Python
#ai-evaluation#ai-safety#benchmarking#llm-evaluation#performance-testing#python
Add Windows CI smoke coverage for the offline demo about 1 month ago
good first issue github_actions
Local-first, auditable evaluation framework for AI quality, safety, compliance evidence, and serving performance.
Python
#ai-evaluation#ai-safety#benchmarking#llm-evaluation#performance-testing#python
Add a vLLM OpenAI-compatible serving contract fixture about 1 month ago
enhancement help wanted python
Local-first, auditable evaluation framework for AI quality, safety, compliance evidence, and serving performance.
Python
#ai-evaluation#ai-safety#benchmarking#llm-evaluation#performance-testing#python