Tag: ai-evaluation
| # | Post Title | Date | User |
| AI evaluation platforms and release gates: what should teams choose? | 1 month ago | NHI Mgmt Group | |
| Multi-turn AI evaluations: what single-turn scoring is missing | 1 month ago | NHI Mgmt Group | |
| AI evals in CI/CD pipelines: are your quality gates enough? | 1 month ago | NHI Mgmt Group | |
| Golden datasets with human review: what AI teams need to know | 1 month ago | NHI Mgmt Group | |
| LLM evaluation beyond Grafana: what teams need to fix | 1 month ago | NHI Mgmt Group | |
| Speech-to-text accuracy is not the whole story for voice agents | 1 month ago | NHI Mgmt Group | |
| AI evals and observability: what practitioners are missing | 1 month ago | NHI Mgmt Group | |
| Compound AI systems: are final-output checks enough for teams? | 2 months ago | NHI Mgmt Group | |