Skip to content

Emmanuel-Rono/LLM-Regression-evaluation

Folders and files

NameName
Last commit message
Last commit date

Latest commit

 

History

4 Commits
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 

About

An agentic framework for semantic regression testing of large language models, designed to detect behavioral drift across model versions using tolerance-based oracles instead of brittle string assertions.

Stars

Watchers

Forks

Releases

No releases published

Packages

 
 
 

Contributors

Languages