Skip to content

koswadi/indonesian-ai-response-evaluation-benchmark

Folders and files

NameName
Last commit message
Last commit date

Latest commit

 

History

1 Commit
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 

Repository files navigation

About

Human-annotated benchmark dataset for evaluating Indonesian AI responses using pairwise comparison, preference ranking, error analysis, and rubric-based scoring. Designed for LLM evaluation, RLHF workflows, AI alignment research, NLP benchmarking, and response quality assessment across diverse real-world prompts and domains.

License

Stars

0 stars

Watchers

0 watching

Forks

Releases

No releases published

Packages

 
 
 

Contributors