Faquad-nli needs to be updated
pinned👍 1
1
#18 opened about 1 year ago
by
nicholasKluge
Multi-metric evaluation framework for Portuguese LLMs — cost + accuracy + hallucination
#22 opened 3 months ago
by
vigneshwar234
Testing considering other LLMs
1
#21 opened about 1 year ago
by
jpedrofonseca
Possible issue with evaluation scores of Falcon-H1 Models
❤️🧠 3
3
#20 opened about 1 year ago
by
rcojocaru
Add OpenAI open-source model
1
#19 opened about 1 year ago
by
arthurcavalcant