AI Infra Bench evaluates frontier AI models on real-world AI infrastructure engineering tasks, starting with vLLM and covering CPU and GPU workloads including bugs, features, performance changes, refactors, and tests.
Open source. Open possibilities.
Discover quality open-source projects, submit projects anonymously, and claim and edit your own project.
A little curiosity. A world of open source.
THE FIRST COLLECTIONHermesMade provides six CLI tools addressing common AI user pain points: censorship risk analysis, model quality monitoring, API cost optimization, local LLM deployment, code inspection, and task cost estimation.
A community-maintained catalog of Bangla (Bengali) NLP resources — 712 papers, 63 datasets, 20 models, and 9 tools across 13 tasks — built as a static Astro site with strict data verification and shareable filters.
Open LLM hallucination benchmark regarding the BNCC (Brazil's National Common Curricular Base). It measures how much models invent official codes and texts using an auditable methodology with raw data and CI.