Skip to content
#

llm-evaluations

Here are 7 public repositories matching this topic...

Language: All
Filter by language

A hyperlocal job board purpose-built for India's Tier-2/3 cities — employers post geotagged jobs, candidates upload resumes that an LLM auto-parses into skill vectors, Elasticsearch matches candidates to nearby jobs within a configurable radius, and job alerts arrive on WhatsApp in the candidate's regional language.

  • Updated Sep 26, 2026
  • Java

Evaluation suite for a banking RAG assistant — faithfulness, hallucination & relevance scoring with CI regression gates (DeepEval · RAGAS · pytest)

  • Updated Oct 8, 2026
  • Python

benchmarking jailbreak-dataset fine-tuning-tools llm-evaluation official high-priority benchmarking task evaluation checkpoint shared memory sandbox API endpoints structured data schemas collaborative environments

  • Updated Sep 9, 2026

Production-ready enterprise RAG and structured data extraction pipeline using LangGraph, Instructor, and Pydantic with automated CI evaluation metrics.

  • Updated Aug 28, 2026
  • Python

Add this topic to your repo

To associate your repository with the llm-evaluations topic, visit your repo's landing page and select "manage topics."

Learn more