Pinned Loading
-
golden-dataset-studio
golden-dataset-studio PublicTool to create golden data set and compare with generated LLM answer and assess the same.
Python
-
rag-eval-project
rag-eval-project PublicRAG pipeline with LLM evaluation using DeepEval, Ollama/Groq judge models, and Langfuse observability.
Python
-
self-healing-schema-sentinel
self-healing-schema-sentinel PublicSelf-healing schema validation for JSON APIs — detects field drift and auto-generates updated pytest tests using LLMs.
HTML
-
-
-
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.