Local RAG evaluation framework for compressed trace retrieval, grounding, citation support, multi-step traces, latency profiling, and agentic efficiency.
-
Updated
Jun 28, 2026 - Python
Local RAG evaluation framework for compressed trace retrieval, grounding, citation support, multi-step traces, latency profiling, and agentic efficiency.
High-performance C++20 order book engine with REST API, React web terminal, LOBSTER replay, and online ML pipeline.
Autonomy ML project for driving data, failure mining, policy learning, safety metrics, and latency benchmarking.
Add a description, image, and links to the latency-profiling topic page so that developers can more easily learn about it.
To associate your repository with the latency-profiling topic, visit your repo's landing page and select "manage topics."