← Back to List
⚠
Notice:This resource is provided by a third-party author. Please review the code with AI tools or manually before use to ensure security and compatibility.
Pythonrelari-ai/continuous-eval

continuous-eval

Data-Driven Evaluation for LLM-Powered Applications

70.2/100
★ 517Forks: 38
View on GitHub →Homepage →
Loading report...

Similar Projects

haystack

94

Open-source AI orchestration framework for building context-engineered, production-ready LLM applications. Design modular pipelines and agent workflows with explicit control over retrieval, routing, memory, and generation. Built for scalable agents, RAG, multimodal applications, semantic search, and conversational systems.

Python★ 26.6K

deepeval

88

The LLM Evaluation Framework

Python★ 18.4K

LLM-Engineers-Handbook

61

The LLM's practical guide: From the fundamentals to deploying advanced LLM and RAG apps to AWS using LLMOps best practices

Python★ 5.3K

future-agi

81

Open-source, end-to-end platform for evaluating, observing, and improving LLM and AI agent applications. Tracing · Evals · Simulations · Datasets · Gateway · Guardrails. Self-hostable. Apache 2.0.

Python★ 2.1K
← Back to List