ParseBench - A Document Parsing Benchmark for AI Agents
OpenCompass is an LLM evaluation platform, supporting a wide range of models (Llama3, Mistral, InternLM2,GPT-4,LLaMa2, Qwen,GLM, Claude, etc) over 100+ datasets.
LLM-Driven Extraction of Unstructured Data — Built for API Deployments & ETL Pipeline Workflows
ExtractThinker is a Document Intelligence library for LLMs, offering ORM-style interaction for flexible and powerful document workflows.
Open-source benchmark for browser AI agents on daily tasks.