Back to List
Notice:This resource is provided by a third-party author. Please review the code with AI tools or manually before use to ensure security and compatibility.
Pythonharbor-framework/terminal-bench-science

terminal-bench-science

Terminal-Bench-Science: Evaluating AI agents on research workflows across scientific domains

73.6/100
544Forks: 316
View on GitHubHomepage →
Loading report...

Similar Projects

AutoGPT

96

AutoGPT is the vision of accessible AI for everyone, to use and to build on. Our mission is to provide the tools, so that you can focus on what matters.

Python187.2K

PaddleOCR

86

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.

Python89.1K

OpenMontage

80

World's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent skill and production-knowledge files. Turn your AI coding assistant into a full video production studio.

Python56.6K

agents

83

Multi-harness agentic plugin marketplace for Claude Code, Codex, Cursor, OpenCode, GitHub Copilot, and Google Antigravity

Python39.5K
Back to List