Back to List
Notice:This resource is provided by a third-party author. Please review the code with AI tools or manually before use to ensure security and compatibility.
Pythonrun-llama/ParseBench

ParseBench

ParseBench - A Document Parsing Benchmark for AI Agents

77.2/100
565Forks: 101
View on GitHubHomepage →
Loading report...

Similar Projects

opencompass

87

OpenCompass is an LLM evaluation platform, supporting a wide range of models (Llama3, Mistral, InternLM2,GPT-4,LLaMa2, Qwen,GLM, Claude, etc) over 100+ datasets.

Python7.4K

unstract

88

LLM-Driven Extraction of Unstructured Data — Built for API Deployments & ETL Pipeline Workflows

Python7.2K

LightCompress

56

[EMNLP 2024 & AAAI 2026] A powerful toolkit for compressing large models including LLMs, VLMs, and video generative models.

Python747

ClawBench

77

Open-source benchmark for browser AI agents on daily tasks.

Python673
Back to List