Back to List
Notice:This resource is provided by a third-party author. Please review the code with AI tools or manually before use to ensure security and compatibility.
PythonGoodStartLabs/AI_Diplomacy

AI_Diplomacy

Frontier Models playing the board game Diplomacy.

54.2/100
687Forks: 96
View on GitHub
Loading report...

Similar Projects

tau2-bench

81

τ-Bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains

Python1.7K

InferenceX

73

Open Source Continuous Inference Benchmark Research Platform — Kimi K2.7-Code, MiniMax M3, DeepSeekv4, GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72 & soon™ TPUv6e/v7/Trainium2/3 | 开源持续推理基准研究平台 — Kimi K2.7-Code、MiniMax M3、DeepSeekv4、GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72,即将推出™ TPUv6e/v7/Trainium2/3

Python1.3K

meta-agents-research-environments

70

Meta Agents Research Environments is a comprehensive platform designed to evaluate AI agents in dynamic, realistic scenarios. Unlike static benchmarks, this platform introduces evolving environments where agents must adapt their strategies as new information becomes available, mirroring real-world challenges.

Python531

hermes-agent

91

The agent that grows with you

Python220.0K
Back to List