P
PawBench
agentscope-ai/PawBench
A benchmark for evaluating LLM × harness performance.
★97stars
Python
Apache-2.0
Updated: 1d ago
📋 Project at a Glance
Tap to expand
What's this?A , built with Python open-source project in the Security & Governance category, core strengths: agent/benchmark
Who made it?Maintained by agentscope-ai team, 97⭐ on GitHub, #300 out of 413 in Security & Governance
Why does it exist?With the growing demand for agent and benchmark in Security & Governance, agentscope-ai created PawBench as a streamlined solution.
What can it do?Key use cases: harness, hermes, llm
How to install with AI?Use an AI coding assistant (Claude Code, Cursor, Copilot) to automatically set up pip dependencies and virtual env. Follow the README — the AI handles the rest.
🔗 github.com/agentscope-ai/PawBench | 官网 https://agentscope-ai.github.io/PawBench/
🔗 github.com/agentscope-ai/PawBench | 官网 https://agentscope-ai.github.io/PawBench/
Topics
agentbenchmarkharnesshermesllmopenclawqwenpaw