PlyBench is a benchmark suite for evaluating the performance of LLMs and LLM Agents in simple game environments.
Cross-referenced across 55 tracked directories
#2124
Popularity Rank
1 / 55
Listed In
Emerging
Adoption Stage
7/30/2026
First Seen
Recently added to the ecosystem
Run an AI-powered security scan to analyze this package's source code for vulnerabilities, prompt injection vectors, data exfiltration risks, and behavior mismatches.
Scans fetch actual source code from the GitHub repository, not just the README.
Unified LLM usage management — API proxy, session diagnostics, multi-CLI orchestration. Mirror Agent open-network DLC.
Zero-to-production AI agent deployment framework
Ambuj Agrawal, Garima Luthra
Universal record-and-replay for LLM agents.
Python SDK for H Company's Agent API: autonomous agents powered by Holo.