Cyber_analyst-round1 / tests /test_closed_loop_runtime.py

Commit History

feat: enhance CyberSecurity_OWASP observation model with scenario prompt, improve GRPO batch configuration validation, and add scenario grouping for adaptive difficulty curriculum
632c145

Humanlearning commited on

feat: enhance scenario authoring and caching mechanisms, update action submission terminology, and improve reward configuration for CyberSecurity_OWASP environment
be8eade

Humanlearning commited on

feat: implement RL environment server with training infrastructure and Modal integration
6abc8c5

Humanlearning commited on