Open-Source Library for Fully Cooperative Multi-LLM Reinforcement Learning
-
Updated
Sep 7, 2026 - Python
Open-Source Library for Fully Cooperative Multi-LLM Reinforcement Learning
EMNLP 2026 Findings benchmark for LLM social reasoning across dynamic interaction, social deduction, deception, and information uncertainty; maintained HRD code and data.
Structured behavioral labeling and trust-curve analysis for replay-faithful AI conversation logs.
Benchmark and evaluation code for the ICLR 2026 Workshop paper 'Residual Drift Dominates Contradiction in Multi-Turn Constraint Reasoning'. 1,020 Z3-validated multi-turn constraint problems across seating, scheduling, and logic-grid domains.
To associate your repository with the multi-turn-reasoning topic, visit your repo's landing page and select "manage topics."