Public datasets for memory-grounded and command-line AI-agent evaluation. Research artifacts only; not commercial Klik releases or product validation.
Chengyi Xu
ChengyiX
AI & ML interests
None yet
Recent Activity
new activity about 12 hours ago
ChengyiX/CLI-Bench:Living Memory then Proactive Execution: state-diff recovery evidence new activity about 20 hours ago
ChengyiX/KLIK-Bench:Living Memory then Proactive Execution: evaluate the boundary new activity about 20 hours ago
ChengyiX/context-run-game:Living Memory then Proactive Execution: play the decision boundaryOrganizations
None yet