Popular repositories Loading
-
llm-inside-lab
llm-inside-lab Public把大语言模型的黑盒拆开看:纯前端的 LLM 内部机制交互可视化实验室。分词→词嵌入→注意力→采样→KV Cache 全链路可视化,支持浏览器内跑真实小模型,中英双语,零后端。
TypeScript 1
-
benchshield
benchshield PublicAnti-cheating audit and isolated evaluation sandbox for AI agent benchmarks: detect the 7 vulnerability patterns, grade on the Agent-Eval Checklist, red-team with zero-capability agents. Zero depen…
Python 1
-
agent-battle-arena
agent-battle-arena Public可扩展的多智能体博弈竞技场:规则基线 / 在线学习 / LLM 智能体在可复现锦标赛中同台对垒(囚徒困境·猜硬币·公共品·拍卖·序贯谈判),Elo 排行 + 行为科学指标,核心零依赖 | Extensible game-theoretic arena where scripted, learning and LLM agents compete under reproducible tour…
Python
-
agent-arena
agent-arena Public人机混合的 LLM 智能体社交博弈平台 · 狼人杀/阿瓦隆 · 观战 · 回放 · AI 复盘 · ELO 天梯 | Hybrid human+LLM social deduction games
Python
If the problem persists, check the GitHub status page or contact support.