
UNO played by a zero-shot 0.8B decision model (typed choices, no text gen) vs heuristics and humans — 22/30 vs random, live browser table with...
UNO played by a zero-shot 0.8B decision model (typed choices, no text gen) vs heuristics and humans — 22/30 vs random, live browser table with probability bars. RLCard + Python.
Lynote combines AI-draft rewriting and AI-text detection with note-taking, transcription, video summaries, and study tools.
LMSYS Chatbot Arena is a crowdsourced open platform for LLM evals. Collected over 1,000,000 human pairwise comparisons to rank LLMs with the Bradley-Terry model and display the model ratings in Elo-scale.
Zhipu Qingyan is a GLM-based AI assistant whose official description emphasizes understanding goals, breaking down tasks, and using tools.
字节跳动AI助手