Topic · Education

Best quizzes and assessments skills for Claude Code, Codex and other agents.

Skills that write quizzes, exam questions and grading rubrics.
skills
285
official
12

Quizzes and assessments skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Quizzes and assessments skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Teaches the agent to set up and run DeepTutor from the command line: chat and capabilities, knowledge bases, partners, memory, sessions, notebooks and the server or Web app.

HKUDS/DeepTutor41k—~2.8kAutomated safety check: PassApache-2.0today
2

Runs a 10-question quiz across five areas to place a learner in the AI Engineering from Scratch curriculum, so they skip what they already know.

rohitg00/ai-engineering-from-scratch66k—~2kAutomated safety check: PassMITyesterday
3

Turns a codebase into an interactive single-page HTML course for non-technical learners, with scroll modules, animated diagrams, quizzes and plain-English code translations.

zarazhangrui/codebase-to-course5.7k—~4.4kAutomated safety check: PassNo licence6 mo ago
4

Quizzes you on a completed phase of the AI Engineering from Scratch course, taking a phase number or name and mapping it to that phase's directory.

rohitg00/ai-engineering-from-scratch66k—~2.1kAutomated safety check: PassMITyesterday
5

Turns books, PDFs, slides and web pages into a source-grounded knowledge base and an interactive learning page in English or Chinese, with quizzes, relationship maps and reusable methodology notes.

dmoshehun-prog/learn-from-materials916—~7.9kAutomated safety check: PassMITyesterday
6

Maps the structure of an OpenMAIC stage document so an agent can find the right path, read it and patch quizzes, widgets, actions and project pages without guessing.

THU-MAIC/OpenMAIC40k—~2.4kAutomated safety check: PassMITtoday
7

Provide qualitative-first, evidence-traceable developmental review of scholarly works and audit low-stakes research-assessment rubrics with optional local quality controls.

K-Dense-AI/claude-scientific-writer2.4k2 repos~2.9kAutomated safety check: NotesMIT8 days ago
8

This skill should be used when building agent evaluation systems: deterministic checks, regression suites, multi-dimensional rubrics, quality gates, production monitoring, baseline comparison, and…

guanyang/open-agent-hub9752 repos~4.2kAutomated safety check: PassMITyesterday
9

Guides a learner through one of four independent Claude certification tracks with onboarding, lessons, practice labs, mock exams and remediation.

rohitg00/ai-engineering-from-scratch66k—~3kAutomated safety check: PassMITyesterday
10

Quizzes you on the notes in an Obsidian StudyVault, tracks proficiency per concept and drills weak areas in four-question rounds.

bevibing/tutor-skills1.3k—~1.4kAutomated safety check: PassMIT7 mo ago
11

Quizzes you on Claude Code in a quick or deep mode, scores your level across 10 topics and recommends what to learn next, in Chinese.

lhfer/claude-howto-zh-cn2.3k—~1.7kAutomated safety check: PassMIT2 mo ago
12

Designs a review-and-practice lesson around an independent first attempt, targeted feedback, supported practice, a fresh independent check and a next step.

THU-MAIC/OpenMAIC40k—~1.1kAutomated safety check: PassMITtoday
13

Builds a Verifiers (PrimeIntellect) variant of an RL environment.

adithya-s-k/FineEnvs4431 repo~2.3kAutomated safety check: PassApache-2.0today
14

Authors deterministic and LLM rubric graders for skillgrade evaluations.

mgechev/skillgrade720—~972Automated safety check: PassMIT2 days ago
15

Train strict project ownership from repo-local docs/ai memory and central LLM Wiki project entities.

tudoumashu/ai-memory-skillpack422—~1.8kAutomated safety check: PassMIT1 mo ago
16

Teaches the next lesson of the AI Engineering from Scratch curriculum in the terminal, quizzes you at the end and records your progress.

rohitg00/ai-engineering-from-scratch66k—~2.2kAutomated safety check: PassMITyesterday
17

SkillsBench task authoring — walk a contributor from idea to submission-ready task following CONTRIBUTING.md and the task-implementation rubric.

benchflow-ai/benchflow353—~4.5kAutomated safety check: PassApache-2.02 days ago
18

Runs an interactive quiz in a quick or comprehensive mode, scores your skill level by topic and generates a personalized learning path with practice projects.

FlorianBruniaux/claude-code-ultimate-guide6.1k—~2.3kAutomated safety check: PassCC-BY-SA-4.0yesterday
19

Universal quality bar and final audit rubric for any agent system prompt.

mastra-ai/mastra29k—~2kAutomated safety check: PassUnknowntoday
20

Runs a quick or deep quiz on Claude Code skills, scores ten feature areas and generates a personalized learning path with prioritized next steps.

luongnv89/claude-howto42k—~5.5kAutomated safety check: PassMIT8 days ago
21
21.Commerce EvalsOfficial

Authoring and running behavioral evals for a shopping or merchant agent, covering the case shape, authoring rules, code graders and judges, the run pattern, and poisoned fixtures.

anthropics/commerce-agents3.2k—~1.8kAutomated safety check: PassApache-2.06 days ago
22

Critique a data graphic against a nine-criterion rubric derived from Tufte's VDQI — score it, name the chartjunk species present, compute the lie factor, compare against the book's named-failure…

gnurio/tufte-vdqi-plugin309—~1.6kAutomated safety check: PassNo licence2 mo ago
23

Make 100 versions of Claude fight to the death over one task.

Jakeschincariol/arena-skill340—~4.8kAutomated safety check: PassMIT11 days ago
24

Author a new CORAL task — the three pieces that must line up (task.yaml, seed/, a packaged grader/), the coral init → coral validate → smoke-test loop, and how to pick a grader pattern (stdout…

Human-Agent-Society/CORAL1k—~2.2kAutomated safety check: PassApache-2.01 mo ago
25

A skill your agent uses when converting an existing benchmark, rubric, verifier, task YAML/JSON, or domain check into SkillEvaluator BYOG/BYOT custom evaluation.

NVIDIA/SkillEvaluator5481 repo~2.1kAutomated safety check: PassApache-2.0today
26

Interactive wizard: guided questions with multiple-choice options about subscriptions, then outputs a ready-to-paste capabilitytiers YAML + fixed agent model assignments.

yohey-w/multi-agent-shogun1.4k—~3.1kAutomated safety check: PassMIT2 mo ago
27

Compile a repository's instruction files (AGENTS.md, CLAUDE.md and friends) into an Abide rubric, then validate and calibrate it.

coldteadotai/abide564—~3.3kAutomated safety check: PassMIT3 days ago
28

SkillsBench task PR review — classifies the task track (standard / research / multimodal), runs static policy checks against the track-specific rubric, benchmarks the task across oracle plus Claude…

benchflow-ai/benchflow353—~4.5kAutomated safety check: NotesApache-2.02 days ago
29

Evaluate one design or a user-approved maturity-mapped batch through a transparent evidence-based rubric.

SeanJ1ang/design-judge-skills712—~3.1kAutomated safety check: PassApache-2.01 mo ago
30

Audit a website against the Forter Agentic Readiness Guide. An agent skill from forter/agentic-readiness-guide.

forter/agentic-readiness-guide106—~4.6kAutomated safety check: PassUnknown16 days ago
31
31.Perf ReviewOfficial

Performance-overhead review of a code diff / branch / PR for the dd-trace-java tracer.

DataDog/dd-trace-java736—~3.6kAutomated safety check: NotesApache-2.0today
32

Quizzes a learner on one lesson of the Claude Code tutorial with ten questions, scores the answers and points out weak spots.

luongnv89/claude-howto42k—~2.3kAutomated safety check: PassMIT8 days ago
33

Add fast, local, typed decisions to any project with Laya, an open-source non-generative decision model (pip install laya).

wdobry/laya-playground182—~3.2kAutomated safety check: PassMIT17 days ago
34

Write prompts, system instructions, agent directives, slash commands, and skill descriptions using two stacked layers — outcome-first (define the destination, success criteria, stopping condition)…

kingbootoshi/directional-prompting143—~2.3kAutomated safety check: PassMIT4 mo ago
35

Comprehensive OSINT methodology for external red-team operations and authorized attack-surface assessments.

elementalsouls/Claude-OSINT2.8k—~8.7kAutomated safety check: NotesMIT1 mo ago
36

GAN-style iterative improvement loop for any text artifact. An agent skill from crimeacs/auto-improve.

crimeacs/auto-improve135—~651Automated safety check: PassMIT2 mo ago
37

考公AI导师 — a tutor for the Chinese civil service exam (公务员考试), covering 行测 (aptitude), 申论 (essay), and 面试 (structured interview).

KeWang0622/kaogong-skill161—~1.2kAutomated safety check: PassMIT29 days ago
38

Review code changes yourself, locally, with OrcaCode Review's severity contract and merge gate — no GitHub Action, no OrcaRouter account, no API key.

Continuum-AI-Corp/Orca-Code-Review173—~3.5kAutomated safety check: PassMIT17 days ago
39

A complete methodology for turning scanned or image-based learning materials into high-quality desktop, web, or mobile practice products.

parz0val0/scan-to-practice108—~2kAutomated safety check: PassMIT1 mo ago
40

Build investor-ready pitch scripts in multiple formats (10-min, 5-min, 2-min, 1-min elevator, investor email).

ferdinandobons/startup-skill1.2k—~6.4kAutomated safety check: PassMIT3 mo ago
41
41.Review PROfficial

Review a specific vscode-containers pull request on demand from the CLI (or any interactive agent), the way a Container Tools maintainer would.

microsoft/vscode-containers141—~900Automated safety check: PassUnknowntoday
42

Find and read academic papers (S2 + arXiv). An agent skill from EvoScientist/EvoSkills.

EvoScientist/EvoSkills475—~6.3kAutomated safety check: NotesApache-2.07 days ago
43

Expert guide for the NotebookLM CLI (nlm) and MCP server - interfaces for Google NotebookLM.

iusztinpaul/ai-research-os-workshop1791 repo~6.9kAutomated safety check: PassMIT3 mo ago
44

Organize computed metrics into a tiered evaluation rubric with leading, lagging, and quality indicators.

kayba-ai/agentic-context-engine2.6k—~2.2kAutomated safety check: PassApache-2.014 days ago
45

A skill your agent uses when performing App Store Optimization work with OpenASO MCP data: ASO audits, keyword research, metadata optimization, screenshot strategy, review analysis, competitor…

hubab1/OpenASO177—~946Automated safety check: PassMIT1 mo ago
46

cheat-on-money skill 的反诈核心。对某个具体的兼职/副业/项目做反诈 rubric 打分,实时上网查证负面与真实收款证据,输出 高危/存疑/可行 判定 + 验证第一步 + 诚实收入预期。触发词:"XX靠谱吗"/"这个项目是不是骗局"/"帮我查下这个兼职"/"money verify"/"验证机会"/"这个能信吗"。

XBuilderLAB/cheat-on-money769—~503Automated safety check: NotesNo licence3 mo ago
47

Paperclip Vision is an AI-powered founder interview skill that extracts strategic direction from a company founder and produces two ready-to-use documents: VISION.md (the company constitution) and…

aronprins/paperclip-vision101—~3.7kAutomated safety check: NotesNo licence6 mo ago
48

Run the weak-agent adversarial test harness against docx-cli.

kklimuk/docx-cli216—~6.1kAutomated safety check: NotesMIT12 days ago

Questions, answered from the data.

What is the best quizzes and assessments skill?

DeepTutor CLI from HKUDS/DeepTutor ranks first of the 285 quizzes and assessments skills listed here, with the highest score: its repository has 41k GitHub stars, its SKILL.md loads about 2.8k tokens and it passes the automated safety check with no findings. Next come AI Engineering Placement Quiz and Codebase to Course.

Which quizzes and assessments skills are official?

12 of the 285 quizzes and assessments skills are official, published by the vendor's own GitHub organization: Commerce Evals, Create Custom Grader, Perf Review, Review PR, PR Review and 7 more.

How are these skills ranked?

By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.