Quality

Code review, testing, and verification skills.

All Quality (1897 found)

πŸ“¦
2026/07/20

AAMAS Artifact Evaluation

Guides the packaging of multiagent code, game definitions, opponent sets, and logs for AAMAS artifact evaluation, ensuring reviewers can inspect and re-run interaction claims like emergent cooperation or convergence to equilibrium.
Quality
974114
πŸ”
2026/07/20

AAMAS Experiments

Guides the design and audit of multi-agent experiments for AAMAS, focusing on interaction claims, opponent selection, equilibrium metrics, and reproducibility. Ensures experiments probe strategic interactions rather than single-agent performance.
Quality
974114
πŸ”
2026/07/20

AAMAS Reproducibility

Audit and strengthen reproducibility evidence for multi-agent interaction claims in AAMAS papers, covering game definitions, opponent sets, seeds, uncertainty, and artifact consistency.
Quality
974114
πŸ“¦
2026/07/20

ACL Artifact Evaluation

Use when packaging code, datasets, prompts, model outputs, or annotation materials for an ACL submission, covering anonymized supplements, Responsible NLP checklist items, licensing, data statements, and post-acceptance release.
Quality
974114
πŸ§ͺ
2026/08/12

ACL Experiments

ACL Experiments helps design or audit experiments for ACL papers, including baselines, multi-dataset and multilingual evaluation, significance testing, human evaluation, contamination checks, ablations, and error analysis. Use it when preparing evidence for an NLP research claim.
Quality
974125
πŸ›‘οΈ
2026/08/12

CCS Artifact Evaluation

CCS Artifact Evaluation helps package accepted security research artifacts for ACM badges and committee review. It covers reproducible setup, turnkey attack or defense demos, safe handling, and justifying withheld files.
Quality
974125
πŸ›‘οΈ
2026/08/12

CCS Experiments

CCS Experiments helps audit ACM CCS attack and defense evaluations for claim-to-evidence fit, adaptive attackers, baselines, and measured overhead. Use it when preparing or reviewing experiments for a security paper submission.
Quality
974125
πŸ”Ž
4w ago

Codex Review

Codex Review runs an independent code review with the OpenAI Codex CLI and saves a severity-prioritised report. Use it for PRs, commits, uncommitted changes, or whole-app checks to catch bugs, regressions, security, and testing gaps.
Quality
974102
🧩
4w ago

Fork Discipline

Fork Discipline audits multi-client codebases for blurred core/client boundaries, hardcoded client checks, config replacement bugs, scattered client code, and missing extension points. It produces a boundary map, violation report, and refactoring plan.
Quality
974102
πŸ“±
4w ago

Responsiveness Check

Responsiveness Check tests a website across viewport widths in one browser session. It screenshots breakpoints and reports overflow, stacking, navigation changes, and the exact widths where layouts shift or break.
Quality
974102
πŸ§ͺ
4w ago

ASC Crash Triage

ASC Crash Triage fetches and summarizes TestFlight crashes, beta feedback, and performance diagnostics from App Store Connect. Use it when you need to investigate crashes, hangs, disk writes, launches, or beta tester reports for an app or build.
Quality
97155
πŸ—ΊοΈ
4w ago

Map Territory

Map Territory checks live code or data when docs, tests, metrics, or assumptions conflict with observed behavior. Use it to confirm whether the current territory matches the map before making a decision.
Quality
969133
PreviousPage 72 of 159Next