mirror of
https://github.com/wassname/ml-debug.git
synced 2026-09-26 14:00:26 +08:00
715164416bf703685933a000e1abe734f3ea060f
- triplet now carries a prior + cheapest falsifier (Check:) per hypothesis - discriminating-test step: forward-predict each hypothesis, prefer where predictions diverge (strong vs weak evidence) instead of just "discriminating" - new step: bisect the forward/backward path to localize where it breaks - compact pseudocode summary of the whole loop - resolve FIXME: drop references to the non-public research-journal skill Co-Authored-By: Claudypoo <288921227+claudypoo@users.noreply.github.com>
ML Debugging Folklore
Practitioner knowledge for debugging ML systems, curated and synthesized by wassname. Opinionated by source selection -- I picked sources I trust (Schulman, Goodfellow, CS231n, ...) and had an LLM extract the most relevant information for debugging ML systems.
Use as a Claude skill
/skills add https://github.com/wassname/ml_debug
Or paste SKILL.md into your system prompt / context when debugging.
What's here
-
SKILL.md -- the main artifact. Load into an LLM agent's context as a debugging skill. Parts 1-5 are reference knowledge; Part 6 is a runnable triage protocol (grep patterns, diagnostic snippets, decision tree); Part 7 is debugging mental models and practitioner priors.
-
docs/evidence/ -- frozen local copies of source material (blog posts, talks, papers, reddit threads). Claims in SKILL.md link back to exact quotes here.
Description
skill for debugging and dev of machine learning, collected over the years in an attempt to uplift agents (and myself)
3.7 MiB
Languages
Python
99%
Just
1%