AI Tools
11 articles
Claude Code Skills Get Stale. Audit Them Quarterly.
A proposed quarterly audit for AI skills, hooks, and memory entries, with hypothetical failure paths to convert into regression tests.
What Agents Actually Do (And What They Don't)
A conceptual framework for specification-bounded research agents. Its linked timing example is article-reported, not a public workflow benchmark.
What We Mistake for AI Capability
A hypothesis about how task tolerance and specification shape perceived AI capability, with a controlled comparison still to run.
The Data We Forgot We Had: A Tagging System
Tag datasets by the questions they can answer, not just what they contain. A question-first system makes dormant research data discoverable when new questions…
Antigravity Research Graphics Workflow
A practical workflow for using Antigravity image generation to explore a figure layout while keeping the final quantitative graphic reproducible in code.
Cleaning a Research Codebase: An Article-Reported Example
A workflow for mapping dependencies and reorganizing research code. The 47-to-15 script count and timing claims are not publicly reproduced.
Robust API Collection: Pagination, Limits, Retries
Collecting location data from Google Places API at scale requires handling rate limits, pagination, and failure recovery.
Grocery Store Classifier Results Under Review
The article reports 94% balanced accuracy and 94% and 96% spot-check rates; those metrics and the feature-importance values are not publicly reproduced.
Copy-Paste vs. Agent Coding: An Illustrative Comparison
An illustrative comparison of copy-paste and repository-aware coding. The workflows and SNAP-transit correlation are hypothetical, not benchmarks or findings.
Methods-to-Code with AI: An Article-Reported Workflow
An article-described methods-to-code workflow; the 408 by 4,847 example, output, and timing comparison are not publicly reproduced.
One Context File: A Workflow for Persistent Project Context
A workflow for documenting research conventions in CLAUDE.md. Time savings and prevented-error counts are article-reported, not independently timed.