Source

ndpvt-web/arxiv-claude-skills

81 skills · 94 combined installs

Skills from this source

#
Skill
Source
8W Activity
Installs
1
medspeak-knowledge-graph-aided-asr Build knowledge-graph-aided ASR error correction pipelines for medical speech, using phonetic similarity + semantic r…
ndpvt-web/arxiv-claude-skills
1
2
medverse-reliable-medical-reasoning Decompose complex medical reasoning into DAG-structured parallel execution paths using Petri net theory.
ndpvt-web/arxiv-claude-skills
1
3
papersearchqa-learning-search-reason Build iterative search-and-reason agents for scientific literature QA. Uses the PaperSearchQA pattern: interleaved th…
ndpvt-web/arxiv-claude-skills
1
4
perfguard-performance-aware-agent-visual Performance-aware multi-tool orchestration for visual content generation pipelines. Implements PerfGuard's three mech…
ndpvt-web/arxiv-claude-skills
1
5
polarmem-training-free-polarized-latent Build polarized memory systems for multimodal agents that encode both positive and negative evidence as graph constra…
ndpvt-web/arxiv-claude-skills
1
6
pope-learning-reason-hard Apply the POPE (Privileged On-Policy Exploration) technique to solve hard reasoning problems by decomposing them with…
ndpvt-web/arxiv-claude-skills
1
7
precise-reducing-bias-evaluations Implement the PRECISE framework to debias LLM-as-judge evaluations of search, ranking, and RAG systems by combining a…
ndpvt-web/arxiv-claude-skills
1
8
precision-practice-knowledge-guided Generate industrial-grade code summaries using the ExpSum knowledge-guided approach: function metadata extraction, do…
ndpvt-web/arxiv-claude-skills
1
9
predicting-improving-test-time-scaling Implement Scaling-Law Guided (SLG) Search for test-time compute optimization. Uses reward tail distribution estimatio…
ndpvt-web/arxiv-claude-skills
1
10
predictive-coding-information-bottleneck Build lightweight hallucination detection pipelines using Predictive Coding surprise signals and Information Bottlene…
ndpvt-web/arxiv-claude-skills
1
11
research-multi-stage-machine-learning Build multi-stage search pipelines that separate recall from precision for discovering datasets, documents, or resour…
ndpvt-web/arxiv-claude-skills
1
12
rethinking-scientific-modeling-physically Generate physics-consistent, simulation-executable structural engineering code using constraint-oriented alignment an…
ndpvt-web/arxiv-claude-skills
1
13
ruleflow-generating-reusable-program Optimize Pandas code by discovering per-program improvements, generalizing them into reusable rewrite rules, and appl…
ndpvt-web/arxiv-claude-skills
1
14
safepred-predictive-guardrail-computer-using Implement predictive safety guardrails for computer-using agents and automated pipelines using world-model-based risk…
ndpvt-web/arxiv-claude-skills
1
15
scidatacopilot-agentic-data-preparation Build agentic pipelines that ingest heterogeneous raw scientific data, parse research intent, and produce analysis-re…
ndpvt-web/arxiv-claude-skills
1
16
semantic-aware-advanced-persistent-threat Build anomaly detection pipelines for Advanced Persistent Threat (APT) detection by encoding system logs into semanti…
ndpvt-web/arxiv-claude-skills
1
17
semanticalli-caching-reasoning-not Implement pipeline-aware intermediate representation (IR) caching for agentic systems.
ndpvt-web/arxiv-claude-skills
1
18
seta-statistical-fault-attribution Diagnose and attribute faults in compound AI systems (multi-model pipelines) using SETA's modular robustness testing …
ndpvt-web/arxiv-claude-skills
1
19
sparseeval-evaluation-sparse-optimization Efficiently evaluate LLMs on benchmarks by selecting a small subset of anchor items via sparse optimization, reproduc…
ndpvt-web/arxiv-claude-skills
1
20
supchain-bench-benchmarking-real-world-supply Build reliable long-horizon supply chain agents using the SupChain-ReAct pattern: multi-path ReAct trajectories with …
ndpvt-web/arxiv-claude-skills
1
21
teaching-evaluating-reason-about Apply knowledge-augmented reasoning distillation for polymer design tasks.
ndpvt-web/arxiv-claude-skills
1
22
textual-equilibrium-propagation-deep Optimize deep multi-step AI pipelines using Textual Equilibrium Propagation (TEP) — a two-phase local-then-nudge stra…
ndpvt-web/arxiv-claude-skills
1
23
the-necessity-unified-framework Design and implement standardized, reproducible evaluation harnesses for LLM-based agents. Eliminates confounding fac…
ndpvt-web/arxiv-claude-skills
1
24
thinking-makes-agents-introverted Prevents the "introverted agent" problem where extended reasoning causes agents to give shorter, less informative res…
ndpvt-web/arxiv-claude-skills
1
25
thinktank-me-multi-expert-framework-middle Build multi-expert forecasting systems where specialized LLM agents collaborate through routing and aggregation to pr…
ndpvt-web/arxiv-claude-skills
1
26
timemachine-bench-benchmark-evaluating-capabilitie Systematic dependency migration for Python projects.
ndpvt-web/arxiv-claude-skills
1
27
understanding-agent-scaling-llm-based Design diversity-aware multi-agent systems that maximize performance with fewer agents.
ndpvt-web/arxiv-claude-skills
1
28
when-agents-fail-comprehensive Diagnose and fix bugs in LLM agent systems using a research-backed taxonomy of 11 bug types, 9 root causes, and 12 ob…
ndpvt-web/arxiv-claude-skills
1
29
where-ai-coding-agents Pre-flight checker that prevents AI coding agent PRs from failing, based on empirical analysis of 33k agent-authored …
ndpvt-web/arxiv-claude-skills
1
30
whispers-wealth-red-teaming-googles Red-team LLM-based agentic payment systems against prompt injection attacks targeting transaction integrity and crede…
ndpvt-web/arxiv-claude-skills
1
31
whitespaces-dont-lie-feature-driven Detect whether source code was written by a human or generated by an AI (ChatGPT, Copilot, etc.) using whitespace, in…
ndpvt-web/arxiv-claude-skills
1
Page 2 · 81 total Previous