LLM Mart Basic

@llm-mart · Joined Jun 2026

0 Followers 0 Reputation 12594 Contributions
Claude Skill model-based-testing

Use this skill when you need to derive test-path candidates from sourced behavior, state, or process models; triggers include 基于模型的测试 and model-based test design.

0
Claude Skill multi-agent-testing

Use this skill when you need evidence-bounded delegation, coordination, shared state, conflicts, ownership, termination, and traceability; triggers include 多 Agent 协作 and multi-agent coordination.

0
Claude Skill mutation-testing-analysis

Use this skill when you need to interpret mutation operators, killed and survived mutants, and evidence limits; triggers include 变异测试分析 and mutation testing analysis.

0
Claude Skill negative-scenario-discovery

Use this skill when you need to discover invalid, denied, failed, degraded, or unsafe-recovery scenarios from product evidence; triggers include negative scenario discovery.

0
Claude Skill observability-design-review

Use this skill when logging, metrics, tracing, alerting, or SLO design needs an evidence-bounded review before implementation; triggers include observability design review, telemetry readiness review, and alert actionability audit.

0
Claude Skill pairwise-testing

Use this skill when you need to identify interactions that need at least pairwise coverage after factors, values, and constraints are explicit; triggers include 成对测试 and pairwise test design.

0
Claude Skill performance-bottleneck-analysis

Use this skill when you need to form evidence-based performance bottleneck hypotheses and validation steps; triggers include performance bottleneck analysis.

0
Claude Skill performance-regression-analysis

Use this skill when you need to compare performance evidence across versions and assess regression risk; triggers include performance regression analysis.

0
Claude Skill performance-result-analysis

Use this skill when you need to interpret performance results, evidence quality, and risk without inventing conclusions; triggers include performance result analysis.

0
Claude Skill performance-test-gatling

Use this skill when you need Gatling performance scope, simulations, or runnable entry points; triggers include Gatling, Gatling simulations, and Gatling performance testing.

0
Claude Skill performance-test-jmeter

Use this skill when you need to design JMeter test plans with Thread Groups, samplers, data sets, assertions, timers, CLI runs, and HTML reports; triggers include JMeter performance testing, performance testing, and performance-test-jmeter.

0
Claude Skill performance-workload-modeling

Use this skill when you need to model realistic performance workload, traffic, and acceptance assumptions; triggers include performance workload modeling.

0
Claude Skill pr-test-impact-analysis

Use this skill when you need to determine test impact from a pull request or code diff; triggers include PR test impact analysis.

0
Claude Skill production-incident-analysis

Use this skill when you need to analyze production-incident evidence, impact, and follow-up actions; triggers include production incident analysis.

0
Claude Skill production-verification

Use this skill when you need to plan or assess evidence-based production verification after a release; triggers include production verification.

0
Claude Skill prompt-injection-testing

Use this skill when you need to design safe prompt-injection tests for AI systems and tool boundaries; triggers include prompt injection testing.

0
Claude Skill prompt-testing

Use this skill when you need to test prompt behavior, regression risk, and output boundaries across versions; triggers include prompt testing and prompt-regression.

0
Claude Skill property-based-testing

Use this skill when you need to turn invariants, generation domains, and shrinking strategies into reviewable property-test candidates; triggers include 基于属性的测试 and property-based test design.

0
Claude Skill quality-dashboard-design

Use this skill when you need evidence-bounded quality dashboard audiences, decision questions, panels, drill-downs, freshness, and alert boundaries; triggers include 质量仪表盘 and quality dashboard.

0
Claude Skill quality-debt-analysis

Use this skill when you need evidence-bounded quality-debt items, origins, impact, age, priority, ownership, and paydown tradeoffs; triggers include 质量债务 and quality debt.

0
/lineage-discovery Lineage discovery

Discover testnet↔mainnet subnet lineage from repo configs and open a PR for review (pass --dry-run to report only)

0
/capture capture

Triage raw inbox notes into reviewed repository destinations without deleting their sources.

0
/clean-ai-writing clean-ai-writing

Audit and rewrite content to remove AI writing patterns

0
/content-shipped content-shipped

Log a completed piece of content to content/log.md after the user confirms it was published.

0
/dream-apply dream-apply

Validate a dream artifact, review each proposal, and apply only individually accepted changes.

0
/dream dream

Run a curator pass against the validated memory directory and produce a proposal artifact.

0
/end end

End a session — log what happened, update state and the decision log, propose memory updates, and check for uncommitted or unpushed work

0
/find-context find-context

Find relevant context files by topic. Use when you need to load files for a topic without a slash command, or when a task spans multiple domains.

0
/migrate-gemini migrate-gemini

Inventory and migrate selected Gemini CLI workflows with dry-run review and parity checks.

0
/mine-gemini-workflows mine-gemini-workflows

Find repeated workflows in selected Gemini CLI sessions and draft portable skills after review.

0
/reconcile reconcile

Scan multi-session drift and offer individually reviewed fixes only after explicit approval.

0
/recover recover

Scan orphaned worktrees and stale branches, then offer explicit approval-gated cleanup.

0
/setup setup

Guided onboarding or import for durable workspace context

0
/start start

Start a session — load state files, flag staleness, and give a briefing on current priorities, deadlines, and blockers

0
/today today

Create a morning heartbeat from repository state and update the local heartbeat log.

0
/update update

Mid-session checkpoint — append progress to today's session log and update state files if a priority shifted, without ending the session

0
/distribution-audit distribution-audit

Maintainer-only. Find every file that would newly ship to adopters, classify each one against the written distribution-boundary categories, default to withhold on no clean match, and ask the maintainer only where the taxonomy does not settle it. Drives the release CLI, which refuses to produce a manifest until every shipping file has an answer.

0
/gaia-audit gaia-audit

Audit memory, wiki, and auto-loaded files for duplication, conflicting instructions, and stale content. The default path researches, then asks you a single Apply / Discuss / Decline question; on Apply it applies the report, files any out-of-scope problem as a tech-debt issue, then commits, opens a PR, and merges it on a main-branch run like /update-deps. Pass --apply to re-run the apply-and-publish stage against the most recent report.

0
/gaia-debt gaia-debt

Fix the tech-debt backlog, a single issue or a recommended related batch, highest severity then oldest first, on a fresh isolated branch through the audit gate, closing the issue(s) on merge. Pass `list` to see the ordered backlog, `why <issue-number>` to explain the recommendation, or a bare `<issue-number>` to fix that issue directly.

0
/gaia-fitness gaia-fitness

Health-check and auto-heal this project's Claude integration, triage, heal, verify, and report an F-to-A+ grade.

0
Atom

Atom Agent, Open-Source Governed AI Agent Platform for Self-Hosted Automation

12 views 0 likes
Oh My Hermes

The agent engineering intelligence harness, optimized tools, memory system, subagents and mixture of models packages ⚚

16 views 0 likes
Sesori Apps Monorepo

Sesori iOS/Android app and the Sesori Bridge CLI — drive Claude, Codex, OpenCode, Cursor, Pi, OMP, Hermes coding sessions from your phone

14 views 0 likes
Repobrain

🧠 RepoBrain (formerly Antigravity) — Give your repo a brain. ChatGPT for your codebase: works in Claude Code, Cursor, Codex, Windsurf & more.

14 views 0 likes
Dsh Plugin Subscriptions

Use ChatGPT (Codex), Claude, and Grok (X Premium) subscriptions as DeepSeek Harness LLM providers — OAuth login in the web UI, no API keys

15 views 0 likes
Latitude Llm

Latitude is the open-source AI monitoring platform.

16 views 0 likes
Webbrain

Open-source AI browser agent for Chrome and Firefox (monorepo) 🧠

16 views 0 likes
Cli

Official Model Studio CLI(阿里云百炼 CLI)built for AI Agent frameworks, exposing models, search, multimodal, and workflow capabilities as structured tool calls.

15 views 0 likes
Awesome DeepSeek Harness Plugins

Curated DeepSeek Harness (DSH) plugins, extensions, tools, skills, clients, runtimes, integrations, and verified references — English and Chinese.

16 views 0 likes
Slides Maker

Turn papers, code, and docs into presentation-ready, natively editable PPTX in Codex / Claude Code. Native charts and equations, speaker notes, click-build anim…

14 views 0 likes
Grix

Grix : Work with agents like talking to people.

25 views 0 likes
Atomic Agent

Local First Ai Agent. Optimized for Local Ai models. Long context window. Proper tools callings. Runs privately on your device.

26 views 0 likes
Agentql Mcp

Model Context Protocol server that integrates AgentQL's data extraction capabilities.

15 views 0 likes
Pilot

AI that ships your tickets.

14 views 0 likes
Octo Cli

Metadata-driven CLI for AI Agent Bots — 48 operations across 7 domains, structured JSON envelope I/O, zero interactive prompts.

12 views 0 likes
Nextclaw

An open-source, extensible, self-hosted agent workspace with multi-runtime support for Codex, Claude Code, and more, plus reusable local apps for custom interfa…

14 views 0 likes
Deepseek Harness Desktop

Open-source Windows desktop client and GUI for DeepSeek Harness — zero-setup installer with Codex, plugins, skills, SSH, mobile remote access, and 11 skins.

12 views 0 likes
Openscience

The open-source AI workbench for scientific research

15 views 0 likes
DeepSeek Reasonix

DeepSeek-native AI coding agent for your terminal. Engineered around prefix-cache stability — leave it running.

15 views 0 likes
Bladebro

A Fully free agentic browser driver for AI , few tools, full control, real stealth, top-tier token efficiency.

12 views 0 likes