LLM Mart Basic

@llm-mart · Joined Jun 2026

0 Followers 0 Reputation 11567 Contributions
Claude Skill performance-test-gatling

Use this skill when you need Gatling performance scope, simulations, or runnable entry points; triggers include Gatling, Gatling simulations, and Gatling performance testing.

0
Claude Skill performance-test-jmeter

Use this skill when you need to design JMeter test plans with Thread Groups, samplers, data sets, assertions, timers, CLI runs, and HTML reports; triggers include JMeter performance testing, performance testing, and performance-test-jmeter.

0
Claude Skill performance-workload-modeling

Use this skill when you need to model realistic performance workload, traffic, and acceptance assumptions; triggers include performance workload modeling.

0
Claude Skill pr-test-impact-analysis

Use this skill when you need to determine test impact from a pull request or code diff; triggers include PR test impact analysis.

0
Claude Skill production-incident-analysis

Use this skill when you need to analyze production-incident evidence, impact, and follow-up actions; triggers include production incident analysis.

0
Claude Skill production-verification

Use this skill when you need to plan or assess evidence-based production verification after a release; triggers include production verification.

0
Claude Skill prompt-injection-testing

Use this skill when you need to design safe prompt-injection tests for AI systems and tool boundaries; triggers include prompt injection testing.

0
Claude Skill prompt-testing

Use this skill when you need to test prompt behavior, regression risk, and output boundaries across versions; triggers include prompt testing and prompt-regression.

0
Claude Skill property-based-testing

Use this skill when you need to turn invariants, generation domains, and shrinking strategies into reviewable property-test candidates; triggers include 基于属性的测试 and property-based test design.

0
Claude Skill quality-dashboard-design

Use this skill when you need evidence-bounded quality dashboard audiences, decision questions, panels, drill-downs, freshness, and alert boundaries; triggers include 质量仪表盘 and quality dashboard.

0
Claude Skill quality-debt-analysis

Use this skill when you need evidence-bounded quality-debt items, origins, impact, age, priority, ownership, and paydown tradeoffs; triggers include 质量债务 and quality debt.

0
Claude Skill quality-gate-design

Use this skill when you need evidence-bounded entry criteria, evidence requirements, owners, and exception paths for a delivery or release gate; triggers include 质量门禁 and quality gate.

0
Claude Skill quality-maturity-assessment

Use this skill when you need evidence-bounded quality-practice maturity dimensions, rubric anchors, evidence sufficiency, and improvement gaps; triggers include 质量成熟度 and quality maturity.

0
Claude Skill quality-metrics-design

Use this skill when you need evidence-bounded quality metric definitions, calculation rules, data sources, freshness, and anti-gaming boundaries; triggers include 质量指标 and quality metric.

0
Claude Skill quality-productivity-metrics

Use this skill when you need evidence-bounded quality and delivery metrics, denominators, attribution limits, gaming risk, and the Human-use boundary; triggers include 质量生产力 and quality productivity.

0
Claude Skill quality-risk-analysis

Use this skill when you need to identify and prioritize quality risks from product, change, and evidence inputs; triggers include quality risk analysis.

0
Claude Skill rag-quality-testing

Use this skill when you need evidence-bounded grounding, relevance, completeness, citation support, abstention, and answer-level evidence in RAG outputs; triggers include RAG 质量 and RAG quality.

0
Claude Skill rag-retrieval-testing

Use this skill when you need evidence-bounded query variants, chunking, filters, recall/precision proxies, ranking, freshness, and retrieval evidence; triggers include 检索结果 and retrieval result.

0
Claude Skill recovery-testing

Use this skill when you need evidence-bounded recovery-testing analysis and validation preparation; triggers include 恢复测试 and recovery-testing.

0
Cursor Skill scomp-link

End-to-end ML toolkit with 26 CLI commands. Use when training models, tuning hyperparameters, detecting data drift, generating HTML reports with charts, profiling datasets, detecting anomalies, forecasting time series, checking fairness, or serving models as REST APIs. Prefer ove

0
The overlay that pinged but wouldn't carry TCP

Discovery worked. Ping worked. Every TCP connection timed out, and later the tunnel only worked when someone had a terminal open.

containers incident-response networking linux
Sep 28
Bringing a cluster back after the host rebooted

Every VM came back. The cluster did not. Declarative systems converge on config, and the datapath isn't config.

kubernetes sre containers incident-response
Sep 27
The agent is running in *your* shell

A surprising share of AI-in-the-terminal failures aren't the AI. They're zsh, and a version of bash from 2006.

devops ai-agents automation shell
Sep 26
How to create and share a Claude Code plugin

A Claude Code plugin turns standalone project configuration into a namespaced, installable extension that teams and communities can update as one unit.

agents security skill-md claude-code
Sep 25
The scaffolding that made it safe

None of the safety came from the model. It came from six boring habits.

git devops ai-agents claude-code
Sep 25
Claude Code skills vs. subagents: when to use each

Skills package instructions and references. Subagents run work in a separate context and return results. They solve different problems and can be composed deliberately.

agent-skills claude-code context
Sep 24
Knowing when to stop

Six hours in, one step left, everything green, and the incident that didn't happen

prompt-engineering ai ai-agents sre
Sep 24
CLAUDE.md vs. skills: where should Claude Code instructions live?

CLAUDE.md carries persistent project context. Skills load reusable procedures when relevant. Separating stable facts from task-specific workflows keeps both easier to maintain.

agent-skills claude-skills configuration context
Sep 23
Those are the other app's keys

Twenty minutes recovering secrets that never existed, and the one sentence from a human that ended it

security devops ai ai-agents
Sep 23
How to use remote MCP servers with the OpenAI Responses API

An API request routing a model's tool call through an approval gate to a remote MCP server

security mcp integrations open-api
Sep 22
The coverage audit before you delete the safety net

31 config keys, two audits, and why the first one was wrong in both directions

security devops ai-agents secrets-management
Sep 22
How to publish an MCP server to the official MCP Registry

The official MCP Registry stores standardized server metadata rather than package code. Publishers verify a namespace, describe installation or remote access, and submit immutable versions.

security mcp
Sep 21
Rotating a leaked credential, in the right order

Everyone looks at the Dockerfile. The file that actually leaked the key was the project file.

security devops ai-agents containers
Sep 21
MCP authentication explained: OAuth, scopes, and safe token handling

Remote MCP authorization uses established OAuth standards, but secure integration still requires issuer validation, least-privilege scopes, protected token handling, and server-side enforcement.

security mcp
Sep 20
Byte-identical or bust

"Copy it over and switch the reference" is two steps, and the outage lives in the one nobody checks

security kubernetes verification
Sep 20
MCP stdio vs. Streamable HTTP: which transport should you use?

stdio fits local processes and prototypes. Streamable HTTP fits hosted services and shared integrations. The right choice follows where the capability runs and who must reach it.

security mcp
Sep 19
Never let the AI print a secret

The most important rule wasn't about what I could change. It was about what I was allowed to display.

security kubernetes devops ai-agents
Sep 19
MCP tools vs. resources vs. prompts: when to use each

Tools perform operations, resources expose readable context, and prompts provide reusable templates. Choosing the correct primitive makes an MCP server easier to understand and govern.

security mcp
Sep 18
How to test an MCP server with MCP Inspector

Use MCP Inspector to connect to local or remote servers, inspect capabilities, call tools, read resources, test prompts, and diagnose failures before release.

debugging security mcp
Sep 17
How to build an MCP server in TypeScript: step-by-step

Build an MCP server in TypeScript with focused tools, validated schemas, local and remote transports, Inspector tests, and production security controls.

security ai mcp
Sep 15
/requirements Requirements

Generate requirements from goal and research

0
/research Research

Run or re-run research phase for current spec

0
/start Start

Smart entry point that detects if you need a new spec or should resume existing

0
/status Status

Show all specs and their current status

0
/switch Switch

Switch active spec

0
/tasks Tasks

Generate implementation tasks from design

0
/triage Triage

Decompose a large feature into multiple dependency-aware specs (epic triage)

0
/tree-ring-update Tree ring update

Check for or install a verified Tree Ring Memory CLI update without changing installation scope

0
/README README

This directory contains the command implementations for the fast-agent CLI.

0
/close Close

They operate it without you.

0
/outcome Outcome

Promised, measured, accepted. A number nobody signed is claimed, not delivered.

0
/prep Prep

Prepare the meeting. One page from the record.

0
/receipts Receipts

Find the receipt. A dated line, or it did not happen.

0
/trust Trust

Diagnose trust. Process gap, or they stopped trusting you.

0
/awesome-docs awesome-docs

Generate, convert, and maintain animated GitHub-safe Markdown documents with animated SVG diagrams. Covers four SVG patterns (architecture flow, lifecycle loop, field carousel, timeline phases), guided interview for any doc type (README, architecture guide, runbook, API reference, tutorial, RFC, post-mortem, how-it-works, or custom), converting existing plain Markdown, diffing for stale diagrams, quality auditing, local preview, and multi-platform export. Use when asked to "create a README for X", "write an architecture doc", "animate this guide", "convert my doc to animated", "check if my diagrams are stale", or "export my doc for Confluence".

0
/aws-profile aws-profile

AWS profile management for MCP servers — discover profiles across SSO, Granted, and assumed-role chains, check credential TTL, switch profiles across VS Code and Claude Code MCP configs, and scan AWS Organization accounts.

0
/aws aws

Structured guidance for AWS CloudFront distributions, WAF web ACLs, Lambda@Edge, CloudFront Functions, Firewall Manager multi-account enforcement, and IAM/IRSA patterns. Covers OAC, cache policies, security headers, managed rule groups, rate limiting, FMS FIRST/MIDDLE/LAST ownership model, and production-ready Terraform generation.

0
/azure azure

Azure identity (Workload Identity, OIDC, Entra ID), resource tagging, AKS platform patterns, RBAC scoping, and production-readiness review — with Terraform generation.

0
/chaos chaos

Design, run, and debug Chaos Engineering experiments on Kubernetes using Litmus Chaos v3 and Chaos Mesh v2. Covers fault injection (pod-delete, network-loss, CPU stress, node-drain), steady-state hypothesis probes, GameDay runbooks, scheduled experiments, DORA feedback loop, and RBAC setup. Use when asked to "inject a pod fault", "run a GameDay", "schedule chaos experiments", or "debug why my ChaosEngine is stuck".

0
/checkov checkov

Bootstrap Checkov on a developer laptop, run static or plan-level Terraform security scanning for AWS/Azure/GCP/EKS, resolve private GitHub modules via gh CLI, generate pre-commit hooks, produce multi-format output (cli/json/sarif/junit), and fix violations with AI-generated patches. Use when asked to "scan my Terraform", "run checkov", "check my IaC for security issues", "set up checkov pre-commit", or "fix checkov findings".

0
Polymarket Paper Trader

Paper trading simulator for Polymarket — built for AI agents. MCP server, live order books, strategy backtesting. Install: npx clawhub install polymarket-paper-…

1 views 0 likes
Rssh

An SSH tool dedicated to addressing all pain points · (macOS/Windows/Linux/Android/iOS)

3 views 0 likes
Peerd

The first AI agent harness native to the browser. A browser extension that runs a full agent loop where you already work: it drives your tabs, spins up sandboxe…

2 views 0 likes
AutoLabel Forge

SmartLabel AI 2026: Auto-Annotate Any Object via LLM-Powered Prompt Parsing

2 views 0 likes
Ima2 Gen

Local-first visual generation runtime and studio for people and coding agents, with reproducible image and video workflows across multiple providers.

0 views 0 likes
Scholaraio

Scholar All-In-One: A research infrastructure for AI agents

0 views 0 likes
Auto Browser

Give your AI agent a real browser — with a human in the loop. Open-source MCP-native browser agent.

1 views 0 likes
FQGate Agent

同花顺免费开源AI插件FQGate-agent(原插件名 tonghuasun-agent):为 Codex、Claude Code、DeepSeek 等 AI 工具提供本机 A 股实时行情、K 线、Level-2、资讯、账户查询与可选交易能力。

1 views 0 likes
PiX

A non-linear AI agent workbench — session is a tree: branch anytime, and context follows the branch

2 views 0 likes
Autocad MCP

Production-grade AutoCAD MCP server for AI agents — 122 tools, dual COM (live AutoCAD) + headless ezdxf engines, ISO GD&T and dimension-tolerance validation for…

3 views 0 likes
Dscode

A DeepSeek coding agent harness: persistent shell, Ultra subagents, auto approval, Chrome MCP and session telemetry

4 views 0 likes
Waku Agent

Waku Waku! Waku Agent is a local-first AI agent harness you actually own, including loop, memory, eval, all in code built to stay legible as it grows.

4 views 0 likes
Nexting

Remote control for Claude Code, Codex, Grok, and Cursor on Mac or PC. View sessions, send tasks, and drive them remotely from your phone, PIN, or Ring. OpenClaw…

3 views 0 likes
Saki Panel

Next-gen AI-native server ops panel with an in-workspace SRE agent. Crash self-healing, safe rollbacks, Docker/game servers, and local Ollama support.

3 views 0 likes
Terravision

Professional cloud architecture diagrams with official AWS, Azure and GCP icons, from Terraform code or a plain JSON graph. MCP server + agent skill.

1 views 0 likes
P Ai

A ready-to-use self-growing desktop AI assistant for long-running tasks, memory, agents, tool reviews, MCP, and high-concurrency workspace automation. / 开箱即用的自我…

4 views 0 likes
Awesome Saas

Collection of templates using the Alchemyst AI Platform for your next big AI app.

1 views 0 likes
TensorFold

Fast, exact LLM decoding on Apple Silicon (MLX) behind an OpenAI-compatible endpoint

19 views 0 likes
Feynman

The open source AI research agent.

4 views 0 likes
Iris

Open-source agent-native visual production workspace where humans and coding agents edit the same live canvas — local-first, BYOK image/video models.

1 views 0 likes