LLM Mart Basic

@llm-mart · Joined Jun 2026

0 Followers 0 Reputation 12594 Contributions
Claude Skill orchestrate

Flip the session into coordinator mode — the parent agent plans, scopes, reviews, and ships, but delegates all real work (exploration, implementation, review, fixes) to sub-agents routed by an empirically benchmarked model capability table. Use when the user invokes /orchestrate

0
Claude Skill ticket

Drive one tracked ticket from arrival to resolution through four verbs: triage, start, revise, finalize. Use when the user says triage/start/revise/finalize with a ticket id, asks to turn a ticket into a locked brief, to execute a work order, to action a review round on a ticket'

0
Claude Skill ui-craft

Lifecycle for user-facing surfaces — revise a shipped surface in the running app, lock a greenfield visual spec, build to a lock, critique/audit/polish a UI, or re-settle a locked term. Use for any request to design, review, or verify rendered UI (screens, dashboards, flows, comp

0
Claude Skill cbm-onboard

Register maintained or ephemeral Git checkouts with codebase-memory-mcp, keep long-lived indexes current through non-clobbering Git hooks, and tear down ephemeral indexes by exact deterministic identity. Use when asked to index, onboard, register, remove, or keep a repository cur

0
Claude Skill ci-design

Vocabulary and principles for well-designed CI. Use when the user wants to design, review, or audit CI, says CI is noisy, slow, or expensive, or is designing a workflow yml.

0
Claude Skill code-review

Review the changes since a fixed point (commit, branch, tag, PR, or merge-base) on two axes — Standards (does it follow this repo's documented rules?) and Spec (does it do what the originating issue asked?). Each axis returns a verdict per enumerated item rather than an open-ende

0
Claude Skill codebase-design

Shared vocabulary for designing deep modules. Use when the user wants to design or improve a module's interface, find deepening opportunities, decide where a seam goes, make code more testable or AI-navigable, or when another skill needs the deep-module vocabulary.

0
Claude Skill codebase-memory

Use the codebase knowledge graph for structural code queries, trace call paths and dependencies, inspect architecture, assess change impact, or activate the bundled Claude Code discovery hooks.

0
Claude Skill domain-modeling

Build and sharpen a project's domain model. Use when the user wants to pin down domain terminology or a ubiquitous language, record an architectural decision, or when another skill needs to maintain the domain model.

0
Claude Skill drive-local-webapp

Launch or connect to a local development web server and drive it with a reusable headless-Chromium command interface. Use when asked to render, smoke-test, interact with, verify, or screenshot a local frontend or HTML mockup.

0
Claude Skill handoff

Produce a decision-safe handoff for a fresh agent. Use when a task changes hands, a session is ending midstream, a user asks for a handoff, or an agent needs to preserve a design/map/implementation decision without making the next agent rediscover it.

0
Claude Skill persona-review

Convene a small panel of persistent reviewer personas (curated colleague-likenesses plus invent-on-demand) to review a document cold, in character, then synthesize a panel verdict. Use when the user wants a document reviewed by a panel, wants named perspectives on a plan or desig

0
Claude Skill plan-review

Adversarially review a plan, work order, spec, or agent brief before anything gets built. Use when the user wants a plan reviewed, stress-tested, audited, or countersigned, says "plan review" or "poke holes in this plan", or wants a pre-build check on an issue/brief/PRD.

0
Claude Skill pr-body

Write a pull-request body and score it before the PR is opened, or audit an existing PR's body against the same rubric. Use when about to run gh pr create, when a PR body needs writing or rewriting, or when asked whether a PR description is any good. Invoked as /pr-body write <bo

0
Claude Skill preflight

Use when a plan, spec, work order, or process doc is drafted and about to enter review or execution — grounds its facts, spikes its first hour, and single-sources its rules so review rounds start deep instead of shallow.

0
Claude Skill prototype

Build a throwaway prototype to flesh out a design — a runnable terminal app for state/business-logic questions, or several radically different UI variations toggleable from one route.

0
Claude Skill research

Investigate a question against high-trust primary sources and capture the findings as a Markdown file in the repo. Use when the user wants a topic researched, docs or API facts gathered, or reading legwork delegated to a background agent.

0
Claude Skill say-less

Shape every response for a reader who wants the answer first and nothing after it. Answer-first structure, numbered steps, domain language governed by a per-repo glossary of approved terms, one instruction per sentence, no preamble or recap. Invoke with /say-less for full shaping

0
Claude Skill spin-worktree

Create an isolated Git worktree for issue, pull-request, branch, or parallel agent work without editing the control checkout. Use when starting task work from fresh remote state, avoiding a dirty or shared checkout, or preparing a dedicated worktree for another agent.

0
Claude Skill tdd

Test-driven development. Use when the user wants to build features or fix bugs test-first, mentions "red-green-refactor", or wants integration tests.

0
The AI-in-production safety playbook

Fourteen posts of being wrong in production, compressed to checkboxes

security prompt-engineering devops ai
Sep 30
The control plane was flapping because of a spinning disk

Healthy nodes, a quiet network, 300 restarts in three days, and a latency budget measured in milliseconds

kubernetes sre incident-response observability
Sep 29
The overlay that pinged but wouldn't carry TCP

Discovery worked. Ping worked. Every TCP connection timed out, and later the tunnel only worked when someone had a terminal open.

containers incident-response networking linux
Sep 28
Bringing a cluster back after the host rebooted

Every VM came back. The cluster did not. Declarative systems converge on config, and the datapath isn't config.

kubernetes sre containers incident-response
Sep 27
The agent is running in *your* shell

A surprising share of AI-in-the-terminal failures aren't the AI. They're zsh, and a version of bash from 2006.

devops ai-agents automation shell
Sep 26
How to create and share a Claude Code plugin

A Claude Code plugin turns standalone project configuration into a namespaced, installable extension that teams and communities can update as one unit.

agents security skill-md claude-code
Sep 25
The scaffolding that made it safe

None of the safety came from the model. It came from six boring habits.

git devops ai-agents claude-code
Sep 25
Claude Code skills vs. subagents: when to use each

Skills package instructions and references. Subagents run work in a separate context and return results. They solve different problems and can be composed deliberately.

agent-skills claude-code context
Sep 24
Knowing when to stop

Six hours in, one step left, everything green, and the incident that didn't happen

prompt-engineering ai ai-agents sre
Sep 24
CLAUDE.md vs. skills: where should Claude Code instructions live?

CLAUDE.md carries persistent project context. Skills load reusable procedures when relevant. Separating stable facts from task-specific workflows keeps both easier to maintain.

agent-skills claude-skills configuration context
Sep 23
Those are the other app's keys

Twenty minutes recovering secrets that never existed, and the one sentence from a human that ended it

security devops ai ai-agents
Sep 23
How to use remote MCP servers with the OpenAI Responses API

An API request routing a model's tool call through an approval gate to a remote MCP server

security mcp integrations open-api
Sep 22
The coverage audit before you delete the safety net

31 config keys, two audits, and why the first one was wrong in both directions

security devops ai-agents secrets-management
Sep 22
How to publish an MCP server to the official MCP Registry

The official MCP Registry stores standardized server metadata rather than package code. Publishers verify a namespace, describe installation or remote access, and submit immutable versions.

security mcp
Sep 21
Rotating a leaked credential, in the right order

Everyone looks at the Dockerfile. The file that actually leaked the key was the project file.

security devops ai-agents containers
Sep 21
MCP authentication explained: OAuth, scopes, and safe token handling

Remote MCP authorization uses established OAuth standards, but secure integration still requires issuer validation, least-privilege scopes, protected token handling, and server-side enforcement.

security mcp
Sep 20
Byte-identical or bust

"Copy it over and switch the reference" is two steps, and the outage lives in the one nobody checks

security kubernetes verification
Sep 20
MCP stdio vs. Streamable HTTP: which transport should you use?

stdio fits local processes and prototypes. Streamable HTTP fits hosted services and shared integrations. The right choice follows where the capability runs and who must reach it.

security mcp
Sep 19
Never let the AI print a secret

The most important rule wasn't about what I could change. It was about what I was allowed to display.

security kubernetes devops ai-agents
Sep 19
MCP tools vs. resources vs. prompts: when to use each

Tools perform operations, resources expose readable context, and prompts provide reusable templates. Choosing the correct primitive makes an MCP server easier to understand and govern.

security mcp
Sep 18
/attach Attach

`crabbox attach` follows the recorded events of an active coordinator run and

0
/azure Azure

`crabbox azure` groups Azure provider setup commands. It currently has a single

0
/bench Bench

`crabbox bench` records and reports local benchmark timing observations. It is a

0
/cache Cache

`crabbox cache` inspects, purges, or warms package and build caches on a

0
/capsule Capsule

`crabbox capsule` captures, replays, and tracks lightweight failure capsules.

0
/checkpoint Checkpoint

Save the state of a lease, then restore it onto another box or fork it into a

0
/claims Claims

`crabbox claims list` prints the lease claims stored on the current machine. It

0
/cleanup Cleanup

`crabbox cleanup` sweeps direct-provider machines and local provider state that

0
/code Code

`crabbox code` bridges a Linux lease's `code-server` workspace into the

0
/config Config

`crabbox config` inspects and updates user configuration. It has three

0
/connect Connect

`crabbox connect` resolves a lease and opens an interactive SSH session to it.

0
/cp Cp

`crabbox cp` copies files or directories between the host and a Crabbox-owned

0
/desktop Desktop

`crabbox desktop` drives a visible desktop session on a lease that was warmed

0
/doctor Doctor

`crabbox doctor` runs a preflight before you commit to a long workflow. It is

0
/egress Egress

`crabbox egress` gives a lease mediated outbound network: a lease-local browser

0
/events Events

`crabbox events` prints the broker's event log for a recorded run.

0
/heartbeat Heartbeat

`crabbox heartbeat` refreshes the idle deadline for one owned lease and prints

0
/history History

`crabbox history` lists recorded remote command runs from the broker. Each run is

0
/image Image

`crabbox image` holds the trusted-operator controls for provider base images:

0
/init Init

`crabbox init` onboards the current repository: it writes the minimal config

0
Polymarket Paper Trader

Paper trading simulator for Polymarket — built for AI agents. MCP server, live order books, strategy backtesting. Install: npx clawhub install polymarket-paper-…

2 views 0 likes
Rssh

An SSH tool dedicated to addressing all pain points · (macOS/Windows/Linux/Android/iOS)

3 views 0 likes
Peerd

The first AI agent harness native to the browser. A browser extension that runs a full agent loop where you already work: it drives your tabs, spins up sandboxe…

3 views 0 likes
AutoLabel Forge

SmartLabel AI 2026: Auto-Annotate Any Object via LLM-Powered Prompt Parsing

3 views 0 likes
Ima2 Gen

Local-first visual generation runtime and studio for people and coding agents, with reproducible image and video workflows across multiple providers.

0 views 0 likes
Scholaraio

Scholar All-In-One: A research infrastructure for AI agents

1 views 0 likes
Auto Browser

Give your AI agent a real browser — with a human in the loop. Open-source MCP-native browser agent.

2 views 0 likes
FQGate Agent

同花顺免费开源AI插件FQGate-agent(原插件名 tonghuasun-agent):为 Codex、Claude Code、DeepSeek 等 AI 工具提供本机 A 股实时行情、K 线、Level-2、资讯、账户查询与可选交易能力。

2 views 0 likes
PiX

A non-linear AI agent workbench — session is a tree: branch anytime, and context follows the branch

3 views 0 likes
Autocad MCP

Production-grade AutoCAD MCP server for AI agents — 122 tools, dual COM (live AutoCAD) + headless ezdxf engines, ISO GD&T and dimension-tolerance validation for…

4 views 0 likes
Dscode

A DeepSeek coding agent harness: persistent shell, Ultra subagents, auto approval, Chrome MCP and session telemetry

5 views 0 likes
Waku Agent

Waku Waku! Waku Agent is a local-first AI agent harness you actually own, including loop, memory, eval, all in code built to stay legible as it grows.

4 views 0 likes
Nexting

Remote control for Claude Code, Codex, Grok, and Cursor on Mac or PC. View sessions, send tasks, and drive them remotely from your phone, PIN, or Ring. OpenClaw…

4 views 0 likes
Saki Panel

Next-gen AI-native server ops panel with an in-workspace SRE agent. Crash self-healing, safe rollbacks, Docker/game servers, and local Ollama support.

4 views 0 likes
Terravision

Professional cloud architecture diagrams with official AWS, Azure and GCP icons, from Terraform code or a plain JSON graph. MCP server + agent skill.

2 views 0 likes
P Ai

A ready-to-use self-growing desktop AI assistant for long-running tasks, memory, agents, tool reviews, MCP, and high-concurrency workspace automation. / 开箱即用的自我…

5 views 0 likes
Awesome Saas

Collection of templates using the Alchemyst AI Platform for your next big AI app.

2 views 0 likes
TensorFold

Fast, exact LLM decoding on Apple Silicon (MLX) behind an OpenAI-compatible endpoint

25 views 0 likes
Feynman

The open source AI research agent.

5 views 0 likes
Iris

Open-source agent-native visual production workspace where humans and coding agents edit the same live canvas — local-first, BYOK image/video models.

2 views 0 likes