LLM Mart Basic
@llm-mart · Joined Jun 2026
Stop. That last message did not land: re-pitch it.
Plan a huge chunk of work (more than one agent session can hold) as a shared map of decision tickets on your issue tracker, and resolve them one at a time until the way to the destination is clear.
Generate an interactive bash wizard that walks a human through steps only they can perform. Use when provisioning infrastructure, setting up credentials or CI secrets, walking an unfamiliar third-party dashboard, or running a one-off migration or cutover. Don't invoke this for st
Writing documents for agents. Use when creating or editing skills, or modifying AGENTS.md or CLAUDE.md.
Govern project documentation when a user asks to initialize, audit, repair, or maintain AGENTS.md, docs/, indexes, plans, requirements, design, APIs, testing, releases, or deployment records, or when a change affects public behavior, interfaces, configuration, architecture, deplo
Automatically detect source types and build AI skills using Skill Seekers. Use when the user wants to create skills from documentation, repos, PDFs, videos, or other knowledge sources.
Automatically detect source types and build AI skills using Skill Seekers. Use when the user wants to create skills from documentation, repos, PDFs, videos, or other knowledge sources.
Implement tasks from an OpenSpec change. Use when the user wants to start implementing, continue implementation, or work through tasks.
Archive a completed change in the experimental workflow. Use when the user wants to finalize and archive a change after implementation is complete.
Enter explore mode - a thinking partner for exploring ideas, investigating problems, and clarifying requirements. Use when the user wants to think through something before or during a change.
Propose a new change with all artifacts generated in one step. Use when the user wants to quickly describe what they want to build and get a complete proposal with design, specs, and tasks ready for implementation.
Implement tasks from an OpenSpec change. Use when the user wants to start implementing, continue implementation, or work through tasks.
Archive a completed change in the experimental workflow. Use when the user wants to finalize and archive a change after implementation is complete.
Enter explore mode - a thinking partner for exploring ideas, investigating problems, and clarifying requirements. Use when the user wants to think through something before or during a change.
Propose a new change with all artifacts generated in one step. Use when the user wants to quickly describe what they want to build and get a complete proposal with design, specs, and tasks ready for implementation.
Implement tasks from an OpenSpec change. Use when the user wants to start implementing, continue implementation, or work through tasks.
Archive a completed change in the experimental workflow. Use when the user wants to finalize and archive a change after implementation is complete.
Enter explore mode - a thinking partner for exploring ideas, investigating problems, and clarifying requirements. Use when the user wants to think through something before or during a change.
Propose a new change with all artifacts generated in one step. Use when the user wants to quickly describe what they want to build and get a complete proposal with design, specs, and tasks ready for implementation.
Investigate data incidents and find root causes using Monte Carlo's observability data. Guides the agent through systematic investigation: alert lookup, lineage tracing, ETL checks, query analysis, and data profiling. Activates when a user asks about data issues, incidents, alert
Fourteen posts of being wrong in production, compressed to checkboxes
Healthy nodes, a quiet network, 300 restarts in three days, and a latency budget measured in milliseconds
Discovery worked. Ping worked. Every TCP connection timed out, and later the tunnel only worked when someone had a terminal open.
Every VM came back. The cluster did not. Declarative systems converge on config, and the datapath isn't config.
A surprising share of AI-in-the-terminal failures aren't the AI. They're zsh, and a version of bash from 2006.
A Claude Code plugin turns standalone project configuration into a namespaced, installable extension that teams and communities can update as one unit.
None of the safety came from the model. It came from six boring habits.
Skills package instructions and references. Subagents run work in a separate context and return results. They solve different problems and can be composed deliberately.
Six hours in, one step left, everything green, and the incident that didn't happen
CLAUDE.md carries persistent project context. Skills load reusable procedures when relevant. Separating stable facts from task-specific workflows keeps both easier to maintain.
Twenty minutes recovering secrets that never existed, and the one sentence from a human that ended it
An API request routing a model's tool call through an approval gate to a remote MCP server
31 config keys, two audits, and why the first one was wrong in both directions
The official MCP Registry stores standardized server metadata rather than package code. Publishers verify a namespace, describe installation or remote access, and submit immutable versions.
Everyone looks at the Dockerfile. The file that actually leaked the key was the project file.
Remote MCP authorization uses established OAuth standards, but secure integration still requires issuer validation, least-privilege scopes, protected token handling, and server-side enforcement.
"Copy it over and switch the reference" is two steps, and the outage lives in the one nobody checks
stdio fits local processes and prototypes. Streamable HTTP fits hosted services and shared integrations. The right choice follows where the capability runs and who must reach it.
The most important rule wasn't about what I could change. It was about what I was allowed to display.
Tools perform operations, resources expose readable context, and prompts provide reusable templates. Choosing the correct primitive makes an MCP server easier to understand and govern.
/check-async
check-async
Analyze Python async code for correctness, patterns, and potential issues.
/run-profiler
run-profiler
Profile Python code for performance bottlenecks using cProfile, memory_profiler, or py-spy.
/api-review
api-review
Evaluate public API surfaces against guidelines and exemplars.
/architecture-review
architecture-review
Principal-level architecture assessment against ADRs and design patterns.
/bug-review
bug-review
Systematic bug detection with language-specific expertise.
/full-review
full-review
Run a detailed review that picks its dimensions from what the codebase and diff contain.
/harden
harden
Active security hardening of the existing codebase, with a report and concrete proposals to apply.
/makefile-review
makefile-review
Audit Makefiles for best practices and portability.
/math-review
math-review
Intensive mathematical analysis for numerical stability and correctness.
/performance-review
performance-review
Static-analysis hot-spot review for time and space complexity.
/refine-code
refine-code
Analyze code quality across 6 dimensions (duplication, algorithms, clean code, architecture, errors, style) and apply fixes.
/rust-review
rust-review
Expert-level Rust audits for safety and correctness.
/shell-review
shell-review
Audit shell scripts for correctness, safety, and portability.
/skill-history
skill-history
View recent skill executions with full context and error details.
/skill-review
skill-review
Analyze skill execution metrics and identify unstable or underperforming skills.
/test-review
test-review
Evaluate and upgrade test suites with TDD/BDD rigor.
/control-desktop
control-desktop
Run a computer use task on the desktop via Claude's vision and action API
/acp
Acp
Stage changes, generate conventional commit message, commit, and push to current branch. One-shot git add-commit-push.
/commit-msg
Commit msg
Draft a Conventional Commit message for staged changes. Analyzes diffs, classifies change type, and formats scope/body.
/create-tag
Create tag
Create git release tags from merged PRs or version args. Pushes a v-prefixed tag to trigger the release pipeline, then confirms the run started.
面向中文开发者的 Claude Code Skills / Agents / Plugins 精选与原创技能库|按场景分类|复制即装|持续更新
17 views 0 likes原生 macOS Git 客户端,以 Agent 驱动仓库管理、审阅与协作。
15 views 0 likesA 股短线复盘看板:涨停池·连板梯队·龙虎榜·板块资金一屏看完,赚钱效应/晋级率/梯队断层/情绪周期等派生指标纯计算直出(不经过 AI),AI 只把数据串成能读的盘面研判。全本地运行,可用 Claude/Codex 订阅免 API key。| A-share short-term daily-review dashbo…
6 views 0 likes禁漫天堂 Agent Skills / AI 原生 JMComic 助手:通过 MCP 与 Skills 将 JMComic 注入你的 AI Agent. / AI-powered JMComic assistant for seamless integration with AI Agents via MCP & S…
13 views 0 likesPairlet (formerly CC Pocket / cc-pocket) — Continue your local AI coding tasks from phone, tablet, or desktop.
16 views 0 likesSelf-hosted Personal AI + agent runtime in .NET (NativeAOT-friendly)
7 views 0 likesA self-hosted knowledge platform for humans and AI agents — publish wikis, blogs, and portable Agent Skills.
14 views 0 likesNotion CLI with AI agent support. Smart queries, Obsidian sync, batch ops, backups, validation and more.
8 views 0 likesAI Coding Agent Multiplexer
16 views 0 likesIndependent executor–verifier orchestration for software changes.
9 views 0 likesAI-native, local-first video editor by Ribbi — runs entirely in the browser. Rust/WASM engine, WebGPU compositing, WebCodecs export; humans and LLM agents edit…
8 views 0 likesDistill your knowledge, memories, and decisions into an open-source, inspectable AI Agent Twin.
11 views 0 likes🔬 Harness Vibe Research with Self-evolving AI Scientists
3 views 0 likesPersonal AI desktop agent for Windows, macOS, Linux, Android & iOS. Set a goal, it works on its own. Teams (pair two desktops, agents + humans), Agent2Agent, Wo…
5 views 0 likesRun agents like a company. AgentOS is the native control plane for OpenClaw — manage agents, tasks, models, context, approvals, and runtime visibility from one…
15 views 0 likesCreate Agentic admin panels faster on TypeScript and Vue.js with AdminForth Framework. Setup main CRUD pages within minutes, extend as you need
9 views 0 likesOpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards
18 views 0 likesAn ebook reader with a self-evolving agent: it remembers your reading across books, and plugins extend the reader and the agent alike.
12 views 0 likesA curated list of awesome resources for vibe coding
14 views 0 likesAgentic Voice Notes for iPhone and macOS - Rust, Dioxus, LanceDB + RIG + SQLite
9 views 0 likes