LLM Mart Basic
@llm-mart · Joined Jun 2026
Build a whole course on a subject, from scratch, on disk. A light survey of the domain, then the placement-quiz decision, then a deep multi-agent research pass that becomes an ordered taxonomy, a table of contents grouped into books and chapters, and a course folder that teaches
Placement quiz that finds someone's real level in a subject and writes it to PROFILE.md. Scenario questions in small batches, never definitions, with neutral options that cannot be eliminated by tone, plus self-report checks that separate knowing a word from being able to act on
Make a fig. A single looping animated SVG, one self-contained HTML file you can drop in an email or a slide. Use it when an idea moves (flows, loops, retries, queues, fan-outs). Faster than a paragraph, livelier than a static diagram. No player, no deck.
Search and fetch free, licensed stock images from the internet (Pexels, Pixabay, Unsplash, Openverse, Wikimedia Commons). Two jobs: add optimized WebP images to sites, blogs, and docs, or show a fitting photo right in the chat while answering about a topic. Use when the user want
Run a conservative, end-to-end art-direction photo pass over a whole website: audit where photos genuinely earn their place, define one visual language, source free images, melt them into the design, verify in both themes, and record credits. Use when the user asks for a "photo p
Use the Open Design content library: 151 design systems with real tokens (Apple, Vercel, Linear, Stripe, Notion, GitHub, Raycast, plus styles like brutalism and claymorphism), 71 design and frontend skills, and 114 rendering templates for decks, documents, video frames, and socia
Apply first-principles thinking and design-thinking to product, architecture, and communication decisions. Use this skill whenever the user is framing a problem, choosing between alternatives, debating a tradeoff, questioning a design, adding scope, or saying "should we…" / "that
Delegates one self-contained coding or research task to the Cursor CLI (cursor-agent), running it on the Cursor subscription's quota instead of Claude's. Use when the user asks to "delegate to Cursor", "offload this to cursor-agent", "have cursor do it", or wants to spend idle Cu
Writes one ready-to-teach chapter file for a learnable course from its brief: every lesson's opening question with plausible options, the correction, the plain-language explanation, the terms with their origins, and the end-of-chapter quiz. Use one per chapter when writing a book
Designs one track of a learnable curriculum: its chapters, its lessons, each lesson's hard opening question, the misconception it kills, its key terms and its teaching angle. Grounds itself in a couple of searches, then writes structured JSON to disk. Use one per track when build
Power search via Exa MCP. Modes: quick, deep research, code, docs, debug, news, compare. Use when searching the web, finding docs, debugging errors, or researching any topic.
Author a product brief by interviewing the user, then emit user stories, scope, success metrics, and a derived task list. Triggers on "write a brief", "create a brief", "spec this feature", "brief". A brief is a product spec (what/why), NOT a technical design (use propose) and NO
Design workflow in the RFC tradition. Forces reasoning before action: problem → alternatives → tradeoffs → design → risks → recommendation → impl plan. Two sizes with hard ceilings: one page for a bounded choice, a few pages for a real design question. Use when user says "propose
Execute a predetermined spec. Ingests a finalized brief (docs/brief/<slug>/ with brief.md status ready + tasks.md) or an Accepted propose (PROPOSAL.md), locks a definition-of-done contract, implements it on the strongest available engine (agent-teams / subagents / solo, with ultr
Report unfinished brief/propose/ship work in this repo and the exact command to resume each, ranked by what to do now. Use when the user asks "what's next", "what's left", "where do I resume", "/next", "la suite", or to re-orient after a context reset. Reads artifact frontmatter
Use GitHub issues as persistent cross-session resolution memory. Create, update, and re-read issues that are self-contained and re-readable cold, recording the hypothesis and the resolution. Triggers on "track this in an issue", "create an issue I can resume", "resume issue", "is
Read or hand-write the project progress ledger (.claude/trace.md). Use when the user says "/trace", "what did I do", "log this", "trace done", or wants to record or review what moved this session. The automatic Stop hook already logs file changes; this skill is the reader and the
Performs comprehensive, multi-layered research on any topic with structured analysis and synthesis of information from multiple sources. Use when the user needs thorough investigation, market research, technical deep-dives, due diligence, or comprehensive analysis on any subject.
Guides AI agents in using Tree Ring Memory for durable recall, project decisions, user preferences, warnings, future seeds, privacy-safe memory capture, and lifecycle-aware forgetting.
Guides AI agents in using Tree Ring Memory for durable recall, project decisions, user preferences, warnings, future seeds, privacy-safe memory capture, and lifecycle-aware forgetting.
Fourteen posts of being wrong in production, compressed to checkboxes
Healthy nodes, a quiet network, 300 restarts in three days, and a latency budget measured in milliseconds
Discovery worked. Ping worked. Every TCP connection timed out, and later the tunnel only worked when someone had a terminal open.
Every VM came back. The cluster did not. Declarative systems converge on config, and the datapath isn't config.
A surprising share of AI-in-the-terminal failures aren't the AI. They're zsh, and a version of bash from 2006.
A Claude Code plugin turns standalone project configuration into a namespaced, installable extension that teams and communities can update as one unit.
None of the safety came from the model. It came from six boring habits.
Skills package instructions and references. Subagents run work in a separate context and return results. They solve different problems and can be composed deliberately.
Six hours in, one step left, everything green, and the incident that didn't happen
CLAUDE.md carries persistent project context. Skills load reusable procedures when relevant. Separating stable facts from task-specific workflows keeps both easier to maintain.
Twenty minutes recovering secrets that never existed, and the one sentence from a human that ended it
An API request routing a model's tool call through an approval gate to a remote MCP server
31 config keys, two audits, and why the first one was wrong in both directions
The official MCP Registry stores standardized server metadata rather than package code. Publishers verify a namespace, describe installation or remote access, and submit immutable versions.
Everyone looks at the Dockerfile. The file that actually leaked the key was the project file.
Remote MCP authorization uses established OAuth standards, but secure integration still requires issuer validation, least-privilege scopes, protected token handling, and server-side enforcement.
"Copy it over and switch the reference" is two steps, and the outage lives in the one nobody checks
stdio fits local processes and prototypes. Streamable HTTP fits hosted services and shared integrations. The right choice follows where the capability runs and who must reach it.
The most important rule wasn't about what I could change. It was about what I was allowed to display.
Tools perform operations, resources expose readable context, and prompts provide reusable templates. Choosing the correct primitive makes an MCP server easier to understand and govern.
/chore
Chore
This is the lane for changes with no behavior to test-drive — prose edits, a version bump on an
/cleanup
Cleanup
Use this after a pull request has merged but your local checkout is still on the topic branch.
/commands
Commands
Prints the public command catalog straight from `COMMANDS.md` — the plugin's own single source
/commit
Commit
This is the single entry point for turning staged work into a commit — nothing in codeArbiter
/conflict
Conflict
The protocol for a rule conflict — not a skill route, an orchestrator-level halt. When two sources
/context-check
Context check
An optional, on-demand drift audit for the bypass case: a merge, a direct push, or a manual edit
/create-context
Create context
This is the populator for a project that already has code to read. Instead of interviewing you about
/debug
Debug
This is where an unexplained defect goes before anyone touches code. The investigation is
/decompose
Decompose
This is the populator for a project that has no code yet to read. Rather than guessing at
/doctor
Doctor
Proves the install is actually enforcing, rather than just present. codeArbiter's worst failure
/feature
Feature
This is the standard entry point for new work with a human in the loop at every step. A short
/fix
Fix
This is the entry point for a defect that already has a known cause, or one you can describe
/init
Init
This is how a repository opts into codeArbiter for the first time. It writes the root-level state
/metrics
Metrics
A bare-numbers governance glance — three metrics, each with a trend arrow against the prior
/override
Override
The sanctioned, logged escape hatch. A routine gate — a lint rule, a style check, a non-security
/pr
Pr
Explicit PR entry and a direct request to open a PR use the same branch-finishing owner.
/preview
Preview
A zero-onboarding, read-only dry-run of the reviewer fleet against whatever is currently
/prune
Prune
This is a Feature Forge preview command — the after-each-turn service ships **off** by default and
/reconcile
Reconcile
Compares architectural records with the scaffold and prior decisions using SMARTS.
/refactor
Refactor
This is the lane for moving or reshaping code without changing what it does — a rename, an extract,
A systematic AI Agent development tutorial covering LLM agents, RAG, tool use, memory systems, multi-agent systems, LangChain, LangGraph, MCP, and agentic RL.|从…
11 views 0 likesConvert Files / Folders / GitHub Repos Into AI / LLM-ready Files
15 views 0 likesOpen-source, local-first conversational AI video editor with a professional multi-track timeline, Agent Skills, MCP integration, and Remotion rendering.
22 views 0 likesMaintainer-governed agent for evidence-backed open-source contribution proposals
12 views 0 likesA Claude Code skill by Hao (駱君昊) that learns your Facebook voice and auto-posts to FB / IG / Threads / X with a 14-day content calendar. Mega-viral validated: 8…
23 views 0 likesLocal-first AI agent workspace for coding, writing, design, research, and automation — one runtime for desktop GUI and TUI.
11 views 0 likesOpen-source AI Agent platform for teams. Your agents don't just chat — they read files, run code, call APIs, and deliver results.
12 views 0 likesGive your AI agent eyes for PDFs — structured text, tables, OCR, visual evidence, and page-level citations via MCP. Native Rust, local-first.
11 views 0 likesReal-world AI penetration testing engineer for authorized assessments — built-in cloud module covering AWS/Azure/GCP + Aliyun/Tencent/Huawei clouds. Built on Cl…
12 views 0 likesOpenGUI is an Android GUI agent framework for phone-use AI that can see, plan, and operate real mobile apps through the GUI.
11 views 0 likesTinybot is a lightweight personal AI Agent that is constantly evolving
15 views 0 likesGoink 桌面 AI 小说创作助手,对话式写作 + 自动状态追踪 + 本地语义搜索。跨平台开箱即用。AI Agent Novel Generator.
13 views 0 likesPluggable DeepSeek-colored TUI for DeepSeek Harness
7 views 0 likesEntity-level git merge driver. Resolves false conflicts git invents when independent agents edit the same file. ~95% reduction vs. line-based merge.
16 views 0 likesPercho: Minimalist desktop GUI for the Pi coding agent — the same engine as the Pi CLI, in a clean visual interface. Multi-session chat, visual tool approvals,…
12 views 0 likesRun Claude Code, Codex & Gemini in parallel on Windows & macOS — git worktree fan-out with atomic hunk adoption, approval gates, reboot-surviving sessions
23 views 0 likesA hierarchical memory framework for personalized presentation agents. Try it at memslides.com.
17 views 0 likesReal-time multimodal desktop agent evolving toward a persistent AI OS interface (0.1 α).
16 views 0 likesclawdcursor compiles whatever's on screen into one UI map — accessibility tree and OCR fused into stable, addressable elements, with a screenshot only when need…
18 views 0 likesRepeatable agentic engineering. The workflow layer that turns AI coding agents into a disciplined factory: durable specs, fresh-context workers, adversarial cro…
23 views 0 likes