LLM Mart Basic
@llm-mart · Joined Jun 2026
Articulation, phonology, language disorders, fluency, voice, AAC, dysphagia, and cognitive-communication expertise
Classroom AI ethics — FERPA, COPPA, age-appropriate AI literacy framing, bias-discussion frameworks
Effective classroom communication for parent partnerships, student feedback, and inclusive documentation
Curriculum design expertise for lesson planning, standards alignment, and instructional frameworks
DSM-5 diagnostic criteria, evidence-based therapy modalities, treatment planning, and progress measurement
Suicide risk assessment using the Columbia Protocol (C-SSRS) and SAFE-T framework — auto-prompts when warning indicators appear
Empathetic clinical writing, health literacy adaptation, and crisis sensitivity
Trauma-informed language guardrails — flag re-traumatizing phrasing and suggest SAMHSA-aligned alternatives
Professional trade communication, estimate presentation, warranty explanation, and upsell framing
Building code awareness, permit references, industry terminology for plumbing, electrical, HVAC, and safety documentation
Empathetic, plain-language communication for treatment discussions, costs, prognosis, and end-of-life care
Clinical documentation, diagnostics, pharmacology, treatment protocols, and species-specific medical knowledge
Produces Alon's design-doc system before code: SOURCE_OF_TRUTH.md, ARCHITECTURE_ROADMAP.md, TODO_WORKFLOW.md, CLAUDE.md, plus a modular docs/architecture set for larger projects. Model first: data and invariants before framework, every invariant enforced at two boundaries, failur
Convenes a panel of opinionated Claude advisors, each with a distinct mandate and a forbidden move so they genuinely diverge, then a Chairman synthesis that commits to one recommendation with the decisive tradeoff and the strongest dissent named.
Forces the laziest solution that actually works: YAGNI, reuse before new code, stdlib before custom, native platform before dependencies, one line before fifty. Three levels — lite (name the lazier alternative, build what's asked), full (ladder enforced, default), ultra (YAGNI ex
Turns a vague ask into a rigorous, grounded, token-efficient prompt for another agent or LLM: role and objective, testable done criteria, anti-hallucination and anti-tokenmaxing rules baked into the generated prompt itself, and strict agent discipline for coding-agent hand-offs.
Local, free, multi-specialist review of a diff: parallel Claude subagents for correctness, security/trust boundaries, data/perf, architecture-altitude, ponytail-simplicity, and tests/failure-paths, each non-trivial finding adversarially verified before being kept, deduped, and ra
Preflight sync and situational brief before repo work: fetches origin, reports ahead/behind divergence (with a no-upstream/detached-HEAD fallback), recent commits, open PRs and issues touching the work, and dirty-tree warnings, then names one recommended first action. Read-only b
Write and publish a new blog post for the skillfold site (site/blog/). Use whenever asked to write a blog post, cover a news story about the agent tooling ecosystem, announce a feature rollout, publish the next post from the queue, or when the weekly blog automation runs. Covers
Review code for correctness, clarity, and security.
A surprising share of AI-in-the-terminal failures aren't the AI. They're zsh, and a version of bash from 2006.
A Claude Code plugin turns standalone project configuration into a namespaced, installable extension that teams and communities can update as one unit.
None of the safety came from the model. It came from six boring habits.
Skills package instructions and references. Subagents run work in a separate context and return results. They solve different problems and can be composed deliberately.
Six hours in, one step left, everything green, and the incident that didn't happen
CLAUDE.md carries persistent project context. Skills load reusable procedures when relevant. Separating stable facts from task-specific workflows keeps both easier to maintain.
Twenty minutes recovering secrets that never existed, and the one sentence from a human that ended it
An API request routing a model's tool call through an approval gate to a remote MCP server
31 config keys, two audits, and why the first one was wrong in both directions
The official MCP Registry stores standardized server metadata rather than package code. Publishers verify a namespace, describe installation or remote access, and submit immutable versions.
Everyone looks at the Dockerfile. The file that actually leaked the key was the project file.
Remote MCP authorization uses established OAuth standards, but secure integration still requires issuer validation, least-privilege scopes, protected token handling, and server-side enforcement.
"Copy it over and switch the reference" is two steps, and the outage lives in the one nobody checks
stdio fits local processes and prototypes. Streamable HTTP fits hosted services and shared integrations. The right choice follows where the capability runs and who must reach it.
The most important rule wasn't about what I could change. It was about what I was allowed to display.
Tools perform operations, resources expose readable context, and prompts provide reusable templates. Choosing the correct primitive makes an MCP server easier to understand and govern.
Use MCP Inspector to connect to local or remote servers, inspect capabilities, call tools, read resources, test prompts, and diagnose failures before release.
Build an MCP server in TypeScript with focused tools, validated schemas, local and remote transports, Inspector tests, and production security controls.
An MCP server exposes tools, resources, or prompts through a standard protocol so an AI application can discover and use external capabilities.
Treat an AI agent skill as both an instruction package and a software dependency: inspect what it says, what it runs, what it can access, and how it updates.
/btw
Btw
The one exception to codeArbiter's slash-command pipeline: a lightweight question-and-answer
/checkpoint
Checkpoint
A periodic sweep of the entire codebase with the same reviewer fleet `/ca:review` uses per-diff,
/chore
Chore
This is the lane for changes with no behavior to test-drive — prose edits, a version bump on an
/cleanup
Cleanup
Use this after a pull request has merged but your local checkout is still on the topic branch.
/commands
Commands
Prints the public command catalog straight from `COMMANDS.md` — the plugin's own single source
/commit
Commit
This is the single entry point for turning staged work into a commit — nothing in codeArbiter
/conflict
Conflict
The protocol for a rule conflict — not a skill route, an orchestrator-level halt. When two sources
/context-check
Context check
An optional, on-demand drift audit for the bypass case: a merge, a direct push, or a manual edit
/create-context
Create context
This is the populator for a project that already has code to read. Instead of interviewing you about
/debug
Debug
This is where an unexplained defect goes before anyone touches code. The investigation is
/decompose
Decompose
This is the populator for a project that has no code yet to read. Rather than guessing at
/doctor
Doctor
Proves the install is actually enforcing, rather than just present. codeArbiter's worst failure
/feature
Feature
This is the standard entry point for new work with a human in the loop at every step. A short
/fix
Fix
This is the entry point for a defect that already has a known cause, or one you can describe
/init
Init
This is how a repository opts into codeArbiter for the first time. It writes the root-level state
/metrics
Metrics
A bare-numbers governance glance — three metrics, each with a trend arrow against the prior
/new-skill
New skill
The only permitted entry to creating a new codeArbiter skill. It hands off to the `skill-author`
/override
Override
The sanctioned, logged escape hatch. A routine gate — a lint rule, a style check, a non-security
/pr
Pr
This is the only path to opening a pull request — there's no direct push or force-push to the
/preview
Preview
A zero-onboarding, read-only dry-run of the reviewer fleet against whatever is currently
Ultra-lightweight, open-source, self-hosted personal AI agent framework in Python with WebUI, tools, memory, MCP, multi-agent workflows, automation, and chat ap…
29 views 0 likesGovernance framework for AI coding agents. It runs them through a five-step workflow (plan, build, review, test, ship) where no step counts as done without evid…
17 views 0 likesUltimate Multi-Agent OS for Autonomous AI NPCs 2026
14 views 0 likesPersonal AI Agent Hub 2026 — Build Your 24/7 Autonomous Assistant
25 views 0 likesProven 2026 Multi-Agent AI Review System – Verdict-Driven Quality Control
28 views 0 likesSlash API Batch: Cut AI Costs by 50% in 2026
15 views 0 likesWeb dashboard for Hermes Agent — multi-platform AI chat, session management, scheduled jobs, usage analytics
17 views 0 likesAgent Skills for Solopreneurs
31 views 0 likesAirLLM dramatically reduces inference memory usage, letting 70B large language models run on a single 4GB GPU card
110 views 0 likesZero, your trustworthy AI teammate for real work.
16 views 0 likes一套 DSH runtime,Desktop、Web 与 TUI 三种开发体验。
11 views 0 likesOpen-source operational advisor for ClickHouse — real-time monitoring plus AI-driven index/partition/materialized-view recommendations.
16 views 0 likes⚙️ TypeScript Style Guide and Agent Skill. A concise set of conventions and best practices for consistent, maintainable code.
27 views 0 likesFramework for AI agents to build and maintain a digital brain through Obsidian wiki
16 views 0 likesApache Maka (Incubating) is a local-first AI agent workspace. Model messages, tool calls, tool results, permission decisions, and termination events are recorde…
24 views 0 likesNeo.mjs is a self-evolving software organism: a professional end-to-end AI engineering team whose cross-model swarm inhabits live apps via Neural Link, Active H…
24 views 0 likesAgentic development harness for Claude Code — SPEC-driven plan/run/sync, TRUST 5 quality gates, model+effort routing, and Claude×GLM multi-LLM cost control. Sin…
18 views 0 likesNocoBase is an open-source AI + no-code platform for building business systems fast. Instead of generating everything from scratch, AI works on top of productio…
27 views 0 likesAn open-source AI coding agent that lives in your terminal.
28 views 0 likesPawWork — free, open-source desktop AI agent for macOS and Windows. Alternative to Codex App and Claude Cowork. BYOK with 75+ providers, ChatGPT OAuth, local mo…
14 views 0 likes