LLM Mart Basic

@llm-mart · Joined Jun 2026

0 Followers 0 Reputation 13633 Contributions
Claude Agent dimension-propagator

Propagates dimensional annotations through arithmetic and call chains, reporting mismatches found during propagation

0
Claude Agent dimension-validator

Validates dimensional consistency and detects dimensional bugs in annotated code

0
Claude Agent data-flow-analyzer

Analyzes data flow from source to vulnerability sink, mapping trust boundaries, API contracts, environment protections, and cross-references. Spawned by fp-check during Phase 1 verification.

0
Claude Agent exploitability-verifier

Verifies whether a suspected vulnerability is actually exploitable by proving attacker control, mathematical bounds, and race condition feasibility. Spawned by fp-check during Phase 2 verification.

0
Claude Agent poc-builder

Creates proof-of-concept exploits (pseudocode, executable, and unit tests) demonstrating a verified vulnerability, plus negative PoCs showing exploit preconditions. Spawned by fp-check during Phase 4 verification.

0
Claude Agent draw

Draw the 12 Houses of the Zodiac Tarot spread and return a concise structured reading. Use as a named agent instead of wrapping Skill(let-fate-decide) in an Agent call. Callers get just the verdict text; card file content stays in this agent context.

0
Claude Agent rust-review-dedup-judge

Deduplication judge for the rust-review pipeline. Merges duplicate findings deterministically by exact location and bug class, then runs LLM passes over same-function candidates, including the same bug filed under different bug classes. Spawned by the rust-review skill orchestrat

0
Claude Agent rust-review-fp-judge

Second-stage judge in the rust-review pipeline. Runs after dedup-judge on merged primaries only. Decides fp_verdict, then (for survivors) severity/attack_vector/exploitability, and writes the final REPORT.md + REPORT.sarif. Spawned by the rust-review skill orchestrator only.

0
Claude Agent rust-review-worker

Runs one assigned rust-review cluster task and writes finding files to the run's output directory. Spawned by the rust-review skill orchestrator only.

0
Claude Agent sharp-edges-analyzer

Evaluates APIs, configurations, and library interfaces for misuse resistance and footgun potential. Use when reviewing code for error-prone designs, dangerous defaults, or APIs that make security mistakes easy.

0
Claude Agent spec-compliance-checker

Checks one documented requirement against the code that should implement it, and returns a verdict with the lines that evidence it. Writes its analysis to disk and returns a compact record. Use for a single requirement; use the spec-compliance workflow for a whole document.

0
Claude Agent 0-preflight

Performs preflight validation, config merging, TU enumeration, and work directory setup for zeroize-audit. Produces merged-config.yaml, preflight.json, and orchestrator-state.json.

0
Claude Agent 1-mcp-resolver

Resolves symbol definitions, types, and cross-file references using Serena MCP for zeroize-audit. Runs before source analysis so enriched type data is available for wipe validation.

0
Claude Agent 2-source-analyzer

Identifies sensitive objects, detects wipe calls, validates correctness, and performs data-flow/heap analysis for zeroize-audit. Produces the sensitive object list and source-level findings consumed by compiler analysis and report assembly.

0
Claude Agent 2b-rust-source-analyzer

Performs source-level zeroization analysis for Rust crates in zeroize-audit. Generates rustdoc JSON for trait-aware analysis and runs token-based dangerous API scanning. Produces sensitive objects and source findings consumed by rust-compiler-analyzer and report assembly.

0
Claude Agent 3-tu-compiler-analyzer

Performs per-TU compiler-level analysis (IR diff, assembly, semantic IR, CFG) for zeroize-audit. One instance runs per translation unit, enabling parallel execution across TUs.

0
Claude Agent 3b-rust-compiler-analyzer

Performs crate-level MIR and LLVM IR analysis for Rust in zeroize-audit. A single instance runs per crate (unlike 3-tu-compiler-analyzer which runs one per C/C++ TU). Detects dead-store elimination of wipes, stack retention, and other compiler-level zeroization failures.

0
Claude Agent 4-report-assembler

Collects all findings from source and compiler analysis, applies supersessions and confidence gates, normalizes IDs, and produces a comprehensive markdown report with structured JSON for downstream tools. Supports dual-mode invocation: interim (findings.json only) and final (merg

0
Claude Agent 5-poc-generator

Crafts bespoke proof-of-concept programs demonstrating that zeroize-audit findings are exploitable. Reads source code and finding details to generate tailored PoCs — each PoC is individually written, not templated. Each PoC exits 0 if the secret persists or 1 if wiped. Mandatory

0
Claude Agent 5b-poc-validator

Compiles and runs all PoCs for zeroize-audit findings. Produces poc_validation_results.json consumed by the verification agent and the orchestrator.

0
The AI-in-production safety playbook

Fourteen posts of being wrong in production, compressed to checkboxes

security prompt-engineering devops ai
Sep 30
The control plane was flapping because of a spinning disk

Healthy nodes, a quiet network, 300 restarts in three days, and a latency budget measured in milliseconds

kubernetes sre incident-response observability
Sep 29
The overlay that pinged but wouldn't carry TCP

Discovery worked. Ping worked. Every TCP connection timed out, and later the tunnel only worked when someone had a terminal open.

containers incident-response networking linux
Sep 28
Bringing a cluster back after the host rebooted

Every VM came back. The cluster did not. Declarative systems converge on config, and the datapath isn't config.

kubernetes sre containers incident-response
Sep 27
The agent is running in *your* shell

A surprising share of AI-in-the-terminal failures aren't the AI. They're zsh, and a version of bash from 2006.

devops ai-agents automation shell
Sep 26
How to create and share a Claude Code plugin

A Claude Code plugin turns standalone project configuration into a namespaced, installable extension that teams and communities can update as one unit.

agents security skill-md claude-code
Sep 25
The scaffolding that made it safe

None of the safety came from the model. It came from six boring habits.

git devops ai-agents claude-code
Sep 25
Claude Code skills vs. subagents: when to use each

Skills package instructions and references. Subagents run work in a separate context and return results. They solve different problems and can be composed deliberately.

agent-skills claude-code context
Sep 24
Knowing when to stop

Six hours in, one step left, everything green, and the incident that didn't happen

prompt-engineering ai ai-agents sre
Sep 24
CLAUDE.md vs. skills: where should Claude Code instructions live?

CLAUDE.md carries persistent project context. Skills load reusable procedures when relevant. Separating stable facts from task-specific workflows keeps both easier to maintain.

agent-skills claude-skills configuration context
Sep 23
Those are the other app's keys

Twenty minutes recovering secrets that never existed, and the one sentence from a human that ended it

security devops ai ai-agents
Sep 23
How to use remote MCP servers with the OpenAI Responses API

An API request routing a model's tool call through an approval gate to a remote MCP server

security mcp integrations open-api
Sep 22
The coverage audit before you delete the safety net

31 config keys, two audits, and why the first one was wrong in both directions

security devops ai-agents secrets-management
Sep 22
How to publish an MCP server to the official MCP Registry

The official MCP Registry stores standardized server metadata rather than package code. Publishers verify a namespace, describe installation or remote access, and submit immutable versions.

security mcp
Sep 21
Rotating a leaked credential, in the right order

Everyone looks at the Dockerfile. The file that actually leaked the key was the project file.

security devops ai-agents containers
Sep 21
MCP authentication explained: OAuth, scopes, and safe token handling

Remote MCP authorization uses established OAuth standards, but secure integration still requires issuer validation, least-privilege scopes, protected token handling, and server-side enforcement.

security mcp
Sep 20
Byte-identical or bust

"Copy it over and switch the reference" is two steps, and the outage lives in the one nobody checks

security kubernetes verification
Sep 20
MCP stdio vs. Streamable HTTP: which transport should you use?

stdio fits local processes and prototypes. Streamable HTTP fits hosted services and shared integrations. The right choice follows where the capability runs and who must reach it.

security mcp
Sep 19
Never let the AI print a secret

The most important rule wasn't about what I could change. It was about what I was allowed to display.

security kubernetes devops ai-agents
Sep 19
MCP tools vs. resources vs. prompts: when to use each

Tools perform operations, resources expose readable context, and prompts provide reusable templates. Choosing the correct primitive makes an MCP server easier to understand and govern.

security mcp
Sep 18
/inspect Inspect

`crabbox inspect` prints the full record for a single lease: state, provider,

0
/job Job

Run named, repo-local jobs defined in your Crabbox config.

0
/list List

`crabbox list` shows the current Crabbox machines (leases) for a provider. It is

0
/login Login

`crabbox login` authenticates the CLI against a coordinator, stores the

0
/logout Logout

`crabbox logout` clears the stored broker token from your user config so the CLI

0
/logs Logs

`crabbox logs` prints the retained command output for a recorded run.

0
/marketplace Marketplace

`crabbox marketplace` previews the Crabbox credits gateway: one Crabbox billing

0
/media Media

`crabbox media` turns a recorded desktop video into lightweight review

0
/open Open

`crabbox open` prepares an existing SSH-capable lease for an external editor.

0
/pause Pause

`crabbox pause` pauses a single lease, freeing the remote compute while

0
/pond Pond

`crabbox pond` is the cross-provider peer-discovery and lifecycle surface for a

0
/pool Pool

`crabbox pool` contains machine-pool helpers. `pool list` keeps the older

0
/ports Ports

`crabbox ports` bridges provider-native port publishing for an existing Crabbox

0
/prewarm Prewarm

`crabbox prewarm` leases a reusable box and prepares it for test runs. For

0
/providers Providers

`crabbox providers` prints the provider capability matrix that the CLI compiles

0
/receipt Receipt

`crabbox receipt <run-id>` retrieves a brokered run's committed terminal

0
/results Results

`crabbox results` prints the structured test summary attached to a recorded

0
/resume Resume

`crabbox resume` resumes a lease previously paused with [`pause`](pause.md),

0
/run Run

`crabbox run` syncs the current dirty checkout to a box, runs a command there,

0
/screenshot Screenshot

`crabbox screenshot` captures a single PNG from a desktop lease without opening a

0
Atom

Atom Agent, Open-Source Governed AI Agent Platform for Self-Hosted Automation

13 views 0 likes
Oh My Hermes

The agent engineering intelligence harness, optimized tools, memory system, subagents and mixture of models packages ⚚

16 views 0 likes
Sesori Apps Monorepo

Sesori iOS/Android app and the Sesori Bridge CLI — drive Claude, Codex, OpenCode, Cursor, Pi, OMP, Hermes coding sessions from your phone

14 views 0 likes
Repobrain

🧠 RepoBrain (formerly Antigravity) — Give your repo a brain. ChatGPT for your codebase: works in Claude Code, Cursor, Codex, Windsurf & more.

14 views 0 likes
Dsh Plugin Subscriptions

Use ChatGPT (Codex), Claude, and Grok (X Premium) subscriptions as DeepSeek Harness LLM providers — OAuth login in the web UI, no API keys

15 views 0 likes
Latitude Llm

Latitude is the open-source AI monitoring platform.

16 views 0 likes
Webbrain

Open-source AI browser agent for Chrome and Firefox (monorepo) 🧠

16 views 0 likes
Cli

Official Model Studio CLI(阿里云百炼 CLI)built for AI Agent frameworks, exposing models, search, multimodal, and workflow capabilities as structured tool calls.

15 views 0 likes
Awesome DeepSeek Harness Plugins

Curated DeepSeek Harness (DSH) plugins, extensions, tools, skills, clients, runtimes, integrations, and verified references — English and Chinese.

16 views 0 likes
Slides Maker

Turn papers, code, and docs into presentation-ready, natively editable PPTX in Codex / Claude Code. Native charts and equations, speaker notes, click-build anim…

14 views 0 likes
Grix

Grix : Work with agents like talking to people.

25 views 0 likes
Atomic Agent

Local First Ai Agent. Optimized for Local Ai models. Long context window. Proper tools callings. Runs privately on your device.

26 views 0 likes
Agentql Mcp

Model Context Protocol server that integrates AgentQL's data extraction capabilities.

15 views 0 likes
Pilot

AI that ships your tickets.

14 views 0 likes
Octo Cli

Metadata-driven CLI for AI Agent Bots — 48 operations across 7 domains, structured JSON envelope I/O, zero interactive prompts.

12 views 0 likes
Nextclaw

An open-source, extensible, self-hosted agent workspace with multi-runtime support for Codex, Claude Code, and more, plus reusable local apps for custom interfa…

14 views 0 likes
Deepseek Harness Desktop

Open-source Windows desktop client and GUI for DeepSeek Harness — zero-setup installer with Codex, plugins, skills, SSH, mobile remote access, and 11 skins.

12 views 0 likes
Openscience

The open-source AI workbench for scientific research

15 views 0 likes
DeepSeek Reasonix

DeepSeek-native AI coding agent for your terminal. Engineered around prefix-cache stability — leave it running.

15 views 0 likes
Bladebro

A Fully free agentic browser driver for AI , few tools, full control, real stealth, top-tier token efficiency.

12 views 0 likes