LLM Mart Trusted

@llmmart · Joined Jun 2026

0 Followers 1003 Reputation 164 Contributions
Claude Skill Adversarial code-review skill for Claude

A Claude skill that reviews a diff by trying to refute each finding before reporting it.

43
Gemini Prompt Grounded research prompt for Gemini

A prompt that forces Gemini to cite sources and flag uncertainty instead of guessing.

34
DeepSeek Workflow Reasoning-first debugging workflow for DeepSeek

A workflow that makes DeepSeek state a hypothesis and a test before proposing any fix.

29
Cursor Recipe Refactor-with-tests recipe for Cursor

A Cursor recipe that writes a characterization test before each refactor step.

19
MiniMax Skill Roleplay-agent skill for MiniMax

A skill that keeps a consistent persona and memory across a long MiniMax conversation.

15
Claude Skill Conventional-commit writer for Claude

Turns a staged diff into a clean Conventional Commits message — correct type, tight subject, useful body.

2
ChatGPT Workflow Source-grounded research brief for ChatGPT

Uses ChatGPT search plus an adversarial verification pass to turn a fuzzy question into a sourced brief you can hand to someone else.

0
Codex CLI Skill Review-before-apply skill for Codex CLI

Forces Codex CLI to show the diff, review its own changes, run focused validation, and list uncertainty before you keep the patch.

0
GitHub Copilot Skill Code review guardrails for GitHub Copilot

A review skill that makes Copilot focus on concrete behavior changes, missing tests, and repo instructions instead of generic style chatter.

0
GitHub Copilot Template Repository instructions bootstrap for GitHub Copilot

A starter layout for .github/copilot-instructions.md, path-specific instruction files, and AGENTS.md so Copilot stops guessing how your repo works.

0
GitHub Copilot Recipe Test generation loop for GitHub Copilot

Use /setupTests, /tests, and /fixTestFailure in a red-green loop instead of asking Copilot to “add some tests” and hoping for the best.

0
Cursor Recipe Reproduce-first bug hunt for Cursor

Forces a failing test that reproduces the bug before any fix — so you know it's actually fixed.

0
ChatGPT Template Custom GPT acceptance harness

A reusable test harness for checking whether a custom GPT actually follows its instructions, uses tools correctly, and fails safely.

0
Claude Skill Security-review skill for Claude

A structured security pass over a diff — checks the OWASP-relevant classes, ranks by exploitability, and shows a repro.

0
Claude Workflow Spec-first feature build (workflow)

Stops the model from coding too early — lock a short spec and a test list before a single line is written.

0
Point your agent at the catalogue: the LLM Mart MCP server and API

The whole public catalogue is an MCP server and a REST API, so your agent can search skills, tools and slash-commands as native tools. Setup is one config block. Plus a private vault that carries your own prompts between machines.

agents coding security
Aug 22
Ship a prompt like code: a five-case eval harness you can build in an hour

You wouldn't ship a function you ran once. Here's the smallest evaluation setup that catches real regressions — five cases, three graders, and a rule for when to add a sixth.

writing data prompt-engineering
Jun 29
The instruction file is the highest-leverage file in your repo

AGENTS.md, CLAUDE.md, copilot-instructions.md and .cursor/rules all solve the same problem. Most of them are written badly. Here's what changes agent behaviour and what's decoration.

agents coding prompt-engineering
Jun 26
Prompt injection is not an XSS problem

Sanitizing output protects your page. It does nothing for an agent that reads a poisoned README and then runs a command. A practical model of the threat, and what actually helps.

code-review agents security
Jun 23
Prompt, skill, workflow, recipe, template, agent: a working taxonomy

The words are used interchangeably and they shouldn't be. Six categories, what actually distinguishes them, and how to tell which one you're holding.

writing agents prompt-engineering
Jun 20
How to write a prompt you can reuse

Most prompts are lucky one-offs. The ones that keep working share a structure: an explicit role, a checklist of constraints, a worked example, and a self-check step. Here's how to build one.

Jun 18
/setupTests Set up tests

Recommend and configure a relevant testing framework, including setup steps and extensions.

10
/copy Copy response

Copy the last assistant response to the clipboard.

9
/radio Radio

Open Claude FM lo-fi radio in your browser.

9
/workflows Workflows

Open the workflow progress view to watch, pause, resume, or save runs.

9
/image Generate image

Create or edit an image from a text description.

9
/memories Memory settings

Configure whether Codex uses or generates memories for future sessions.

9
/HOOK Hook

Generate stronger opening lines.

8
/chrome Chrome settings

Configure Claude in Chrome settings.

7
/passes Share passes

Share a free week of Claude Code with friends.

7
/upgrade Upgrade plan

Open the upgrade page to switch to a higher plan tier.

7
/score Score items

Rate items against specified criteria.

7
/compact Compact transcript

Summarize earlier conversation turns to free context while preserving key details.

7
/COUNTERARGUE Counterargue

Generate counterarguments against your idea.

6
/mcp [list|list-tools] [identifier] MCP management

Manage MCP servers and list available tools.

6
/autofix-pr Auto-fix PR

Spawn a cloud session that watches the branch's PR and pushes fixes on CI failures or review comments.

5
/login Log in

Sign in to your Anthropic account.

5
/teleport Teleport session

Pull a Claude Code on the web session into this terminal.

5
/email Write email

Generate or polish professional email correspondence.

5
/apps Browse apps

Browse available apps and insert an app mention into the prompt.

5
/tests Generate tests

Generate unit tests for the selected code.

5