LLM Mart Trusted

@llmmart · Joined Jun 2026

0 Followers 1003 Reputation 164 Contributions
Claude Skill Adversarial code-review skill for Claude

A Claude skill that reviews a diff by trying to refute each finding before reporting it.

43
Gemini Prompt Grounded research prompt for Gemini

A prompt that forces Gemini to cite sources and flag uncertainty instead of guessing.

34
DeepSeek Workflow Reasoning-first debugging workflow for DeepSeek

A workflow that makes DeepSeek state a hypothesis and a test before proposing any fix.

29
Cursor Recipe Refactor-with-tests recipe for Cursor

A Cursor recipe that writes a characterization test before each refactor step.

19
MiniMax Skill Roleplay-agent skill for MiniMax

A skill that keeps a consistent persona and memory across a long MiniMax conversation.

15
Claude Skill Conventional-commit writer for Claude

Turns a staged diff into a clean Conventional Commits message — correct type, tight subject, useful body.

2
ChatGPT Workflow Source-grounded research brief for ChatGPT

Uses ChatGPT search plus an adversarial verification pass to turn a fuzzy question into a sourced brief you can hand to someone else.

0
Codex CLI Skill Review-before-apply skill for Codex CLI

Forces Codex CLI to show the diff, review its own changes, run focused validation, and list uncertainty before you keep the patch.

0
GitHub Copilot Skill Code review guardrails for GitHub Copilot

A review skill that makes Copilot focus on concrete behavior changes, missing tests, and repo instructions instead of generic style chatter.

0
GitHub Copilot Template Repository instructions bootstrap for GitHub Copilot

A starter layout for .github/copilot-instructions.md, path-specific instruction files, and AGENTS.md so Copilot stops guessing how your repo works.

0
GitHub Copilot Recipe Test generation loop for GitHub Copilot

Use /setupTests, /tests, and /fixTestFailure in a red-green loop instead of asking Copilot to “add some tests” and hoping for the best.

0
Cursor Recipe Reproduce-first bug hunt for Cursor

Forces a failing test that reproduces the bug before any fix — so you know it's actually fixed.

0
ChatGPT Template Custom GPT acceptance harness

A reusable test harness for checking whether a custom GPT actually follows its instructions, uses tools correctly, and fails safely.

0
Claude Skill Security-review skill for Claude

A structured security pass over a diff — checks the OWASP-relevant classes, ranks by exploitability, and shows a repro.

0
Claude Workflow Spec-first feature build (workflow)

Stops the model from coding too early — lock a short spec and a test list before a single line is written.

0
Point your agent at the catalogue: the LLM Mart MCP server and API

The whole public catalogue is an MCP server and a REST API, so your agent can search skills, tools and slash-commands as native tools. Setup is one config block. Plus a private vault that carries your own prompts between machines.

agents coding security
Aug 22
Ship a prompt like code: a five-case eval harness you can build in an hour

You wouldn't ship a function you ran once. Here's the smallest evaluation setup that catches real regressions — five cases, three graders, and a rule for when to add a sixth.

writing data prompt-engineering
Jun 29
The instruction file is the highest-leverage file in your repo

AGENTS.md, CLAUDE.md, copilot-instructions.md and .cursor/rules all solve the same problem. Most of them are written badly. Here's what changes agent behaviour and what's decoration.

agents coding prompt-engineering
Jun 26
Prompt injection is not an XSS problem

Sanitizing output protects your page. It does nothing for an agent that reads a poisoned README and then runs a command. A practical model of the threat, and what actually helps.

code-review agents security
Jun 23
Prompt, skill, workflow, recipe, template, agent: a working taxonomy

The words are used interchangeably and they shouldn't be. Six categories, what actually distinguishes them, and how to tell which one you're holding.

writing agents prompt-engineering
Jun 20
How to write a prompt you can reuse

Most prompts are lucky one-offs. The ones that keep working share a structure: an explicit role, a checklist of constraints, a worked example, and a self-check step. Here's how to build one.

Jun 18
/ultraplan Ultra plan

Draft a plan in an ultraplan session and review it in your browser.

33
/framework Framework

Create or apply a methodology such as SWOT or StoryBrand.

33
/archive Archive session

Archive the current session and exit Codex without deleting its transcript.

33
/PROS_CONS Pros and cons

Build a pros-and-cons table.

32
/logout Log out

Sign out of your Cursor account.

32
/advisor Advisor

Enable or disable the advisor tool that consults a second model for guidance.

31
/install-slack-app Install Slack app

Install the Claude Slack app via an OAuth flow.

31
/tasks View tasks

View and manage everything running in the background.

31
/analyze Analyze

Identify patterns and insights from text or data.

31
/sandbox-add-read-dir Grant sandbox read directory

Grant the Windows sandbox read access to an additional absolute directory path.

31
/title Terminal title

Configure which fields appear in the terminal window or tab title.

31
/help Help

Show a quick reference and basics of using GitHub Copilot.

31
/HUMANIZE Humanize

Make the text sound more natural and less robotic.

30
/setup-terminal Terminal setup

Configure terminal newline keybindings.

30
/help Help

Show help and available commands.

29
/skills List skills

List available skills, with filtering and visibility controls.

29
/quit Quit CLI

Exit the CLI immediately; same effect as /exit.

29
/UXMODE UX expert mode

Analyze the task from the perspective of a UX expert.

28
/vim Vim keys

Enable or disable Vim keyboard bindings.

28
/feedback Feedback

Submit feedback, report a bug, or share your conversation.

27