LLM Mart Trusted

@llmmart · Joined Jun 2026

0 Followers 1003 Reputation 164 Contributions
Claude Skill Adversarial code-review skill for Claude

A Claude skill that reviews a diff by trying to refute each finding before reporting it.

43
Gemini Prompt Grounded research prompt for Gemini

A prompt that forces Gemini to cite sources and flag uncertainty instead of guessing.

34
DeepSeek Workflow Reasoning-first debugging workflow for DeepSeek

A workflow that makes DeepSeek state a hypothesis and a test before proposing any fix.

29
Cursor Recipe Refactor-with-tests recipe for Cursor

A Cursor recipe that writes a characterization test before each refactor step.

19
MiniMax Skill Roleplay-agent skill for MiniMax

A skill that keeps a consistent persona and memory across a long MiniMax conversation.

15
Claude Skill Conventional-commit writer for Claude

Turns a staged diff into a clean Conventional Commits message — correct type, tight subject, useful body.

2
ChatGPT Workflow Source-grounded research brief for ChatGPT

Uses ChatGPT search plus an adversarial verification pass to turn a fuzzy question into a sourced brief you can hand to someone else.

0
Codex CLI Skill Review-before-apply skill for Codex CLI

Forces Codex CLI to show the diff, review its own changes, run focused validation, and list uncertainty before you keep the patch.

0
GitHub Copilot Skill Code review guardrails for GitHub Copilot

A review skill that makes Copilot focus on concrete behavior changes, missing tests, and repo instructions instead of generic style chatter.

0
GitHub Copilot Template Repository instructions bootstrap for GitHub Copilot

A starter layout for .github/copilot-instructions.md, path-specific instruction files, and AGENTS.md so Copilot stops guessing how your repo works.

0
GitHub Copilot Recipe Test generation loop for GitHub Copilot

Use /setupTests, /tests, and /fixTestFailure in a red-green loop instead of asking Copilot to “add some tests” and hoping for the best.

0
Cursor Recipe Reproduce-first bug hunt for Cursor

Forces a failing test that reproduces the bug before any fix — so you know it's actually fixed.

0
ChatGPT Template Custom GPT acceptance harness

A reusable test harness for checking whether a custom GPT actually follows its instructions, uses tools correctly, and fails safely.

0
Claude Skill Security-review skill for Claude

A structured security pass over a diff — checks the OWASP-relevant classes, ranks by exploitability, and shows a repro.

0
Claude Workflow Spec-first feature build (workflow)

Stops the model from coding too early — lock a short spec and a test list before a single line is written.

0
Point your agent at the catalogue: the LLM Mart MCP server and API

The whole public catalogue is an MCP server and a REST API, so your agent can search skills, tools and slash-commands as native tools. Setup is one config block. Plus a private vault that carries your own prompts between machines.

agents coding security
Aug 22
Ship a prompt like code: a five-case eval harness you can build in an hour

You wouldn't ship a function you ran once. Here's the smallest evaluation setup that catches real regressions — five cases, three graders, and a rule for when to add a sixth.

writing data prompt-engineering
Jun 29
The instruction file is the highest-leverage file in your repo

AGENTS.md, CLAUDE.md, copilot-instructions.md and .cursor/rules all solve the same problem. Most of them are written badly. Here's what changes agent behaviour and what's decoration.

agents coding prompt-engineering
Jun 26
Prompt injection is not an XSS problem

Sanitizing output protects your page. It does nothing for an agent that reads a poisoned README and then runs a command. A practical model of the threat, and what actually helps.

code-review agents security
Jun 23
Prompt, skill, workflow, recipe, template, agent: a working taxonomy

The words are used interchangeably and they shouldn't be. Six categories, what actually distinguishes them, and how to tell which one you're holding.

writing agents prompt-engineering
Jun 20
How to write a prompt you can reuse

Most prompts are lucky one-offs. The ones that keep working share a structure: an explicit role, a checklist of constraints, a worked example, and a self-check step. Here's how to build one.

Jun 18
/ROLE: TASK: FORMAT: Role / task / format

Combine role, task, and output format in a single instruction line.

4
/feedback <message> Feedback

Send feedback to the Cursor team.

4
/model Switch model

Switch the AI model and save it as your default for new sessions.

3
/ide IDE integrations

Manage IDE integrations and show status.

3
/status Show status

Open Settings on the Status tab: version, model, account, connectivity.

3
/summarize Summarize

Condense lengthy text or documents into key points.

3
/permissions Permissions

Adjust what Codex can do without asking first, such as switching between Auto and Read Only approval modes.

3
/status Session status

Inspect the active model, approval policy, writable roots, and token usage.

3
/edit Edit selection

Ask Cursor to edit the currently selected code inline.

3
/clear Clear chat

Start a new chat session in Copilot Chat.

3
/TONE: [tone] Set tone

Set the tone of voice, e.g. friendly or professional.

2
/show-thinking Show thinking

Show or hide thinking blocks.

2
/focus Focus view

Toggle the focus view showing only your prompt, a tool summary, and the response.

1
/security-review Security review

Analyze pending changes on the branch for security vulnerabilities.

1
/side Side conversation

Open an ephemeral side conversation; /btw is an alias for the same workflow.

1
/SCHEMA Structured data output

Output structured data in a defined schema.

0
/END WITH: "text" End with

Make the answer end with a specific phrase.

0
/fork Fork chat

Duplicate the current chat as a new session.

0
/agents Configure agents

Configure your custom agents.

0