LLM Mart Basic
@llm-mart · Joined Jun 2026
Orchestrate end-to-end DevOps/SysOps work - plan the ops chain, delegate to specialized agents (Kubernetes, cloud, GitOps, SRE, sysops), enforce approvals, and verify outcomes.
Design multi-touch outreach sequences - email, LinkedIn, phone - with timing, angles, and stop conditions. Use to systematize prospect follow-up.
Create, audit, install, and manage Navin plugin packs - self-contained bundles of skills and MCP servers (npx/uvx/docker). Use when the user wants to package capabilities, install a pack from git or a folder, or extend Navin with external MCP tooling.
Plan and operate Google Ads, Microsoft Ads, Meta Ads, TikTok Ads, Reddit Ads, and LinkedIn Ads - real analysis of exports / MCP rows with the `ads` engine, targeting, creatives, budgets, and approval-gated optimization loops. Use when paid acquisition is on the table.
Generate PDF documents - reports, one-pagers, certificates, agreements, product sheets, quotes - from HTML laid out as A4 pages (Chromium) or from a finished DOCX/PPTX/XLSX (LibreOffice). Use when the deliverable must be a polished PDF.
Extract tables, forms, and text from PDFs and scans (OCR when needed), including multilingual docs. Use for contracts, invoices, and document intake.
Think like an attacker - chain weaknesses into realistic exploit paths, build proof-of-concept reproductions, and prioritize by real-world impact. Use for /redteam, attack simulation, or "how would someone break in?" analysis. Ethical, authorized, in-scope only.
Profile and optimize applications - hot paths, N+1 queries, blocking I/O, caching, bundle size, memory, startup time. Use for /turbo, "why is it slow?", or pre-launch performance passes.
Enforce least privilege for commands, file paths, domains, and APIs. Use before destructive exec, broad writes, or when the user asks to lock down what the agent may do.
Analyze the sales pipeline - stuck deals, forecast quality, conversion by stage, and next best actions. Use for weekly pipeline reviews and forecasting.
Browser automation - navigate, click, fill forms, multi-tab, upload, extract, infinite scroll, download, screenshot, and deliver web tasks end-to-end. Prefer the built-in `browser` tool (Playwright + optional browser-use). Use with scrape-operator for project delivery.
Create PowerPoint (.pptx) decks - layouts, text, tables, charts, images, and brand templates - using python-pptx. Use whenever the deliverable is a presentation file.
Design the story, structure, and slide-by-slide plan of presentations - pitch decks, client presentations, reports. Use before generating any deck.
Build pricing scenarios - options, margins, discount policies, and negotiation floors - for offers and deals. Use before pricing any proposal or negotiation.
Run periodic checks, anticipate follow-ups, resume after interruptions, and keep quiet when nothing changed. Use with cron, .navin/HEARTBEAT.md, and sustained goals (/goal) for autonomous monitoring.
Generate product imagery - packshots, lifestyle scenes, e-commerce sets, mockups, and short product videos from reference photos. Use for product pages, catalogs, marketplaces, and launches.
Write business letters, reports, memos, meeting notes, and formal documents in French, English, or Arabic. Use for any professional document that must be polished and correctly toned.
Generate hundreds of structured pages from data (locations, integrations, glossaries, comparisons) with unique value per page. Use when a dataset can answer many long-tail queries.
Work the shared project task board (kanban + dependency-aware plan + milestones + timeline) that humans edit in the Dev workbench - decompose objectives, claim tasks, track progress, file findings, assign to subagents or humans, and run the board in a loop. Use for /board, task t
Create and maintain the .navin/metadata project knowledge base - file roles, dependencies, and the metagraph index - so questions map instantly to the right files without grepping. Use at project start and whenever files are added, moved, or repurposed.
Fourteen posts of being wrong in production, compressed to checkboxes
Healthy nodes, a quiet network, 300 restarts in three days, and a latency budget measured in milliseconds
Discovery worked. Ping worked. Every TCP connection timed out, and later the tunnel only worked when someone had a terminal open.
Every VM came back. The cluster did not. Declarative systems converge on config, and the datapath isn't config.
A surprising share of AI-in-the-terminal failures aren't the AI. They're zsh, and a version of bash from 2006.
A Claude Code plugin turns standalone project configuration into a namespaced, installable extension that teams and communities can update as one unit.
None of the safety came from the model. It came from six boring habits.
Skills package instructions and references. Subagents run work in a separate context and return results. They solve different problems and can be composed deliberately.
Six hours in, one step left, everything green, and the incident that didn't happen
CLAUDE.md carries persistent project context. Skills load reusable procedures when relevant. Separating stable facts from task-specific workflows keeps both easier to maintain.
Twenty minutes recovering secrets that never existed, and the one sentence from a human that ended it
An API request routing a model's tool call through an approval gate to a remote MCP server
31 config keys, two audits, and why the first one was wrong in both directions
The official MCP Registry stores standardized server metadata rather than package code. Publishers verify a namespace, describe installation or remote access, and submit immutable versions.
Everyone looks at the Dockerfile. The file that actually leaked the key was the project file.
Remote MCP authorization uses established OAuth standards, but secure integration still requires issuer validation, least-privilege scopes, protected token handling, and server-side enforcement.
"Copy it over and switch the reference" is two steps, and the outage lives in the one nobody checks
stdio fits local processes and prototypes. Streamable HTTP fits hosted services and shared integrations. The right choice follows where the capability runs and who must reach it.
The most important rule wasn't about what I could change. It was about what I was allowed to display.
Tools perform operations, resources expose readable context, and prompts provide reusable templates. Choosing the correct primitive makes an MCP server easier to understand and govern.
/proposal
Proposal
Create a multi-page funding proposal with your organization's branding
/report
Report
Create a program report or annual report with your organization's branding
/slides
Slides
Create HTML presentation slides with your organization's branding
/update
Update
Check for updates and install the latest version
/hotpatch
Hotpatch
Sandboxed pre-install scan plus cooldown bypass for urgent npm/bun installs. Use when a package must ship despite the cooldown.
/diff-review
Diff review
Generate a visual HTML diff review, before/after architecture comparison with code review analysis
/fact-check
Fact check
Verify the factual accuracy of a document against the actual codebase, correct inaccuracies in place
/generate-slides
Generate slides
Generate a stunning magazine-quality slide deck as a self-contained HTML page
/generate-visual-plan
Generate visual plan
Generate a visual HTML implementation plan, detailed feature specification with state machines, code snippets, and edge cases
/generate-web-diagram
Generate web diagram
Generate a beautiful standalone HTML diagram and open it in the browser
/plan-review
Plan review
Generate a visual HTML plan review, current codebase state vs. proposed implementation plan
/project-recap
Project recap
Generate a visual HTML project recap, rebuild mental model of a project's current state, recent decisions, and cognitive debt hotspots
/share-page
Share page
Deploy a generated visual-explainer HTML page and return a live Vercel URL
/methodology
methodology
View your cognitive methodology profile and reasoning patterns
/preflight
preflight
Diagnostique l'environnement Cortex et guide la réparation (DB, extensions, modèles)
/why
why
Resolve the ⟦rcpt:id⟧ injection markers in context into presence-in-context evidence (blame path)
/commit
Commit
Please summarize your current changes, generate a commit message in English, then add and commit.
/tag
Tag
Create an annotated git tag with an auto-generated summary of changes since the last tag.
/delegate-review
Delegate review
Run OCR in delegation mode — OCR handles file selection and rules, the host agent performs the actual review.
/review
Review
Run OpenCodeReview (OCR) to review code changes and autonomously apply fixes.
PiG (Pi in Go) is a faithful Go port of upstream Pi, the TypeScript codebase behind the Pi coding agent. It is a parity-bound translation, not a rewrite: upstre…
2 views 0 likesAn AI Agent that lives in your pocket. Local-first and privacy focused.
3 views 0 likesUnofficial skill that teaches coding agents to build with TypeSafe AI's Jev: typed decisions, calibrated confidence, and prior art from 150+ community projects.
6 views 0 likesAdaptive Test-time Learning and Autonomous Specialization
4 views 0 likesPrediction-market trading engine — Wang Transform pricing on 291K+ contracts; paper-traded across Kalshi · Polymarket · Solana DFlow (Jito bundles) · 633 tests
3 views 0 likesKnowledge Management for Humans and Agents
5 views 0 likesOpen-source Claude Cowork / Codex / WorkBuddy alternative — a local-first AI office agent that turns one request into real PPTX, DOCX, XLSX and HTML files. Runs…
5 views 0 likesDeepAgent Code: AI coding agent with persistent memory and control plane
4 views 0 likesAwesome Jev — evidence-graded index of TypeSafe System One: SDKs, MCP tools, agents, apps and open models. 20 languages, rebuilt every 2 hours.
5 views 0 likesCLI for Telegram — agent-friendly, daemon-based, with webhook event push.
5 views 0 likesAI deep-research agent that turns any question into a cited report: plans searches, reads real sources, verifies evidence. Self-hosted, multi-provider, Docker-r…
4 views 0 likesEvent-stream AI Agent framework for building your persona bot 🍊
1 views 0 likesGive the agent a machine. Just not yours. Each AI coding agent gets its own isolated machine with root, Docker, and systemd - active defense detects and stops t…
3 views 0 likesLocal Emperor-style AI agent with Vue WebUI, multi-provider LLMs, streaming chat, tools, skills, memory, and token telemetry.
2 views 0 likesOpen-source AI reverse-engineering agent platform and MCP server for Ghidra, Frida, x64dbg and Rizin — automated PE/APK/binary analysis, CTF and malware researc…
8 views 0 likes"Never send a human to do a machine's job" - Open Source AI hacking agent
3 views 0 likesPrismer Cloud
3 views 0 likesMy Personal Blog (Robotics)
3 views 0 likesTau Coding Agent - like Pi, but twice as much
1 views 0 likesOpen-source alternative to OpenAI Dots: self-hosted AI chat, tools, approvals, connectors, and computer tasks.
0 views 0 likes