Format prompts for different LLM providers with chat templates and HNSW-powered context retrieval
Install
npx skills add https://github.com/ruvnet/ruflo --skill chat-formatSKILL.md
Chat Format
Format prompts for multi-provider LLM inference with context retrieval.
When to use
When preparing prompts for different LLM providers (Claude, GPT, Gemini, Ollama) or building RAG pipelines with HNSW-powered context retrieval.
Steps
- Format chat — call
mcp__plugin_ruflo-core_ruflo__ruvllm_chat_formatwith messages and target provider - Create HNSW index — call
mcp__plugin_ruflo-core_ruflo__ruvllm_hnsw_createfor context retrieval - Add documents — call
mcp__plugin_ruflo-core_ruflo__ruvllm_hnsw_addto index documents - Route query — call
mcp__plugin_ruflo-core_ruflo__ruvllm_hnsw_routeto find relevant context - Check status — call
mcp__plugin_ruflo-core_ruflo__ruvllm_statusfor provider availability
Supported providers
- Anthropic (Claude) — native format
- OpenAI (GPT) — chat completion format
- Google (Gemini) — generative AI format
- Ollama — local model format
- Cohere — generate/chat format
Related skills
find-skillsvercel-labs3.6MHelps users discover and install agent skills when they ask questions like "how do I do X", "find a skill for X", "is there a skill that can...", or express interest in extending capabilities. This skill should be used when the user is looking for functionality that might exist as an installable skill.handoffmattpocock883KCompact the current conversation into a handoff document for another agent to pick up.microsoft-foundrymicrosoft618KBuild, deploy, evaluate, optimize, fine-tune, and manage Microsoft Foundry agents, models, and resources end to end. USE FOR: foundry, azd ai agent, azd provision/deploy, hosted agent scaffold/develop/run/deploy/troubleshoot, prompt agent create, create agent, update agent, add tool to agent, invoke agent, agent.yaml, agent insights, pull agent insights, evaluate agent, batch eval, continuous eval, continuous monitoring, agent CI/CD, optimize prompt, improve prompt, prompt optimizer, optimize agcavemanjuliusbrussee544KUltra-compressed communication mode that cuts output tokens while keeping technical accuracy. Levels: lite, full, ultra and the wenyan variants. Use for /caveman, "caveman mode", "talk like caveman", "be brief" or "less tokens".