Works with the AI tools you already use

    Claude CodeClaude CodeCursorCursorCodex CLICodex CLIGitHub CopilotGitHub CopilotGemini CLIGemini CLI+20 more

    model routing

    1

    A multi-tier routing system that manages local and remote models across diverse GPU hardware and scheduled renders.

    Free

    0 installsSecurity scanned
    model routing

    model routing

    Example session with this skill installed

    Show me the current routing roster and verify if the data-analyst profile is correctly assigned to the MAIN-NODE ornith:9b instance.

    • Read your context and instructions
    • Compiled the model routing
    • Generated the data export

    The lab's model routing matrix is active across three physical nodes. The data-analyst profile is correctly pointed to the ornith:9b model hosted on MAIN-NODE (local IP) to preserve GPU-NODE VRAM for pipeline reasoning.

    Current Profile Roster

    ProfileActive ModelHostTierContext
    coderqwen3.8:27b-132kGPU-NODET2192K
    researcherqwen3.8:27b-132kGPU-NODET2192K
    architectornith:35b-q4_K_MGPU-NODELOCAL65K
    builderornith:9bMAIN-NODET2.564K
    data-analystornith:9bMAIN-NODEMID64K
    reviewerqwen3.5:4b-q4_K_MAUX-NODEAUX64K
    copywriterqwen3.5:4b-q4_K_MAUX-NODEAUX64K
    devopsqwen3.5:4b-q4_K_MAUX-NODEAUX64K

    Verification Status:

    • MAIN-NODE Connectivity: Verified (HTTP 200).
    • GPU-NODE Status: Available (Outside of 18:00 render window).
    • AUX-NODE Health: Active (Auxiliary tasks routed via 2b-aux).

    Next steps

    1. Run hermes fallback list to ensure OpenRouter remains the active T3 failover for the coder profile.
    2. Monitor GPU-NODE VRAM during the 18:00 NZT render slot to ensure qwen3.8 unloads properly.
    3. Validate the data-analyst CSV extraction performance using the native /api/chat endpoint.

    model-routing.csv

    CSV · data export

    Generated

    Example file from a real run - the skill writes it into your workspace.

    Connects securely to your tools. The creator never sees your data.

    About this skill

    The problem

    Running advanced agentic workflows across local hardware often leads to VRAM fragmentation, context window collisions, and unmanaged API costs. Developers struggle to maintain high-performance local coding models while simultaneously running GPU-intensive tasks like ComfyUI renders.

    What it does

    • Routes tasks through a three-tier system: Local Routine (27B), Agentic Coder (9B), and Remote Premium fallback.
    • Enforces render-exclusive windows to prevent OOM errors during scheduled GPU tasks.
    • Satisfies Hermes 64K context floors by managing baked Modelfile parameters across different node specs.
    • Offloads lightweight internal operations (triage, title gen, approvals) to auxiliary nodes to save primary GPU compute.

    Frameworks & tools

    Ollama, OpenRouter, Hermes, RTX 3060 Dual-GPU (24GB), Windows WSL, and Python-based agent environments.

    Why this beats prompting it yourself

    Manual model selection fails during concurrent agent runs and hardware-heavy render cycles. This skill handles the complex logic of VRAM fitting, port-forwarding across network nodes, and specific OpenAI-compatible header requirements for local LLMs.

    Use cases

    • Executing repo-level bug fixes on a secondary node while the main GPU renders high-resolution video.
    • Maintaining a 192K context window for routine health checks without burning through paid API tokens.
    • Scaling auxiliary agent functions to low-spec hardware (4GB VRAM) using specific context bakes.
    • Falling back to high-reasoning remote models only when local tiers reach saturation.

    Known limitations

    Ollama /v1 endpoints ignore runtime context parameters, requiring context floors to be baked directly into Modelfiles. SSH access to auxiliary nodes is restricted to HTTP API calls only.

    How to install

    Works the same in every agent - Claude, Cursor, Codex, Copilot and 20+ more.

    ~30 seconds
    1. 1

      Download the ZIP

      Free skills download straight away. Paid skills unlock right after purchase.

    2. 2

      Unzip into your skills folder

      Every agent reads skills from one folder on your machine. Drop the unzipped folder in there.

    3. 3

      Ask your agent to use it

      Restart the agent if it was already running. It picks the skill up automatically - no config needed.

    Skills folder by agent

    Click the path to copy it. Create the folder if it does not exist yet.

    Reviews

    No reviews yet

    Be one of the first to try it. Every listed skill passes our trust checks below.

    Security scanned

    Passed our 8-point scan before listing

    Fresh listing

    Recently published to Agensi

    Free forever

    No account required to browse

    Trust & safety

    Security scanned

    Verified clean 6 days ago

    • Free to download with an account

    Listed6 days ago

    What's inside

    Frequently Asked Questions