Favicon of llama-cpp

llama-cpp Skill

AI Agent SkillPythonOpen source

llama.cpp local GGUF inference + HF Hub model discovery. Published by NousResearch in hermes-agent.

Decision snapshot

Is this a fit?

Best for

Research, Includes SKILL.md, Reusable instructions

Works with

Compatibility not yet detected.

Access

Permission behavior not yet detected.

Setup

Copy skill directory

Project health

22 days ago · MIT license

Considerations

No specific cautions were detected. Review the source and requested permissions before installing.

What is llama-cpp Skill?

llama.cpp local GGUF inference + HF Hub model discovery. Published by NousResearch in hermes-agent. This profile combines repository metadata with install, compatibility, and usage signals so developers can quickly decide whether it fits their agent workflow before opening the source repository.

Trust signal
95/100
Maintenance signal
90/100
Adoption signal
100/100

Automated repository signals based on public metadata such as recency, license, installation evidence, and adoption. These are not a security audit or endorsement. See how SkillIndex evaluates profiles.

Key capabilities

  • Includes SKILL.md support
  • Reusable instructions support
  • Research
  • Research use cases

Declared skill metadata

  • Declared author: Orchestra Research
  • Declared license: MIT
  • Source file: skills/mlops/inference/llama-cpp/SKILL.md

These fields retain source and confidence evidence from the indexed SKILL.md.

Compatibility and setup

Copy skill directory
  • Install or run with Copy skill directory

When to use llama-cpp Skill

  • Use it for research.

Built with

PythonCopy skill directory

Editorial notes

Source

  • Creator: NousResearch
  • Repository: NousResearch/hermes-agent
  • Skill file: skills/mlops/inference/llama-cpp/SKILL.md

What it does

llama.cpp local GGUF inference + HF Hub model discovery.

Skill instructions

llama.cpp + GGUF Use this skill for local GGUF inference, quant selection, or Hugging Face repo discovery for llama.cpp. When to use - Run local models on CPU, Apple Silicon, CUDA, ROCm, or Intel GPUs - Find the right GGUF for a specific Hugging Face repo - Build a llama-server or llama-cli command from the Hub - Search the Hub for models that already support llama.cpp - Enumerate available .gguf files and sizes for a repo - Decide between Q4/Q5/Q6/IQ variants for the user's RAM or VRAM Model Discovery workflow Prefer URL workflows before asking for hf, Python, or custom scripts. 1. Search for candidate repos on the Hub: - Base: https://huggingface.co/models?apps=llama.cpp&sort=trending - Add search=<term for a model family - Add numparameters=min:0,max:24B or similar when the user has size constraints 2. Open the repo with the llama.cpp local-app view: - https://huggingface.co/<repo?local-app=llama.cpp 3. Treat the local-app snippet as the source of truth when it is visible: - copy th

Verified compatibility and discovery

Frequently asked questions

What is llama-cpp?

llama-cpp is a open-source AI agent skill with Copy skill directory. llama.cpp local GGUF inference + HF Hub model discovery.

Who is llama-cpp best for?

llama-cpp is best for reusing agent instructions, scripts, and references, research workflows.

How do I install llama-cpp?

Install or run llama-cpp using Copy skill directory. Check llama-cpp for the latest setup command.

Is llama-cpp actively maintained?

llama-cpp may need a closer maintenance check before production use.

Share:

Stars
232,138
Forks
46,253
Last commit
22 days ago
Last verified
Aug 19, 2026
Metadata fetched
Aug 19, 2026
Repository age
1 year
License
MIT

Project health auto-fetched from the source repository.

Maintain this resource?

Review this source-backed profile, send a correction with evidence, or link to it from your documentation. Claims verify your relationship to the project; profile facts still require source evidence and editorial review.

Alternatives to llama-cpp