MCPMarket
Sell SkillsPower Your AgentsConnect
  1. Home
  2. Servers
  3. Llama2 WebUI

Llama2 WebUI

liltom-ethbyliltom-eth
•
1.9k
•
API Development
Developer Tools
Data Science & ML

Provides a Gradio web interface for locally running various Llama 2 models on GPU or CPU across different operating systems.

Related MCPs

View more
  • neondatabase-labs

    Neon

    Enables natural language interaction with the Neon Management API and databases through the Model Context Protocol.

  • mendableai

    Firecrawl

    Empowers LLMs with advanced web scraping capabilities for content extraction, crawling, and search functionalities.

  • GLips

    Figma Context

    Provides AI coding agents with simplified Figma layout information via the Model Context Protocol.

Related Skills

View all
  • openclaw

    Diagram Maker & Visualizer

    Generates professional SVG, HTML, and Excalidraw diagrams for software architecture, system flows, and educational concepts.

  • openclaw

    GH Issues Auto-Fixer

    Automates the end-to-end GitHub issue lifecycle by spawning sub-agents to implement code fixes, open pull requests, and resolve review comments.

  • openclaw

    Discord Integration

    Manages Discord operations including messaging, reactions, and channel management directly through Claude.

MCPMarket

Discover MCP servers that connect MCP clients like Claude and Cursor to your favorite tools. Browse the MCP Market to get started.

Browse

  • MCP Search
  • MCP Servers
  • MCP Clients
  • Agent Skills
  • MCP Market Hub
  • Categories
  • What is an MCP server?
  • Model Context Protocol

Rankings

  • Top MCPs Today
  • Top Agent Skills Today
  • Top 100 Agent Skills
  • Top 100 MCP Servers

About

  • News
  • Submit
  • Contact

© 2026 MCP Market. All rights reserved.·Privacy·Terms

Llama2 WebUI offers a user-friendly Gradio web interface designed for seamless local execution of Llama 2 models. It supports a wide range of Llama 2 variants, including 7B, 13B, 70B, GPTQ, GGML, and GGUF, and integrates with various backends like transformers, bitsandbytes, AutoGPTQ, and llama.cpp for optimized GPU or CPU inference. Developers can leverage `llama2-wrapper` as a powerful local Llama 2 backend for building generative agents and applications, or utilize its OpenAI-compatible API for broader integration. The tool is compatible with Linux, Windows, and Mac, making it accessible for diverse development and experimental setups.

Key Features

01Multiple model backends including transformers, bitsandbytes, AutoGPTQ, and llama.cpp.
02Offers an OpenAI-compatible API for Llama 2 models, enabling use with existing clients.
031,958 GitHub stars
04Provides `llama2-wrapper` for seamless integration as a local Llama 2 backend for generative agents/apps.
05Cross-platform compatibility for running on GPU or CPU across Linux, Windows, and Mac.
06Supports all Llama 2 models (7B, 13B, 70B, GPTQ, GGML, GGUF, CodeLlama) with 8-bit and 4-bit inference.

Use Cases

01Developing and integrating generative AI applications using Llama 2 as a local backend.
02Benchmarking Llama 2 model performance on various local hardware configurations.
03Running Llama 2 models locally for chat or code completion via a web-based user interface.