All topics

llm-webui

Projects tagged with llm-webui on GitHub.

77 projects

open-webui logo

open-webui

Python
100

User-friendly AI Interface (Supports Ollama, OpenAI API, ...)

150.6k+9040Jan 21, 1970
browser-use logo

browser-use

Python
100

🌐 Make websites accessible for AI agents. Automate tasks online with ease.

111.6k+7640Jan 21, 1970
llm-space logo

llm-space

TypeScript
95

A desktop app to prototype agent ideas, inspect every harness step, replay failures, and evaluate performance, all in one place. Local-first, cloud-ready for managed agents.

1.7k+1300Jan 21, 1970
awesome-llm-apps logo

awesome-llm-apps

Python
100

100+ AI Agents, Agent Skills and RAG Apps - Free and Open Source.

134.7k+1.5k0Jan 21, 1970
local-llm logo

local-llm

Shell
70

Everything I know about running LLMs locally

1.8k+200Jan 21, 1970
llama.cpp logo

llama.cpp

C++
95

LLM inference in C/C++

126.4k+1.1k0Jan 21, 1970
apex-inference-chip logo

apex-inference-chip

Python
68

An inference chip design that runs a real LLM (Qwen2.5-0.5B) on FPGA — one transformer decoder layer in RTL, every silicon value bit-exact against a golden model. 0.56 tok/s measured, a 140× climb, full evidence trail.

661+490Jan 21, 1970
ctrlb-decompose logo

ctrlb-decompose

Rust
77

LLM-ready reasoning surface over logs

290+10Jan 21, 1970
Investbrain logo

Investbrain

PHP
82

Smart LLM-enabled investment tracker that consolidates and monitors market performance across your different brokerages

9240Jan 21, 1970
chorus logo

chorus

TypeScript
82

Multi-LLM peer review for code decisions. Bring your own CLI; Chorus convenes 2-4 other LLMs to review the work before you ship.

5270Jan 21, 1970
pi logo

pi

TypeScript
100

AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI

98.8k0Jan 21, 1970
shard logo

shard

Python
72

Pipeline-parallel LLM inference across GPUs on separate machines.

447+40Jan 21, 1970
voidllm logo

voidllm

Go
72

Privacy-first LLM proxy and AI gateway - load balancing, multi-provider routing, API key management, usage tracking, rate limiting. Self-hosted. Zero knowledge of your prompts.

128+10Jan 21, 1970
bindwidth logo

bindwidth

JavaScript
72

Evidence-aware on-prem LLM inference sizing and TCO calculator

123+30Jan 21, 1970
tokenspeed logo

tokenspeed

Python
90

TokenSpeed is a speed-of-light LLM inference engine.

2.0k+560Jan 21, 1970
macos-harness logo

macos-harness

Python
77

The simplest, thinnest harness that gives an LLM complete freedom to control a Mac.

7890Jan 21, 1970
MLX-LoRA-Studio logo

MLX-LoRA-Studio

Swift
72

A native Mac App for LLM fine-tuning on Apple Silicon — fully on-device, fully open source.

258+10Jan 21, 1970
trueforge logo

trueforge

TypeScript
90

The open-source agent harness - the runtime layer that turns an LLM into a working agent.

4.4k0Jan 21, 1970
inference-school logo

inference-school

Swift
77

A hands-on Swift and Metal course for building LLM inference from first principles on Apple silicon, with 48 guided lessons, runnable exercises, a native macOS Studio, and a complete companion book.

185+10Jan 21, 1970
vomit logo

vomit

Go
68

Clean up Claude's token vomit with a separate LLM. Save your tokens, Opus is hopeless

1810Jan 21, 1970
Photo-agents logo

Photo-agents

Python
72

Autonomous self-evolving agents. Vision-grounded layered memory and self-written skills for LLM agents that operate your computer.

6980Jan 21, 1970
feeds.fun logo

feeds.fun

Python
67

News reader with tags, scoring, and LLM

391+50Jan 21, 1970
harness logo

harness

Swift
72

AI-driven user testing for iOS Simulator, macOS apps, and web apps. Write a goal in plain language; an LLM agent drives the UI and reports friction. macOS 14+, Swift 6.

343+200Jan 21, 1970
deja-vu logo

deja-vu

Go
82

Search your past AI coding sessions — Claude Code, Codex, Cursor and 17 more. Indexes the session history they already wrote to disk, including months from before you installed it, and recalls it in any of them. No LLM, no embeddings, one local Go binary.

746+640Jan 21, 1970
Soup logo

Soup

Python
90

Fine-tune LLMs from one YAML. Layer streaming trains an 8B model on a 4 GB laptop GPU.

3.0k+2.0k0Jan 21, 1970
ollama logo

ollama

Go
95

Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.

179.9k+5300Jan 21, 1970
langchain logo

langchain

Python
100

The agent engineering platform.

145.4k+4600Jan 21, 1970
deltafin logo

deltafin

Rust
72

Run full Kimi K3 on a single device. And an OpenAI-compatible API server for local chat and coding agents.

781+160Jan 21, 1970
ratel logo

ratel

TypeScript
77

Context engineering for AI agents. ~80% fewer tokens. Fix tool overload. Skills and memory with in-process BM25 and semantic retrieval. Progressive Disclosure. No vector DB.

428+50Jan 21, 1970
spec-ptc logo

spec-ptc

Python
72

Speculative programmatic tool calling (sPTC) for harnesses like RLM, CodeAct, etc.

1820Jan 21, 1970
market-pilot logo

market-pilot

JavaScript
63

Evidence-grounded market research prototype with traceable AI workflows.

830Jan 21, 1970
jarvis logo

jarvis

Shell
68

Jarvis meta-repository: shared docs, configs, and setup scripts

66+20Jan 21, 1970
LocalAI logo

LocalAI

Go
100

LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.

48.8k+1240Jan 21, 1970
unlazy logo

unlazy

JavaScript
95

Anti-laziness skill for AI agents. Core: the Depth Tree method, which splits a task N layers deep and gives every leaf the full time budget of the whole task, so effort multiplies with depth. Grounded in 2025-2026 research on model laziness, underthinking and premature completion.

2.8k+6420Jan 21, 1970
OptMem logo

OptMem

Python
72

Permanent memory for AI agents. A 426-token prompt, a script, plug and play.

1.5k+2000Jan 21, 1970
waggle logo

waggle

Rust
67

Attributed, resolvable artifact references for agent handoffs — a ~30-byte token instead of pasted context. MCP-native; the reference layer for the agent-harness world.

6800Jan 21, 1970
appless logo

appless

TypeScript
72

What if your phone had no apps

387+50Jan 21, 1970
rome logo

rome

TypeScript
77

Rome is the agentic OS.

3870Jan 21, 1970
evonic logo

evonic

Python
82

Open Agentic AI Platform - The home your agents deserve

3120Jan 21, 1970
AI-Engineering-Lab logo

AI-Engineering-Lab

Jupyter Notebook
75

A free, self-paced 24-week AI engineering course: Python, machine learning, LLMs, RAG, fine-tuning, agents and MCP, Azure and Vertex and Bedrock, and Databricks. 43 runnable notebooks, one continuous case study. MIT licensed, no signup. By Zorost Intelligence AI Lab.

2150Jan 21, 1970
lora-speedrun logo

lora-speedrun

Python
73

Speedrunning LoRA fine-tuning: frozen task, frozen hardware, public wall-clock leaderboard. modded-nanogpt for fine-tuning.

147+20Jan 21, 1970
CodeJury logo

CodeJury

Python
72

Terminal-first, knowledge-grounded multi-agent software delivery pipeline: scope requirements, implement changes, run tests, and gate pull requests with deterministic QA and ensemble code review.

1420Jan 21, 1970
ambient-context logo

ambient-context

Rust
73

A menu bar app that keeps a written record of what you worked on.

1400Jan 21, 1970
ECC logo

ECC

JavaScript
100

The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.

245.6k+2.1k0Jan 21, 1970
transformers logo

transformers

Python
95

🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.

164.7k+2690Jan 21, 1970
MoneyPrinterTurbo logo

MoneyPrinterTurbo

Python
100

利用 AI 大模型和自动化工作流,根据主题或关键词一键生成高清短视频。Generate HD short videos from a topic or keyword with an automated AI workflow.

116.5k+7.1k0Jan 21, 1970
ponytail logo

ponytail

JavaScript
100

Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.

115.5k+9.3k0Jan 21, 1970
ragflow logo

ragflow

Go
95

RAGFlow is a leading open-source Retrieval-Augmented Generation (RAG) engine that fuses cutting-edge RAG with Agent capabilities to create a superior context layer for LLMs

89.7k+4770Jan 21, 1970