llm-agent
Projects tagged with llm-agent on GitHub.
77 projects
llm-space
TypeScriptA desktop app to prototype agent ideas, inspect every harness step, replay failures, and evaluate performance, all in one place. Local-first, cloud-ready for managed agents.
Photo-agents
PythonAutonomous self-evolving agents. Vision-grounded layered memory and self-written skills for LLM agents that operate your computer.
harness
SwiftAI-driven user testing for iOS Simulator, macOS apps, and web apps. Write a goal in plain language; an LLM agent drives the UI and reports friction. macOS 14+, Swift 6.
macos-harness
PythonThe simplest, thinnest harness that gives an LLM complete freedom to control a Mac.
forsy-trace-skill
PythonOpen skill for capturing AI agent work as structured traces.
ai-copywriter
PythonAn AI copywriter that uses real copywriting skills + real marketing knowledge with human tone.
ambient-context
RustA menu bar app that keeps a written record of what you worked on.
awesome-llm-apps
Python100+ AI Agents, Agent Skills and RAG Apps - Free and Open Source.
pi
TypeScriptAI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI
trueforge
TypeScriptThe open-source agent harness - the runtime layer that turns an LLM into a working agent.
hermes-agent
PythonThe agent that grows with you
labs-OO-Agents
PythonNVIDIA Object Oriented Agents: the Pythonic way to build AI Agents.
langchain
PythonThe agent engineering platform.
ratel
TypeScriptContext engineering for AI agents. ~80% fewer tokens. Fix tool overload. Skills and memory with in-process BM25 and semantic retrieval. Progressive Disclosure. No vector DB.
unlazy
JavaScriptAnti-laziness skill for AI agents. Core: the Depth Tree method, which splits a task N layers deep and gives every leaf the full time budget of the whole task, so effort multiplies with depth. Grounded in 2025-2026 research on model laziness, underthinking and premature completion.
OptMem
PythonPermanent memory for AI agents. A 426-token prompt, a script, plug and play.
waggle
RustAttributed, resolvable artifact references for agent handoffs — a ~30-byte token instead of pasted context. MCP-native; the reference layer for the agent-harness world.
rome
TypeScriptRome is the agentic OS.
evonic
PythonOpen Agentic AI Platform - The home your agents deserve
AI-Engineering-Lab
Jupyter NotebookA free, self-paced 24-week AI engineering course: Python, machine learning, LLMs, RAG, fine-tuning, agents and MCP, Azure and Vertex and Bedrock, and Databricks. 43 runnable notebooks, one continuous case study. MIT licensed, no signup. By Zorost Intelligence AI Lab.
ECC
JavaScriptThe agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.
browser-use
Python🌐 Make websites accessible for AI agents. Automate tasks online with ease.
tokenspeed
PythonTokenSpeed is a speed-of-light LLM inference engine.
sample-specship
ShellSpec-driven autonomous engineering workflow for AI coding agents: recon → plan → build → validate → ship — with TDD, adversarial validation, and anti-slop quality gates. Packaged as a Kiro Power.
CodeJury
PythonTerminal-first, knowledge-grounded multi-agent software delivery pipeline: scope requirements, implement changes, run tests, and gate pull requests with deterministic QA and ensemble code review.
dify
TypeScriptBuild Agentic workflows, RAG pipelines, with rich AI model and tool support on one collaborative workspace. Deploy on cloud, VPC, or self-hosted, so teams move from prototype to production without rebuilding the stack.
ponytail
JavaScriptMakes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.
ragflow
GoRAGFlow is a leading open-source Retrieval-Augmented Generation (RAG) engine that fuses cutting-edge RAG with Agent capabilities to create a superior context layer for LLMs
nopus
TypeScriptDeterministic prose checks for clearer coding-agent responses
Jixu
TypeScriptDurable single-Agent Harness for TypeScript: recoverable Threads, context continuity, explicit side effects, and a native TUI.
omnigent
PythonOmnigent is an open-source AI agent framework and meta-harness: orchestrate Claude Code, Codex, Cursor, Pi, and custom agents — swap harnesses without rewriting, enforce policies and sandboxing, and collaborate in real time from any device.
gemma-chat
TypeScriptLocal AI chat + coding agent for Apple Silicon, powered by Gemma 4 via MLX / Supports Ollama
Flawless
PythonAI SRE AgenticOps for Kubernetes and cloud infrastructure.
deja-vu
GoSearch your past AI coding sessions — Claude Code, Codex, Cursor and 17 more. Indexes the session history they already wrote to disk, including months from before you installed it, and recalls it in any of them. No LLM, no embeddings, one local Go binary.
loop.js
TypeScriptA loop engineering framework — state a Goal; Rounds run until a skeptical, read-only Verify agent settles it.
openwiki
TypeScriptOpenWiki is a CLI that writes and maintains agent documentation for your codebase.
deltafin
RustRun full Kimi K3 on a single device. And an OpenAI-compatible API server for local chat and coding agents.
Flowise
TypeScriptBuild AI Agents, Visually
CarWatch
PythonYour car as a chat-room agent: Raspberry Pi 5 + dashcam + local AI. CodeWatch's sibling for the garage.
directional-prompting
Outcome-first plus directional language. A two-layer skill for writing prompts, agent directives, and skill descriptions. Works in Claude Code and Codex CLI.
klaatcode
TypeScriptOpen-source AI coding agent for the terminal. Claude Code-grade accuracy with smart model routing — uses the right AI model for each task, cutting costs 10x. Supports Claude, GPT, Gemini, DeepSeek & more.
local-llm
ShellEverything I know about running LLMs locally
llama.cpp
C++LLM inference in C/C++
apex-inference-chip
PythonAn inference chip design that runs a real LLM (Qwen2.5-0.5B) on FPGA — one transformer decoder layer in RTL, every silicon value bit-exact against a golden model. 0.56 tok/s measured, a 140× climb, full evidence trail.
shard
PythonPipeline-parallel LLM inference across GPUs on separate machines.
ctrlb-decompose
RustLLM-ready reasoning surface over logs
chorus
TypeScriptMulti-LLM peer review for code decisions. Bring your own CLI; Chorus convenes 2-4 other LLMs to review the work before you ship.
Investbrain
PHPSmart LLM-enabled investment tracker that consolidates and monitors market performance across your different brokerages