OpenRadar

Tag

#llm

26 open-source projects filed under this tag.

01 Python

aisuite

Andrew Ng's aisuite is a unified Python library for building with multiple LLM providers — one API for OpenAI, Anthropic, Google, Mistral, and more, plus an Agents API with toolkits and MCP support.

#ai#llm#python#agents
14.3k 1.5k Read
02 Go

bifrost

Bifrost is a high-performance Go AI gateway unifying 23+ LLM providers behind one OpenAI-compatible API with semantic caching, MCP support, and sub-15µs overhead.

#ai-gateway#llm#go#infrastructure
5.6k 724 Read
03 TypeScript

context7

Context7 is an MCP-powered documentation platform that feeds version-specific library docs directly into your AI coding assistant — no more hallucinated APIs.

#developer-tools#mcp#llm#documentation
56.7k 2.7k Read
04 Go

crabtrap

Open-source HTTP proxy by Brex that uses LLM-as-a-judge to secure AI agents in production. Static rules + AI policy evaluation with full audit trail.

#security#ai-agents#proxy#go
673 51 Read
05 C

ds4

DwarfStar is antirez's standalone DeepSeek V4 inference engine — run a 284B MoE model locally on your Mac with Metal or CUDA, with an integrated coding agent.

#local-inference#deepseek#llm#metal
12.8k 1.1k Read
06 Python

forge

Python framework that makes self-hosted LLM tool-calling reliable — guardrails boost 8B models from 5% to 84% accuracy on agentic tasks.

#ai-agents#llm#guardrails#tool-calling
2.1k 144 Read
07 TypeScript

freellmapi

FreeLLMAPI stacks free tiers from 16 LLM providers behind one OpenAI-compatible endpoint — ~1.7B tokens/month with smart routing and automatic failover.

#llm#ai-infrastructure#openai-compatible#proxy
7.8k 1.3k Read
08 Rust

goose

Open source AI agent by Block with 48k+ stars — Rust-based desktop, CLI, and API for code, workflows, and automation with MCP support and 70+ extensions.

#ai-agent#mcp#open-source#rust
48.8k 5.1k Read
09 Python

headroom

Headroom is a context compression layer for AI agents that reduces token usage by 60-95% while preserving answer quality. Library, proxy, MCP server.

#ai#token-optimization#context-engineering#mcp
18.4k 1.2k Read
10 TypeScript

improve

Shadcn's agent skill audits your codebase with an expensive model and writes execution plans for cheaper AI models — 5K stars in its first week.

#ai-agents#code-audit#developer-tools#llm
5.2k 191 Read
11 Python

kvarn

KVarN is a vLLM KV-cache quantization backend from Huawei that delivers 3-5x more context capacity with FP16-level accuracy — one flag, no calibration.

#llm#quantization#vllm#inference
366 18 Read
12 Go

lathe

Lathe is a Go CLI that generates hands-on technical tutorials using LLMs with a local UI, designed for developers who learn by doing rather than watching AI code for them.

#developer-tools#learning#llm#tutorials
644 16 Read
13 TypeScript

llm-wiki

LLM Wiki is a cross-platform desktop app that turns your documents into a persistent, interlinked knowledge base — built on Karpathy's llm-wiki pattern with Tauri, React, and sigma.js.

#knowledge-base#llm#desktop-app#tauri
10.3k 1.3k Read
14 Rust

llmtrim

A Rust-based local proxy that compresses LLM API requests in real-time, cutting costs by 66% without changing your AI tool's answers. Works with Claude Code, Codex, Cursor, and more.

#developer-tools#llm#cost-optimization#rust
72 4 Read
15 Python

lmcache

LMCache is a KV cache management layer that slashes LLM inference latency by reusing computed attention states across requests, sessions, and GPU clusters.

#llm#inference#kv-cache#vllm
8.7k 1.3k Read
16 TypeScript

mastra

Mastra is a TypeScript framework for building AI agents and workflows — from the team behind Gatsby, backed by Y Combinator W25.

#ai-agents#typescript#workflows#llm
25k 2.2k Read
17 Rust

mnemo

Mnemo is a local-first AI memory sidecar built in Rust — persistent knowledge graph, entity extraction, and semantic retrieval for any LLM, zero cloud dependency.

#ai-memory#knowledge-graph#llm#rust
193 6 Read
18 Python

opensquilla

Token-efficient microkernel AI agent with local model router, persistent memory, layered sandbox, and support for 20+ LLM providers — same budget, higher intelligence density.

#ai-agent#llm#token-optimization#developer-tools
3.8k 299 Read
19 TypeScript

openui

OpenUI is an open standard for generative UI — a streaming-first language and React runtime that lets LLMs render structured interfaces with 67% fewer tokens than JSON.

#generative-ui#react#llm#streaming
6.9k 520 Read
20 Python

skillopt

SkillOpt by Microsoft trains reusable natural-language skills for frozen LLM agents using trajectory-driven edits and validation-gated updates — no fine-tuning required.

#ai-agents#llm#optimization#microsoft
5.8k 564 Read
21 JavaScript

system-prompts-leaks

Extracted system prompts from 50+ AI models and tools including Claude, ChatGPT, Gemini, Codex, Cursor, and Copilot. 43K+ stars on GitHub.

#ai#prompt-engineering#llm#open-source
43.2k 7.2k Read
22 TypeScript

tencentdb-agent-memory

TencentDB Agent Memory gives AI agents layered long-term memory with Mermaid symbolic compression, cutting token usage by 61% and boosting task success by 51%.

#ai-agents#memory#llm#typescript
5.1k 440 Read
23 Rust

tensorzero

TensorZero is an open-source Rust LLMOps platform unifying LLM gateway, observability, evaluation, and optimization with sub-millisecond latency overhead.

#llm#ai-infrastructure#gateway#observability
11.6k 920 Read
24 Python

tokenspeed

TokenSpeed is an open-source LLM inference engine built for agentic workloads — 580 tokens/sec on Qwen3.5-397B with TensorRT-LLM performance and vLLM-level ease of use.

#llm#inference#ai-agents#gpu
1.4k 142 Read
25 TypeScript

toon

TOON is a compact, human-readable JSON encoding that cuts LLM token usage by ~40% while improving accuracy. A must-know data format for AI-powered apps.

#data-format#llm#serialization#tokenization
24.5k 1.1k Read
26 Python

whichllm

Find the best local LLM for your hardware with real benchmark rankings. Auto-detect GPU, CPU, and RAM to get evidence-based model recommendations from HuggingFace.

#llm#cli#local-ai#developer-tools
4k 221 Read