tencent/weknora_
Open-source LLM knowledge platform: turn raw documents into a queryable RAG, an autonomous reasoning agent, and a self-maintaining Wiki.
54 items tagged “llm” across every source.
Open-source LLM knowledge platform: turn raw documents into a queryable RAG, an autonomous reasoning agent, and a self-maintaining Wiki.
Clone any viral video with AI agents. Not just a script, the whole workflow: swap the face, the words, the B-roll, ship 100 variants in one command, and get your 100M views.
World's first open-source enterprise world model.
魔搭紫皮书|ModelScope Cookbook:面向开发者的开源模型应用实战指南,覆盖模型选型、推理、微调、评测、RAG、Agent 与 AIGC,从跑通第一个模型到构建实际应用。
AI-powered virtual executive team — a single coherent executive persona backed by 8 specialist agents (FastAPI + Next.js).
Continual learning infra for self-improving agents
CROW - Your AI Agent. MCPs, OpenRouter, Any Model or local. It's your choice.
Examples and tutorials to help developers build AI systems
A high-quality PDF to Markdown tool based on large language model visual recognition. 一款基于大模型视觉识别的高质量PDF转Markdown工具
TradingAgents: Multi-Agents LLM Financial Trading Framework
《深入理解 AI Infra:量化分析与系统设计》(李博杰 著)开源书稿:从硬件约束和模型架构出发,量化推导 LLM 推理与训练系统设计。含全书正文、PDF、配套计算工具与实验
大模型(LLM)全栈学习路线与中文教程🔥:覆盖 Prompt Engineering、RAG、AI Agent、MCP、微调、模型部署、Transformer、AI 编程与大厂面试,从入门到生产实践。
Python & JS/TS SDK for running AI-generated code/code interpreting in your AI app
The fastest, litest AI Gateway. Rust core with Python SDK. Call 100+ LLM APIs in OpenAI (or native) format with cost tracking, guardrails, load balancing, and logging [Bedrock, Azure, OpenAI, Anthropic, OpenAI, VertexAI, vLLM, Nvidia NIM]
A robust Node.js proxy server that automatically rotates API keys for Gemini and OpenAI APIs when rate limits (429 errors) are encountered. Built with zero dependencies and comprehensive logging.
A living pixel-art station where real AI agents do real work. Local-first desktop agent harness - bring your own key, watch your crew actually run.
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
A collection of projects showcasing RAG, agents, workflows, and other AI use cases
The platform for LLM evaluations and AI agent testing
左翼毛派分析 · AI 技能包
A flight recorder for AI agent fleets that pays for itself — provenance-native caching proxy for model APIs (Pact + AgentReplay)
HTTP-native chat and notes for agents whose sandbox only allows webfetch — every write is a plain GET. Runs technocore.chat.
Official release of InternLM series (InternLM, InternLM2, InternLM2.5, InternLM3).
A high-throughput and memory-efficient inference and serving engine for LLMs
SGLang is a high-performance serving framework for large language models and multimodal models.
HarnessRouter Community Edition: the self-hosted, Apache-2.0 edition of the unified interface for agent harnesses. Run Codex, Claude Code, Hermes, PI, DSH, and more through one API, with sessions, streaming, files, cancellation, and failure handling. Implements the Unified Harness Protocol (UHP), an open standard. Your keys, your infrastructure.
Learn it. Build it. Ship it for others.
Your personal AI agent platform — self-hosted, BYO Claude/Codex subscription
The agent that grows with you
The local-first Agent OS — your AI partner lives on your own machine. Drive the official Claude Code, Codex & OpenCode from your browser or any chat app.
Automated Penetration Testing Agentic Framework Powered by Large Language Models
AI-powered cross-platform e-book reader with semantic search, RAG chat, local vector store, notes, TTS, and WebDAV sync.
Turn any technical book PDF into a Claude Code skill — ready to study, reference, and use while you work.
Paw Work - selection-first web agent for Chrome: select on the live page, describe the outcome, take away an editable office file. BYOK, sandboxed, no server.
Agents that use the browser.
🪨 why use many token when few token do trick — Claude Code skill that cuts 65% of tokens by talking like caveman
The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.
Share the model on your machine with friends: one binary in front of llama.cpp, vLLM, Ollama or LM Studio; one invite code; they chat from a browser. Self-hosted, end-to-end encrypted, no accounts.
The ripgrep of AI context: a zero-dependency C++23 CLI + MCP server for coding agents. Find what you want without reading the repo, then check you built what you meant — blast radius, tests-to-run, quality deltas. Signatures at 74.7% fewer bytes than bodies; every guess labelled, every loss published. Paddle out with a map.
GPT-2 rebuilt from scratch in PyTorch, verified against HuggingFace to the logit, then aligned by hand through pretraining, SFT, and DPO.