codex — Coding agent and developer workflows
openai/codex
Lightweight coding agent that runs in your terminal
ToolAI.io · GitHub चैनल
LLM, एजेंट, MCP, RAG और कोडिंग विषयों में सक्रिय ओपन-सोर्स AI रिपॉज़िटरीज़ को स्टार्स और सत्यापित प्रोजेक्ट जानकारी के साथ ट्रैक करें।
सार्वजनिक रिपॉज़िटरी का डेटा
openai/codex
Lightweight coding agent that runs in your terminal
raullenchai/rapid-mlx
The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.
theopenco/llmgateway
Route, manage, and analyze your LLM requests across multiple providers with a unified API interface.
sgl-project/sglang
SGLang is a high-performance serving framework for large language models and multimodal models.
deepspeedai/deepspeed
DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
ctxrs/ctx
Search the coding agent history already on your machine
vllm-project/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
hmbown/codewhale
Open-source, community-driven agent harness
matrixorigin/matrixone
AI-native HTAP database with Git-for-Data and built-in vector search, serving as the data and memory backbone for intelligent agents and applications.
taichuy/1flowbase
An open-source AI gateway that allows local agent clients to publish fusion-style multi-model workflows as OpenAI and Claude-compatible virtual models, providing full observability into traces, tokens, latency, and costs.
NVIDIA/TensorRT-LLM
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way
hal0ai/hal0
An open-source, self-hosted home AI inference platform designed to turn a Linux box into an OpenAI-compatible inference appliance, with native optimization for AMD Strix Halo hardware.
पृष्ठ 5 / 22 · 253 प्रोजेक्ट