memra
avifenesh/memra
A from-scratch LLM inference engine built in Rust and CUDA, specifically optimized for RTX 5090 (sm_120a) and H100 (sm_90a) architectures, delivering exactness-gated performance without relying on frameworks like ggml.
ToolAI.io · GitHub চ্যানেল
LLM, এজেন্ট, MCP, RAG এবং কোডিং বিষয়ের সক্রিয় ওপেন-সোর্স AI রিপোজিটরিগুলো স্টার ও যাচাইকৃত প্রজেক্ট তথ্যসহ ট্র্যাক করুন।
পাবলিক রিপোজিটরির তথ্য
avifenesh/memra
A from-scratch LLM inference engine built in Rust and CUDA, specifically optimized for RTX 5090 (sm_120a) and H100 (sm_90a) architectures, delivering exactness-gated performance without relying on frameworks like ggml.
bionic-gpt/bionic-gpt
Bionic is a Rust-based, on-premise alternative to ChatGPT designed for private generative AI deployments. It provides a familiar chat interface, local or remote model access, team controls, retrieval-augmented assistants, data integrations, observability, and Kubernetes-oriented scaling.
mozilla-ai/any-llm
any-llm is a Python SDK for communicating with multiple LLM providers through one interface. It supports switching providers and models with minimal code changes while using official provider SDKs when available.
paulduvall/ai-development-patterns
A comprehensive collection of AI development patterns for building software with AI assistance, organized by implementation maturity and development lifecycle phases. Includes Foundation, Development, and Operations patterns with practical examples and anti-patterns.
openai/tiktoken
tiktoken is a fast BPE tokeniser for use with OpenAI's models
embabel/embabel-agent
Embabel is an open-source JVM framework for building strongly typed agentic applications that combine LLM interactions, regular code, tools, and domain models. It supports dynamic planning and can be used from Kotlin or Java.
kaito-project/aikit
🏗️ Fine-tune, build, and deploy open-source LLMs easily!
0xlazai/alith
A simple, composable, and high-performance AI agent framework designed for Web3 and Crypto, enabling developers to build, deploy, and manage on-chain AI agents with multi-language support and LazAI Gateway integration.
mtrnix/metronix-memory
Self-hosted memory infrastructure for AI agents featuring MCP-native integration, hybrid RAG, a temporal knowledge graph, and an ontology layer, designed for local-model friendliness and durable, agent-scoped context.
sno-ai/llmix
A production LLM call layer for AI agents and tools that wraps existing provider SDKs with config-driven model presets, caching, resilience patterns, and key rotation across Python, TypeScript, and Rust.
wukongim/wukongim
More than just IM 不只是即时通讯(IM)
stackitcloud/rag-template
A template for building AI chatbots and document management systems using Retrieval-Augmented Generation (RAG), vector search, and FastAPI, designed for deployment on Kubernetes.
পৃষ্ঠা 9 / 22 · 253টি প্রজেক্ট