jan
janhq/jan
A local AI chat application for running models on a personal computer.
ToolAI.io · GitHub چینل
GitHub پر حال ہی میں شامل کیے گئے اوپن سورس AI پروجیکٹس دریافت کریں، جن کے ساتھ ریپوزٹری کی معلومات، سیٹ اپ نوٹس، لائسنس اور متعلقہ وسائل بھی موجود ہیں۔
عوامی ریپوزٹری کا ڈیٹا
janhq/jan
A local AI chat application for running models on a personal computer.
ray-project/ray
A distributed computing runtime for machine learning training, tuning and model serving.
mudler/LocalAI
A local inference engine for self-hosting models and AI services.
ultralytics/yolov5
A PyTorch-based object detection project with training, inference and export workflows.
ultralytics/ultralytics
Training and inference tools for object detection, segmentation and pose estimation models.
ollama/ollama
A tool for running and managing large language models locally.
MoonshotAI/checkpoint-engine
Checkpoint-engine is a simple middleware to update model weights in LLM inference engines
microsoft/edgeai-for-beginners
This course is designed to guide beginners through the exciting world of Edge AI, covering fundamental concepts, popular models, inference techniques, device-specific applications, model optimization, and the development of intelligent Edge AI agents.
PaddlePaddle/Paddle-Lite
PaddlePaddle High Performance Deep Learning Inference Engine for Mobile and Edge (飞桨高性能深度学习端侧推理引擎)
facebookresearch/sam2
The repository provides code for running inference with the Meta Segment Anything Model 2 (SAM 2), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
NVIDIA/TensorRT
NVIDIA® TensorRT™ is an SDK for high-performance deep learning inference on NVIDIA GPUs. This repository contains the open source components of TensorRT.
NVIDIA/Model-Optimizer
A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downstream deployment frameworks like TensorRT-LLM, TensorRT, vLLM, etc. to optimize inference speed.
صفحہ 1 / 4 · 41 پروجیکٹس