whisper — Speech recognition and transcription
openai/whisper
Robust Speech Recognition via Large-Scale Weak Supervision
ToolAI.io · Kanal GitHub
Indeks berbasis fakta untuk proyek pengembangan LLM, agen, MCP, RAG, dan AI, dengan lisensi, pengaturan, unduhan, tangkapan layar, dan sumber daya terkait untuk setiap entri.
Data repositori publik
openai/whisper
Robust Speech Recognition via Large-Scale Weak Supervision
huggingface/transformers
A model-definition framework for state-of-the-art machine learning models across text, vision, audio, and multimodal domains, supporting both inference and training.
hacksider/Deep-Live-Cam
A computer vision tool for real-time face replacement and video processing.
opencv/opencv
An open-source library for image processing, video analysis and computer vision.
tesseract-ocr/tesseract
An open-source OCR engine for converting text in images into machine-readable text.
hiyouga/LlamaFactory
A unified fine-tuning toolkit for large language and vision-language models.
ultralytics/ultralytics
Training and inference tools for object detection, segmentation and pose estimation models.
ultralytics/yolov5
A PyTorch-based object detection project with training, inference and export workflows.
deepfakes/faceswap
An open-source toolkit for face processing, model training and video conversion.
ageitgey/face_recognition
Face detection and recognition interfaces for Python and the command line.
photoprism/photoprism
A self-hostable photo and video library with automatic labels, face organization, and metadata search.
naptha/tesseract.js
Tesseract OCR for browsers and Node.js using WebAssembly, suitable for image-to-text features in JavaScript apps.
Halaman 1 / 3 · 28 proyek