whisper — Speech recognition and transcription
openai/whisper
Robust Speech Recognition via Large-Scale Weak Supervision
ToolAI.io · GitHub 채널
GitHub의 전체 ToolAI 오픈 소스 AI 프로젝트 디렉터리를 주제, 언어 및 라이선스별로 살펴보세요.
공개 리포지토리 데이터
openai/whisper
Robust Speech Recognition via Large-Scale Weak Supervision
huggingface/transformers
A model-definition framework for state-of-the-art machine learning models across text, vision, audio, and multimodal domains, supporting both inference and training.
hacksider/Deep-Live-Cam
A computer vision tool for real-time face replacement and video processing.
opencv/opencv
An open-source library for image processing, video analysis and computer vision.
tesseract-ocr/tesseract
An open-source OCR engine for converting text in images into machine-readable text.
hiyouga/LlamaFactory
A unified fine-tuning toolkit for large language and vision-language models.
ultralytics/ultralytics
Training and inference tools for object detection, segmentation and pose estimation models.
ultralytics/yolov5
A PyTorch-based object detection project with training, inference and export workflows.
deepfakes/faceswap
An open-source toolkit for face processing, model training and video conversion.
ageitgey/face_recognition
Face detection and recognition interfaces for Python and the command line.
photoprism/photoprism
A self-hostable photo and video library with automatic labels, face organization, and metadata search.
naptha/tesseract.js
Tesseract OCR for browsers and Node.js using WebAssembly, suitable for image-to-text features in JavaScript apps.
1 / 3페이지 · 프로젝트 28개