whisper — Speech recognition and transcription
openai/whisper
Robust Speech Recognition via Large-Scale Weak Supervision
ToolAI.io · Canal de GitHub
Explora el directorio completo de ToolAI de proyectos de AI de código abierto en GitHub, organizado por tema, lenguaje y licencia.
Datos del repositorio público
openai/whisper
Robust Speech Recognition via Large-Scale Weak Supervision
huggingface/transformers
A model-definition framework for state-of-the-art machine learning models across text, vision, audio, and multimodal domains, supporting both inference and training.
hacksider/Deep-Live-Cam
A computer vision tool for real-time face replacement and video processing.
opencv/opencv
An open-source library for image processing, video analysis and computer vision.
tesseract-ocr/tesseract
An open-source OCR engine for converting text in images into machine-readable text.
hiyouga/LlamaFactory
A unified fine-tuning toolkit for large language and vision-language models.
ultralytics/ultralytics
Training and inference tools for object detection, segmentation and pose estimation models.
ultralytics/yolov5
A PyTorch-based object detection project with training, inference and export workflows.
deepfakes/faceswap
An open-source toolkit for face processing, model training and video conversion.
ageitgey/face_recognition
Face detection and recognition interfaces for Python and the command line.
photoprism/photoprism
A self-hostable photo and video library with automatic labels, face organization, and metadata search.
naptha/tesseract.js
Tesseract OCR for browsers and Node.js using WebAssembly, suitable for image-to-text features in JavaScript apps.
Página 1 / 3 · 28 proyectos