12 Weeks, 24 Lessons, AI for All!
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
OpenVINO™ is an open source toolkit for optimizing and deploying AI inference
Ultralytics YOLO26, YOLO11, YOLOv8 — object detection, instance segmentation, semantic segmentation, image classification, pose estimation, object tracking
Implement a ChatGPT-like LLM in PyTorch from scratch, step by step
Faster Whisper transcription with CTranslate2
We write your reusable computer vision tools. 💜
Open Source Computer Vision Library
🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
Open standard for machine learning interoperability
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
ONNX Runtime: cross-platform, high performance ML inferencing and training accelerator
Translate manga/image 一键翻译各类图片内文字 https://cotrans.touhou.ai/ (no longer working)
This repository contains a hand-curated resources for Prompt Engineering with a focus on Generative Pre-trained Transformer (GPT), ChatGPT, PaLM etc
High-performance In-browser LLM Inference Engine
(⌐■_■) - Deep Reinforcement Learning instrumenting bettercap for WiFi pwning.
Fast and accurate AI powered file content types detection
StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models
Fast inference engine for Transformer models
A curated list of awesome things related to artificial intelligence tools
Train Large Language Models on MLX.
A desktop application that extracts YouTube playlist transcripts and enhances them using Google’s Gemini AI models. The output is a book in any language you want.