Week of Jul 27 to Aug 2, 2026
Tools of the Week
Every Monday this page reshuffles and puts a different tool forward from each of the 27 categories in the directory. It is a way to surface the good work that never reaches the top of a ratings list. 27 picks are live right now, and the next rotation lands Monday, August 3.
Reconstructs 3D shape, texture, and layout of objects from a single cluttered real-world photo.
Multi-agent framework for building and studying societies of communicating LLM agents at scale.
Virtual human pose generation and animation from images.
Open-source CLI and IDE tool for source-controlled AI code review checks enforceable in CI.
Framework for building production-ready AI application services.
High-performance numerical computing library by Google with auto-differentiation.
Diffusion tool that relights image foregrounds using text prompts or background images.
Open-source ML monitoring framework for data drift and model quality.
Transformer-based text-to-audio model from Suno
Transformer model from Meta AI that tracks any point through a video, jointly and through occlusions.
Computer vision annotation tool by Intel for image and video labeling.
Library for accessing and sharing ML datasets
Multi-GPU inference engine for diffusion transformers using hybrid parallelism techniques.
Distilled diffusion models enabling image generation in 1-4 steps.
Open-weight LLM by Meta available in 8B, 70B, and 405B parameter sizes.
Platform for serving LLM, embedding, speech, and image models behind one OpenAI-compatible API.
Framework for generating synthetic data and AI feedback through composable LLM pipelines.
Music source separation library with pretrained models that split songs into 2, 4, or 5 stems.
Framework for computing dense vector representations of sentences and paragraphs.
Ready-to-use OCR library supporting 80+ languages with simple Python API.
Managed vector database for machine learning applications
Multilingual model covering speech recognition, emotion recognition, and audio event detection.
Zero-shot voice cloning TTS from ByteDance using a 0.45B diffusion transformer with bilingual output.
Cross-platform desktop LLM client with 300+ assistants, MCP support, and document processing.
Open embedding and reranker model series in 0.6B to 8B sizes covering more than 100 languages.
Open-source video generation model by Genmo with state-of-the-art motion quality.
Few-shot voice cloning and TTS using GPT and SoVITS architectures.