Dev
GitHub repos gaining traction - what high-signal users are starring and what's climbing the board, captured daily and enriched from GitHub. Raw material for spotting new tech and patterns worth building on.
2,235
repos tracked
175
surfaced this week
194
created < 30d
Python
top language
12 repos
-
π€ Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
-
Sub-real-time MiniMax H3 on Hopper: 13.506 s for a 14.375 s 768p video with audio on 8xH100. Four runtime patches removing 9.72 s of non-model overhead from FastVideo.
-
DimCut is a novel editing interaction design that folds the 1D timeline into multiple rows, integrating text, audio, and visuals β multidimensional information at a glance.
-
SOTA Open Source TTS
-
Native macOS and iOS audiobook highlighting with on-device alignment, Readwise, and Marginalia.
-
AI turns documents or topics into real, native PowerPoint decksβwith native shapes, transitions and animations, data-backed charts and tables on demand, audio narration from speaker notes, and support for your own .pptx templates. Β· by Hugo He
-
250+ Fine-tuning & RL Notebooks for text, vision, audio, embedding, TTS models.
-
Stable Audio LoRA Trainer of salty goodness
-
A powerful, hackable FPGA-based audio multitool for Eurorack.
-
Fun-ASR speech recognition models, with native Hugging Face Transformers support for Fun-ASR-Nano and separate FunASR, vLLM and llama.cpp deployment paths.
-
"it just vibes your audio into text bro" - π¦ββ¬