Dev
GitHub repos gaining traction - what high-signal users are starring and what's climbing the board, captured daily and enriched from GitHub. Raw material for spotting new tech and patterns worth building on.
2,235
repos tracked
175
surfaced this week
194
created < 30d
Python
top language
12 repos
-
A high-throughput and memory-efficient inference and serving engine for LLMs
-
Open Source Continuous Inference Benchmark Research Platform — Kimi K3 2.8T, MiniMax M3, DeepSeekv4, GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72 & soon™ TPUv6e/v7/Trainium2/3 | 开源持续推理基准研究平台 — Kimi K2.7-Code、MiniMax M3、DeepSeekv4、GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72,即将推出™ TPUv6e/v7/Trainium2/3
-
Monkey patch, test, and trace Python by attaching bindings to call sites, without modifying the code being observed. Built on wrapt.
-
Turn your PC, Mac, or Linux box into an AI server. LLM inference, chat UI, voice, agents, workflows, RAG, and image generation.
-
Official AMD catalog of AI agent skills. Empower your AI agents with AMD's optimized SW stack.
-
Apple Silicon (Metal) backend for Triton: write standard @triton.jit kernels on your Mac GPU. The same source runs bit-identical on NVIDIA and AMD, so you develop kernel logic locally and rent a datacenter GPU only for the perf pass.
-
LMCache: Supercharge Your LLM with the Fastest KV Cache Layer
-
★ 1 neubig/playbooksOfficial repository of AMD Playbooks
-
UCCL is an efficient communication library for GPUs, covering collectives, P2P (e.g., KV cache transfer, RL weight transfer), and EP (e.g., GPU-driven)
-
Smartthings plugin for StreamDeck
-
FastCrest Tether: the OSS edge-to-cloud AI deploy CLI. Optimize, verify, deploy across Jetson, RTX, Apple Silicon, AMD. Hybrid edge-cloud inference with parity certs.