Pinned Loading
-
infer-lab
infer-lab PublicLLM inference engine internals: PagedAttention, RadixAttention, continuous batching, speculative decoding - kernels to distributed systems
Python
-
turbovec
turbovec PublicFast, from-scratch ANN vector search engine — SIMD-accelerated C++ (HNSW) core with a Python-first API. 100% recall at ~11k QPS.
Python
-
OmniSeg-Audio-Pipeline
OmniSeg-Audio-Pipeline PublicHigh-performance multimodal pipeline synchronizing Meta's SAM 2 (Vision) and MIT's AST (Audio). Features O(1) resource management, temporal video slicing, and automated JSON metadata orchestration …
Python
-
Fast-Prototyping-of-a-GenAI-App-with-Streamlit
Fast-Prototyping-of-a-GenAI-App-with-Streamlit Public🤖 Streamlit-based GenAI application featuring RAG and Snowflake integration for advanced data ingestion, cleaning, and AI-driven insights.
Python 1
If the problem persists, check the GitHub status page or contact support.