DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning Paper • 2501.12948 • Published Jan 22, 2025 • 458
view article Article Mixture of Experts (MoEs) in Transformers +5 ariG23498, pcuenq, merve, IlyasMoutawwakil, ArthurZ, sergiopaniego, Molbap • Feb 26 • 173
view article Article Native-speed vLLM transformers modeling backend hmellor, lysandre • 23 days ago • 64
UniClawBench: A Universal Benchmark for Proactive Agents on Real-World Tasks Paper • 2607.08768 • Published 22 days ago • 34
PixWorld: Unifying 3D Scene Generation and Reconstruction in Pixel Space Paper • 2607.05373 • Published 25 days ago • 66
Image Classification Models Collection LiteRT image-classification models from litert-community. • 89 items • Updated 5 days ago • 7
Google Tensor Collection LiteRT Models that can run on Google Tensor • 5 items • Updated 5 days ago • 9
ABC-Bench Collection Evaluating Agentic Backend Coding Capabilities in Real-World Development Scenarios • 4 items • Updated 20 days ago • 5
MOSS Transcribe Collection A unified multimodal large language model for end-to-end speaker-attributed, time-stamped transcription. • 4 items • Updated 20 days ago • 13
MOSS Transcribe Diarize: Accurate Transcription with Speaker Diarization Paper • 2601.01554 • Published Jan 4 • 65
view article Article Hugging Face and Cerebras bring Gemma 4 to real-time voice AI +2 A-Mahla, andito, lvwerra, vyassaurabh • 30 days ago • 89
view article Article Run AI workloads on any cloud, store on Hugging Face: zero-egress storage with SkyPilot +3 njha, michaelvll, hopechong, XciD, julien-c • 24 days ago • 27
SenseNova-SI Collection Scaling Spatial Intelligence with Multimodal Foundation Models • 16 items • Updated May 12 • 24
SenseNova-Vision Collection Vision as Unified Multimodal Generation • 5 items • Updated 20 days ago • 31
Nemotron-Labs-Audex Collection Unified Audio Intelligence Without Regressing on Text Intelligence • 4 items • Updated 14 days ago • 8