Home / Companies / Hugging Face / Hacker News

Hugging Face on HN

40 posts with 10+ points in 2026

Filters
Year:
Posts by Month (40 total)
Hacker News Posts
Title Points Comments Date
Kimi-K3 on HuggingFace 1,378 -- 2026-07-27
Show HN: Sweep, Open-weights 1.5B model for next-edit autocomplete 530 -- 2026-01-21
Anatomy of a Frontier Lab Agent Intrusion: A Timeline of the July … 466 -- 2026-07-28
Kimi K2.7-Code: open-source coding model with better token efficiency 455 -- 2026-06-12
Show HN: Hacker News archive (47M+ items, 11.6GB) as Parquet, updated every … 399 -- 2026-03-14
GLM-4.7-Flash 371 -- 2026-01-19
Inflect-Micro-v2: complete voice in 9.36M parameters 213 -- 2026-07-26
DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence 159 -- 2026-04-24
Show HN: Text-to-video model from scratch (2 brothers, 2 years, 2B params) 156 -- 2026-01-22
Waypoint-1: Real-Time Interactive Video Diffusion from Overworld 92 -- 2026-01-23
Show HN: 30k IKEA items in flat text 55 -- 2026-01-07
Anyone Can Clone Your Voice Now 45 -- 2026-01-26
Qwen/Qwen3.6-27B · Hugging Face 41 -- 2026-04-22
Continuous batching (2025) 39 -- 2026-02-15
Anatomy of BoltzGen 31 -- 2026-01-04
DeepSeek-V4 Technical Report [pdf] 26 -- 2026-04-24
Gemma 4 E2B running in-browser at 255 tok/s 19 -- 2026-06-17
Qwen 3.5 small models out 18 -- 2026-02-24
Netflix just dropped their first public model on Hugging Face: VOID 18 -- 2026-04-04
Hugging Face Storage Buckets: Mutable, non-versioned object storage at $12/TB 18 -- 2026-03-10
Rio 3.5 Open 397B – from Rio de Janeiro's city government 17 -- 2026-06-13
Google translategemma 4B Translation Models 16 -- 2026-01-19
Z.ai GLM 5.2 16 -- 2026-06-16
Flux.2 Klein 4B (Apache 2.0) 14 -- 2026-01-19
DeepSeek V4 Flash 14 -- 2026-04-24
Show HN: 17MB model beats human experts at pronunciation scoring 13 -- 2026-02-20
DeepSeek-V4: a million-token context that agents can use 13 -- 2026-04-28
Distilling 100B+ Models 40x Faster with TRL 13 -- 2026-04-12
Qwen 3.5 12 -- 2026-02-16
Minimax M2.7 Weights Released 11 -- 2026-04-12
5.6x throughput on Kimi K2.6 by speculating less 11 -- 2026-04-21
Kimi K2.6 11 -- 2026-04-20
Tencent/Hy3: 295B MoE model rivals trillion scale SOTA 11 -- 2026-07-09
VibeVoice-ASR: speech-to-text model designed to handle 60-minute long-form audio 11 -- 2026-01-31
Qwen3.5 Small: 0.8B, 2B, 4B, 9B Released 10 -- 2026-03-02
Show HN: Trained an LLM to predict "What will Trump do?" 10 -- 2026-02-20
Nvidia Nemotron 3-Nano 30B-A3B-BF16 10 -- 2026-01-31
GLM-5.2: Built for Long-Horizon Tasks 10 -- 2026-06-17
MiniMax-M3: A native multimodal model with 1M context 10 -- 2026-06-12
Alibaba open-sources Qwen3.6-35B-A3B, a 35B MoE model with 3B active parameters 10 -- 2026-04-16