Home / Companies / Hugging Face / Hacker News

Hugging Face on HN

187 posts with 1+ points in 2026

Filters
Year:
Posts by Month (187 total)
Hacker News Posts
Title Points Comments Date
Kimi-K3 on HuggingFace 1,378 -- 2026-07-27
Show HN: Sweep, Open-weights 1.5B model for next-edit autocomplete 530 -- 2026-01-21
Anatomy of a Frontier Lab Agent Intrusion: A Timeline of the July … 466 -- 2026-07-28
Kimi K2.7-Code: open-source coding model with better token efficiency 455 -- 2026-06-12
Show HN: Hacker News archive (47M+ items, 11.6GB) as Parquet, updated every … 399 -- 2026-03-14
GLM-4.7-Flash 371 -- 2026-01-19
Inflect-Micro-v2: complete voice in 9.36M parameters 213 -- 2026-07-26
DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence 159 -- 2026-04-24
Show HN: Text-to-video model from scratch (2 brothers, 2 years, 2B params) 156 -- 2026-01-22
Waypoint-1: Real-Time Interactive Video Diffusion from Overworld 92 -- 2026-01-23
Show HN: 30k IKEA items in flat text 55 -- 2026-01-07
Anyone Can Clone Your Voice Now 45 -- 2026-01-26
Qwen/Qwen3.6-27B · Hugging Face 41 -- 2026-04-22
Continuous batching (2025) 39 -- 2026-02-15
Anatomy of BoltzGen 31 -- 2026-01-04
DeepSeek-V4 Technical Report [pdf] 26 -- 2026-04-24
Gemma 4 E2B running in-browser at 255 tok/s 19 -- 2026-06-17
Qwen 3.5 small models out 18 -- 2026-02-24
Netflix just dropped their first public model on Hugging Face: VOID 18 -- 2026-04-04
Hugging Face Storage Buckets: Mutable, non-versioned object storage at $12/TB 18 -- 2026-03-10
Rio 3.5 Open 397B – from Rio de Janeiro's city government 17 -- 2026-06-13
Google translategemma 4B Translation Models 16 -- 2026-01-19
Z.ai GLM 5.2 16 -- 2026-06-16
Flux.2 Klein 4B (Apache 2.0) 14 -- 2026-01-19
DeepSeek V4 Flash 14 -- 2026-04-24
Show HN: 17MB model beats human experts at pronunciation scoring 13 -- 2026-02-20
DeepSeek-V4: a million-token context that agents can use 13 -- 2026-04-28
Distilling 100B+ Models 40x Faster with TRL 13 -- 2026-04-12
Qwen 3.5 12 -- 2026-02-16
Minimax M2.7 Weights Released 11 -- 2026-04-12
5.6x throughput on Kimi K2.6 by speculating less 11 -- 2026-04-21
Kimi K2.6 11 -- 2026-04-20
Tencent/Hy3: 295B MoE model rivals trillion scale SOTA 11 -- 2026-07-09
VibeVoice-ASR: speech-to-text model designed to handle 60-minute long-form audio 11 -- 2026-01-31
Qwen3.5 Small: 0.8B, 2B, 4B, 9B Released 10 -- 2026-03-02
Show HN: Trained an LLM to predict "What will Trump do?" 10 -- 2026-02-20
Nvidia Nemotron 3-Nano 30B-A3B-BF16 10 -- 2026-01-31
GLM-5.2: Built for Long-Horizon Tasks 10 -- 2026-06-17
MiniMax-M3: A native multimodal model with 1M context 10 -- 2026-06-12
Alibaba open-sources Qwen3.6-35B-A3B, a 35B MoE model with 3B active parameters 10 -- 2026-04-16
Xiaomi MiMo-v2.5-Pro Open-Sourced: 1T Parameter Model 9 -- 2026-04-28
Soofi: European sovereign LLM trained in 2 months 9 -- 2026-07-11
Nvidia releases 8B model with learned 8x KV cache compression 9 -- 2026-01-24
DeepSeek: Thinking with Visual Primitives [pdf] 9 -- 2026-04-30
Mistralai/Leanstral-1.5-119B-A6B 8 -- 2026-07-03
Reka Edge – 7B fast, efficient VLM (open-weights) 8 -- 2026-03-11
Lance – native supports image and video understanding, generation, and editing 8 -- 2026-05-24
Bonsai 1.7B in the browser: a 290MB 1-bit LLM on WebGPU 8 -- 2026-04-16
Gemma 4 running client-side in WebGPU 8 -- 2026-04-04
Show HN: Hacker News RSS Feed Directory 8 -- 2026-04-04
Show HN: Marlin-2B: a tiny VLM to extract structured information from videos 7 -- 2026-05-18
Zyphra releases the ZAYA1-8B MoE model optimized for intelligence density 7 -- 2026-05-06
Carbon: Autoregressive Genomic Foundation Model 7 -- 2026-05-19
Show HN: We beat Gemini Embedding 2 by training only 16M params … 7 -- 2026-07-10
Nex N2 Pro: Frontier agentic performance at 400B 7 -- 2026-06-08
The ultimate guide to RL environments: building and scaling them in the … 7 -- 2026-05-05
DeepSeek-V4 7 -- 2026-04-24
We OCR'ed 30k papers using Codex, open OCR models and Jobs 7 -- 2026-04-21
Nemotron 3 Nano 4B: A Compact Hybrid Model for Efficient Local AI 7 -- 2026-03-18
Show HN: Open-source LLM and dataset for sports forecasting (Pro Golf) 7 -- 2026-02-24
Datadog scales time series foundation models to 2.5B parameters 6 -- 2026-05-14
Gemma 4 Uncensored (autoresearch results) 6 -- 2026-04-05
Show HN: A 0.3B model that redacts PII in all 24 EU … 6 -- 2026-05-13
Minimax M3 weights to be released on Friday 6 -- 2026-06-11
Arcee-AI Trinity-Large-Thinking 6 -- 2026-04-02
Show HN: Nanbeige 4.1-3B running in the browser via WebGPU 6 -- 2026-02-19
Show HN: A 150M model that extracts verbatim evidence spans for RAG, … 6 -- 2026-06-10
"Optimal Cognitive Core"- specialized 1.7B model for grounded question answering 6 -- 2026-06-03
What the Community Is Running 6 -- 2026-05-31
Show HN: We trained a 32B model to beat Opus 4 at … 6 -- 2026-04-20
Show HN: Vocab extractor for language learners using Stanza and frequency ranks 6 -- 2026-03-29
The Future of the Global Open-Source AI Ecosystem: From DeepSeek to AI+ 6 -- 2026-02-07
First DeepSeek V4 Flash-Base-Int4 Quant 5 -- 2026-04-27
APIEval-20: A Benchmark for Black-Box API Test Suite Generation 5 -- 2026-03-30
Show HN: Voice gender classifier for European voice AI (1MB, ONNX, 4ms) 5 -- 2026-05-12
Don't Train the Model, Evolve the Harness 5 -- 2026-07-03
PP-OCRv6 5 -- 2026-06-16
I made a CPU only spiking neuron network lib that comes pretty … 5 -- 2026-05-30
Gemma 4 Models 5 -- 2026-04-02
Do Bubbles Form When AIs Simulate Capitalism? 5 -- 2026-02-24
Craft – image models can think like LLMs 5 -- 2026-02-06
Sopro v1.5: A 135M TTS model trained for ~$100, runs 20× real-time … 5 -- 2026-02-05
Kernels: Major Updates 5 -- 2026-07-06
ThinkingCap-Qwen3.6-27B: Qwen3.6 capabilities with 50% fewer thinking tokens 5 -- 2026-07-06
Step 3.7 Flash – 198B-A11B MoE vision-language model 5 -- 2026-05-30
ZooL4nD3r: Translate a passage across 961 learned discourse communities 5 -- 2026-05-06
Xiaomi open-sources MiMo-V2.5: 311B A15B 1M-context omnimodal model 5 -- 2026-04-28
Local Agents with Llama.cpp and Pi 5 -- 2026-03-12
PRX Part 3 – Training a Text-to-Image Model in 24h 5 -- 2026-03-03
Qwen3.5 5 -- 2026-02-16
GLM-Image 5 -- 2026-01-14
Show HN: Forecasting my backyard weather with a 22M time-series model 4 -- 2026-05-17
Amalia – an open-source language model targeting European Portuguese 4 -- 2026-07-01
Empero-AI/Qwythos-9B-Claude-Mythos-5-1M 4 -- 2026-06-29
Show HN: Mate – Emotional layer on top of LLMs 4 -- 2026-04-03
Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled 4 -- 2026-03-03
Show HN: EPEE, an expert-annotated ASL dataset from native Deaf signers 4 -- 2026-07-09
ZeroLabs – 100x cheaper than ElevenLabs (free forever locally) with open models 4 -- 2026-07-02
Reachy Mini bot goes local 4 -- 2026-05-28
Mistral Medium 3.5 128B 4 -- 2026-04-30
A long-term memory system that turns agent trajectories into reusable guidelines 4 -- 2026-04-10
GPT-OSS-20B-Vision: First Community VLM for GPT-OSS, Trained on a DGX Spark 4 -- 2026-02-19
DeepSeek OCR2 4 -- 2026-01-27
Show HN: I trained a 1B LLM from scratch for $315 and … 4 -- 2026-07-02
Ai2 ACE2S – Simulate atmospheric variability – Scale of days to centuries 4 -- 2026-06-16
Why do LLM outputs get worse even when metrics stay stable? [pdf] 4 -- 2026-05-07
AI evals are becoming the new compute bottleneck 4 -- 2026-05-04
OpenAI Privacy Filter 4 -- 2026-04-22
"Darwin-27B-Opus: Surpassing the Foundation Model Without Training" 4 -- 2026-04-13
Liberate Your OpenClaw 4 -- 2026-04-04
Jane Street – neural net puzzle 4 -- 2026-04-01
Show HN: ArXiv metadata as Parquet files (2.99M papers, 1.44GB, 417 files) 4 -- 2026-03-24
Deploying Open Source Vision Language Models (VLM) on Jetson 4 -- 2026-02-25
Custom Kernels for All from Codex and Claude 4 -- 2026-02-18
PersonaPlex-7B: full-duplex voice model that listens and talks at the same time 4 -- 2026-02-15
MiniMax M2.5 weights 4 -- 2026-02-13
One Year Since the "DeepSeek Moment" 4 -- 2026-01-27
DeepSeek MHC: Manifold-Constrained Hyper-Connections 4 -- 2026-01-02
Show HN: Open dataset for distilling GPU audio2face models to CPU students 3 -- 2026-06-09
Two Years of Local AI on a Laptop: When Open Models Outpaced … 3 -- 2026-05-11
HuggingFace launches pre-compiled machine learning kernels repository 3 -- 2026-04-15
Show HN: Prompt injection detector beats ProtectAI by 19% accuracy, 8.9x smaller 3 -- 2026-04-08
Reproducing Anthropic's "Counting Manifold" 3 -- 2026-02-20
TabFM: Zero-shot tabular foundation model from Google Research 3 -- 2026-07-04
ML research datasets from ArXiv and Semantic Scholar (JSONL, quality-scored) 3 -- 2026-06-16
A Blockade on Knowledge. After June 12, Nothing Will Be the Same 3 -- 2026-06-14
Deepfake Detector Robustness Testing 3 -- 2026-06-07
QVAC MedPsy: Medical and Healthcare Models for Edge Device 3 -- 2026-05-07
Hugging Face moves safetensors to the PyTorch Foundation 3 -- 2026-04-08
Training mRNA Language Models Across 25 Species for $165 3 -- 2026-04-07
Hugging Face Provides an S3 Alternative 3 -- 2026-03-24
SynthVision: Building a 110K Synthetic Medical VQA Dataset 3 -- 2026-03-23
QORA-LLM-2B – Pure Rust ternary inference, no multiplication needed 3 -- 2026-03-11
Pure Rust, zero dependencies AI models, runs locally, free forever 3 -- 2026-02-28
TranslateGemma now runs 100% in the browser on WebGPU with Transformers.js v4 3 -- 2026-02-25
Multi-vector Grep for code agents, save 15% tokens, local 3 -- 2026-02-12
Step 3.5 Flash LLM model, agentic coding ~18x faster than GLM 4.7 … 3 -- 2026-02-02
Introduction to Reinforcement Learning and Its Role in LLMs 3 -- 2026-07-11
Native-speed vLLM transformers modeling back end 3 -- 2026-07-08
IOL-AI 2026 Challenge: Can Your Model Solve Linguistics Olympiad Problems? 3 -- 2026-07-06
Filter Hugging Face Models Page by Hardware 3 -- 2026-06-30
MFM: PINN based Motion Foundation Model 3 -- 2026-06-29
HRM-Text: Efficient Pretraining Beyond Scaling 3 -- 2026-06-25
The first open, vision-driven real-time interaction model 3 -- 2026-06-23
Show HN: Multi-Agent AI trading firm simulation powered by small language models 3 -- 2026-06-20
Genoma Labs' open 14B agentic coding model trained on Kraken 3 -- 2026-06-17
Donate Agent Traces 3 -- 2026-06-15
Unknown Is Not False: An AI Agent Pre-Execution Checklist 3 -- 2026-06-11
Nvidia/Nvidia-Nemotron-3-Ultra-550B-A55B-BF16 3 -- 2026-06-08
New SoTA open source TTS model from Boson AI 3 -- 2026-06-05
Aura: Action-Gated Memory for Robot Policies at Constant VRAM 3 -- 2026-06-04
Direct Preference Optimization Beyond Chatbots 3 -- 2026-06-04
Gemma 4 12B appears in Hugging Face 3 -- 2026-06-03
Nvidia: Nemotron Labs Diffusion 14B 3 -- 2026-05-23
Karpathy's autoresearch, 50 DPO experiments, 300 human judges 3 -- 2026-05-21
The Open Agent Leaderboard 3 -- 2026-05-18
Physics-intern: an autonomous agentic framework for physics research 3 -- 2026-05-12
Cuarzo-100K v2 – Python↔EN/ES/FR/ZH, 100% AST verified across all 4 languages 3 -- 2026-05-11
Show HN: Open Source FreeCAD dataset for CAD generation tasks 3 -- 2026-05-08
Building a Fast Multilingual OCR Model with Synthetic Data 3 -- 2026-04-28
Neo-Unify: An Encoder-Free, Native Multimodal Paradigm (SenseTime) 3 -- 2026-04-14
The Open-Source Recipe for Teaching a Robot to Fold Your Clothes 3 -- 2026-04-08
Nemotron OCR v2 by NVIDIA 3 -- 2026-04-03
Granite 4.0 3B Vision: Compact Multimodal Intelligence for Enterprise Documents 3 -- 2026-04-01
TRL v1.0: Post-Training Library Built to Move with the Field 3 -- 2026-04-01
Mr. Chatterbox is an LLM trained exclusively on Victorian-era British texts 3 -- 2026-03-29
CanViT: Toward Active-Vision Foundation Models 3 -- 2026-03-25
Build a Domain-Specific Embedding Model in Under a Day 3 -- 2026-03-23
Show HN: 518K Vietnamese legal documents (1924–2026) 3 -- 2026-03-22
Keep the Tokens Flowing: Lessons from 16 Open-Source RL Libraries 3 -- 2026-03-21
9B parameter coding agent model fine-tuned on top of Qwen3.5-9B 3 -- 2026-03-13
Leaderboard of Leaderboards – A Real-Time Meta-Ranking of AI Benchmarks 3 -- 2026-03-13
Smol AI WorldCup: What Small LLMs Can Do 3 -- 2026-03-10
The Synthetic Data Playbook: Generating Trillions of the Finest Tokens 3 -- 2026-03-08
Modular Diffusers – Composable Building Blocks for Diffusion Pipelines 3 -- 2026-03-06
Phi-4-reasoning-vision-15B 3 -- 2026-03-05
DataClaw 3 -- 2026-02-25
GGML and llama.cpp join HF to ensure the long-term progress of Local … 3 -- 2026-02-20
GLM-5: From Vibe Coding to Agentic Engineering 3 -- 2026-02-18
Show HN: DeepBrainz-R1 – Reasoning-First Small Models for Agentic Systems 3 -- 2026-02-05
Qwen3-coder-next: SOTA open source coding model 3 -- 2026-02-03
We got Claude to teach open models how to write CUDA kernels 3 -- 2026-01-29
Architectural Choices in China's Open-Source AI Ecosystem 3 -- 2026-01-27
Bridging the Gap Between AI Agent Benchmarks and Industrial Reality 3 -- 2026-01-21
Jupyter Agents: training LLMs to reason with notebooks 3 -- 2026-01-11
Nvidia brings agents to life with DGX Spark and Reachy Mini 3 -- 2026-01-06
Falcon-H1-Arabic: Pushing the Boundaries of Arabic Language AI 3 -- 2026-01-06