|
GPU-rich labs have won: What's left for the rest of us is …
|
88 |
-- |
2025-08-08 |
|
RAG Is Over: RL Agents Are the New Retrieval Stack
|
8 |
-- |
2025-10-17 |
|
Inference.net
|
7 |
-- |
2024-09-21 |
|
Inference.net – Custom AI models in 6 weeks
|
7 |
-- |
2025-09-12 |
|
Arbitraging Down LLM Inference to the Cost of Electricity
|
6 |
-- |
2025-08-28 |
|
Show HN: Project AELLA – Open LLMs for structuring 100M research papers
|
6 |
-- |
2025-11-11 |
|
Show HN: Costco for LLM Tokens
|
6 |
-- |
2024-10-30 |
|
Hybrid-Attention models are the future for SLMs
|
4 |
-- |
2025-11-04 |
|
How much energy does it take to produce an LLM token?
|
3 |
-- |
2025-07-26 |
|
LAION Dataset Explorer
|
2 |
-- |
2025-11-12 |
|
LLM Token Grants for Researchers
|
2 |
-- |
2024-08-28 |
|
Fast Cheap Image Captioning Model Trained on Video Frames
|
1 |
-- |
2025-08-22 |
|
Project OSSAS: Custom LLMs to Process 100M Research Papers
|
1 |
-- |
2025-11-12 |
|
The Economics of Hosting Open Source Models
|
1 |
-- |
2025-07-30 |
|
When to use model distillation in production
|
1 |
-- |
2025-07-24 |
|
Smart Routing Saved Exa 90% on LLM Costs
|
1 |
-- |
2025-07-23 |
|
The Cheapest LLM Call Is the One You Don't Await
|
1 |
-- |
2025-07-21 |