Home / Companies / Fireworks AI / Hacker News

Fireworks AI on HN

32 posts with 1+ points since 2022

Filters
Since:
Posts by Month (32 total)
Hacker News Posts
Title Points Comments Date
Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA 877 450 2026-07-21
Ember-1 550 237 2026-09-27
Fireworks: Function Calling Model and API 53 18 2023-12-21
Fireworks F1: A Breakthrough in Complex Reasoning with Compound AI 17 8 2024-11-18
FireFunction V1 – GPT-4-level function calling model – 4x faster, open weights 7 0 2024-02-22
Fireworks – Announcing our Series D and $1B ARR 5 0 2026-07-16
How we fixed prompt injection for all models on Fireworks 4 1 2026-04-23
Optimizing MiniMax M3 Sparse Attention on Nvidia Blackwell 4 0 2026-07-13
LLM Eval Driven Development with Claude Code 4 0 2025-08-28
How are people training this LLMs? Dont they need lot of money? 4 1 2024-01-19
Matching frontier performance through training and harness engineering 3 0 2026-06-08
Multi-Query Attention Is All You Need 3 0 2023-07-13
Phylo brings frontier AI to more scientists with open models on Fireworks 3 0 2026-09-20
DeepSeek V4 Pro: Validating Frontier Models for Production 3 0 2026-04-28
FireAttention – Serving Mixtral and open-source MoE models at 4x speed vs. … 3 0 2024-01-09
Accelerating Code Completion with Fireworks Fast LLM Inference 2 0 2023-10-11
Deep-Dive into LLM Fine-Tuning 2 0 2026-02-23
Agents Don't Fail on Intelligence, They Fail on Execution 2 0 2026-05-21
Turn Your LLM into a Calibrated Classifier for $2 2 0 2026-02-20
The Benchmark Gap: What It Takes to Ship Kimi K2.5 2 0 2026-02-16
Why Building Mega Clusters Is Wrong 2 0 2026-03-21
Frontier RL Is Cheaper Than You Think 2 0 2026-03-26
How to accurately and interpretably evaluate the quant effect of large models? 1 0 2024-08-09
Serving Open Source Models 4x faster than vLLM by quantizing with ~no … 1 0 2024-03-21
Document Inlining: Crossing the Modality Gap with Compound AI 1 0 2024-12-23
Can DeepSeek R1 Teach Better Than Humans? 1 0 2025-02-05
Fireworks.ai: Language Model Serving with Custom LoRA Fine-Tuned Models 1 0 2023-08-18
Natural Language → SQL with Reinforcement Fine Tuning (RFT) 1 0 2025-08-18
Turning Production Logs into Evaluation Datasets: A Data-Driven Approach 1 0 2026-02-16
DPO, your simplest RL pipeline with two rollouts 1 0 2026-02-18
FireAttention V3: Enabling AMD as a Viable Alternative for GPU Inference 1 0 2024-10-16
GPUs on-demand: Not serverless, not reserved, but some third thing 1 0 2024-06-07