Light
Home
/
Companies
/
OpenPipe
/
Hacker News
OpenPipe on HN
20 posts with 1+ points since 2022
Filters
Min points:
1
10
25
50
100
250
500
Since:
2023
2024
2025
2026
Posts by Month (20 total)
Hacker News Posts
Search:
Title
Points
Comments
Date
Is AI the next crypto? Insights from HN comments
237
367
2023-11-08
Mistral 7B Fine-Tune Optimized
234
103
2023-12-20
Using reinforcement learning and $4.80 of GPU time to find the best …
217
95
2024-10-28
Using GRPO to Beat o1, o3-mini and R1 at “Temporal Clue”
199
55
2025-03-06
Show HN: RULER – Easily apply RL to any agent
81
11
2025-07-11
OpenPipe Mixture of Agents: Outperform GPT-4 at 1/25th the Cost
13
2
2024-06-20
Serverless RL: Faster, Cheaper and More Flexible RL Training
9
3
2025-10-08
PII-Redact – SOTA PII Redaction on Your Laptop
6
1
2025-03-26
Analyzing OpenAI's Reinforcement Fine-Tuning: Less Data, Better Results
4
0
2024-12-30
ART·E: how we built an email research agent that beats o3
3
2
2025-04-29
What we've learned in 3 days of Llama 3
3
0
2024-04-22
Everything I know about reward hacking
3
0
2025-06-12
Fine-Tuning Best Practices: Models
2
0
2024-09-24
Open Deep Research Tutorial – Train a deep research agent to exceed …
2
0
2025-09-02
Mixtral Curious? Comparing Mistral 7B and Mixtral for fine-tuning
1
0
2024-02-29
DPO fine-tuning outperforms SFT
1
0
2024-10-02
OpenPipe
1
0
2024-09-28
Summary-RL
1
0
2025-06-26
LLM Fine-Tuning Best Practices for Training Data Curation
1
2
2024-08-02
S-LoRA: Serving Thousands of Models from One GPU for Fun and Profit
1
0
2024-01-18